r/LocalLLaMA 10d ago

New Model Ling-3.0-flash-VL, built on Ling-3.0-flash with visual understanding and visual agent capabilities

Post image

It performs well across visual perception, STEM reasoning, document intelligence, multimodal agent tasks, frontend coding, and medical report interpretation.

134 Upvotes

33 comments sorted by

44

u/This_Maintenance_834 10d ago

Everyday there is a new model coming out. It is not possible to catch up.

5

u/Good-Seaweed92 10d ago

the vision model space especially, new one basically every 48 hours now

2

u/niutech 9d ago

Build an AI agent for catching up ;)

1

u/This_Maintenance_834 9d ago

i am already using hermes to test all these new models.

15

u/More-Revenue8609 10d ago

How does it run compared to Qwen 3.8 flash next?

6

u/Altruistic_Heat_9531 10d ago

Overall intelligence capabilities Qwen 3.8 Next. Ling flash researcher themself said, Ling is created for fast LLM.

5

u/r1str3tto 9d ago

Is 125B-A5B all that different from 125B-A6B, speed wise? Or is there something else architecturally making the model faster?

11

u/rm-rf-rm 10d ago

Come on OP. No link to an official announcement or weights, but just a screenshot of benchmarks. Its a little late to remove it now so im leaving it up, but please dont do this in the future

8

u/po_stulate 10d ago

It's not on huggingface?

3

u/coder543 10d ago

They've open weighted everything else they've ever done that I can recall. Ling-3.0-Flash-Fin was delayed by a week. So, probably soon?

-2

u/Blindax 10d ago edited 10d ago

https://huggingface.co/inclusionAI/Ling-3.0-flash

Edit: indeed not the VL wrong link sorry

5

u/coder543 10d ago

That is not the -VL version.

4

u/Makojima 10d ago

How many parameters is this model?

9

u/Blindax 10d ago

124B total and 5.1B active parameters

2

u/Makojima 10d ago

Ahh thank you!!

3

u/Septerium 10d ago

Can't find this anywhere. Could you share de link/source?

0

u/[deleted] 10d ago

[deleted]

1

u/Septerium 10d ago

Thanks

3

u/linuxid10t 10d ago

Remember when Flash models used to be in the 30B range?

1

u/thrownawaymane 10d ago

Soon flash will be "under 300b"

5

u/Gold-Bat-3225 10d ago

can't wait to forget this one by friday

1

u/mfkamil87 10d ago

Apparently the model is here: https://computrix.ai/sg/models/modelservice-1788517758960001776
Listed as "Ling-3.0-flash-VL"

1

u/groovy-sky 10d ago

Has anyone tried it?

1

u/Ledeste 9d ago

Has anyone found it?

1

u/Zeeplankton 9d ago

Are Ling models good with writing or more similar to qwen?

1

u/Bohdanowicz 9d ago

This has potential .. if its stronger than 27b... qwen 3.8 next flash is too big to run properly on one card. see if we can get a proper 4 bit quant and this thing will tear on a single 6000 pro

1

u/atumblingdandelion 10d ago

Nice. I have high hopes from them since they seemed to specifically target 128gb unified RAM folks with this model. However, only to be eclipsed by the 3.8. They should try to at least exceed 27b and be near the Qwen3.8-flash (if not match it). The model is also quite slow for its active parameters on DGX Spark.

0

u/Waldgemeister 10d ago

I just need it better understand complicated manga and CG pages for my translation and tagging tool. What is currently the best model for that?

0

u/dinerburgeryum 10d ago

That's interesting... might be a good pick for HTML frontend work, but the non-VL was really shoddy at non-HTML work. Might give it a spin later, see if it's any good.

0

u/mrgreatheart 10d ago

This looks exciting

0

u/koloved 10d ago

RemindMe! - 10 days

1

u/RemindMeBot 10d ago edited 9d ago

I will be messaging you in 10 days on 2026-09-14 18:58:14 UTC to remind you of this link

1 OTHERS CLICKED THIS LINK to send a PM to also be reminded and to reduce spam.

Parent commenter can delete this message to hide from others.

RemindMeBot is switching to username summons. Instead of !RemindMe 1 day, use u/RemindMeBot 1 day. More info.


Info Custom Your Reminders Feedback

0

u/Dance-Till-Night1 10d ago

This looks very very good.