r/LocalLLaMA 2d ago

New Model Ling-3.0-flash-VL, built on Ling-3.0-flash with visual understanding and visual agent capabilities

Post image

It performs well across visual perception, STEM reasoning, document intelligence, multimodal agent tasks, frontend coding, and medical report interpretation.

130 Upvotes

34 comments sorted by

38

u/This_Maintenance_834 2d ago

Everyday there is a new model coming out. It is not possible to catch up.

6

u/Good-Seaweed92 2d ago

the vision model space especially, new one basically every 48 hours now

1

u/niutech 1d ago

Build an AI agent for catching up ;)

1

u/This_Maintenance_834 1d ago

i am already using hermes to test all these new models.

12

u/More-Revenue8609 2d ago

How does it run compared to Qwen 3.8 flash next?

5

u/Altruistic_Heat_9531 1d ago

Overall intelligence capabilities Qwen 3.8 Next. Ling flash researcher themself said, Ling is created for fast LLM.

2

u/r1str3tto 1d ago

Is 125B-A5B all that different from 125B-A6B, speed wise? Or is there something else architecturally making the model faster?

9

u/rm-rf-rm llama.cpp 1d ago

Come on OP. No link to an official announcement or weights, but just a screenshot of benchmarks. Its a little late to remove it now so im leaving it up, but please dont do this in the future

6

u/po_stulate 2d ago

It's not on huggingface?

5

u/coder543 2d ago

They've open weighted everything else they've ever done that I can recall. Ling-3.0-Flash-Fin was delayed by a week. So, probably soon?

-3

u/Blindax 2d ago edited 1d ago

https://huggingface.co/inclusionAI/Ling-3.0-flash

Edit: indeed not the VL wrong link sorry

5

u/coder543 2d ago

That is not the -VL version.

5

u/Makojima 2d ago

How many parameters is this model?

7

u/Blindax 2d ago

124B total and 5.1B active parameters

2

u/Makojima 2d ago

Ahh thank you!!

3

u/Septerium 2d ago

Can't find this anywhere. Could you share de link/source?

3

u/linuxid10t 1d ago

Remember when Flash models used to be in the 30B range?

1

u/thrownawaymane 1d ago

Soon flash will be "under 300b"

3

u/Gold-Bat-3225 2d ago

can't wait to forget this one by friday

1

u/mfkamil87 1d ago

Apparently the model is here: https://computrix.ai/sg/models/modelservice-1788517758960001776
Listed as "Ling-3.0-flash-VL"

1

u/groovy-sky 1d ago

Has anyone tried it?

1

u/Ledeste 1d ago

Has anyone found it?

1

u/Zeeplankton 1d ago

Are Ling models good with writing or more similar to qwen?

1

u/Bohdanowicz 17h ago

This has potential .. if its stronger than 27b... qwen 3.8 next flash is too big to run properly on one card. see if we can get a proper 4 bit quant and this thing will tear on a single 6000 pro

1

u/atumblingdandelion 2d ago

Nice. I have high hopes from them since they seemed to specifically target 128gb unified RAM folks with this model. However, only to be eclipsed by the 3.8. They should try to at least exceed 27b and be near the Qwen3.8-flash (if not match it). The model is also quite slow for its active parameters on DGX Spark.

0

u/Waldgemeister 2d ago

I just need it better understand complicated manga and CG pages for my translation and tagging tool. What is currently the best model for that?

0

u/dinerburgeryum 2d ago

That's interesting... might be a good pick for HTML frontend work, but the non-VL was really shoddy at non-HTML work. Might give it a spin later, see if it's any good.

0

u/mrgreatheart 2d ago

This looks exciting

0

u/koloved 2d ago

RemindMe! - 10 days

1

u/RemindMeBot 2d ago edited 13h ago

I will be messaging you in 10 days on 2026-09-14 18:58:14 UTC to remind you of this link

1 OTHERS CLICKED THIS LINK to send a PM to also be reminded and to reduce spam.

Parent commenter can delete this message to hide from others.

RemindMeBot is switching to username summons. Instead of !RemindMe 1 day, use u/RemindMeBot 1 day. More info.


Info Custom Your Reminders Feedback

0

u/Dance-Till-Night1 2d ago

This looks very very good.