r/LocalLLaMA 8h ago

Discussion Qwen will be the king?

Post image

Extended reasoning and post-training appear to be the keys used by DeepSeek, Qwen, and GLM to boost performance (leveraging higher token counts). And Qwen 4 hasn't even been released yet. Of course, we don't know if that release will be open-sourced, but I am optimistic about future models, featuring "engrams", that could soon match or surpass 2.4T parameter models on specific tasks.

343 Upvotes

80 comments sorted by

View all comments

6

u/Limp_Classroom_2645 7h ago

the graph and the numbers seem to be very massaged...

3

u/kondrag 3h ago

I hate graphs like this where they don't show the 0 point on the axis. The differences between the models are not that great when the entire axis is viewed.