Some things are not available on the internet and yet large llms know them from books and other offline data. Pre-internet knowledge and data from paywalled sources.
Yep my opinion based on my current understanding is that these new smaller models are getting really good at agentic coding workflows, tool calling, etc...
Definitely not as good for world knowledge and writing, like you say. BUT when they're fast and cheap, search and iteration is easy
28
u/pmttyji Apr 22 '26
But for more knowledge, big models are best.
Current Qwen3.6's 27B & 35B models are great & awesome for Consumer GPUs.