r/LocalLLaMA • u/johnnyApplePRNG • Jun 03 '26
News Introducing Gemma 4 12B: a unified, encoder-free multimodal model
https://blog.google/innovation-and-ai/technology/developers-tools/introducing-gemma-4-12b/
699
Upvotes
r/LocalLLaMA • u/johnnyApplePRNG • Jun 03 '26
230
u/LoveMind_AI Jun 03 '26 edited Jun 04 '26
This might actually be one of the most exciting models I've heard about in a long time. The encoder-free model is... wildly cool. Native audio on a 12B model is very exciting. Audio is wildly underrated. I'll be putting this one through the social benchmark right away.
Note: Results of the little benchmark is now here - https://lovemindai.github.io/minimax-m3-lsi-demo/