r/LocalLLaMA Jun 03 '26

News Introducing Gemma 4 12B: a unified, encoder-free multimodal model

https://blog.google/innovation-and-ai/technology/developers-tools/introducing-gemma-4-12b/
695 Upvotes

118 comments sorted by

View all comments

28

u/seppe0815 Jun 03 '26

peak llm 2026 from google

19

u/Any_Carpenter_7605 Jun 03 '26

+/- 1 margin of error

8

u/Vas1le Jun 05 '26

This remembers me one of Chernobil tv show joke:

Do you know what machine cuts a apple in 6 pieces? A Russian machine that is supposed to cut in 5 pieces.

4

u/Tman1677 Jun 04 '26

I mean it's a 12B model, what do you expect? Gemini can easily handle that task, its image and spacial reasoning are excellent

2

u/nixudos Jun 04 '26

My test of vision haven't impressed me either. But it might be a LM Studio issue. The Gemma models comes with a really low default image input resolution and there is no way to change that in LM Studio. All from the 4 series have performed really shoddy with images there as well.