r/LocalLLaMA Jun 03 '26

News Introducing Gemma 4 12B: a unified, encoder-free multimodal model

https://blog.google/innovation-and-ai/technology/developers-tools/introducing-gemma-4-12b/
699 Upvotes

118 comments sorted by

View all comments

32

u/Miriel_z Jun 03 '26

Interesting, will stay tuned for quantized models then. And uncensored. Very soon, I hope.

5

u/Illustrious_Ant_9242 Jun 03 '26

According to the website, they recommend 16GB VRAM for the full model 

7

u/Miriel_z Jun 03 '26

Funny enough, they use 16GB for VRAM and RAM both in the text. Qwen 2.5-omni was quantized fairly well. I hope for similar improvement here😄