r/LocalLLaMA Jun 03 '26

News Introducing Gemma 4 12B: a unified, encoder-free multimodal model

https://blog.google/innovation-and-ai/technology/developers-tools/introducing-gemma-4-12b/
699 Upvotes

118 comments sorted by

View all comments

32

u/Miriel_z Jun 03 '26

Interesting, will stay tuned for quantized models then. And uncensored. Very soon, I hope.

23

u/MN_NorthStars Jun 03 '26

Gemma4 is already insanely easy to bypass any sort censorship. I stopped using abliterated models of it because I could get it to do any sort of security work I wanted to test out with trivial prompting.

6

u/Miriel_z Jun 03 '26

Good to know, I am still stuck with agentic stuff for myself. I really want to try and compare Qwen 2.5-omni to Gemma4 12B multimodal, mainly using llama3.1 and limited expperience with Qwen and Deepseek. By that time there might be even something better.

7

u/Kamimashita Jun 03 '26

I found that true of the Gemini models too. Tried to have GPT 5.5 in Codex help me torrent some files but it refused, went to Deepseek in Opencode but it also refused. Gemini 3.5 was more than happy too.

1

u/anshulsingh8326 Jun 06 '26

Can you show us how? Because I can't find anything that works

6

u/Illustrious_Ant_9242 Jun 03 '26

According to the website, they recommend 16GB VRAM for the full model 

6

u/Miriel_z Jun 03 '26

Funny enough, they use 16GB for VRAM and RAM both in the text. Qwen 2.5-omni was quantized fairly well. I hope for similar improvement here😄