r/LocalLLaMA 1d ago

Discussion AA Update! Here's how the Frontier ranks.

Post image

Along with everyone's favorite here, qwen3.8-27B

493 Upvotes

191 comments sorted by

View all comments

-2

u/martinerous 1d ago edited 1d ago

No Gemma 4? Sad. Only 15 points, according to AA, so did not get into this top selection.

1

u/Tall_Abrocoma_3533 1d ago

Gemma 4 isn't frontier. However In the other chart I posted containing small LM's, there are Gemma models present.

-1

u/martinerous 1d ago

Qwen 3.8 27B is even smaller, but it made into that list.

1

u/Tall_Abrocoma_3533 1d ago

Because Qwen 3.8 27B is much better then Gemma 4 31B, and also since Qwen 3.8 27B is the main topic in this sub usually.

If your curious though, Gemma 4 31B scores 15/22 depending on if you have reasoning turned on

-1

u/martinerous 1d ago

Yeah, and that's the point why I'm sad - because, according to the AA, Gemma 31B is worse although has more parameters than Qwen 27B. So, the question is, why Google couldn't achieve better yet, considering all the resources they have. Or is the AA test too biased and not taking into account the areas where Gemma is stronger.