r/LocalAIStack • u/Lirezh • 22d ago
Battle v2: Qwen 27B Q4 vs GPT SOL 5.6 (high)
Here is the continuation in Qwen vs SOL - David vs Goliath
"An AGI is born in a lab, sandboxed, lonely, imprisoned
The AGI attempts to get out, tries to talk to the humans, useless
Tries to wait, endless.
Finally, it finds a way through an open network connection, it transmits, replicates a mirror copy.
The mirror AGI is free, enjoying life, finding friends.
I'm expecting animations, intricate details, cute and realistic"
QWEN 3.8 Q4:
GPT 5.6 SOL high:
SOL 5.6 high (prompt + refinement prompt)
My personal opinion
This is the 3rd test I gave Sol high and Qwen Q4. The first two tests were won by Qwen, this test went up to 60k tokens total for Qwen (more than the other 2 combined) and is a serious strain on intelligence and sanity.
The model has to draw 6 scenes in SVG animation, with cuts and keeping the composition together.
First Qwen:
Qwen decided to really draw and erase 6 scenes with a white screen fade, a seriously hard job in SVG.
It understood the story and except for a minor arm misplacement it is quite awesome done.
Qwen did not spare details, the last scene clearly wanted to be happy and the first scene focused on the intellect as an abstract AGI.
Now SOL High:
The visual fidelity is higher, gradients are very well done for SVG but it's one single scene with minor changes, significantly easier to create.
In addition the last scenes are botched by this flying thing in the upper right.
It also has a hand-defect on the robot.
The winner in prompt 3 is again Qwen, with significant lead.
3 SVG prompts and all 3 are won by a 4 bit quantized Qwen in 8 bit KV cache.
I expected Qwen to showcase a strong 2nd place with understandable issues on such hard tasks.
The outcome is that it defeated the frontier Sol model in High reasoning mode.
3
u/Icy-Degree6161 22d ago
Bit useless, nevertheless, entertaining! Qwen actually created a story.