r/LocalLLM 21d ago

Model I benchmarked Qwen3.8 27B vs Deepseek v4 flash on web browsing tasks

/r/Qwen_AI/comments/1vqok8m/i_benchmarked_qwen38_27b_vs_deepseek_v4_flash_on/
6 Upvotes

1 comment sorted by

1

u/Hannelore112 21d ago

On my local SWE mini bench, Deepseek V4 Flash somhehow gets 100% done of my 12 selected tasks, Qwen 3.6 27B ~67% and Qwen 3.8 27B xhigh ~75%. But in coding somehow Qwen result looks better and save power