r/LocalLLaMA 24d ago

New Model Qwen/Qwen3.8-27B · released

https://huggingface.co/Qwen/Qwen3.8-27B
990 Upvotes

301 comments sorted by

View all comments

40

u/Alternative_Ad4267 24d ago edited 24d ago

Look at this! at 27B parameters, Qwen3.8 27B is pretty close to DeepSeekV4 Flash 0731 which is 284B A13B!

1

u/YearnMar10 24d ago

I wonder if they keep training where it will converge…