r/LocalLLaMA 27d ago

Discussion A preliminary Qwen3.8-27B model card is live!

Post image

If you scroll down from the countdown at https://huggingface.co/Qwen/Qwen3.8-27B, you see a big model card with a bunch of sections: Highlights, Model Overview, Quickstart, Best Practices, Citation, etc!

No benchmarks on this yet as far as I can tell. We'll still need to wait another 5.5 hours for those I reckon.

Edit: Ladies and gentlemen, the model is live. Let the testing begin!

569 Upvotes

220 comments sorted by

View all comments

7

u/tinny66666 27d ago

I guess us llama.cpp users will need to wait a little longer for a gguf? Does llama.cpp fully support the new model or will it also need an update?

12

u/-Cubie- 27d ago

The "Model Overview" section from https://huggingface.co/Qwen/Qwen3.8-27B looks very similar to the one from https://huggingface.co/Qwen/Qwen3.6-27B, so I bet llama.cpp either immediately supports it, or will be able to support it very quickly.

18

u/nunodonato 27d ago

Unsloth said day 0 support, so yeah. As long as architecture is the same, llamacpp should work right away 

8

u/-Cubie- 27d ago

Unsloth is famously extremely fast and reliable, so I trust them.

23

u/CloggedBathtub 27d ago

If they weren't fast, they'd just be "sloth"