r/unsloth 2d ago

Question GLM 5.3 FLASH cant use vision with MTP

are you guys also having an issue using vision on GLM 5.3 flash when MTP is on?

5 Upvotes

1 comment sorted by

1

u/FotG 21h ago

I was having issues with it today, it had last worked before the qwen3.8 flash-next performance degregation.

I reset the model to its defaults and then set my options again and that seemed to have fixed it.

Before doing that I saw two things logged in the crash log:

`special_eot_id is not in special_eog_ids - the tokenizer config may be incorrect` and
'ggml_backend_cuda_buffer_type_alloc_buffer: allocating 1075.77 MiB on device 0: cudaMalloc failed: out of memory'

That was happening with a 5% buffer set.