New Model based on Qwen 3.8

#23
by jaskfsafdjlk - opened

Hello, are you thinking about doing a ThinkingCap for the upcoming Qwen3.8-27B?

Oof that would be epic!

I too would be interested in knowing 😁

SOS

3.8 takes so long to reason that it’s driving me crazy. Please release it ASAP!

I really hope we get one!!

The deep thinking of 3.8 is terrifyingly strong and the thinking process is terrifyingly long. I am looking forward to your next performance.

+1 for 3.8 thinkingcap version!

IMG_0704

Yes we are, but no promises for now :)
We're still researching the algorithms and working on more diverse models, we have to figure out the priorities and the right approach to take. But it is very useful for us to see the community interest, and it's great that so many people care about reasoning efficiency! Thanks for verifying Qwen3.8 still has some room to improve

im waiting patiently in the corner 😀

I tried that model, it slightly faster than the stock model with the same quality, but it still tend to overthink, it is faster, but it doesnt cut reasoning as good as bottlecapai, still worth a try if you are curous

I tried that model, it slightly faster than the stock model with the same quality, but it still tend to overthink, it is faster, but it doesnt cut reasoning as good as bottlecapai, still worth a try if you are curous

Also with a other chat template?

yes with froggreric chat template, so far i found the best model that can cut overthink on qwen3.8 the best are model by @DavidAU , you can read more here https://huggingface.co/DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-NM-DAU-NEO-MTP-GGUF, this is the best model in my opinion that able to cut overthink while maintaining same quality on Q4/Q6, i dont find Q8 useful as it gives similar accuracy with Q6 on my test

Sign up or log in to comment