It's surprisingly good

#1
by sometimesanotion - opened

This model does a good job of surpassing typically overbaked Qwen 3+ prose. I'm testing it with prompts for tense, probing third-person dialogue between characters suspicious of each other, and this model punches well above its weight. It holds on to details in the scene rather well.

In fact, the only problem I have to report is that the think block markers are getting misplaced by a q5_k_m GGUF, but that's a very common issue with finetuning a thinking model for prose! I bet merges with low-rank LoRAs from a finetune with strong IFEVAL and strong, concise reasoning would help this model.

It's already impressive for its size. Nice work!

Sign up or log in to comment