Text Generation
Transformers
Safetensors
nvidia
unsloth
conversational

Awful performance

#1
by Prezmi - opened

The benchmarks suggest that the model outperforms Qwen 3.6

I run the Q8_0 quant and it just performs awful

Unusable in Copilot chat via custom endpoint.

Claude Code it runs but just full of bugs.

This model is fast but just the reasoning is full of looped functions.

It deletes a line, adds a line and then deletes it again.

Set the Sampling to various options 0.2 temp, 0.6, 1
It all is the same, just rubbish.

Sign up or log in to comment