Why your local LLM feels dumber than it to me

Genuine question : is there something that they are claiming about ~%15 word prediction accuracy? Wow, I didn't realize how much the default quantization in popular runners degrades logic compared to full FP16. I'd be curious to see if this is addressed.). Most of the time when a local model feels dumb its not the quant, its the chat template. a lot of sense to me.