"Dumber" isn't dumb: capability versus response style
A recurring complaint about new assistant models is that they feel 'dumber'. Often what changed is not capability — which keeps rising on benchmarks — but response style: shorter, more hedged, more cautious answers by default.
In favour- Concise can be a virtue: stripping padding and performance leaves the substance.
- Separating capability from style helps you judge a model by what it does, not how chatty it is.
- Concise can also mean evasive: shorter, more generic, more hedged answers are a real complaint, not a feature.
- 'Feels dumber' is a perception; measure the actual task before trusting or dismissing it.
Our takeThe useful distinction: removing padding is good, removing the position is not. Judge the output on the task, not on how it reads.
Source: Commentary on recent assistant model updates
← All Focus posts