
There's a conversation that keeps resurfacing on Hacker News this month, and it goes by a deliberately provocative title: "Models Are Getting Dumber on Purpose." The claim isn't that labs failed to make models smarter. It's that they succeeded — and then made a second, deliberate choice to ship models that think less, because dumber is cheaper, faster, and better for margins.
It sounds like a conspiracy theory until you line up what the labs themselves published this week.
OpenAI announced it hit its "automated research intern" goal. Its researchers now run 3.1 agent-workdays per human workday, and the heaviest users burn through $7,000+ per day on tokens. That's what frontier intelligence costs in 2026: a mid-tier salary, paid daily, in compute.
OpenAI's chief scientist Jakub Pachocki said no lab has solved alignment well enough to keep scaling at maximum speed — and that he hopes voluntary slowdowns become commonplace. The smartest models are, by the lab's own admission, being throttled.
Anthropic, meanwhile, signed agreements for at least 14.8 GW of compute capacity since October, with a projected spend of up to $517B over the next decade. Somebody is buying power plants to feed the smart track.
Read those three together and the "dumber on purpose" thesis stops being paranoid. The labs are running a barbell:
The middle, the "pretty smart, pretty cheap, pretty general" model everyone actually wanted, is being hollowed out. Not by accident. By unit economics.
It helps to be precise, because "dumber" is doing a lot of work here. The everyday models aren't losing knowledge. They're losing effort:
Spend-tracking data from earlier this year told the same story from the demand side: enterprises fleeing the flagship models for cheaper ones, while the flagship's share stalled. Users noticed the quality drift and moved down anyway. The market is collaborating with the enshittification.
The uncomfortable part is that no villain is required. If your costs are dominated by inference, and 95% of your traffic is simple queries, then optimizing the 95% is just good engineering. The frontier model exists to win benchmarks and justify the enterprise contract; the everyday model exists to carry the traffic. Two products, one brand.
The problem is what this does to trust. When a model's intelligence is a dial the provider turns for margin reasons — silently, without a version bump — benchmark scores become marketing and your evals become the only source of truth. "Which model is best?" is no longer a question with a stable answer. "Which model is best this month, for my task, at my volume" is the real question, and it now has to be re-answered continuously.
There's a version of this story that ends fine: the frontier keeps compounding, distillation keeps improving, and next year's "dumbed-down" everyday model is still smarter than this year's flagship. That has been the pattern so far, and it's why the labs can get away with it.
But there's another version, the one the HN thread is really about: the everyday models plateau at "good enough for chat," the frontier retreats behind enterprise pricing, and the gap between what AI can do and what your AI does grows quietly — visible only to the people running evals.
The models aren't getting dumber because the field stalled. They're getting dumber because dumber sells. The only defense is to measure.
What's your experience — noticed your daily driver model getting subtly worse over the past few months? The evals people share in the comments of that HN thread suggest it's not just you.
Carefully selected AI tools to improve your work, study, and live efficiency.
A major breakthrough has been achieved in the core architecture of large-scale models! The release of Kimi Linear marks the first time that linear attention technology has comprehensively surpassed and significantly outperformed the traditional Transformer full-attention model in both performance and efficiency. This "win-win" achievement is expected to significantly reduce the computational barriers and costs for long text processing, complex reasoning, and AI agent applications, potentially changing the competitive landscape of underlying technologies for large-scale models.
Over the past week, the AI community's attention has been drawn to a mysterious model that quietly emerged on the OpenRouter platform—Polaris Alpha. As a direct continuation of yesterday's discussion of the GPT-5.1 leak, this suddenly appearing model brings more technical details and strategic signals worthy of in-depth exploration.
A new paradigm in knowledge acquisition has arrived, this time powered by AI.
Standing at this moment in 2025, when we look back at the development journey of artificial intelligence, we witness how this revolutionary technology has reshaped every aspect of human society. From initial theoretical concepts to today's practical applications, each step forward in AI technology has changed the way we live. Let's revisit this fascinating journey together.
Sponsored byImage to Image AI