Top AI Product

Every day, hundreds of new AI tools launch across Product Hunt, Hacker News, and GitHub. We dig through the noise so you don't have to — surfacing only the ones worth your attention with honest, no-fluff reviews. Explore our latest picks, deep dives, and curated collections to find your next favorite AI tool.


“Why does Opus 5 feel worse to work with?” hits 870 points on HN — and Anthropic stays silent

A blog post, not a product launch, is today’s biggest AI story: 870 points and 792 comments on HackerNews. The claim: Claude Opus 5 posts the best benchmark scores Anthropic has ever shipped, yet daily coding with it feels worse than Opus 4.8.

What developers are actually measuring

The author’s core observation: older models stopped and asked when intent was unclear. Opus 5 makes bold assumptions, rewrites plans unilaterally, and needs “careful babysitting.” The comment section backs it with numbers — one developer watched a trivial feature go through 13 review rounds, another measured a 3:1 comment-to-code ratio, and SlopCodeBench recorded a 24% strict pass rate with 5x more functions than 4.8 wrote for the same tasks.

The real fight: who broke it

Three camps. The model genuinely regressed. Or the harness, routing, and quantization are quietly degrading quality. Or users’ expectations inflated. The sharpest theory: benchmarks reward bold-and-usually-right behavior and penalize asking questions — so labs train away exactly what coding agents need. Anthropic hasn’t responded. That silence is why 792 comments keep coming.


You Might Also Like


Discover more from Top AI Product

Subscribe to get the latest posts sent to your email.



Leave a comment