Top AI Product

Every day, hundreds of new AI tools launch across Product Hunt, Hacker News, and GitHub. We dig through the noise so you don't have to — surfacing only the ones worth your attention with honest, no-fluff reviews. Explore our latest picks, deep dives, and curated collections to find your next favorite AI tool.


DeepSeek V4 Pro 0813 scores 87.9 on Terminal Bench 2.1 while charging $0.87 per million output tokens

DeepSeek pushed its flagship out of preview on August 13 and hit number one on HackerNews the same day — 799 points, 309 comments. It’s a 1.6T-parameter MoE with ~49B active, hybrid attention to keep long-context inference cheap, 32T+ pretraining tokens, 1M context and 384K max output.

The numbers that got it to HN #1

Terminal Bench 2.1 went from 72.1 to 87.9 over the preview. CyberGym 52.7 → 83.3. DeepSWE 12.8 → 62.7. On DeepSeek’s own agent and coding benchmarks it edges past Opus 4.8, and SWE-bench Verified lands at 80.6 — a hair behind Claude. These are vendor-reported, so treat them as a direction, not a ranking. Independent frontend and 3D work still looks weaker than Opus.

What you actually plug it into

It’s a model, not an app: an API you point your agent loop at. DeepSeek’s official endpoint keeps the same deepseek-v4-pro model id, so old integrations inherit the new weights with zero code change. OpenAI-compatible format, also on OpenRouter. $0.435/M in, $0.87/M out, $0.003625/M cached input — roughly an order of magnitude under Opus for long-horizon terminal agents, SWE-bench-style repo work, and million-token codebase reads.

DeepSeek has already warned a big price hike is coming.


You Might Also Like


Discover more from Top AI Product

Subscribe to get the latest posts sent to your email.



Leave a comment