Top AI Product

Every day, hundreds of new AI tools launch across Product Hunt, Hacker News, and GitHub. We dig through the noise so you don't have to — surfacing only the ones worth your attention with honest, no-fluff reviews. Explore our latest picks, deep dives, and curated collections to find your next favorite AI tool.


DeepSeek-V4-Flash-Vision-Exp: multimodal agents near Opus 4.8, at $0.22/M input

DeepSeek shipped its first V4-series vision model on August 21, and HackerNews put it on the front page with 319 points. This is a multimodal model served through the DeepSeek API: text capabilities match DeepSeek-V4-Flash, but on multimodal agent benchmarks it jumps far past its text-only sibling — DeepSeek says close to Claude Opus 4.8.

The price gap is the story

Sparse MoE, 284B total parameters with only 13B active. That’s how you get a 1,048,576-token context window and 384K max output at $0.22/M input and $0.66/M output — a fraction of what frontier multimodal models charge. Images are capped at 384 tokens each, so screenshot-heavy agent loops stay cheap.

API access

Live now as deepseek-v4-flash-vision-exp, also on OpenRouter. Supports Chat Completions, Messages and Responses, with images via base64, URL, or the new Files API (free uploads, reusable file IDs). Built for document and chart understanding, visual QA, and agents that interleave text and images — think an agent reading dashboards or parsing PDFs mid-task.

It’s labeled experimental. If the “Exp” suffix follows DeepSeek’s usual pattern, a stable V4 vision model is coming.


You Might Also Like


Discover more from Top AI Product

Subscribe to get the latest posts sent to your email.



Leave a comment