-
OmniRoute (231+ provider AI gateway) added 1,010 GitHub stars in one day — here’s what it does
OmniRoute is an open-source, MIT-licensed AI gateway: a self-hosted proxy that gives you one OpenAI-compatible endpoint routing to 231+ model providers — Claude, GPT, Gemini, DeepSeek, and 50+ free tiers. It just crossed 10,000 stars, gaining 1,010 in a single day. The problem is real: nobody wants to write a separate integration for every provider,… Continue reading
-
694 stars in one day: Vibe-Trading (HKUDS) wants to do for trading what vibe coding did for code
The HKU lab behind LightRAG and AI-Trader has done it again. Vibe-Trading (HKUDS) jumped 694 GitHub stars in a single day, now around 17K total. This lab is an open-source hit factory at this point. What it is A trading agent that plugs into Claude Code or Codex CLI as a skill. You describe a… Continue reading
-
OpenAI GeneBench-Pro: top AI models fail 70% of real biology tasks
OpenAI dropped a number that stings. On July 1 it released GeneBench-Pro, a benchmark for computational biology agents — and the best model on it, GPT-5.6 Sol at max reasoning, passes only 28.7% (31.5% in Pro mode). The strongest non-OpenAI model, Claude Opus 4.8, gets 16%. Everyone else is worse. What it actually tests Not… Continue reading
-
Claude Fable 5 Global Redeployment (Export Controls Lifted): From US Ban to Worldwide Access in 18 Days
On June 12, the US government did something with no precedent: export controls on a commercial AI model. Claude Fable 5, Anthropic’s frontier LLM, was ordered off-limits to foreign nationals. Anthropic couldn’t verify nationality in real time, so it shut the model down for everyone. June 30, controls lifted. July 1, Fable 5 is back… Continue reading
-
Google Gemini 3.5 Pro (GA) ships with a 2M-token context window — the biggest in any production model
Google Gemini 3.5 Pro (GA) is Google’s new frontier LLM, reachable through the Gemini API and Vertex AI. It solves one problem better than anything shipping today: fitting an absurd amount of context into a single call. The 2M number is the whole story Two million input tokens. Double what Gemini 3.5 Flash handles, and… Continue reading
-
ZCode (Zhipu / z.ai GLM-5.2 coding agent) hit HN’s 1 spot — China’s open answer to Claude Code
Zhipu (now z.ai) just shipped ZCode, a desktop coding agent tuned for its GLM-5.2 model. It topped Hacker News at 260 points. The pitch: this is China’s open-weight camp taking a direct swing at Claude Code. What it actually is ZCode is a desktop ADE — think Claude Code, but built around GLM-5.2 and designed… Continue reading
-
Mistral Leanstral 1.5 writes math proofs that Lean 4 can actually verify
While everyone piles into general-purpose coding agents, Mistral went the other way. On June 30 it shipped Leanstral 1.5, a model built for one narrow, hard thing: automated theorem proving and autoformalization in Lean 4. It replaces the March original, which is now retired. What it actually does Leanstral 1.5 is a 119B-parameter MoE model… Continue reading
-
Google Gemini Omni Flash drops video generation to $0.10 a second
Google shipped Google Gemini Omni Flash on June 30 — the cheap, fast tier of its Omni family, built for one thing: making video by talking to it. Describe a scene in plain English, get a clip. Then keep talking to fix it: “darken the sky,” “add a dog,” “make it slower.” No timeline, no… Continue reading
-
Pluno skips the UI and talks to APIs — 14x faster than Claude’s browser extension
Most browser agents work like a slow human: screenshot the page, find the button, move the cursor, click, wait, screenshot again. Pluno throws that whole loop out. It’s a browser extension that ignores the interface entirely and talks straight to the private APIs sitting underneath web apps like HubSpot, Notion, and Stripe. You describe the… Continue reading
-
ByteDance Seedance 2.5 generates 30 seconds of video in one shot — no stitching
Most AI video models top out at 5-10 seconds, then fake longer clips by gluing segments together — and the seams show. ByteDance’s Seedance 2.5, unveiled at the Volcano Engine FORCE conference, generates a single native 30-second clip in one pass. Character faces, lighting, and motion hold steady the whole way through, because audio and… Continue reading
