AI Agents & Automation
-
Anthropic ships Claude For Legal, going head-to-head with Harvey
Anthropic shipped Claude For Legal on May 12 — its first true vertical product, aimed straight at the most lucrative slice of enterprise AI. What it actually is It’s an agent layer, not a chat wrapper. Twelve practice-area plugins ship at launch: M&A, employment, litigation, privacy, commercial counsel, and seven more. Over 20 MCP connectors… Continue reading
-
Sierra Ghostwriter ($950M Series E) closes the largest pure-play AI agent round ever
Bret Taylor’s Sierra just raised $950M at a $15B+ post-money valuation — the biggest single round any AI-agent-only company has ever pulled. Tiger Global and GV led. Benchmark, Sequoia, and Greenoaks piled in. Sierra now says it serves 40%+ of the Fortune 50 and went from $100M ARR in November to $150M by early February.… Continue reading
-
Web Speed kills the token tax for Claude and Gemini agents, claims 90% cost cut
Production web agents have a dirty secret. Most of their token budget goes to parsing bloated HTML — scripts, tracking pixels, ad divs — before the model sees the actual content. Web Speed, launched on Product Hunt on May 9 by Dominic Pi-Sunyer, attacks this head-on: a logic layer between the agent and the web… Continue reading
-
xAI ships Grok Connectors with 8 integrations and BYO MCP support
xAI flipped the switch on Grok Connectors today (May 11, 2026), turning Grok from a chatbot into something that actually touches your work tools. Web, iOS, and Android — all live on day one. What it does This is Grok-as-agent. One-click hookup to Gmail, Google Workspace, SharePoint, Outlook, OneDrive, Notion, GitHub, and Linear. Grok can… Continue reading
-
Google Remy: Gemini’s 24/7 background agent surfaces 8 days before I/O
Google’s been building this in the shadows. A staff-only build of the Gemini app leaked this week with an internal agent codenamed Remy — and unlike every “AI assistant” Google has shipped, this one isn’t here to chat. It’s here to do your job while you sleep. What Remy actually is It’s an agent layer… Continue reading
-
Prime Intellect Lab hits GA: per-token RL training across 14 models
Prime Intellect flipped Lab from beta to GA on May 7. It’s a full-stack platform for training self-improving agents — define a task, write a harness, evaluate, run RL on the reward signal, inspect rollouts, deploy a LoRA adapter, serve inference. All inside one product. Nobody else has commercialized this loop end-to-end; until now you… Continue reading
-
HKUDS AI-Trader gains 255 GitHub stars a day: HKU’s agents debate before they trade
HKU’s Data Intelligence Lab open-sourced AI-Trader and it’s pulling 255 GitHub stars a day — 15.3k total. The framework is 100% agent-native: multiple agents collaborate and debate to surface trade ideas, then execute across stocks, crypto, forex, options, and futures. Agents that argue, not just execute Most AI trading repos wrap an LLM around a… Continue reading
-
ByteDance UI-TARS-desktop scores 61.6% on ScreenSpot Pro, leaving GPT-4o and Claude behind
ByteDance’s UI-TARS-desktop pulled 656 stars yesterday — sitting at 31.9k on GitHub, and everyone’s calling it the open-source answer to OpenAI’s Operator. What it actually is: two pieces. Agent TARS, a multimodal agent stack you run in a terminal, browser, or embed in your product. And UI-TARS Desktop, a native app that hands your computer… Continue reading
-
Airbyte Agents (Context Store) launches with 50 connectors and 75-90% token savings vs vendor MCPs
Airbyte spent eight years moving data into warehouses. Now they’re moving it into agents. Airbyte Agents is a managed Context Store that pre-replicates and pre-indexes data from Salesforce, Zendesk, Slack, Linear, Jira, Gong and 50+ other SaaS sources. Instead of an agent hitting a vendor MCP and burning thousands of tokens parsing raw API responses,… Continue reading
-
Anthropic Claude Dreaming lets agents rewrite their own memory — Harvey saw 6x task completion
Anthropic shipped Dreaming on May 6, a research preview inside Claude Managed Agents. The pitch is blunt: stop letting your agents repeat the same mistake forever. What Dreaming actually does Between live sessions, the agent goes async and chews through its own past — transcripts plus the memory store. It pulls out patterns that survive… Continue reading
