-
Andrej Karpathy joins Anthropic to build a team using Claude to accelerate pre-training research
Andrej Karpathy announced on May 19 that he’s joined Anthropic — starting a team focused on using Claude itself to accelerate pre-training research. He’s working under pre-training team lead Nick Joseph and started this week. ## The career arc Karpathy co-founded OpenAI, left in 2017 for Tesla (where he led Full Self-Driving and Autopilot), returned… Continue reading
-
andrej-karpathy-skills hits 1,955 daily stars: a CLAUDE.md that stops AI from breaking your code
Forrest Chang’s andrej-karpathy-skills is the #1 trending repo on GitHub today with 1,955 stars in a single day — and over 220,000 combined across his personal account and the multica-ai organization mirror. It’s a single CLAUDE.md file encoding four behavioral rules derived from Andrej Karpathy’s documented frustrations with LLM coding agents. ## The core rules… Continue reading
-
Google Antigravity 2.0 ships standalone desktop app, CLI, SDK, and Managed Agents in the Gemini API
Google turned Antigravity into a full agent-first development platform at I/O 2026. Antigravity 2.0 is now a standalone desktop application built entirely around agent orchestration — plus a CLI, an SDK, Managed Agents in the Gemini API, and enterprise support via the Gemini Enterprise Agent Platform. ## The four surfaces Desktop app: a central home… Continue reading
-
Gemini Omni: Google ships a multimodal video model that takes image, audio, video, and text as input
Google announced Gemini Omni at I/O 2026 — a new model series that combines Gemini’s reasoning capabilities with native video generation. The first release, Gemini Omni Flash, accepts image, audio, video, and text input and outputs video grounded in real-world knowledge that can be easily edited. ## What’s actually new Most video generation models today… Continue reading
-
Google ships Gemini 3.5 Flash at I/O 2026: 4x faster than 3.1 Pro and tuned agentic-first
Google opened Google I/O 2026 yesterday with Gemini 3.5 Flash — a frontier model that combines reasoning with agentic task execution. The headline: 4x faster output tokens per second than other frontier models, while beating Gemini 3.1 Pro on coding, agentic, and multimodal benchmarks. Gemini 3.5 Pro is in internal testing now, with public availability… Continue reading
-
HKUDS’s ViMax orchestrates Director, Screenwriter, Producer, and Video Generator agents for multi-shot AI video
HKUDS released ViMax — a multi-agent video generation framework that combines four specialized roles into one end-to-end pipeline: Director, Screenwriter, Producer, and Video Generator. Input a concept, output a multi-shot video with consistent characters and scenes. ## The agent roles Each agent owns a discrete stage. Screenwriter drafts the script from your concept. Director plans… Continue reading
-
Anthropic opens the official Claude Code plugins directory with Anthropic Verified badges for high-trust extensions
Anthropic just opened claude-plugins-official — a curated directory of high-quality plugins for Claude Code and Claude Cowork, managed directly by Anthropic. The marketplace ships built-in to every Claude Code install. ## What’s inside Two top-level directories: `/plugins` for internally-developed Anthropic plugins (code review, SDK helpers, language server integrations, external-service connectors), and `/external_plugins` for third-party submissions… Continue reading
-
Cursor Composer 2.5 matches Opus 4.7 on SWE-Bench at 1/10th the cost — Kimi K2.5 base with 85% Cursor RL
Cursor shipped Composer 2.5 on May 18 — an in-house coding agent built on the open-source Kimi K2.5 checkpoint from Moonshot AI, then heavily post-trained by Cursor (roughly 85% of total compute budget went into Cursor’s own reinforcement learning and post-training pipeline). The headline: 79.8% on SWE-Bench Multilingual, matching Claude Opus 4.7 and GPT-5.5 at… Continue reading
-
Odyssey ships Starchild-1, the first real-time multimodal world model that generates synchronized audio and video
Odyssey ML announced Starchild-1 on May 17 — the first general world model that autoregressively generates synchronized audio and video in real-time while continuously responding to streaming user input. The kicker: world models until now have been silent. ## What’s actually new Previous world models (Genie, Sora video, Decart’s models) learned visual dynamics from large-scale… Continue reading
-
pixserp gives LLMs one endpoint, 10 answer shapes, $1.50 per 1k requests
pixserp launched on Product Hunt this week — a single API endpoint that returns 10 different answer shapes (web, news, images, places, shopping, flights, hotels, YouTube, transcripts, any URL) so an LLM can pick the right format for the question instead of stitching together five different services. ## The pricing and architecture $1.50 per 1,000… Continue reading
