AI Agents & Automation
-
agency-agents by msitarzewski: 232 specialist agents that turn Claude Code into a full company
A Reddit thread about agent specialization turned into the fastest-climbing repo of the week. msitarzewski’s agency-agents hit 120K GitHub stars, adding nearly 1,800 in a single day. It’s not another prompt pack. What it actually is It’s an open-source roster of 232 specialized agents spread across 16 departments — frontend engineers, backend API experts, security… Continue reading
-
Cursor for iOS turns your phone into a coding-agent dispatch console
Cursor shipped its first native iPhone and iPad app on June 29, in public beta for paying users. It is not a code editor squeezed onto a small screen. It is a remote control for AI agents. What it actually does Pick a repo, pick a model, then describe the task by typing or talking.… Continue reading
-
Meituan’s LongCat-2.0 is a 1.6T open-weight model trained on 50,000 Chinese ASICs — no NVIDIA
The food-delivery company just dropped a frontier-scale LLM. LongCat-2.0 is a Mixture-of-Experts model with 1.6 trillion total parameters, ~48B activated per token, and a 1M-token context window. The whole pretraining run — 35T+ tokens — happened on 50,000+ domestic AI ASICs in superpod clusters. NVIDIA wasn’t in the building. What it actually is This is… Continue reading
-
HackerRank open-sourced its ATS, and the same resume scored 66 to 99
HackerRank just dumped its hiring agent on GitHub (interviewstreet/hiring-agent), and within a day it pulled nearly 3,000 stars and a 789-point Hacker News thread. Then a developer actually ran it — and the thing fell apart in public. What it is and why it broke It’s an AI resume screener. Feed it a PDF, it… Continue reading
-
Herdr puts every coding agent in one terminal — and lets them orchestrate each other
Anyone running three Claude Code sessions and two Cursor agents at once knows the pain: a forest of terminal tabs, no idea which agent is stuck waiting on you and which is still grinding. Herdr is a Rust-written terminal multiplexer built specifically for that mess. It hit the HackerNews front page on June 29 with… Continue reading
-
Strix (open-source AI pentest agents) won’t report a bug until it’s exploited it
Static scanners flood you with maybes. Strix flips the rule: no working proof-of-concept, no finding. It’s an open-source fleet of autonomous AI agents that hack your app like a real attacker would — run the code, poke the endpoints, and actually break in before saying a word. What it actually does Strix isn’t a linter… Continue reading
-
Lyto runs one AI agent across every tab, tool, and chat window you open
Most AI agents are trapped inside one app. Lyto is a Chrome extension that breaks out of that box — it works wherever your browser goes. It opens and closes tabs, scrolls, clicks, fills forms, and touches every DOM element, then ties that to the tools you already live in: Gmail, Sheets, Slack, GitHub. It… Continue reading
-
OpenAI Codex Record and Replay: demo a Mac task once, the agent repeats it forever
OpenAI quietly turned Codex from a coding tool into a general computer-use agent. Record and Replay, shipped June 18 in the Codex macOS app (v26.616), lets you perform a workflow by hand — click here, fill that, hit submit — and Codex watches, then writes the whole thing up as a reusable skill. What it… Continue reading
-
gstack puts Garry Tan’s full Claude Code setup — 23 role-playing agents — in one install
Garry Tan, YC’s president, just open-sourced the exact Claude Code config he codes with. It’s called gstack, and it crossed 117K GitHub stars (+674 today) almost entirely on his name plus one irresistible pitch: clone a billionaire’s dev workflow with a single command. What it actually is gstack isn’t an app. It’s 23 opinionated slash… Continue reading
-
AI Berkshire (ai-berkshire): four investing legends, four agents, one verdict on your stock
Most “AI stock research” tools dump a wall of bullish-and-bearish hedging and call it analysis. AI Berkshire (ai-berkshire) refuses to. It’s an open-source agent framework that runs on Claude Code, and it forces a call: Pass, Conditional, or Grey Zone — with a concrete price range attached. The author claims a real-money +69% in 2024… Continue reading
