14 employees serving 8.9 million developers a month. That ratio is why Theory Ventures just led a $65M Series B into Ollama — Benchmark, 8VC, and Docker founder Solomon Hykes all in — bringing total funding to $88M. The announcement hit the HackerNews front page this week.
What Ollama is
The default runtime for open-weight models. One command, and Llama, Gemma, or DeepSeek runs on your own laptop — no API keys, no data leaving the machine. That simplicity built 67,000+ integrations and got Ollama inside 85% of the Fortune 500, including government, healthcare, and finance. Model labs like Meta, Google DeepMind, and Mistral treat it as a day-zero distribution partner.
The cloud pivot is the real story
Same command now runs bigger models — GLM, DeepSeek, Kimi, MiniMax, Nemotron — on Ollama’s GPUs. Cloud token volume is doubling every month. Local stays free; cloud is the business model.
For builders
Ollama exposes a local REST API and an OpenAI-compatible cloud API. Prototype against a local model, swap the base URL, ship on cloud — zero code changes.
Ollama calls this “AI’s personal computer moment.” Whoever owns the entry point owns open-model distribution.
You Might Also Like
- Openai Just Acquired Promptfoo the 86m ai Security Startup Used by 25 of Fortune 500
- Kimi Webbridge Plugs Claude Code Cursor and Codex Into Your Browser no Cloud Relay
- Lm Studio Bionic Runs Claude Code Style Agents on glm 5 2 and Kimi k2 7
- Cloudrouter Gives Your ai Coding Agent its own Cloud Machine and Thats a big Deal
- Minimax m2 5 Just Dropped and Open Source ai Will Never be the Same

Leave a comment