Open Models Change The Economics of AI
In a Nutshell
Open models are now economically viable for enterprise AI because they solve the primary blocker of cost, with Chinese models dominating cloud usage and a 40% shift already seen at companies like AT&T. Token consumption is exploding 10-150x due to coding agents and coworker-style automation enabled by expanded context windows, driving demand for efficient orchestration layers above the models themselves. The long-term equilibrium is 80-90% open model usage with hybrid local/cloud execution, where frontier models handle only the hardest tasks while open models power the majority of work through specialized routing and coordination systems.
These notes were generated by AI and may contain inaccuracies.
Cost is by far the largest pain point that open models can jump in and solve. Every business has a vision of getting better control over AI and customizing it for their business, and that's really their north star. Cost is something they can solve in the short term, but it then enables them to customize these models for their unique use case.
Early 2024, there was lots of interest in fine-tuning your own custom models. Then it sort of went away and all of that it will just be wasted effort. It'll get stomped by the next model release. It seems like it's coming back now.
Ollama is used by 9 million developers, has 178,000 GitHub stars, and is used by 85% of the Fortune 500. Jeffrey Morgan, co-founder and CEO of Ollama, has a front seat to what models score highest on benchmarks and what developers actually download and keep using.
The biggest thing being seen is a shift to open models, especially in enterprise. This is from a mix of US and Chinese origin models, and it's predominantly driven by coding agents and also AI assistants more coworker cases like OpenClaw and Hermes.
Ollama sits in the token flow of so many tokens and has really good data on what models people are actually using and how it's changing. Ollama started as a way to run open models on MacBook or other hardware including Nvidia, AMD, Intel. Earlier this year, Ollama launched Ollama's cloud.
Sign in to read the full notes
Get access to AI-generated notes, topic timestamps, and more.