ollama retweeted
Closed model, open model, or train your own? 📅 Wed, Oct 7 🕐 5 PM 📍 San Francisco At our next AI Talk, Crusoe’s @ErwanM94707 joins @ollama's @jmorgan and @trajectorylabs's @MichaelElabd to explore this question and more. Register: 🔗 luma.com/lnnlfa87
2
3
27
6,498
You can now add usage credits for paid cloud models without an Ollama subscription. Add usage credits and pay as you go. ollama.com/settings
24
7
136
19,469
Over the last few hours, some requests to the deepseek-v4.1-flash model on Ollama were charged at an incorrect rate, leading to higher usage consumption than expected for certain users. Sorry about this. For anyone effected: - Usage (monthly, or weekly + session on previous plans) has been reset - If extra usage amounts were consumed due to this model, these extra usage amounts have been restored to your account.
61
12
538
52,463
coding agents are moving fast from prototype to production. the infrastructure question is what's left. join @parthsareen from @ollama and @zainhas Hasan from Together AI at @AIconference for a breakout on what it actually takes to build coding agents on open models, and run them at scale.
4
5
25
6,980
ollama retweeted
Many concerns this weekend about slowing down AI, some targeting open models. We need to be responsible. But we can't let this slow down, or worse, prohibit open models. It's the wrong risk. Open models have tremendous power to democratize AI and make it more personal. The larger risk I see with open models is right in front of us: in the last week I've read about how many popular platforms quietly send data to foreign jurisdictions, or worse, sell or train on it to gain an advantage. And this is becoming more and more mainstream. Now more than ever open model vendors must act in the user's best interest, not their own: zero data retention or training, and hosting in the user's region vs sending data overseas. This has been our belief and commitment with @ollama. Nobody needs to slow down open models. We need to distribute and run them in a way users can trust.
7
11
77
8,465
ollama retweeted
Small models are truly capable of the majority of conversational use cases and even a majority of difficult reasoning ones. Amazing work @Avanika15 @JonSaadFalcon + team and great feature in @FT !
dreams do come true 🥹. excited to see our work (w/@JonSaadFalcon, @HazyResearch, john hennessy and @Azaliamirh) feat. in a major way in @FT. the world is becoming increasingly less dependent on centralized cloud ai. we are just getting started 🚀🌖
6
4
34
10,549
Amp users can now use Ollama's cloud models with Amp's new BYOK model routing. No limits or fees for BYOK. Build remote agents, controllable from everywhere!
Amp is now free to use when you bring your own compute and model subscriptions/keys. No more limits or fees for BYOK. ampcode.com/news/free-agent
35
13
249
47,755
DeepSeek-V4.1-Flash is now fully rolled out and available on Ollama's cloud: - Hosted in US & Europe - Zero data retention: prompts and responses are never logged or trained on - Per-token pricing matches the DeepSeek API, including off-peak pricing - Get started with Ollama's Pro, Max, and Team plans, or pay as you go with a free account with no service fees This new model by DeepSeek is more capable, faster, and more cost effective than all prior DeepSeek models including DeepSeek-V4-Pro 🚀.
🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient. 🔹 Introducing the smallest model in our new architecture family, with native visual understanding. 🔹 Designed for greater capability, faster inference, higher throughput, and scaling to larger models. 1/6
66
47
760
66,979
DeepSeek-V4.1-Flash is now rolling out for Pro plan subscribers.
DeepSeek-V4.1-Flash is now being rolled out on Ollama's cloud, starting with Max and Team accounts. We are quickly adding more capacity to roll it out to all subscribers.
37
20
504
45,161
ChatGPT Desktop (the Codex app) can now be configured to use Ollama models. Download or update to Ollama 0.34 to get started.
71
105
1,429
125,765
Customize which models appear in ChatGPT's model selector in settings. Both local and cloud models can be selected.
2
28
8,553
To get started, Download Ollama and enable ChatGPT in Ollama's app: ollama.com/download
3
1
30
7,159
DeepSeek-V4.1-Flash is now being rolled out on Ollama's cloud, starting with Max and Team accounts. We are quickly adding more capacity to roll it out to all subscribers.
🚀 Introducing DeepSeek-V4.1-Flash: smarter, faster, more efficient. 🔹 Introducing the smallest model in our new architecture family, with native visual understanding. 🔹 Designed for greater capability, faster inference, higher throughput, and scaling to larger models. 1/6
30
39
502
88,082
Introducing off-peak hour token rates. DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends! Off-peak pricing will be available soon for more models. DeepSeek models on Ollama's cloud are hosted in the US & Europe with ZDR and fast performance.
99
50
969
104,141
American open weights/source is a national security imperative.
Exciting day for NVIDIA and @huggingface. Open models strengthen safety and cybersecurity, accelerate innovation and diffusion, and enable sovereignty. They allow every developer, startup, university, industry and country to build with, customize and benefit from AI. Thank you @ClementDelangue for coming to me. NVIDIA is going to be a great home for Hugging Face, its community and the future of open models. 🤗 blogs.nvidia.com/blog/nvidia…
42
84
903
79,391
ollama retweeted
It sounds like the shift back toward open weight models that I noticed among the startups in the summer batch is a genuine trend.
🦙 @ollama is used by 9 million developers and 85% of the Fortune 500, giving co-founder and CEO Jeffrey Morgan (@jmorgan) a unique view into which AI models people are actually using and how that’s changing. Right now, the biggest shift he sees is toward open models, driven by coding agents, falling costs, and capabilities that are rapidly catching up to the frontier labs. On Ollama Cloud, that shift has driven a 150x increase in token usage since the start of the year. In this episode of @LightconePod, Jeff joins @garrytan, @snowmaker, @sdianahu, and @harjtaggar to talk about the future of open models and the story behind Ollama, from two years of searching for the right idea to building one of the most widely used AI developer tools in the world. 00:43 — The Shift to Open Models 03:03 — How AI Agents Are Driving Token Usage 05:31 — Are Open Models Catching Up? 08:26 — What Happens When a New Model Launches 11:31 — Ollama as an Operating System for AI 14:05 — The New Opportunities Above the Model Layer 18:19 — Why 80–90% of Enterprise Tokens Could Be Open 20:57 — The Future Is Local and Cloud 26:40 — Why AI Is Coming Back to Your Computer 28:56 — The Coming Era of Unlimited Tokens 32:30 — Do We Still Need a “God Model”? 33:41 — Open Models and Geopolitics 36:14 — The Origins of Ollama 40:36 — Two Years Lost in the Wilderness 42:39 — The Pivot That Changed Everything 47:02 — How Ollama Found a Business Model 49:43 — Why Second-Time Founders Did YC
69
55
693
179,333