No limits or fees for BYOK.
Build remote agents, controllable from everywhere!
No limits or fees for BYOK.
Build remote agents, controllable from everywhere!
We are quickly adding more capacity to roll it out to all subscribers.
We are quickly adding more capacity to roll it out to all subscribers.
We are quickly adding more capacity to roll it out to all subscribers.
DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends!
Off-peak pricing will be available soon for more models.
DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends!
Off-peak pricing will be available soon for more models.
Let’s go
Let’s go
Thank you @finkd
QT: https://twitter.com/i/status/2095232032896946311
Thank you @finkd
QT: https://twitter.com/i/status/2095232032896946311
Based on your feedback, every plan includes a monthly pool of usage credits. (1/5)
Based on your feedback, every plan includes a monthly pool of usage credits. (1/5)
You can try it with models via Ollama (local and cloud):
ollama launch muse
You can try it with models via Ollama (local and cloud):
ollama launch muse
Private.
Fast.
US and Europe hosted.
Zero data retention.
GLM 5.3:
ollama launch claude --model glm-5.3:cloud
GLM 5.3 Flash:
ollama launch opencode --model glm-5.3-flash:cloud (1/2)
Private.
Fast.
US and Europe hosted.
Zero data retention.
GLM 5.3:
ollama launch claude --model glm-5.3:cloud
GLM 5.3 Flash:
ollama launch opencode --model glm-5.3-flash:cloud (1/2)
Try Ollama's GLM-5.3-Flash as we bring it GLM-5.3 online. Super fast.
Claude Code:
ollama launch claude --model glm-5.3-flash:cloud (1/2)
Try Ollama's GLM-5.3-Flash as we bring it GLM-5.3 online. Super fast.
Claude Code:
ollama launch claude --model glm-5.3-flash:cloud (1/2)
You can now easily configure Claude Desktop to seamlessly work with Ollama as a third-party gateway provider.
One toggle. Cloud & local models just work👇
You can now easily configure Claude Desktop to seamlessly work with Ollama as a third-party gateway provider.
One toggle. Cloud & local models just work👇
3B, 8B, 30B parameter open models made for enterprise agents.
This model is free to use and is licensed for both research and commercial usage. (1/2)
3B, 8B, 30B parameter open models made for enterprise agents.
This model is free to use and is licensed for both research and commercial usage. (1/2)
Ollama is exploding in token usage.
Better models, choice of apps/harness, and all private.
We are just getting started. Let's continue this record-breaking streak.
Ollama is exploding in token usage.
Better models, choice of apps/harness, and all private.
We are just getting started. Let's continue this record-breaking streak.
Excited for what’s to come to open models, and the @nvidia team working on Nemotron.
Excited for what’s to come to open models, and the @nvidia team working on Nemotron.
More models coming soon 🫡
We are working on improving Ollama's cloud to be much more transparent on the pricing to show the best performance / $.
Try it with the tools you already use. (1/2)
More models coming soon 🫡
US and Europe-hosted and zero data retention.
We are working on improving Ollama's cloud to be much more transparent on the pricing to show the best performance / $.
Try it with the tools you already use. (1/2)
US and Europe-hosted and zero data retention.
Gemma is one of the most popular open models.
It's been an amazing journey being a close partner! Can't wait to see what's to come! 🎉
Gemma is one of the most popular open models.
It's been an amazing journey being a close partner! Can't wait to see what's to come! 🎉
We are working on improving Ollama's cloud to be much more transparent on the pricing to show the best performance / $.
Try it with the tools you already use. (1/2)
We are working on improving Ollama's cloud to be much more transparent on the pricing to show the best performance / $.
Try it with the tools you already use. (1/2)
RSVP link below! 👇
RSVP link below! 👇
For local only, you can try qwen3.8 that is optimized:
Apple Silicon:
ollama run qwen3.8:27b-mlx
NVIDIA:
ollama run qwen3.8:27b
For local only, you can try qwen3.8 that is optimized:
Apple Silicon:
ollama run qwen3.8:27b-mlx
NVIDIA:
ollama run qwen3.8:27b