You can now add usage credits for paid cloud models without an Ollama subscription. Add usage credits and pay as you go. https://lnkd.in/g-a3w9wN
Ollama
Technology, Information and Internet
Palo Alto, California 261,864 followers
Get up and running with AI models.
About us
Get up and running with large language models.
- Website
-
https://ollama.com
External link for Ollama
- Industry
- Technology, Information and Internet
- Company size
- 11-50 employees
- Headquarters
- Palo Alto, California
- Type
- Privately Held
- Founded
- 2023
- Specialties
- ollama
Locations
-
Primary
Get directions
Palo Alto, California 94301, US
-
Get directions
744 High St
Palo Alto, California 94301, US
Employees at Ollama
Updates
-
Ollama reposted this
Many concerns this weekend about slowing down AI, some targeting open models. We need to be responsible. But we can't let this slow down, or worse, prohibit open models. It's the wrong risk. Open models have tremendous power to democratize AI and make it more personal. The larger risk I see with open models is right in front of us: in the last week I've read about how many popular platforms quietly send data to foreign jurisdictions, or worse, sell or train on it to gain an advantage. And this is becoming more and more mainstream. Now more than ever open model vendors must act in the user's best interest, not their own: zero data retention or training, and hosting in the user's region vs sending data overseas. This has been our belief and commitment with Ollama. Nobody needs to slow down open models. We need to distribute and run them in a way users can trust.
-
DeepSeek-V4.1-Flash is now fully rolled out and available on Ollama's Cloud: - Hosted in US & Europe - Zero data retention: prompts and responses are never logged or trained on - Per-token pricing matches the DeepSeek API, including off-peak pricing - Get started with Ollama's Pro, Max, and Team plans, or pay as you go with a free account with no service fees This new model by DeepSeek is more capable, faster, and more cost effective than all prior DeepSeek models including DeepSeek-V4-Pro.
-
Introducing off-peak token rates. DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends! Off-peak pricing will be available soon for more models. DeepSeek models on Ollama's cloud are hosted in the US & Europe with ZDR and fast performance.
-
Ollama reposted this
Ollama (YC W21) is used by 9 million developers and 85% of the Fortune 500, giving co-founder and CEO Jeffrey Morgan a unique view into which AI models people are actually using and how that’s changing. Right now, the biggest shift he sees is toward open models, driven by coding agents, falling costs, and capabilities that are rapidly catching up to the frontier labs. On Ollama Cloud, that shift has driven a 150x increase in token usage since the start of the year. In this episode of the Lightcone, Jeff joins Garry, Jared, Diana, and Harj to talk about the future of open models and the story behind Ollama, from two years of searching for the right idea to building one of the most widely used AI developer tools in the world. https://lnkd.in/g5syCqa3
-
GLM 5.3 and GLM 5.3 Flash (previously Ox Alpha) models are fully rolled out on Ollama's cloud. Private. Fast. US and Europe hosted. Zero data retention. GLM 5.3: ollama launch claude --model glm-5.3:cloud GLM 5.3 Flash: ollama launch opencode --model glm-5.3-flash:cloud Works with many more harnesses, or create an API key and plug it straight into your own app. GLM 5.3 model page: https://lnkd.in/g4HQa7ZT GLM 5.3 flash model page: https://lnkd.in/g6GDaJ9j
-
-
Ollama v0.33 is here! You can now easily configure Claude Desktop to seamlessly work with Ollama as a third-party gateway provider. One toggle. Cloud & local models just work! All the features that you are used to continue to work! Ollama’s web search comes baked in so you can use open models to run research tasks. Blog post: https://lnkd.in/g3nE7Nia
-
IBM Granite 4.2 is now available on Ollama. 3B, 8B, 30B parameter open models made for enterprise agents. This model is free to use and is licensed for both research and commercial usage. The data curation and training processes were specifically designed for enterprise scenarios and customization, incorporating governance, risk, and compliance (GRC) evaluations alongside IBM’s standard data clearance and document quality review procedures. Model page: https://lnkd.in/gdvRZKtE
Meet Granite 4.2, IBM’s latest family of open models purpose-built for enterprise agentic AI. With new native reasoning capabilities, Granite 4.2 can plan, reason, self-correct and reliably use tools to help automate complex enterprise workflows. Inside the release: ✅ 3B, 8B and 30B models with native thinking capabilities ✅ Advanced coding and software engineering capabilities ✅ Optimized speech models for high-throughput transcription workloads ✅ Flexible deployment across cloud, on-premises and edge environments Build what’s next with Granite 4.2: https://ibm.co/6041Er8dR Available on: Hugging Face, Replicate, DeepInfra, Ollama, LM Studio, AnythingLLM, RadixArk, OpenRouter, Artificial Analysis, CoreWeave Inference, Arena, Unsloth AI
-
Kimi (Moonshot AI) Kimi K3 is rolling out on Ollama's cloud subscriptions. We are working on improving Ollama's cloud to be much more transparent on the pricing to show the best performance / $. Try it with the tools you already use. Claude Code: ollama launch claude --model kimi-k3:cloud OpenCode: ollama launch opencode --model kimi-k3:cloud Model page: https://lnkd.in/gy2tHAQs
-
-
Ollama reposted this
This Thursday, August 20th, we’re bringing together four leaders at the forefront of open models, coding agents, and AI infrastructure: Kyle Kranen, Senior Manager, Dynamo at NVIDIA Yunmo Koo, Founding Engineer at FriendliAI Dongluo Chen, Software Engineer at Ollama Saurya Velagapudi, Principal Engineer at OpenHands Together, they’ll explore what it takes to move AI coding agents to production and how open models are changing the economics of AI-native engineering. Join us for the conversation, followed by audience Q&A and networking. Spots limited! 📍 San Francisco 📅 Aug. 20 🎟️ RSVP → https://luma.com/o9tv62y5
-