Ollama’s cover photo
Ollama

Ollama

Technology, Information and Internet

Palo Alto, California 261,864 followers

Get up and running with AI models.

About us

Get up and running with large language models.

Website
https://ollama.com
Industry
Technology, Information and Internet
Company size
11-50 employees
Headquarters
Palo Alto, California
Type
Privately Held
Founded
2023
Specialties
ollama

Locations

Employees at Ollama

Updates

  • Ollama reposted this

    Many concerns this weekend about slowing down AI, some targeting open models. We need to be responsible. But we can't let this slow down, or worse, prohibit open models. It's the wrong risk. Open models have tremendous power to democratize AI and make it more personal. The larger risk I see with open models is right in front of us: in the last week I've read about how many popular platforms quietly send data to foreign jurisdictions, or worse, sell or train on it to gain an advantage. And this is becoming more and more mainstream. Now more than ever open model vendors must act in the user's best interest, not their own: zero data retention or training, and hosting in the user's region vs sending data overseas. This has been our belief and commitment with Ollama. Nobody needs to slow down open models. We need to distribute and run them in a way users can trust.

  • View organization page for Ollama

    261,864 followers

    DeepSeek-V4.1-Flash is now fully rolled out and available on Ollama's Cloud: - Hosted in US & Europe - Zero data retention: prompts and responses are never logged or trained on - Per-token pricing matches the DeepSeek API, including off-peak pricing - Get started with Ollama's Pro, Max, and Team plans, or pay as you go with a free account with no service fees This new model by DeepSeek is more capable, faster, and more cost effective than all prior DeepSeek models including DeepSeek-V4-Pro.

  • View organization page for Ollama

    261,864 followers

    Introducing off-peak token rates. DeepSeek-V4-Flash and Pro are now half price outside of 12:00 to 18:00 UTC on weekdays (5am-11am pacific), and all day on weekends! Off-peak pricing will be available soon for more models. DeepSeek models on Ollama's cloud are hosted in the US & Europe with ZDR and fast performance.

  • Ollama reposted this

    Ollama (YC W21) is used by 9 million developers and 85% of the Fortune 500, giving co-founder and CEO Jeffrey Morgan a unique view into which AI models people are actually using and how that’s changing. Right now, the biggest shift he sees is toward open models, driven by coding agents, falling costs, and capabilities that are rapidly catching up to the frontier labs. On Ollama Cloud, that shift has driven a 150x increase in token usage since the start of the year. In this episode of the Lightcone, Jeff joins Garry, Jared, Diana, and Harj to talk about the future of open models and the story behind Ollama, from two years of searching for the right idea to building one of the most widely used AI developer tools in the world. https://lnkd.in/g5syCqa3

  • View organization page for Ollama

    261,864 followers

    GLM 5.3 and GLM 5.3 Flash (previously Ox Alpha) models are fully rolled out on Ollama's cloud. Private. Fast. US and Europe hosted. Zero data retention. GLM 5.3: ollama launch claude --model glm-5.3:cloud GLM 5.3 Flash: ollama launch opencode --model glm-5.3-flash:cloud Works with many more harnesses, or create an API key and plug it straight into your own app. GLM 5.3 model page: https://lnkd.in/g4HQa7ZT GLM 5.3 flash model page: https://lnkd.in/g6GDaJ9j

    • No alternative text description for this image
    • No alternative text description for this image
    • No alternative text description for this image
  • View organization page for Ollama

    261,864 followers

    IBM Granite 4.2 is now available on Ollama. 3B, 8B, 30B parameter open models made for enterprise agents. This model is free to use and is licensed for both research and commercial usage. The data curation and training processes were specifically designed for enterprise scenarios and customization, incorporating governance, risk, and compliance (GRC) evaluations alongside IBM’s standard data clearance and document quality review procedures. Model page: https://lnkd.in/gdvRZKtE

    View organization page for IBM Research

    106,655 followers

    Meet Granite 4.2, IBM’s latest family of open models purpose-built for enterprise agentic AI. With new native reasoning capabilities, Granite 4.2 can plan, reason, self-correct and reliably use tools to help automate complex enterprise workflows. Inside the release: ✅ 3B, 8B and 30B models with native thinking capabilities ✅ Advanced coding and software engineering capabilities ✅ Optimized speech models for high-throughput transcription workloads ✅ Flexible deployment across cloud, on-premises and edge environments Build what’s next with Granite 4.2: https://ibm.co/6041Er8dR Available on: Hugging Face, Replicate, DeepInfra, Ollama, LM Studio, AnythingLLM, RadixArk, OpenRouter, Artificial Analysis, CoreWeave Inference, Arena, Unsloth AI

  • Ollama reposted this

    This Thursday, August 20th, we’re bringing together four leaders at the forefront of open models, coding agents, and AI infrastructure: Kyle Kranen, Senior Manager, Dynamo at NVIDIA Yunmo Koo, Founding Engineer at FriendliAI Dongluo Chen, Software Engineer at Ollama Saurya Velagapudi, Principal Engineer at OpenHands Together, they’ll explore what it takes to move AI coding agents to production and how open models are changing the economics of AI-native engineering. Join us for the conversation, followed by audience Q&A and networking. Spots limited! 📍 San Francisco 📅 Aug. 20 🎟️ RSVP → https://luma.com/o9tv62y5

    • No alternative text description for this image

Similar pages

Browse jobs