/Companies

Ollama

YC W21

Palo Alto, US · Founded 2023 · 71 employees on LinkedIn · 26 known investors

Find your way into Ollama

Sign up to see every warm intro you have to Ollama

  • Paths you didn't know you had: your email and LinkedIn already hold routes to the Ollama team. We find them for you
  • 2nd- and 3rd-degree connections: the friend-of-a-friend routes that take hours to manually find through your inbox or LinkedIn
  • Answers now, not in days: momentum is everything in a raise. Skip asking around whether someone knows someone

Days of research, done the moment you sign in, with every route ranked by how warm it is.

Ollama provides a platform for building with and running open-source language models locally and in the cloud. It enables developers to quickly deploy and access models like Claude Code and OpenClaw.

Also known as Infra · Name TBD · Name TBD, Infra

Founders & leadership· Y Combinator alumni (W21)

Ollama was founded in 2023 by Jeffrey Morgan and Michael Chiang.

JMJeffrey Morgan
Jeffrey MorganCEO, Co-Founder
MCMichael Chiang
Michael ChiangCo-Founder
PD
Patrick DevineCTO
MF
Manning FisherDesign Leadformer

Investors · 26

8VCreportedAustin · $500K – $50M

How we know: press coverage · research · ollama.com · Y Combinator's founder network · Not right? Tell us

BenchmarkreportedSan Francisco · $100K – $15M

How we know: a partner's Signal profile · funding news · press coverage · open dataset · investor-profiles · research · ollama.com · open dataset · 5k.vc investor directory · Y Combinator's founder network · Not right? Tell us

Diogo Mónicareported

How we know: Y Combinator's founder network · Not right? Tell us

Enterprise Fundreported

How we know: Y Combinator's founder network · Not right? Tell us

Garage CapitalreportedWaterloo

How we know: research · ollama.com · Y Combinator's founder network · Not right? Tell us

Also in the syndicate · 8

49 PalmsAaron Katzindividual angelsMarianna TesselMichael MontanoQuinn SlackSolomon HykesSpencer Kimball

Funding

SEC filings, press & company announcements

$130M disclosed across 2 of 4 rounds · 2023–2026

Source: company announcements and press reports — follow each round's link for the claim.

Company profile

researched Aug 2026

Jeffrey (Jeff) Morgan and Michael Chiang founded Ollama in 2023 to make running open-weight language models on a laptop feel like using a package manager: `ollama run` downloads a model, applies a Modelfile/runtime configuration and exposes a consistent local API. The MIT-licensed project spread rapidly across macOS, Linux and Windows and became infrastructure behind editors, agent frameworks and tens of thousands of community integrations. Ollama does not train the underlying major models; it packages, quantizes, distributes and executes models from Meta, Google, DeepSeek, Qwen, Mistral, NVIDIA and others, subject to each model's own license. From 2025 it added optional cloud models so users can preserve the same local tools/API while offloading models too large for personal hardware; desktop apps, model sharing and agent launch flows broadened the product. Current individual pricing is Free ($0) and Pro ($20/month or $200/year), with Pro offering larger hosted models, three concurrent cloud models, 50x free usage and private-model upload/sharing; Team/Enterprise offerings are separately described/quoted. Local inference can keep prompts/data on the user's hardware, while cloud use is governed by Ollama's privacy/security terms and the company says cloud does not retain model inputs. In July 2026 Ollama announced a $65m Theory Ventures-led Series B, following a $15m Benchmark-led A; total funding is $88m, implying about $8m of seed/early capital. Investors include Benchmark, Theory, 8VC, YC, Garage, Pace, 49 Palms, GTMfund and prominent infrastructure founders. TechCrunch reported nearly nine million users/developers and a team of roughly 14 around the round. Risks include upstream model license/provenance, insecure exposure of local APIs, malicious model artifacts, resource/driver compatibility, cloud compute economics and competition from LM Studio, llama.cpp, vLLM, Docker Model Runner and hyperscaler/model-provider APIs.

Founding story

Morgan and Chiang saw open models becoming capable but difficult to install, quantize and operate, and created a Docker-like command-line experience that hid hardware/runtime complexity without forcing data into a hosted API.

Business model

Open-source local runtime with freemium hosted open-model cloud and paid individual/team/enterprise services.

Pro annual/monthly subscriptions, cloud compute/usage allocation and private model features; Team/Enterprise subscriptions/contracts, while local runtime and public models remain free.

Traction

July 2026 reporting: nearly 9m users/developers; company says cloud token volume has more than doubled monthly on average. Pricing page cites 40,000+ community integrations. Exact active/paid users, revenue and cloud gross margin are private.

Latest developments

Raised $65m Series B, scaled hosted access to large open models such as GLM, Nemotron, DeepSeek, Kimi and MiniMax, and integrated agent launch workflows while maintaining the free local runtime.

▸Full profile — market position, technology, go-to-market, geography, history, ownership, risks & controversies

Market position

De facto mainstream local-LLM developer runtime, competing with LM Studio, llama.cpp, GPT4All, Jan, vLLM/SGLang, Docker and direct hosted inference APIs.

Exceptionally simple developer experience and one interface spanning private local inference and optional hosted large open models, with a vast integration ecosystem and model-library distribution.

Technology

Cross-platform model runtime built on optimized inference backends, model manifests/Modelfiles and quantized artifacts, local HTTP/OpenAI-compatible APIs, GPU/CPU hardware detection, registry/distribution, cloud routing and desktop/CLI interfaces.

Go-to-market

Open-source GitHub/community and word of mouth, model library/search, easy CLI/desktop onboarding, integrations with IDEs/agent frameworks, free cloud access and paid upgrades.

Developers, researchers, privacy-conscious individuals and teams building agents/apps with open-weight models locally or in hosted cloud.

Individual developers and power users; AI application/agent teams; enterprises needing governed open-model access; educators/researchers and offline/edge users.

Geography

Globally downloaded open-source software and cloud service; headquarters/team in the San Francisco Bay Area with users across all major desktop/server environments.

History

Founded/public project 2023; rapid GitHub adoption and cross-platform releases; ~$8m early capital and $15m Benchmark A; introduced hosted cloud-model preview September 2025; Free/Pro monetization and desktop/agent flows; $65m Theory-led B July 2026, bringing total to $88m.

Ownership

Private; founders/employees and Benchmark, Theory Ventures, 8VC, Y Combinator, Garage Capital, Pace Capital, 49 Palms, GTMfund plus individual infrastructure executives/founders.

Risks & controversies

Third-party model licenses/terms and training-data provenance; untrusted model artifacts and prompt/tool supply-chain risk; local API/network misconfiguration; inconsistent results/hardware compatibility; cloud data/privacy and compute costs; dependence on upstream open-model releases; crowded commoditizing inference ecosystem.

Compiled by commissioned research from 9 cited public sources — announcements, filings, and press listed under research sources below.

Key figures

latest reported
Cloud token growthJul 2026100%
Community integrationsAug 202640,000 integrations minimum
EmployeesJul 202614 employees approximate
Users or developersJul 20269,000,000 users approximate

Company-reported or press-reported figures, each dated to when it was claimed — not independently audited.

Related companies · 10

Hugging FaceHugging Face is a collaboration platform that hosts and provides access to machine learning models, datasets, and applications. It offers both open-source tools for the ML community and paid compute and enterprise solutions for teams building AI applications.
HeliconeHelicone provides infrastructure for AI companies to route, debug, and analyze their applications. The platform serves developers and organizations building AI systems.
NVIDIANVIDIA designs GPUs, AI computing hardware, and software platforms for AI development, data centers, autonomous vehicles, and robotics, serving developers, researchers, and enterprises. Its offerings include AI models, power architecture for AI factories, and compute infrastructure used across industries such as manufacturing, healthcare, and automotive.
TypaTypa is a content creation and management platform powered by AI that helps teams draft, refine, and publish social media posts and written content. The tool serves marketing teams and content creators by automating research and drafting while maintaining brand voice and editorial control.
OpenRouterOpenRouter operates an AI gateway that allows developers to access and compare hundreds of language models from multiple providers in a single platform. The service eliminates vendor lock-in while offering improved pricing, uptime, and reliability for companies and developers.
Hudson Labs (formerly Bedrock AI)Hudson Labs provides AI-powered analysis software for investor relations teams to monitor peer disclosures, prepare earnings Q&A, and track messaging themes. The platform serves institutional investors, buy-side firms, and public company IR departments.
Sepal AISepal AI generates training and evaluation datasets for AI models at scale, working with leading AI research labs and enterprises to improve model performance on real-world tasks.
DataCampDataCamp is a learning platform for teams that teaches data and AI skills through hands-on coursework accessible via web browser and mobile app. It serves enterprise customers and development teams seeking to build technical capabilities.
nCompass TechnologiesnCompass Technologies builds an AI agent and VS Code extension that identifies GPU performance bottlenecks and generates optimization code, helping developers reduce GPU system analysis and implementation time from weeks to days.
LM StudioLM Studio offers Bionic, a locally-run AI agent designed to work with open models for creativity, work, and coding tasks. The product runs natively on local machines rather than in the cloud.

Companies working in the same space as Ollama.

Pricing

as listed Aug 2026
FreeOllamaIndividuals/developers · freemium
$0/month
ProOllamaPower users · subscription
$20/month

Public list pricing as researched from the company's own pricing pages; negotiated and enterprise terms vary.

Legal entities · 2

corporate structure
Ollama Inc.
Ollama Inc.Delaware, United States · active private

In the news

▸Research sources · 9

primary sources listed

9 public sources were cited for this profile; the first-party ones are listed here.

Frequently asked questions

What does Ollama do?
Open-source local AI runtime and hosted open-model platform that lets developers download, package and run language/multimodal models through a simple CLI, desktop apps and API.
Who founded Ollama?
Ollama was founded by Jeffrey Morgan, Michael Chiang in 2023.
Who are Ollama's investors?
Ollama's investors include 49 Palms Ventures, Angel Collective Opportunity Fund, Essence Venture Capital, GTMfund, Leonis Capital, Pace Capital, Rogue Capital, Sunflower Capital Management Fund and 10 more.
How much funding has Ollama raised?
Ollama has disclosed $130M raised across 2 of its 4 known rounds.
Where is Ollama headquartered?
Ollama is headquartered in Palo Alto, US.