Fundraising Fox

Together AI

Unicorn · $8.3B

San Francisco, US · 42 known investors

together.ai

Together AI builds a full-stack platform for production AI systems, serving teams that need to deploy and scale machine learning models reliably. The company provides infrastructure and tools for AI development and deployment.

Also known as Together · Together Computer

AI & Machine LearningCloud ComputingData & Infrastructure

Founders & leadership

VVVipul Ved Prakash
Vipul Ved Prakashin𝕏Founder
CZCe Zhang
Ce Zhangin𝕏Founder
CRChris Ré
Chris Ré
TDTri Dao
Tri Daoin𝕏Founder
PLPercy Liang
Percy Liangin𝕏Founder

Investors · 42

Also in the syndicate · 7

DAMAC CapitalJohn ChambersMarch CapitalProsperity7leadS Ventures (SentinelOne)Scott BanisterSK Telecom

Funding

SEC filings, press & company announcements

$800M disclosed across 1 of 4 rounds · 2023–2026

Source: company announcements and press reports — follow each round's link for the claim.

Valuation · disclosed

Disclosed events
$8.3Bvaluation at Series CJul 2026
filing ↗
$3.3Bvaluation at Series BFeb 2025
filing ↗

Source: SEC prospectus filings, and round valuations the company or its investors disclosed — follow each entry's link for the claim.

Company profile

researched Aug 2026

Together AI operates what it calls an "AI Native Cloud": a full-stack platform spanning inference, model shaping (fine-tuning) and accelerated compute for training and pre-training, built around open-source models. Product lines listed on its site include serverless inference, batch inference (scaling to 30 billion tokens per model), provisioned throughput with reserved capacity and a 99% uptime SLA, dedicated model inference, dedicated container inference for generative media (video, audio, image) workloads, accelerated compute clusters, code sandboxes for AI apps and agents, managed object storage and parallel filesystems with zero egress fees, and fine-tuning services.

Externally the company is characterized as an AI "neocloud" that rents Nvidia GPU clusters and other AI-specific infrastructure. At its February 2025 Series B the company described an AI Acceleration Cloud supporting more than 200 open-source models across chat, image, audio, vision, code and embeddings, powered by a proprietary inference engine and research techniques such as FlashAttention-3 kernels and quantization, and claimed 2-3x faster inference than hyperscaler alternatives. Infrastructure expansion at that time included 200 MW of secured power capacity, deployments of NVIDIA Blackwell GPUs across multiple North American data centers, and a partnership with Hypertec to co-build a cluster of 36,000 NVIDIA GB200 NVL72 GPUs.

The company maintains an in-house research organization, Together Research, publishing work on kernels (FlashAttention-4, Together Kernel Collection, ParallelKernelBench), architectures (Mamba-3, Parcae), inference optimization (speculative decoding, cache-aware prefill-decode disaggregation, consistency diffusion language models) and agents (ThunderAgent, CoderForge, DSGym), often co-authored with researchers at Princeton, Stanford, CMU and Berkeley.

Founding story

The company was founded in 2022. Co-founder and CEO Vipul Ved Prakash previously founded social media search platform Topsy, which he sold to Apple in 2013 for a reported $200+ million. His co-founders include Stanford professor Percy Liang and ETH Zürich / University of Chicago associate professor Ce Zhang; Chris Ré and Tri Dao (creator of FlashAttention, who serves as Chief Scientist) are also listed as founders.

Business model

Together AI sells cloud infrastructure and platform services for AI workloads, including on-demand and serverless inference, committed/provisioned throughput, dedicated deployments, GPU clusters, storage and fine-tuning. It is available directly and through the AWS Marketplace, and offers both self-serve sign-up and enterprise sales contact.

Usage- and capacity-based cloud pricing: token-based pricing for serverless and provisioned inference, committed reserved throughput, dedicated infrastructure deployments and GPU cluster rental, with no long-term commitment on serverless and zero egress fees on managed storage. TechCrunch reported annual bookings of over $1.15 billion as of the company's last quarter before July 2026.

Traction

Over 450,000 AI developers and companies were reported on the platform as of February 2025, with more than 200 open-source models supported. By July 2026 the company reported thousands of paying customers and annual bookings above $1.15 billion, and had raised roughly $1.2 billion across Series A, B and C rounds at an $8.3 billion valuation.

Latest developments

In July 2026 the company announced an $800 million Series C at an $8.3 billion valuation led by Aramco Ventures, with Vista Equity Partners, General Catalyst, Emergence Capital, Nvidia, March Capital, Pegatron and SentinelOne's S Ventures participating. Recent site announcements include a partnership with Y Combinator to deliver a dedicated YC GPU cluster, on-demand B200 availability on Together GPU Clusters, and serving of MiniMax-M3. Research output includes FlashAttention-4, Mamba-3, Aurora, ThunderAgent and CoderForge-Preview, with a stated presence at ICML 2026.

Full profile — market position, technology, go-to-market, geography, history, risks & controversies

Market position

Positioned as a leading AI neocloud focused on open-source models, competing with hyperscaler inference offerings and other neoclouds such as Groq, TensorWave and Upscale AI. TechCrunch frames its growth against an industry-wide tripling of open-source model usage in the past year as buyers seek lower-cost alternatives to closed frontier model tokens.

Emphasis on systems research feeding directly into product (kernels, inference optimization, architectures), a full-stack offering spanning experimentation to large-scale training and serving, support for a large catalog of open-source models with model ownership and privacy controls, and claimed price-performance advantages over hyperscaler inference.

Technology

Proprietary inference engine and the Together Kernel Collection, plus research-derived techniques including FlashAttention (and FlashAttention-3/-4), quantization, Mixture of Agents, Medusa, Sequoia, Hyena and Mamba. Site materials cite 2x faster inference, 60% lower cost through workload-specific optimization, and up to 90% faster pre-training with the Together Kernel Collection. The platform runs on NVIDIA hardware including H100, HGX B200 and GB200 NVL72 systems, and includes code sandbox infrastructure derived from CodeSandbox.

Go-to-market

Combines self-serve developer sign-up (Google, GitHub or SSO on the API console) with an enterprise sales motion ("Contact Sales", Together Enterprise Platform), marketplace distribution via AWS, and ecosystem partnerships such as the Y Combinator dedicated GPU cluster. Technical content marketing through the Together Research blog and open-source contributions (e.g. FlashAttention) supports developer acquisition. Leadership includes a Chief Revenue Officer, VPs of Sales (including EMEA), revenue operations and strategic partnerships.

AI developers, AI-native startups and global enterprises. Named users and customers include Salesforce, Zoom, SK Telecom, Hedra, Cognition, Zomato, Krea, Cartesia, The Washington Post, Cursor, Decagon, Pika Labs, Nexusflow and Voyage AI.

Geography

Deploys GPU clusters across multiple North American data centers, including DeepSeek model deployments in North America with opt-out privacy controls. It has a VP of Sales for EMEA, indicating European commercial coverage.

History

Founded in 2022, Together AI announced a $102.5 million Series A led by Kleiner Perkins (with Nvidia and Emergence Capital) in November 2023, positioning itself as a full-stack cloud for open-source AI. During 2024 it deployed DeepSeek models in North American data centers, launched the Together Enterprise Platform, announced AWS Marketplace availability, partnered with Cartesia on voice AI, and acquired CodeSandbox for code interpretation. In February 2025 it raised a $305 million Series B at a $3.3 billion valuation led by General Catalyst and co-led by Prosperity7, funding NVIDIA Blackwell deployments and a 36,000-GPU cluster with Hypertec. In July 2026 it announced an $800 million Series C at an $8.3 billion valuation led by Aramco Ventures.

Risks & controversies

Sources do not describe specific controversies. Reported context includes competitive pressure in the neocloud segment (Groq, TensorWave, Upscale AI raising large rounds) and reliance on Nvidia GPU supply and large power/data-center commitments. TechCrunch also noted that the company had reportedly sought $1 billion at a $7.5 billion valuation in March 2026 before closing $800 million at $8.3 billion.

Compiled by commissioned research from 8 cited public sources — announcements, filings, and press listed under research sources below.

Key figures

latest reported
Annual bookingsJul 2026$1.1B
Batch inference scaleJan 202630,000,000,000 tokens per model
Developers on platformFeb 2025450,000 developers
HeadcountAug 2026427
Open source models supportedFeb 2025200 models
Paying customersJul 2026Thousands of paying customers
Power capacity securedFeb 2025200 MW
Provisioned Throughput uptime SLAJan 202699%
Valuation (post-money, Series B)Feb 2025$3.3B
Valuation (post-money)Jul 2026$8.3B

Company-reported or press-reported figures, each dated to when it was claimed — not independently audited.

Competitors · 10

by search overlap
Hugging Face2150 shared keywordsHugging Face is a collaboration platform that hosts and provides access to machine learning models, datasets, and applications. It offers both open-source tools for the ML community and paid compute and enterprise solutions for teams building AI applications.
OpenRouter1408 shared keywordsOpenRouter operates an AI gateway that allows developers to access and compare hundreds of language models from multiple providers in a single platform. The service eliminates vendor lock-in while offering improved pricing, uptime, and reliability for companies and developers.
Ollama1041 shared keywordsOllama provides a platform for building with and running open-source language models locally and in the cloud. It enables developers to quickly deploy and access models like Claude Code and OpenClaw.
DataCamp945 shared keywordsDataCamp is a learning platform for teams that teaches data and AI skills through hands-on coursework accessible via web browser and mobile app. It serves enterprise customers and development teams seeking to build technical capabilities.
OpenAI723 shared keywordsOpenAI is an AI research and deployment company focused on developing artificial general intelligence (AGI) with emphasis on safety and beneficial outcomes for humanity.
Deep Infra675 shared keywordsDeepInfra provides inference infrastructure for deploying open-source AI models at scale, built from GPU hardware to API layer. The company serves organizations seeking reliable AI inference without vendor lock-in to proprietary models.
LM Studio547 shared keywordsLM Studio offers Bionic, a locally-run AI agent designed to work with open models for creativity, work, and coding tasks. The product runs natively on local machines rather than in the cloud.
Netezza508 shared keywordsIBM is a global technology company whose business spans enterprise software (including Red Hat, HashiCorp, and Confluent), IT infrastructure such as mainframes, servers, and storage, and IT consulting services. The company is also investing heavily in quantum computing and AI-based enterprise offerings, including its Lightwell open-source software security clearinghouse and the Anderon quantum wafer foundry.
LLM Stats505 shared keywordsllm-stats.com provides open, reproducible evaluation and benchmarking infrastructure for large language models and AI systems. The platform aggregates 200+ benchmarks and enables performance comparisons across frontier AI models to support informed decision-making by researchers, enterprises, and model developers.
Fireworks ai504 shared keywordsFireworks provides AI infrastructure services, leveraging experience from PyTorch, Meta, and Google. The company serves enterprises seeking product innovation through AI.

Companies competing with Together AI for the same Google search keywords, organic and paid, via search-intersection analysis.

Timeline · 13

launches, deals, and filings
Jul 2026
$800M Series C at an $8.3B valuation led by Aramco Ventures

TechCrunch reported the round, noting earlier reporting by The Information in March that the company had sought $1 billion at a $7.5 billion valuation.

$800M source ↗

Jan 2026
Partnership with Y Combinator to deliver a dedicated YC GPU cluster

source ↗

Jan 2026
On-demand NVIDIA B200 instances available on Together GPU Clusters

source ↗

Jan 2026
Serving MiniMax-M3 for inference

source ↗

Feb 2025
Acquisition of CodeSandbox

Together AI acquired CodeSandbox, adding built-in code interpretation capabilities to its platform; the acquisition was cited among milestones in the Series B announcement.

source ↗

Feb 2025
Secured 200 MW of power capacity for North American data center clusters

Deploying optimized clusters of NVIDIA Blackwell GPUs across multiple North American data centers.

source ↗

Feb 2025
$305M Series B led by General Catalyst and co-led by Prosperity7

Funding announced to scale the company's AI acceleration cloud, including a large-scale deployment of NVIDIA Blackwell GPUs.

$305M source ↗

Feb 2025
Kai Mak joins as Chief Revenue Officer and James Zou joins as researcher

source ↗

Feb 2025
Partnership with Cartesia for low-latency voice AI via Sonic model integration

source ↗

Feb 2025
Partnership with Hypertec to co-build a 36,000-GPU NVIDIA GB200 NVL72 cluster

source ↗

Feb 2025
Together Enterprise Platform launched and AWS Marketplace availability announced

source ↗

Feb 2025
Together GPU Clusters with NVIDIA HGX B200 GPUs made available

Announced immediate access to Together GPU Clusters accelerated by NVIDIA HGX B200 GPUs together with the Together Kernel Collection, cited as delivering 90% faster training performance than the previous generation.

source ↗

Nov 2023
Series A of $102.5M led by Kleiner Perkins announced

Together AI announced a $102.5 million Series A led by Kleiner Perkins, with partner Bucky Moore taking a board seat; Nvidia and Emergence Capital were also major contributors.

$102.5M source ↗

Dated company events from announcements, filings, and press; legal rows summarize public dockets and regulator releases.

In the news

Research sources · 8

primary sources listed

8 public sources were cited for this profile; the first-party ones are listed here.

Frequently asked questions

What does Together AI do?
AI neocloud offering a full-stack platform for inference, fine-tuning and GPU compute on open-source models.
Who founded Together AI?
Together AI was founded by Vipul Ved Prakash, Ce Zhang, Chris Ré, Tri Dao, Percy Liang.
Who are Together AI's investors?
Together AI's investors include 137 Ventures, A.Capital Ventures, Brave Capital, Brilliant Phoenix Capital, Cambium Capital, Chapter One, Coatue Management, Draft Ventures Llc and 27 more.
How much funding has Together AI raised?
Together AI has disclosed $800M raised across 1 of its 4 known rounds.
Where is Together AI headquartered?
Together AI is headquartered in San Francisco, US.