White Circle
14 known investors
White Circle provides real-time control and monitoring tools for AI applications, enabling teams to test, control, analyze, and improve AI model behavior across safety, security, and compliance requirements.
Also known as White Circle AI Β· whitecircle
Founders & leadership
Board
Investors Β· 14
Also in the syndicate Β· 11
Funding
SEC filings, press & company announcements$11M disclosed across 1 of 2 rounds Β· 2026
- $11MseedMay 2026 Β· 5 sources
Hummingbird VC (lead), Abstract Ventures, Clement Delangue, David Cramer, Durk Kingma, Factorial, Francois Chollet, Guillaume Lample, Mehdi Ghissassi, Nomads, Olivier Pomel, Romain Huet, Thomas Wolf
Source β
Source: company announcements and press reports β follow each round's link for the claim.
Company profile
researched Aug 2026White Circle is an AI control and safety platform whose software sits between a company's users and its AI models, checking inputs and outputs in real time against customer-defined policies. The product spans four stages of the AI lifecycle: testing AI systems for failure modes, jailbreaks and hallucinations before release; enforcing guardrails across inputs, outputs and tool calls; analyzing how users and models behave over time; and optimizing systems through prompt engineering and model routing driven by live signals.
Capabilities described by the company include blocking unsafe inputs, jailbreak and prompt-injection prevention, PII detection, output validation and instruction-following verification, data-leak and hallucination detection, tool-abuse and token-drain prevention, compliance checks, risk scoring, anomaly detection and custom metrics. Policies can be written in natural language across safety, security, compliance and reliability, varied by user location or jurisdiction, and enforced by blocking, rewriting or escalating risky behavior β for example alerting a security channel, requesting human confirmation, or terminating a session. Every decision is auditable, showing what happened and which policies were triggered. The platform is offered as an API with an accompanying command-line tool (installable via pip as whitecircle_cli), and the company maintains a public trust center and status page.
Founding story
In late 2024, Denis Shilov devised a universal jailbreak prompt that instructed leading AI models to behave like an API endpoint rather than a chatbot with safety rules, causing them to answer requests they were supposed to refuse. He posted it on X, where it went viral overnight and led to an invitation from Anthropic to privately test its models. That experience convinced him that jailbreaks were only one facet of a broader problem: companies integrating AI into workflows had few ways to control model behavior once users interacted with it, which became the premise for White Circle.
Business model
B2B SaaS. White Circle sells an API-based control layer to businesses deploying AI in production, with enterprise-oriented features such as unlimited custom policies, regional policy variation, audit logging and a trust center; the site offers self-serve sign-up alongside a sales-led demo request.
Sold as a software platform/API to business customers; a self-serve "Get Started" path and a "Contact Sales" enterprise path are both offered.
Traction
The company states its platform has processed more than one billion API requests and cites Lovable, a vibe-coding startup, along with several fintech and legal companies, as users. A customer testimonial from CISO Igor Andriushchenko states that White Circle protects millions of AI interactions per day. Its Hugging Face organization lists 11 team members, five public datasets and two spaces.
Latest developments
In May 2026 White Circle disclosed an $11 million seed round to be used for team expansion, product development and customer growth in the U.S., U.K. and Europe, and published KillBench, a study of more than one million experiments across 15 models from OpenAI, Google, Anthropic and xAI examining behavior in forced decisions about human lives. Additional research on distributed training methods, Blackwell B300 kernels and offline GRPO is announced as forthcoming.
βΈFull profile β market position, technology, go-to-market, geography, history, risks & controversies
Market position
Positioned as a third-party AI control layer in the AI security and guardrails segment, arguing that model-provider safety tuning is too general for production use and that vendors have conflicting incentives to judge their own model outputs. Categorized as a seed-stage security and cybersecurity company; comparable companies cited include Pillar Security, Prompt Security, Adaptive Security and Reken AI.
Focuses on enforcing customer-specific behavior at runtime rather than relying on model-level safety training, covering the full lifecycle from pre-deployment testing to real-time enforcement, analytics and optimization, with jurisdiction-aware policies, auditable policy decisions and independent evaluation of guardrail models.
Technology
A runtime enforcement layer for AI applications that inspects prompts and model outputs against configurable policies, with guardrail models for moderation, jailbreak and prompt-injection detection, PII and data-leak detection, hallucination and reliability checking, and tool-abuse prevention, plus optimization features such as dynamic model routing, memory management and context enrichment. Support is claimed for 150+ languages, roughly five-minute API integration and 99.99% API uptime. The company also publishes research on evaluation (CircleGuardBench, KillBench) and has announced forthcoming work on communication-efficient distributed training (Hat-Muon), custom kernels for Blackwell B300 GPUs and offline GRPO at scale.
Go-to-market
Combines developer-led adoption (fast API integration, CLI, public Hugging Face datasets and benchmark spaces, documentation) with direct enterprise sales via demo requests, supported by published research such as CircleGuardBench and KillBench and by a network of prominent AI-industry angel investors.
Companies embedding foundation models and AI agents into their products, including coding and creative tools, fintech, legal, healthcare, travel, government, education and research organizations.
Geography
Headquartered in Paris, France, with a team of about 20 distributed across London, France, Amsterdam and elsewhere in Europe; customer growth is targeted across the U.S., U.K. and Europe.
History
The company emerged from Denis Shilov's late-2024 jailbreak research and published the CircleGuardBench moderation-guard benchmark in May 2025. In May 2026 it announced an $11 million seed round led by Hummingbird VC with participation from Factorial, Nomads, Abstract Ventures and angel investors from OpenAI, Anthropic, DeepMind, Mistral, Hugging Face, Datadog and Sentry, and released the KillBench study of decision-making biases across 15 frontier models.
Risks & controversies
The company's positioning depends on the premise that model providers will not solve safety at the training stage; its founder notes that AI labs bill for tokens even on refused requests and face an "alignment tax" trade-off between safety and performance, and questions whether customers should trust a lab to judge its own model's outputs. Public data on the company is inconsistent across third-party trackers, which report differing founding years, headquarters locations and headcounts.
Compiled by commissioned research from 8 cited public sources β announcements, filings, and press listed under research sources below.
Key figures
latest reportedCompany-reported or press-reported figures, each dated to when it was claimed β not independently audited.
Timeline Β· 3
launches, deals, and filingsWhite Circle announced an $11 million seed round to expand its team, accelerate product development and grow its customer base across the U.S., U.K. and Europe. Backers include Hummingbird VC as lead, plus Factorial, Nomads and Abstract Ventures and angels from OpenAI, Anthropic, Mistral, Hugging Face, DeepMind, Datadog and Sentry.
$11M source β
White Circle's research arm published KillBench, a study running more than one million experiments across 15 AI models from providers including OpenAI, Google, Anthropic and xAI to evaluate decision-making biases in scenarios involving choices about human lives.
White Circle released CircleGuardBench, a benchmark and Hugging Face space for evaluating the safety and accuracy of AI moderation models and LLM guardrails.
Dated company events from announcements, filings, and press; legal rows summarize public dockets and regulator releases.
In the news
βΈResearch sources Β· 8
primary sources listed
- White Circlewhitecircle.ai Β· web
8 public sources were cited for this profile; the first-party ones are listed here.
Frequently asked questions
- What does White Circle do?
- Paris-based White Circle builds a real-time control layer that tests, guards and monitors AI applications in production.
- Who are White Circle's investors?
- White Circle's investors include Abstract Ventures, Hummingbird Ventures, Moonfire Ventures.
- How much funding has White Circle raised?
- White Circle has disclosed $11M raised across 1 of its 2 known rounds.




