Fundraising Fox

General Reasoning

Entrepreneur First '25

San Francisco, US · Founded 2025 · Delaware corporation · 9 employees on LinkedIn · 2 known investors

General Reasoning (GR) is an AI research company that trains models to work, communicate, and coordinate over long time horizons, with a focus on reinforcement learning, multi-agent systems, and environment design. It also publishes open releases such as benchmarks and RL environment tools to support the wider research community.

Also known as General Reasoning, Inc. · GeneralReasoning · GR

Founders & leadership

General Reasoning was founded in 2025 by Ross James Taylor, Ross Taylor, and Chengxi Taylor.

RJ
Ross James TaylorCo-founder
RTRoss Taylor
Ross TaylorinCEORoss Taylor is CEO of General Reasoning, an AI research company focusing on long-horizon reasoning and multi-agent systems. Previously, he led large language model work at Meta AI including Llama 2 and Llama 3, and co-founded Papers with Code, which was acquired by Meta.
CTChengxi Taylor
Chengxi TaylorinCo-founder & PresidentChengxi Taylor is Co-founder and President of General Reasoning, an AI research company focused on reinforcement learning and multi-agent systems. She previously founded four technology companies spanning 3D printing, AR, and AI, and holds an Oxford MBA and CFA qualification.
CW
Chengxi WangNamed on SEC filing

Board

TJ
Thomas Joseph Grady IIBoard director
KG
Kip Gabriel ParkerBoard director

Investors · 2

Reported raises · per SEC filings

Form D private placements

$10.9M disclosed across 1 round · 2025

$10.9MraisedJul 2025 · 8 investors · Other Technology
Rule 506(b)
Officers, directors & promoters on the filing
  • Chengxi WangExecutive Officer, Director
  • Thomas Joseph Grady IIDirector
  • Kip Gabriel ParkerDirector
  • Ross James TaylorExecutive Officer, Director
Offering amount
$10.9M
Amount sold
$10.9M
First sale
Jul 2025
Incorporated
Corporation, Delaware, 2025
Federal exemptions
06b
Full filing on SEC EDGAR ↗

Source: SEC EDGAR Form D. Amounts as filed; amended filings shown once at their latest values.

Company profile

researched Aug 2026

General Reasoning (GR) is an AI research company working on models that can work, communicate and coordinate over long time horizons. Its stated research premise is that current large language models are procedurally strong but unreliable in non-stationary environments resembling real-world conditions. Research domains listed are reinforcement learning, multi-agent systems and environment design. Alongside internal research, the company publishes open releases intended for the wider research community.

Public outputs to date include OpenReward (March 2026), an open resource and API for discovering and running community-built RL environments with autoscaled infrastructure, accompanied by the Open Reward Standard (ORS) specification; KellyBench (April 2026), a benchmark and environment for long-horizon sequential decision making in which every frontier model evaluated lost money over a simulated full Premier League season; and BackSearch (July 2026), a frozen news archive supporting point-in-time web search and fetch for forecasting, backtesting and RL environments.

A third-party vendor directory characterises GR as an open-source RL environment vendor with more than 330 environments reachable through a single API, deployable as managed-hosted, self-hosted or via API, and describes the founding team as previously leading open language model development at Meta.

Founding story

Sources indicate a 2025 founding by a team that previously worked on open language models at Meta; Ross Taylor, co-founder and CEO, is described as ex-Meta AI/FAIR, research lead on Galactica, lead for reasoning on Llama 2 and Llama 3, and a co-founder of Papers with Code (acquired by Meta). Other named directors and staff include Kip Parker, Chengxi Wang, Thomas Grady, Iliyan Zarov and Henry Course. No further founding narrative is provided in the sources.

Business model

The company operates as an AI research lab that pairs internal research on long-horizon model capabilities with open releases. Its distributed product surface — RL environments and infrastructure — is offered across managed-hosted, self-hosted and API deployment models, and its open-source components include an Apache-2.0 licensed harness repository (firehorse) and the Open Reward Standard specification. No pricing, revenue or investor information is disclosed in the available sources.

Not disclosed in the available sources; no pricing, contract or revenue information is published, and a third-party profile lists revenue signals as unknown.

Traction

Three public releases between March and July 2026, 330+ RL environments accessible via the OpenReward API, mainstream press coverage of KellyBench findings in April 2026, and a Hugging Face organisation with two team members and no public models or datasets at the time of capture. A vendor directory lists product maturity as GA (estimated) and headcount at roughly 10.

Latest developments

The most recent publicly dated release is BackSearch on 24 July 2026, a point-in-time news archive with search and fetch for forecasting, backtesting and RL environments. Preceding it were KellyBench (9 April 2026) and OpenReward (24 March 2026). KellyBench results generated press coverage on 10 April 2026 reporting that frontier models from Google, OpenAI, Anthropic and xAI lost money predicting Premier League match outcomes.

Full profile — market position, technology, go-to-market, geography, history, risks & controversies

Market position

A small, early-stage research lab (headcount reported at ~10, within an 11-50 band) categorised by an RL-vendor directory as an open-source RL environment provider, listed alongside Prime Intellect and Good Start Labs as related vendors. The directory rates overall confidence in its profile as medium, notes no disclosed investors or valuation, and leaves revenue signals and security certifications unknown.

The company positions itself around long-horizon and non-stationary settings rather than single-turn benchmarks, combining benchmark design (KellyBench), an open environment standard and registry (OpenReward / Open Reward Standard) and point-in-time data infrastructure (BackSearch). The site cites prior team contributions at Meta — Galactica, Llama 2 and Llama 3 — plus research on thinking tokens, data-scarce training methods and early RL methods for LLM reasoning.

Technology

Reinforcement learning, multi-agent systems and environment design. Artifacts include the Open Reward Standard, an open specification connecting language models to community-built RL environments with 330+ environments behind one API; the KellyBench long-horizon, non-stationary evaluation environment; BackSearch, a frozen point-in-time news archive with search and fetch; and a GitHub organisation hosting environment repositories and the Apache-2.0 licensed firehorse harness.

Go-to-market

Distribution is primarily through open releases and public artifacts: an open environment registry and API (OpenReward), open specifications and repositories, published benchmarks, and a Hugging Face organisation (no public models or datasets listed as of the snapshot). Recruiting is handled through a public careers page listing three open roles.

Teams that need open, portable RL environments and long-horizon agent benchmarks they can run self-hosted or via managed/API hosting, per a third-party buyer analysis; the wider AI research community is also addressed through open releases.

Geography

An operating research hub in Shoreditch, London, United Kingdom, with the legal entity General Reasoning, Inc. registered in the United States; the SEC Form D lists an address at 2261 Market Street Ste 86270, San Francisco, CA 94114. A third-party profile records the company as not distributed/remote (estimated) and lists no other locations.

History

Founded in 2025, with a US legal entity, General Reasoning, Inc., registered in San Francisco and an operating research hub in Shoreditch, London. An SEC Form D filed 2025-07-11 reported $10,904,992 in equity sold. Public releases followed in 2026: OpenReward (March 2026), KellyBench (April 2026) and BackSearch (July 2026), with press coverage of the KellyBench results in April 2026.

Risks & controversies

No controversies are reported in the sources. Data limitations are notable: investors, valuation, revenue and SOC 2 / certification status are all unknown; the reported total raised derives from a single SEC Form D amount sold rather than a lifetime total, with no stage label. Organisations named in connection with OpenReward — NVIDIA, Nebius, Eigent, OpenAI and Meta — are described by the third-party profile as self-claimed environment providers/contributors on the company's own page, not verified paying customers. Separately, several unrelated entities share similar names (a personal technical blog at generalreasoning.com and a compliance-workflow company at genreason.com), creating identification risk.

Compiled by commissioned research from 8 cited public sources — announcements, filings, and press listed under research sources below.

Key figures

latest reported
HeadcountJun 202610 people
Open rolesJun 20263 roles
Researchers named on siteJun 20266 people
Rl environments availableJun 2026330 environments
Total raisedJun 2026$10.9M

Company-reported or press-reported figures, each dated to when it was claimed — not independently audited.

Competitors · 2

by search overlap

Companies competing with General Reasoning for the same Google search keywords, organic and paid, via search-intersection analysis.

Timeline · 4

launches, deals, and filings
Jul 2026
Introducing BackSearch

A frozen news archive with point-in-time search and fetch, allowing the web to be searched as it was on any past date, aimed at forecasting, backtesting and RL environments.

source ↗

Apr 2026
Press coverage of KellyBench football-betting results

News articles reported that models from Google, OpenAI, Anthropic and xAI struggled to predict Premier League scores over a full season and lost money in the benchmark, citing General Reasoning research.

source ↗

Apr 2026
Introducing KellyBench

Benchmark/environment for long-horizon sequential decision making in non-stationary settings; every frontier model evaluated lost money over a full Premier League season.

source ↗

Mar 2026
Introducing OpenReward

Open resource for discovering and experimenting with community RL environments, offering simple API endpoints and autoscaled infrastructure; accompanied by the Open Reward Standard (ORS), an open specification for connecting language models to community-built RL environments, with 330+ environments accessible through one API.

source ↗

Dated company events from announcements, filings, and press; legal rows summarize public dockets and regulator releases.

Legal entities · 1

corporate structure
General ReasoningDelaware

Research sources · 8

primary sources listed

8 public sources were cited for this profile; the first-party ones are listed here.

Frequently asked questions

What does General Reasoning do?
AI research lab building open RL environments, benchmarks and infrastructure for training and evaluating long-horizon agents.
Who founded General Reasoning?
General Reasoning was founded by Ross James Taylor, Ross Taylor, Chengxi Taylor in 2025.
Who are General Reasoning's investors?
General Reasoning's investors include Air Street Capital, Entrepreneur First.
How much funding has General Reasoning raised?
General Reasoning has disclosed $10.9M raised across 1 round.
Where is General Reasoning headquartered?
General Reasoning is headquartered in San Francisco, US.