NeuReality
Founded 2019 · 78 employees on LinkedIn · 18 known investors
NeuReality builds software and networking infrastructure that helps AI teams improve GPU and XPU utilization, throughput, and cost per token for production AI inference. Its offerings include an inference operating system for orchestrating AI workloads and a purpose-built AI networking chip for moving data across distributed accelerator infrastructure.
Also known as NeuReality Ltd.
Founders & leadership
NeuReality was founded in 2019 by Moshe Tanach, Tzvika Shmueli, and Yossi Kasus.
Investors · 18
Also in the syndicate · 9
Funding
SEC filings, press & company announcements$55M disclosed across 2 of 4 rounds · 2022–2026
- $20MraisedMar 2024 · 3 sources
Alumni Venture Group, Cardumen Capital, Cleveland Avenue, European Innovation Council (EIC) Fund, Glory Ventures, OurCrowd, Varana Capital, XT Hi-Tech
Source ↗ - $35MSeries ADec 2022 · 2 sources
Cardumen Capital (lead), OurCrowd (lead), Samsung Ventures (lead), Varana Capital (lead), XT Hitech (lead), Cleveland Avenue, Glory Ventures, Korean Investment Partners, SK Hynix, SK hynix Inc., StoneBridge
Source ↗
Source: company announcements and press reports — follow each round's link for the claim.
Company profile
researched Aug 2026NeuReality is an AI inference infrastructure company founded in 2019 in Israel that develops purpose-built silicon, systems and software intended to raise the utilization of AI accelerators in data centers and on-premises deployments. Its first-generation product line centers on the NR1 chip, described as a 7nm "AI-CPU" and Network Addressable Processing Unit (NAPU) that replaces conventional host CPU and NIC architecture, combining an ARM-based CPU, media/DSP processors, integrated network functions, a neural network engine and hardware-based AI-Hypervisor IP on a single die. The chip is packaged as the NR1-M module, a full-height double-wide PCIe card, and inside the NR1-S appliance, which pairs several NR1 modules with third-party AI accelerators in 1:1, 1:2 or 1:4 configurations and has been demonstrated with Qualcomm, AMD and IBM accelerators.
Alongside the hardware, the company offers a full-stack software layer organized into model, pipeline and service layers, with an SDK exposing three APIs (toolchain, provisioning and inference) covering development, deployment and serving of AI pipelines. The stack targets generative AI, agentic AI, RAG, computer vision, natural language processing and speech recognition workloads, runs on ARM server-class standard Linux, integrates with Kubernetes orchestration, NFS and S3 storage, and connects to common AI frameworks and MLOps backends. More recently the company has positioned its portfolio around "token factories": NR-NEXUS, an inference operating system providing orchestration, routing, observability and policy controls across heterogeneous GPU and XPU environments, and the NR2 AI-SuperNIC, scale-out networking silicon specified at 1.6 Tbps throughput with Ultra Ethernet Consortium support and in-network compute for low-latency east-west data movement in distributed clusters.
Founding story
NeuReality was founded in 2019 by Moshe Tanach, Tzvika Shmueli and Yossi Kasus. Tanach had been a director of engineering at Marvell and Intel and AVP R&D at DesignArt-Networks (acquired by Qualcomm); Shmueli was VP of back-end infrastructure at Mellanox Technologies and VP of engineering at Habana Labs; Kasus was senior director of engineering at Mellanox and head of VLSI at EZChip. CTO Lior Khermosh, previously co-founder and chief scientist of ParallelM and a fellow at PMC Sierra, is also part of the leadership team.
Business model
NeuReality sells inference infrastructure as a combination of silicon (NR1 AI-CPU, NR2 AI-SuperNIC), integrated modules and appliances, and an accompanying software stack and APIs. For NR-NEXUS it offers three commercial paths: a serverless API billed per token consumed, a fully managed dedicated deployment with enterprise SLAs, and an annual license for customers running managed or owned infrastructure on private capacity.
Stated deployment paths include pay-per-token consumption on a serverless API, managed dedicated capacity with enterprise SLAs, and annual software licensing for self-owned or managed infrastructure, in addition to sales of NR1 modules and appliances.
Traction
First cloud computing and financial services customers were reported to be running NR1 appliances on site, showing up to 6x better total cost per AI mega-token versus traditional CPU/GPU setups. The NR1-S appliance has been publicly demonstrated with Qualcomm, AMD and IBM accelerators, and the company announced collaborations with IBM, AMD and Lenovo. Cumulative disclosed funding reached $70 million as of March 2024.
Latest developments
The company has shifted its public positioning toward "token factories," unveiling the NR-NEXUS inference operating system and the NR2 AI-SuperNIC scale-out networking chip, publishing a white paper on AI networking bottlenecks, and appointing a former Google AI director to take the inference operating system to market. An October 2025 technical interview with CEO Moshe Tanach outlined the move from the first-generation NR1 AI-CPU to the NR2 AI-SuperNIC for east-west scale-out communication in large disaggregated GPU clusters.
▸Full profile — market position, technology, go-to-market, geography, history, risks & controversies
Market position
NeuReality describes itself as an AI inference specialist competing with the conventional CPU-plus-NIC host architecture used to front AI accelerators, and reports collaboration with AI ecosystem partners and customers including IBM, AMD and Lenovo, plus a partnership with TSMC. Press coverage frames it as an Israeli AI inferencing chip startup that had raised $70 million in total by March 2024.
The company positions itself as an open, accelerator-agnostic inference vendor whose AI-CPU complements rather than competes with GPUs and other accelerators, and which co-designs software with silicon rather than shipping hardware alone. Claimed advantages over CPU-reliant inference systems include 50-90% performance gains, up to 15x greater energy efficiency, 6.5x more AI token output, and up to 6x better total cost per AI mega-token in customer appliances.
Technology
Core technology is an AI-centric architecture that moves data-path and pipeline functions from software into hardware. The 7nm NR1 chip integrates an ARM-based CPU, media processors, integrated NIC/network functions, a built-in neural network engine and virtualization, orchestrated by the hardware AI-Hypervisor IP, with AI-over-Fabric technology and AI-pipeline offload; the company says the design supports heterogeneous accelerators including GPUs, FPGAs and ASICs, offers linear pipeline processing and linear scalability versus a partitioned CPU/NIC/PCI-switch host, and lifts accelerator utilization from under 50% toward nearly 100%. The NR2 AI-SuperNIC adds scale-out networking silicon with 1.6 Tbps throughput, UEC support, deterministic low latency and in-network compute, while NR-NEXUS provides the software control plane for serving, routing, utilization, latency and cost governance.
Go-to-market
Direct engagement with enterprise and data center customers supported by developer self-service (API access sign-up and SDK), ecosystem partnerships with accelerator, systems and foundry vendors, and visibility through conference keynotes and technical interviews (AI Summit New York, SC23, TSMC Europe Technology Symposium, Generative AI Summit). Commercially it offers serverless, managed and licensed deployment options, and has recruited a former Google AI director to bring its inference operating system to market.
Data center and cloud operators, enterprise platform and MLOps teams, and near-edge on-premises deployments requiring high-volume inference; early customers cited are in cloud computing and financial services, and the company also targets hyperscalers and less technical enterprises seeking simpler AI deployment.
Geography
Israel-based company with customer-facing activity in the United States and Europe, including appearances at events in New York, Amsterdam and SC23.
History
Founded in 2019, the company raised a seed round in early 2020 and a $35 million Series A announced in December 2022 that brought total funding to $48 million, intended to fund first deployments of its inference solutions in 2023. In 2023 it presented at the TSMC Europe Technology Symposium and showed the NR1-S appliance paired with third-party accelerators at SC23. In March 2024 it raised a further $20 million, lifting total funding to $70 million, earmarked for wider deployment of the NR1-M system. Subsequently the company introduced the NR2 AI-SuperNIC for scale-out networking and unveiled the NR-NEXUS inference operating system, and hired a former Google AI director to lead its market rollout.
Risks & controversies
Performance, utilization and cost-per-token improvements are company-stated figures rather than independently verified benchmarks. As a fabless silicon startup, the company depends on foundry partners such as TSMC, on third-party accelerator ecosystems, and on continued financing to move successive chip generations into volume deployment.
Compiled by commissioned research from 8 cited public sources — announcements, filings, and press listed under research sources below.
Key figures
latest reportedCompany-reported or press-reported figures, each dated to when it was claimed — not independently audited.
Competitors · 6
by search overlapCompanies competing with NeuReality for the same Google search keywords, organic and paid, via search-intersection analysis.
Timeline · 8
launches, deals, and filingsSecond-generation networking silicon addressing east-west scale-out communication in large disaggregated GPU clusters, with 1.6 Tbps throughput, ultra-low latency, UEC support and in-network compute; discussed in a CEO interview published 30 October 2025.
NeuReality appointed a former Google AI director to steer its inference operating system into the market.
NR-NEXUS provides orchestration, routing, observability and policy controls across heterogeneous GPU and XPU inference environments, offered via serverless API, managed deployment or annual license.
Round supported by the European Innovation Council Fund, Varana Capital, Cleveland Avenue, XT Hi-Tech and OurCrowd, with Cardumen Capital, Glory Ventures and Alumni Venture Group participating; proceeds to accelerate deployment of the NR1-M module to more customers.
$20M source ↗
CEO Moshe Tanach presented the NeuReality platform and its partnership with TSMC at the TSMC 2023 Europe Technology Symposium Innovation Zone in Amsterdam.
Series A led by Samsung Ventures, Cardumen Capital, Varana Capital, OurCrowd and XT Hi-Tech, with SK Hynix, Cleveland Avenue, Korea Investment Partners, StoneBridge and Glory Ventures participating; brought total funding to $48 million and was earmarked for deploying inference solutions in 2023.
$35M source ↗
The company reported close collaboration with AI ecosystem partners and customers including IBM, AMD and Lenovo.
Dated company events from announcements, filings, and press; legal rows summarize public dockets and regulator releases.
In the news
▸Research sources · 8
primary sources listed
- NeuRealityneureality.ai · web
8 public sources were cited for this profile; the first-party ones are listed here.
Frequently asked questions
- What does NeuReality do?
- Israeli AI inference infrastructure company building an inference operating system, AI-CPU and AI networking silicon for GPU/XPU clusters.
- Who founded NeuReality?
- NeuReality was founded by Moshe Tanach, Tzvika Shmueli, Yossi Kasus in 2019.
- Who are NeuReality's investors?
- NeuReality's investors include Cardumen Capital, Cleveland Avenue, Gefen Capital, OurCrowd, Varana, XT Venture Capital, Samsung Ventures, Stonebridge and 1 more.
- How much funding has NeuReality raised?
- NeuReality has disclosed $55M raised across 2 of its 4 known rounds.










