Fundraising Fox

NeuReality

Founded 2019 · 78 employees on LinkedIn · 18 known investors

NeuReality builds software and networking infrastructure that helps AI teams improve GPU and XPU utilization, throughput, and cost per token for production AI inference. Its offerings include an inference operating system for orchestrating AI workloads and a purpose-built AI networking chip for moving data across distributed accelerator infrastructure.

Also known as NeuReality Ltd.

Founders & leadership

NeuReality was founded in 2019 by Moshe Tanach, Tzvika Shmueli, and Yossi Kasus.

MTMoshe Tanach
Moshe TanachinCo-founder & CEO
TSTzvika Shmueli
Tzvika ShmueliinCo-Founder & COO
YKYossi Kasus
Yossi KasusinCo-Founder & VP VLSI
LKLior Khermosh
Lior KhermoshinCTO
IAItay Avilevich
Itay AvilevichinVP R&D
GS
Gaurav ShahinVP Business Development
IFIrit Fastovskiy
Irit FastovskiyinVP Business Operations and Chief of Staff

Investors · 18

Also in the syndicate · 9

Alumni Venture GroupEuropean Innovation Council (EIC) FundGlory VenturesKorea Investment PartnersKorean Investment PartnersSK HynixSK hynix Inc.Varana CapitalleadXT Hitechlead

Funding

SEC filings, press & company announcements

$55M disclosed across 2 of 4 rounds · 2022–2026

Source: company announcements and press reports — follow each round's link for the claim.

Company profile

researched Aug 2026

NeuReality is an AI inference infrastructure company founded in 2019 in Israel that develops purpose-built silicon, systems and software intended to raise the utilization of AI accelerators in data centers and on-premises deployments. Its first-generation product line centers on the NR1 chip, described as a 7nm "AI-CPU" and Network Addressable Processing Unit (NAPU) that replaces conventional host CPU and NIC architecture, combining an ARM-based CPU, media/DSP processors, integrated network functions, a neural network engine and hardware-based AI-Hypervisor IP on a single die. The chip is packaged as the NR1-M module, a full-height double-wide PCIe card, and inside the NR1-S appliance, which pairs several NR1 modules with third-party AI accelerators in 1:1, 1:2 or 1:4 configurations and has been demonstrated with Qualcomm, AMD and IBM accelerators.

Alongside the hardware, the company offers a full-stack software layer organized into model, pipeline and service layers, with an SDK exposing three APIs (toolchain, provisioning and inference) covering development, deployment and serving of AI pipelines. The stack targets generative AI, agentic AI, RAG, computer vision, natural language processing and speech recognition workloads, runs on ARM server-class standard Linux, integrates with Kubernetes orchestration, NFS and S3 storage, and connects to common AI frameworks and MLOps backends. More recently the company has positioned its portfolio around "token factories": NR-NEXUS, an inference operating system providing orchestration, routing, observability and policy controls across heterogeneous GPU and XPU environments, and the NR2 AI-SuperNIC, scale-out networking silicon specified at 1.6 Tbps throughput with Ultra Ethernet Consortium support and in-network compute for low-latency east-west data movement in distributed clusters.

Founding story

NeuReality was founded in 2019 by Moshe Tanach, Tzvika Shmueli and Yossi Kasus. Tanach had been a director of engineering at Marvell and Intel and AVP R&D at DesignArt-Networks (acquired by Qualcomm); Shmueli was VP of back-end infrastructure at Mellanox Technologies and VP of engineering at Habana Labs; Kasus was senior director of engineering at Mellanox and head of VLSI at EZChip. CTO Lior Khermosh, previously co-founder and chief scientist of ParallelM and a fellow at PMC Sierra, is also part of the leadership team.

Business model

NeuReality sells inference infrastructure as a combination of silicon (NR1 AI-CPU, NR2 AI-SuperNIC), integrated modules and appliances, and an accompanying software stack and APIs. For NR-NEXUS it offers three commercial paths: a serverless API billed per token consumed, a fully managed dedicated deployment with enterprise SLAs, and an annual license for customers running managed or owned infrastructure on private capacity.

Stated deployment paths include pay-per-token consumption on a serverless API, managed dedicated capacity with enterprise SLAs, and annual software licensing for self-owned or managed infrastructure, in addition to sales of NR1 modules and appliances.

Traction

First cloud computing and financial services customers were reported to be running NR1 appliances on site, showing up to 6x better total cost per AI mega-token versus traditional CPU/GPU setups. The NR1-S appliance has been publicly demonstrated with Qualcomm, AMD and IBM accelerators, and the company announced collaborations with IBM, AMD and Lenovo. Cumulative disclosed funding reached $70 million as of March 2024.

Latest developments

The company has shifted its public positioning toward "token factories," unveiling the NR-NEXUS inference operating system and the NR2 AI-SuperNIC scale-out networking chip, publishing a white paper on AI networking bottlenecks, and appointing a former Google AI director to take the inference operating system to market. An October 2025 technical interview with CEO Moshe Tanach outlined the move from the first-generation NR1 AI-CPU to the NR2 AI-SuperNIC for east-west scale-out communication in large disaggregated GPU clusters.

Full profile — market position, technology, go-to-market, geography, history, risks & controversies

Market position

NeuReality describes itself as an AI inference specialist competing with the conventional CPU-plus-NIC host architecture used to front AI accelerators, and reports collaboration with AI ecosystem partners and customers including IBM, AMD and Lenovo, plus a partnership with TSMC. Press coverage frames it as an Israeli AI inferencing chip startup that had raised $70 million in total by March 2024.

The company positions itself as an open, accelerator-agnostic inference vendor whose AI-CPU complements rather than competes with GPUs and other accelerators, and which co-designs software with silicon rather than shipping hardware alone. Claimed advantages over CPU-reliant inference systems include 50-90% performance gains, up to 15x greater energy efficiency, 6.5x more AI token output, and up to 6x better total cost per AI mega-token in customer appliances.

Technology

Core technology is an AI-centric architecture that moves data-path and pipeline functions from software into hardware. The 7nm NR1 chip integrates an ARM-based CPU, media processors, integrated NIC/network functions, a built-in neural network engine and virtualization, orchestrated by the hardware AI-Hypervisor IP, with AI-over-Fabric technology and AI-pipeline offload; the company says the design supports heterogeneous accelerators including GPUs, FPGAs and ASICs, offers linear pipeline processing and linear scalability versus a partitioned CPU/NIC/PCI-switch host, and lifts accelerator utilization from under 50% toward nearly 100%. The NR2 AI-SuperNIC adds scale-out networking silicon with 1.6 Tbps throughput, UEC support, deterministic low latency and in-network compute, while NR-NEXUS provides the software control plane for serving, routing, utilization, latency and cost governance.

Go-to-market

Direct engagement with enterprise and data center customers supported by developer self-service (API access sign-up and SDK), ecosystem partnerships with accelerator, systems and foundry vendors, and visibility through conference keynotes and technical interviews (AI Summit New York, SC23, TSMC Europe Technology Symposium, Generative AI Summit). Commercially it offers serverless, managed and licensed deployment options, and has recruited a former Google AI director to bring its inference operating system to market.

Data center and cloud operators, enterprise platform and MLOps teams, and near-edge on-premises deployments requiring high-volume inference; early customers cited are in cloud computing and financial services, and the company also targets hyperscalers and less technical enterprises seeking simpler AI deployment.

Geography

Israel-based company with customer-facing activity in the United States and Europe, including appearances at events in New York, Amsterdam and SC23.

History

Founded in 2019, the company raised a seed round in early 2020 and a $35 million Series A announced in December 2022 that brought total funding to $48 million, intended to fund first deployments of its inference solutions in 2023. In 2023 it presented at the TSMC Europe Technology Symposium and showed the NR1-S appliance paired with third-party accelerators at SC23. In March 2024 it raised a further $20 million, lifting total funding to $70 million, earmarked for wider deployment of the NR1-M system. Subsequently the company introduced the NR2 AI-SuperNIC for scale-out networking and unveiled the NR-NEXUS inference operating system, and hired a former Google AI director to lead its market rollout.

Risks & controversies

Performance, utilization and cost-per-token improvements are company-stated figures rather than independently verified benchmarks. As a fabless silicon startup, the company depends on foundry partners such as TSMC, on third-party accelerator ecosystems, and on continued financing to move successive chip generations into volume deployment.

Compiled by commissioned research from 8 cited public sources — announcements, filings, and press listed under research sources below.

Key figures

latest reported
Accelerator utilization with NR1 vs traditional CPU/NICJan 2025from under 50% to nearly 100% (company-stated)
Claimed cost per AI mega-token improvement in customer appliancesJan 2025up to 6x better total cost per AI mega-token vs traditional CPU/GPU setups
Claimed performance gain vs CPU-reliant inference systemsJan 202550-90% performance gains, up to 15x energy efficiency, 6.5x more AI token output
HeadcountAug 202678
NR1 chip process nodeJan 20257nm
NR2 AI-SuperNIC throughputJan 20251.6 Tbps
Total disclosed funding to dateMar 2024$70M

Company-reported or press-reported figures, each dated to when it was claimed — not independently audited.

Competitors · 6

by search overlap
Netezza12 shared keywordsIBM is a global technology company whose business spans enterprise software (including Red Hat, HashiCorp, and Confluent), IT infrastructure such as mainframes, servers, and storage, and IT consulting services. The company is also investing heavily in quantum computing and AI-based enterprise offerings, including its Lightwell open-source software security clearinghouse and the Anderon quantum wafer foundry.
NVIDIA12 shared keywordsNVIDIA designs GPUs, AI computing hardware, and software platforms for AI development, data centers, autonomous vehicles, and robotics, serving developers, researchers, and enterprises. Its offerings include AI models, power architecture for AI factories, and compute infrastructure used across industries such as manufacturing, healthcare, and automotive.
Nebius9 shared keywordsNebius operates an AI-focused cloud platform offering GPU infrastructure, MLOps tooling, and managed and serverless inference for AI training and deployment. It serves AI developers and enterprises, providing custom hardware with non-virtualized NVIDIA GPUs and InfiniBand for workloads ranging from experiments to global-scale environments.
Arm8 shared keywordsArm is a semiconductor company that designs and licenses processor IP and compute platforms used from edge devices to AI data centers, spanning markets such as data centers, automotive, and general computing. It is a publicly traded company majority-owned by SoftBank Group.
CoreWeave7 shared keywordsCoreWeave operates an AI-native cloud platform built to run large-scale AI workloads, combining GPU infrastructure, tooling, and support. It targets organizations developing and deploying complex AI models.
Cloudflare Turnstile6 shared keywordsCloudflare provides a global cloud network platform delivering security, performance, and development services through sixty-plus integrated services including SASE, application security, and full-stack development infrastructure.

Companies competing with NeuReality for the same Google search keywords, organic and paid, via search-intersection analysis.

Timeline · 8

launches, deals, and filings
Oct 2025
NR2 AI-SuperNIC introduced for scale-out AI networking

Second-generation networking silicon addressing east-west scale-out communication in large disaggregated GPU clusters, with 1.6 Tbps throughput, ultra-low latency, UEC support and in-network compute; discussed in a CEO interview published 30 October 2025.

source ↗

Jan 2025
Former Google AI director hired to lead inference operating system go-to-market

NeuReality appointed a former Google AI director to steer its inference operating system into the market.

source ↗

Jan 2025
NeuReality unveils NR-NEXUS inference operating system for AI token factories

NR-NEXUS provides orchestration, routing, observability and policy controls across heterogeneous GPU and XPU inference environments, offered via serverless API, managed deployment or annual license.

source ↗

Mar 2024
NeuReality raises $20M, bringing total funding to $70M

Round supported by the European Innovation Council Fund, Varana Capital, Cleveland Avenue, XT Hi-Tech and OurCrowd, with Cardumen Capital, Glory Ventures and Alumni Venture Group participating; proceeds to accelerate deployment of the NR1-M module to more customers.

$20M source ↗

Jul 2023
TSMC partnership highlighted at TSMC Europe Technology Symposium

CEO Moshe Tanach presented the NeuReality platform and its partnership with TSMC at the TSMC 2023 Europe Technology Symposium Innovation Zone in Amsterdam.

source ↗

Dec 2022
NeuReality announces $35M Series A

Series A led by Samsung Ventures, Cardumen Capital, Varana Capital, OurCrowd and XT Hi-Tech, with SK Hynix, Cleveland Avenue, Korea Investment Partners, StoneBridge and Glory Ventures participating; brought total funding to $48 million and was earmarked for deploying inference solutions in 2023.

$35M source ↗

Dec 2022
Collaboration with IBM, AMD and Lenovo

The company reported close collaboration with AI ecosystem partners and customers including IBM, AMD and Lenovo.

source ↗

Jan 2020
Seed round raised

The company raised its seed round in early 2020.

source ↗

Dated company events from announcements, filings, and press; legal rows summarize public dockets and regulator releases.

In the news

Research sources · 8

primary sources listed

8 public sources were cited for this profile; the first-party ones are listed here.

Frequently asked questions

What does NeuReality do?
Israeli AI inference infrastructure company building an inference operating system, AI-CPU and AI networking silicon for GPU/XPU clusters.
Who founded NeuReality?
NeuReality was founded by Moshe Tanach, Tzvika Shmueli, Yossi Kasus in 2019.
Who are NeuReality's investors?
NeuReality's investors include Cardumen Capital, Cleveland Avenue, Gefen Capital, OurCrowd, Varana, XT Venture Capital, Samsung Ventures, Stonebridge and 1 more.
How much funding has NeuReality raised?
NeuReality has disclosed $55M raised across 2 of its 4 known rounds.