NeuReality CEO Moshe Tanach participated in a full conversation on the SemiWiki.com Semiconductor Insiders Podcast featuring Daniel Nenni.
Public source
Publisher name
Public post
The gross margin pyramid is upside down. Normally SaaS runs 90% gross margin and hardware companies run 50 to 60%. Right now NVIDIA takes 75% on AI chips, while many Gen…
Company
NeuReality
Transforming Heterogeneous AI Infrastructure into a Production Token Factory.
- Industry
- Semiconductor Manufacturing
- Location
- Tel Aviv, IL
- Company size
- 51–200 employees
About NeuReality
AI infrastructure has a hidden problem: the network and orchestration layers. As models scale to trillions of parameters and inference demand explodes, two bottlenecks emerge: how data moves between GPUs and how workloads are managed across them. The industry added more GPUs, scaled clusters, optimized models. But utilization still hovers around 50-70%. The compute is there, idle, burning watts. The bottleneck isn't the silicon. It's how data moves and how work gets distributed. Traditional networking was built for general-purpose workloads, not AI's east-west traffic and microsecond-sensitive synchronization. Traditional orchestration treats GPUs as generic compute, blind to the demands of prefill, decode, and model synchronization. Every GPU cycle wasted waiting is money and energy lost. We asked: What if the network wasn't just faster, but intelligent? What if orchestration understood AI workloads natively? NR-NEXUS is an inference operating system for large-scale production. Hardware-agnostic, it unifies fragmented open-source frameworks into a single production platform, running across hyperscale clouds, GPU clusters, and emerging XPUs. NR2 AI-SuperNIC eliminates data-movement bottlenecks limiting GPU utilization. It executes the networking data path in hardware with no CPUs in the critical path, integrates in-network compute to offload communication operations, and supports open Ethernet-based networking. Together, they transform distributed GPU and XPU clusters into high-throughput token factories. The result: GPUs at near-100% utilization. Inference scales without adding racks. Energy consumption drops. This isn't incremental optimization. It's rethinking the data path and control plane so AI infrastructure matches AI ambition. For our customers: maximum performance from existing hardware. Lower cost, lower power, lower latency, higher throughput. NeuReality is headquartered in Tel Aviv with offices across North America and Europe.
See moreLatest activity
Latest activity from NeuReality
2 signals
Discover more
Similar signals
Similar public activity from other companies.
Presence & Recognition
Cerebras
Cerebras went deep at Hot Chips 2026 on the technology behind CS-4 and shared a glimpse into CS-5 and CS-6.
Presence & Recognition
Phison
Phison US President and GM Michael Wu participated in an AI Advantage Podcast episode discussing AI infrastructure, storage innovation, leadership, and the defining moments shaping Phison's journey.
Presence & Recognition
SemiAnalysis
SemiAnalysis sat down with Sarah Chieng at Cerebras Supernova to talk chips.
Presence & Recognition
Dolphin Semiconductor
Dolphin Semiconductor featured Chief Marketing Officer Emmanuel Sambuis discussing product families including power management, monitoring, audio, and explaining semiconductor IP strategy
Presence & Recognition
Advantest