Podcast EP363: The Impact of an AI Factory on Production Deployment with NeuReality’s Moshe Tanach
Daniel is joined by Moshe Tanach, co-founder and CEO of NeuReality. Prior to founding the company, he held senior engineering leadership roles at Marvell and Intel, where he led complex wireless and networking products from architecture through mass production. He also served as AVP of R&D at DesignArt Networks (later acquired by Qualcomm), where he led development of 4G base station technologies.
Daniel explores the challenges of moving AI inference into large-scale production with Moshe, who describes two primary challenges: data movement and workload orchestration. Moshe explains that as AI moves into production, performance and cost increasingly depend on the systems surrounding the accelerators. Items such as resource allocation, routing, and orchestration must be addressed to balance latency, utilization, and cost. Moshe describes in some detail the software and networking infrastructure NeuReality delivers to help AI teams get more from the GPU and XPU systems they already deploy with higher utilization, greater throughput, and lower cost.
He advances the concept of an “AI factory” that optimizes resources and architectures to allow AI deployment at scale. You can learn more about this unique company here.
The views, thoughts, and opinions expressed in these podcasts belong solely to the speaker, and not to the speaker’s employer, organization, committee or any other group or individual.



Synopsys Wins the AI Silicon Brain Drain