Neil Movva Discusses AI Inference on Podcast
Podcast episode features Neil Movva discussing AI inference layers, chips, power, and spending.
Patrick O'Shaughnessy posted on X announcing an Invest Like the Best episode with Neil Movva, founder of Sail Research. O'Shaughnessy described Movva as having started at Nvidia working on GPUs and now running Sail Research. The episode covers AI inference from software to chips and power, including latency versus throughput, chip pricing, data center strategies, Nvidia lore, and Movva's quoted view that token consumption today differs from dot-com era speculation because users consume tokens immediately. Replies on X largely praised the detailed explanations, while some questioned speculative risks in data center builds and AI funding.
Neil Movva (@neilmovva) started his career at Nvidia, working on GPUs and kernels, and has an unusually deep understanding of inference, from software to chips to power. We spend a lot of time on each of those layers, how they connect, and where the important tradeoffs are.…