Curated by
Code Story
Code Story is a podcast featuring founders, tech leaders, CTO's, CEO's, and software architects, reflecting on their human story in creating world changing, disruptive digital products.
stacklist.com/codestory
Code Story is a podcast featuring founders, tech leaders, CTO's, CEO's, and software architects, reflecting on their human story in creating world changing, disruptive digital products.
More in Code Story S13 Episodes: Web Dev, AI & Tools
See all 14 →More from Code Story
See all stacks →S13 Bonus The Enterprise AI Chip War Rethinking LLM Silicon Inference With Vasanth Mohan Director Of
Code Story podcast episode featuring Vasanth Mohan from SambaNova discussing enterprise AI chip design and LLM inference optimization. The conversation covers hardware heterogeneity, latency budgets, disaggregated inference, and the economic factors driving enterprise adoption of self-hosted AI infrastructure versus proprietary APIs.
Built for AI agentsACO · 9043 tokens
Summary
Code Story podcast episode featuring Vasanth Mohan from SambaNova discussing enterprise AI chip design and LLM inference optimization. The conversation covers hardware heterogeneity, latency budgets, disaggregated inference, and the economic factors driving enterprise adoption of self-hosted AI infrastructure versus proprietary APIs.
Tags
ai-inference · llm-silicon · enterprise-ai · hardware-optimization · chip-architecture · sambanova · podcast
Key entities
Vasanth Mohan (person, 0.95) · Noah Labhart (person, 0.85) · SambaNova (organization, 0.98) · Code Story (organization, 0.95) · LLM (technology, 0.92) · GPU (technology, 0.88) · ASIC (technology, 0.85) · inference (concept, 0.95) · disaggregated-inference (concept, 0.88) · prefill-decode (concept, 0.85)
Classification
transcript · language en · status final
Provenance
claude-haiku-4-5 via @stacklist/be@0.1.0, confidence 0.85, 24 Sep 2026