Curated by

avatar

Code Story

stacklist.com/codestory

Code Story is a podcast featuring founders, tech leaders, CTO's, CEO's, and software architects, reflecting on their human story in creating world changing, disruptive digital products.

More in Code Story S13 Episodes: Web Dev, AI & Tools

See all 14 →

More from Code Story

See all stacks →

S13 Bonus The Enterprise AI Chip War Rethinking LLM Silicon Inference With Vasanth Mohan Director Of

Code Story podcast episode featuring Vasanth Mohan from SambaNova discussing enterprise AI chip design and LLM inference optimization. The conversation covers hardware heterogeneity, latency budgets, disaggregated inference, and the economic factors driving enterprise adoption of self-hosted AI infrastructure versus proprietary APIs.

View card
Built for AI agentsACO · 9043 tokens

Summary

Code Story podcast episode featuring Vasanth Mohan from SambaNova discussing enterprise AI chip design and LLM inference optimization. The conversation covers hardware heterogeneity, latency budgets, disaggregated inference, and the economic factors driving enterprise adoption of self-hosted AI infrastructure versus proprietary APIs.

Tags

ai-inference · llm-silicon · enterprise-ai · hardware-optimization · chip-architecture · sambanova · podcast

Key entities

Vasanth Mohan (person, 0.95) · Noah Labhart (person, 0.85) · SambaNova (organization, 0.98) · Code Story (organization, 0.95) · LLM (technology, 0.92) · GPU (technology, 0.88) · ASIC (technology, 0.85) · inference (concept, 0.95) · disaggregated-inference (concept, 0.88) · prefill-decode (concept, 0.85)

Classification

transcript · language en · status final

Provenance

claude-haiku-4-5 via @stacklist/be@0.1.0, confidence 0.85, 24 Sep 2026