1,000 tokens per second: the case against predicting one word at a time (Kumar, VP of Engineering at Inception Labs)
Episode Date: August 24, 2026What if the entire LLM industry has been solving language generation the slow way — one token at a time?In this episode of The Infra Pod, hosts Tim Ch...
