Podcasts

Filtered to KV cache · clear filter

Browse by topic

Reinforcement learning 11AI agents 10recursive self-improvement 7AI safety 6Diffusion models 6formal verification 5Agentic AI 4AI ethics 3AI infrastructure 3Code generation 3Continual learning 3Human-AI collaboration 3Mechanistic interpretability 3multi-agent systems 3Open Source AI 3Scaling laws 3Synthetic data generation 3Venture capital 3world models 3Agent architecture 2Computational complexity 2Context engineering 2Continuous learning 2Data generation 2Developer experience 2Drug discovery 2Expert parallelism 2Foundation models 2GPU infrastructure 2KV cache 2

Matching episodes

The Inference Engineering Masterclass — Philip Kiely & Ali Taha, Baseten

Latent Space: The AI Engineer Podcast · 3 AUG 2026 · 101 min

In this episode, Philip Kiely and Ali Taha from Baseten discuss the complexities and innovations in inference engineering for large AI models. They cover topics including model deployment, speculative decoding, quantization, hardware optimi…

Reiner Pope – The math behind how LLMs are trained and served

Dwarkesh Podcast · 29 APR 2026 · 134 min

Reiner Pope delivers a blackboard lecture on the mathematical and hardware principles behind training and serving large language models. He explains how batch size, sparsity, and various parallelism strategies (expert, pipeline) impact late…