Firehose

Filtered to tagged “Speculative decoding” · clear filters

All PeopleCompaniesPapersPodcastsHacker News

Browse by tag

29 JUL 2026 · Paper

This paper studies how lossy verification schemes can improve the efficiency of speculative decoding in large language models, but may also degrade generation quality. Practitioners may care about understanding the trade-offs between speed and quality when using these schemes.

8 JUL 2026 · Podcast · Latent Space: The AI Engineer Podcast

In this episode, Akshat Bubna, CTO of Modal, discusses the evolution of AI infrastructure tailored for agent experience, highlighting Modal's journey from a runtime platform to a specialized cloud for AI workloads. They explore challenges i…