This episode features Zico Kolter and Matt Fredrikson from Grey Swan discussing AI security challenges, particularly around adversarial attacks and indirect prompt injection in large language models and AI agents. They explain their approac…
Firehose
Filtered to tagged “Mechanistic interpretability” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives
Browse by tag
artificial intelligence 89continual learning 27AI 23AI safety 13reinforcement learning 13agentic coding 12open-weight models 12AI agents 10machine learning 9AI ethics 8cybersecurity 8existential risk 8language models 8natural language processing 8ethics 7Reinforcement learning 6Diffusion models 5large language models 5multi-agent systems 5open-source 5recursive self-improvement 5robotics 5security 5software development 5Agentic AI 4artificial general intelligence 4mathematics 4Recursive self-improvement 4agents 3AI infrastructure 3
This episode features Biohub co-founders Mark Zuckerberg and Priscilla Chan, along with Head of Science Alex Rives, discussing their ambitious goal to cure, prevent, and manage all disease by the century's end, now accelerated by AI. They d…
Alex Rives, Head of Science at Biohub, discusses ESM-C, a fourth-generation protein language model that leverages the "Bitter Lesson" of scaling with massive metagenomic datasets to achieve unprecedented protein structure prediction and des…