This episode features Zico Kolter and Matt Fredrikson from Grey Swan discussing AI security challenges, particularly around adversarial attacks and indirect prompt injection in large language models and AI agents. They explain their approac…
Firehose
Filtered to tagged “Mechanistic interpretability” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives
Browse by tag
artificial intelligence 87continual learning 32AI 24reinforcement learning 14agentic coding 13AI safety 13open-weight models 13AI agents 10existential risk 9AI ethics 8cybersecurity 8ethics 7language models 7machine learning 7natural language processing 6open-source 6Reinforcement learning 6security 6artificial general intelligence 5Diffusion models 5recursive self-improvement 5robotics 5software development 5Agentic AI 4large language models 4mathematics 4multi-agent systems 4Recursive self-improvement 4agentic AI 3agents 3
This episode features Biohub co-founders Mark Zuckerberg and Priscilla Chan, along with Head of Science Alex Rives, discussing their ambitious goal to cure, prevent, and manage all disease by the century's end, now accelerated by AI. They d…
Alex Rives, Head of Science at Biohub, discusses ESM-C, a fourth-generation protein language model that leverages the "Bitter Lesson" of scaling with massive metagenomic datasets to achieve unprecedented protein structure prediction and des…