
In this episode, Katherine Forrest and Scott Caravello explore how Anthropic researchers built a tool to observe the newly discovered “global workspace” inside its models, the company’s experiments to assess this workspace’s effects on model behavior, and what the discovery could mean for interpretability and AI safety. For the sources referenced in this episode, please see the links below: Anthropic: Verbalizable Representations Form a Global Workspace in Language Models IBM: What Anthropic’s J-space research means for the future of AI ## Learn More About Paul, Weiss’s Artificial Intelligence practice: https://www.paulweiss.com/industries/artificial-intelligence
Podzilla Summary coming soon
Sign up to get notified when the full AI-powered summary is ready.
Free forever for up to 3 podcasts. No credit card required.

Another Step Forward: Mathematics Breakthroughs and Self-Improving AI

Year of the Robot: Fall 2026 Updates

Bill Gates on the “Turbulent AI Era”: Sounding the Alarm and Preparing for Society’s Transition

Signed by the Machine: AI Watermarking and Transparency
Free AI-powered recaps of Paul, Weiss Waking Up With AI and your other favorite podcasts, delivered to your inbox.
Free forever for up to 3 podcasts. No credit card required.