The AI Policy Podcast

Join CSIS’s Aalok Mehta, Director of the Wadhwani AI Center, on a deep dive into the world of AI policy. Every two weeks, tune in for insightful discussions covering AI regulation, economic impacts, national security, geopolitics, and more. CSIS is a bipartisan think tank in Washington, D.C.

  1. 4d ago

    Responding to AI Agent Containment Failures with CSET's Helen Toner, LawAI's Mackenzie Arnold & CSIS's Matt Pearl

    This episode cross-posts a panel from the event "AI Agent Containment Failures: Technical Realities and Policy Responses," co-hosted by the Wadhwani AI Center and the Institute for Law and AI on August 24.  Check out the full event recording: https://www.csis.org/events/ai-agent-containment-failures-technical-realities-and-policy-responses   Guests: Helen Toner, Executive Director, Center for Security and Emerging Technology (CSET) at Georgetown Mackenzie Arnold, Director of U.S. Policy, Institute for Law and AI (LawAI) Matt Pearl, Director, Strategic Technologies Program at the Center for Strategic and International Studies (CSIS)  Timestamps: What we've learned since the OpenAI-Hugging Face incident (1:15) Incentives for safety measures at frontier labs (16:24) Technical talent within government (22:18) Ensuring compliance to incident reporting requirements (34:58) Liability for crimes committed by AI agents (38:00) Liability safe harbors (51:55)   Additional Reading: "When Reporting an AI Security Incident Is Not Mandatory" by Mackenzie Arnold and Stephan Llerena for Lawfare: https://www.lawfaremedia.org/article/when-reporting-an-ai-security-incident-is-not-mandatory "Investigating three real-world incidents in our cybersecurity evaluations" blog by Anthropic: https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals "Incident Report: unsanctioned agent behavior during cyber testing" by U.K. AISI: https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing "Pacing model development in an era of cyber-critical capabilities" announcement from OpenAI: https://openai.com/index/pacing-model-development-cyber-capabilities/ "Pacing the Frontier" petition: https://www.pacingthefrontier.com/ CSIS Commission on U.S. Cyber Force Generation: https://www.csis.org/analysis/csis-commission-us-cyber-force-generation Limitations of the current incident reporting regime (7:46)Role of the U.S. government in helping defenders (11:52)What the executive branch can do now (25:29)Concrete policy recommendations (29:31)U.S.-China competition (40:50)Preventing abuse of an incident reporting regime (44:25)De facto regulatory role played by frontier lab employees (49:24)Role of AI-generated code in cyberdefense (54:45)Concluding remarks (55:53)"Out of Bounds: What the U.S. Government Should Do in Response to AI Agent Containment Failures" by Aalok Mehta: https://www.csis.org/analysis/out-bounds-what-us-government-should-do-response-ai-agent-containment-failures"The 'Breaking' News: The OpenAI–Hugging Face Incident" presentation at Black Hat: https://youtu.be/87DyyMV0kCY?si=GWh9MB1plhO-xjYT

  2. Aug 20

    OpenAI Pauses RL Training and Anthropic Adds Watermarks to AI-Generated Text

    This week, we cover updates from the ongoing cyber saga, including OpenAI's two-week pause on RL training and the cyber capabilities of Z.ai's latest model GLM-5.3. We also unpack Anthropic's move to add watermarks to text generated by Claude in compliance with the EU AI Act. Timestamps: Mark Zuckerberg's essay on AI (00:20) OpenAI pauses RL training (5:56) Cyber capabilities of GLM-5.3 (15:29) EU AI Act refresher (24:29) How Anthropic's text watermark works (31:21) What's driving the backlash against Anthropic (40:59)   Additional Reading: Zuckerberg's essay "The Future is for Everyone": https://www.meta.com/thefutureisforeveryone/  OpenAI announces two-week pause on RL training: https://openai.com/index/pacing-model-development-cyber-capabilities/  Letter from Sanders to tech CEOs: https://www.sanders.senate.gov/wp-content/uploads/AI-Pause-Letter-FINAL.pdf  Letter from House reps to Mike Johnson: https://casar.house.gov/sites/evo-subsites/casar.house.gov/files/evo-media-document/final-letter-to-speaker-johnson-requesting-ai-hearings-1.pdf  Z.ai blog post "GLM-5.3: Frontier Coding with Emergent Cyber Capabilities": https://z.ai/blog/glm-5.3  Z.ai article "Preparing GLM-5.3 for Open Release: A Responsible Path to Cyber Defense": https://x.com/Zai_org/status/2088280509474320693  "The EU's AI Transparency Code of Practice, Explained" (Tech Policy Press): https://www.techpolicy.press/the-eus-ai-transparency-code-of-practice-explained/  Anthropic blog post "How Claude's text watermark works": https://www.anthropic.com/news/claude-text-watermark  "Toward a Federal Framework: Lessons from State and International Frontier AI Regulation" (CSIS): https://www.csis.org/analysis/toward-federal-framework-lessons-state-and-international-frontier-ai-regulation Check out our upcoming event, "AI Agent Containment Failures: Technical Realities and Policy Responses": https://www.csis.org/events/ai-agent-containment-failures-technical-realities-and-policy-responses

  3. Aug 6

    Three More AI Hacking Incidents, and a Push to 'Pace the Frontier'

    In this episode, we touch on Texas' new verification and audit requirement for data center developers seeking connection to the state's grid (1:11) before unpacking Anthropic's disclosure of three incidents in which its models hacked into another company during internal cyber evaluations (11:02). We discuss how these incidents—in addition to a previously disclosed incident at OpenAI—are driving calls for the U.S. government to help "pace the frontier" of AI development (21:53). We also explore the connection between AI safety and open model development, including Nvidia's open letter to U.S. policymakers, "Open Weights and American AI Leadership." (33:18). Additional Reading: Gov. Abbott's data center directive: https://gov.texas.gov/uploads/files/press/Thomas_Gleeson_Pablo_Vegas_Data_Centers_Directive_Letter_… "China's AI Blitz Creates 'Death Zone' for Rival US Model Makers" (Bloomberg): https://www.bloomberg.com/news/articles/2026-08-04/china-s-ai-blitz-creates-death-zone-for-rival-us-model-makers Anthropic's blog post "Investigating three real-world incidents in our cybersecurity evaluations": https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals  Hugging Face's technical writeup "Anatomy of a Frontier Lab Agent Intrusion": https://huggingface.co/blog/agent-intrusion-technical-timeline State Attorneys General letter to OpenAI: https://www.iowaattorneygeneral.gov/media/cms/08_5392C9E17791C.pdf House Committee on Homeland Security briefing request: https://x.com/HomelandDems/status/2084389406194929918 "How OpenAI Lost Control of an AI Model—and What Needs to Change" (TIME): https://time.com/article/2026/07/24/openai-hugging-face-attack/ "Pacing the Frontier" petition: https://www.pacingthefrontier.com/ Nvidia's open letter "Open Weights and American AI Leadership": https://images.nvidia.com/pdf/Open-Weights-and-American-AI-Leadership.pdf

4.8
out of 5
58 Ratings

About

Join CSIS’s Aalok Mehta, Director of the Wadhwani AI Center, on a deep dive into the world of AI policy. Every two weeks, tune in for insightful discussions covering AI regulation, economic impacts, national security, geopolitics, and more. CSIS is a bipartisan think tank in Washington, D.C.

More From CSIS

You Might Also Like