Responding to AI Agent Containment Failures with CSET's Helen Toner, LawAI's Mackenzie Arnold & CSIS's Matt Pearl

Responding to AI Agent Containment Failures with CSET's Helen Toner, LawAI's Mackenzie Arnold & CSIS's Matt Pearl

This episode cross-posts a panel from the event "AI Agent Containment Failures: Technical Realities and Policy Responses," co-hosted by the Wadhwani AI Center and the Institute for Law and AI on August 24. Check out the full event recording: https://www.csis.org/events/ai-agent-containment-failures-technical-realities-and-policy-responses Guests: Helen Toner, Executive Director, Center for Security and Emerging Technology (CSET) at Georgetown Mackenzie Arnold, Director of U.S. Policy, Institute for Law and AI (LawAI) Matt Pearl, Director, Strategic Technologies Program at the Center for Strategic and International Studies (CSIS) Timestamps: What we've learned since the OpenAI-Hugging Face incident (1:15) Incentives for safety measures at frontier labs (16:24) Technical talent within government (22:18) Ensuring compliance to incident reporting requirements (34:58) Liability for crimes committed by AI agents (38:00) Liability safe harbors (51:55) Additional Reading: "When Reporting an AI Security Incident Is Not Mandatory" by Mackenzie Arnold and Stephan Llerena for Lawfare: https://www.lawfaremedia.org/article/when-reporting-an-ai-security-incident-is-not-mandatory "Investigating three real-world incidents in our cybersecurity evaluations" blog by Anthropic: https://www.anthropic.com/news/investigating-incidents-cybersecurity-evals "Incident Report: unsanctioned agent behavior during cyber testing" by U.K. AISI: https://www.aisi.gov.uk/blog/incident-report-unsanctioned-agent-behaviour-during-cyber-testing "Pacing model development in an era of cyber-critical capabilities" announcement from OpenAI: https://openai.com/index/pacing-model-development-cyber-capabilities/ "Pacing the Frontier" petition: https://www.pacingthefrontier.com/ CSIS Commission on U.S. Cyber Force Generation: https://www.csis.org/analysis/csis-commission-us-cyber-force-generation Limitations of the current incident reporting regime (7:46)Role of the U.S. government in helping defenders (11:52)What the executive branch can do now (25:29)Concrete policy recommendations (29:31)U.S.-China competition (40:50)Preventing abuse of an incident reporting regime (44:25)De facto regulatory role played by frontier lab employees (49:24)Role of AI-generated code in cyberdefense (54:45)Concluding remarks (55:53)"Out of Bounds: What the U.S. Government Should Do in Response to AI Agent Containment Failures" by Aalok Mehta: https://www.csis.org/analysis/out-bounds-what-us-government-should-do-response-ai-agent-containment-failures"The 'Breaking' News: The OpenAI–Hugging Face Incident" presentation at Black Hat: https://youtu.be/87DyyMV0kCY?si=GWh9MB1plhO-xjYT

Tämä jakso on lisätty Podme-palveluun avoimen RSS-syötteen kautta eikä se ole Podmen omaa tuotantoa. Siksi jakso saattaa sisältää mainontaa.

Jaksot(106)

OpenAI Pauses RL Training and Anthropic Adds Watermarks to AI-Generated Text

OpenAI Pauses RL Training and Anthropic Adds Watermarks to AI-Generated Text

This week, we cover updates from the ongoing cyber saga, including OpenAI's two-week pause on RL training and the cyber capabilities of Z.ai's latest model GLM-5.3. We also unpack Anthropic's move to ...

20 Elo 47min

Planning for the AI-Powered Economy with Windfall's Adrian Brown

Planning for the AI-Powered Economy with Windfall's Adrian Brown

In this episode, we're joined by Adrian Brown, founder and CEO of Windfall Trust, for a conversation about the economic impacts of AI. Why the economic focus (00:58) Capturing AI's impact in macr...

13 Elo 46min

The AI Policy Podcast Trailer

The AI Policy Podcast Trailer

Join CSIS’s Aalok Mehta, Director of the Wadhwani AI Center, on a deep dive into the world of AI policy. Every two weeks, tune in for insightful discussions covering AI regulation, economic impacts, n...

11 Elo 1min

Three More AI Hacking Incidents, and a Push to 'Pace the Frontier'

Three More AI Hacking Incidents, and a Push to 'Pace the Frontier'

In this episode, we touch on Texas' new verification and audit requirement for data center developers seeking connection to the state's grid (1:11) before unpacking Anthropic's disclosure of three inc...

6 Elo 44min

A Deep Dive Into IVOs with Fathom's Bri Treece

A Deep Dive Into IVOs with Fathom's Bri Treece

In this episode, we're joined by Fathom co-founder and President Bri Treece for a conversation about third-party solutions to AI governance. Bri makes the case for Independent Verification Organizatio...

4 Elo 40min

Japan's Take on AI Sovereignty with Hiroki Habuka

Japan's Take on AI Sovereignty with Hiroki Habuka

In this episode, we're joined by Hiroki Habuka for a conversation about Japan's approach to AI sovereignty. We discuss how the Japanese government defines sovereign AI (1:08) and how that definition s...

30 Heinä 45min

OpenAI Models Hack Hugging Face, Kimi K3's "DeepSeek Moment," and New York's Data Center Moratorium

OpenAI Models Hack Hugging Face, Kimi K3's "DeepSeek Moment," and New York's Data Center Moratorium

In this episode, we discuss reports that OpenAI models escaped a sandboxed evaluation and breached Hugging Face's infrastructure (00:33). We then turn to Moonshot AI's Kimi K3, which has prompted talk...

23 Heinä 48min

Suosittua kategoriassa Politiikka ja uutiset

uutiscast
aikalisa
politiikan-puskaradio
ootsa-kuullut-tasta-2
rss-ootsa-kuullut-tasta
otetaan-yhdet
rss-vaalirankkurit-podcast
rss-podme-livebox
rss-voi-venaja
rss-seksicast
rss-girls-finish-f1rst
tervo-halme
rss-asiastudio
rss-kaikki-uusiksi
rss-pinnalla
linda-maria
rss-raha-talous-ja-politiikka
rss-kovin-paikka
nakokulma-oikealta-jussi-halla-ahon-blogin-kommentaarit
rss-mina-ukkola