AI Red Teaming in 2024 and Beyond

AI Red Teaming in 2024 and Beyond

Host Caleb Sima and Ashish Rajan caught up with experts Daniel Miessler (Unsupervised Learning), Joseph Thacker (Principal AI Engineer, AppOmni) to talk about the true vulnerabilities of AI applications, how prompt injection is evolving, new attack vectors through images, audio, and video and predictions for AI-powered hacking and its implications for enterprise security.

Whether you're a red teamer, a blue teamer, or simply curious about AI's impact on cybersecurity, this episode is packed with expert insights, practical advice, and future forecasts. Don’t miss out on understanding how attackers leverage AI to exploit vulnerabilities—and how defenders can stay ahead.


Questions asked:

(00:00) Introduction

(02:11) A bit about Daniel Miessler

(02:22) A bit about Rez0

(03:02) Intersection of Red Team and AI

(07:06) Is red teaming AI different?

(09:42) Humans or AI: Better at Prompt Injection?

(13:32) What is a security vulnerability for a LLM?

(14:55) Jailbreaking vs Prompt Injecting LLMs

(24:17) Whats new for Red Teaming with AI?

(25:58) Prompt injection in Multimodal Models

(27:50) How Vulnerable are AI Models?

(29:07) Is Prompt Injection the only real threat?

(31:01) Predictions on how prompt injection will be stored or used

(32:45) What’s changed in the Bug Bounty Toolkit?

(35:35) How would internal red teams change?

(36:53) What can enterprises do to protect themselves?

(41:43) Where to start in this space?

(47:53) What are our guests most excited about in AI?


Resources

Daniel's Webpage - Unsupervised Learning

Joseph's Website

Det här avsnittet är hämtat från ett öppet RSS-flöde och publiceras inte av Podme. Det kan innehålla reklam.

Avsnitt(61)

Black Hat 2026: Why Threat Researchers Are Hoarding Zero-Days

Black Hat 2026: Why Threat Researchers Are Hoarding Zero-Days

Is the AI "vulnpocalypse" already here? According to Casey Ellis, Founder of Bugcrowd and pioneer of Disclose.io, we aren't quite in an apocalypse yet, we're actually in a "slopdemic." The cost of dis...

15 Sep 48min

Why Prompt Filters Fail & How to Explain AI Risk to the Board | Cezary Piekarski, Standard Chartered.

Why Prompt Filters Fail & How to Explain AI Risk to the Board | Cezary Piekarski, Standard Chartered.

Is the cybersecurity industry repeating the same mistakes with prompt injection that it made with buffer overflows decades ago? As attackers iterate through 50 to 60 prompt filter bypasses daily, atte...

2 Sep 37min

Why 95% of AI Projects Fail: Model Risk & AI Governance | Sandip Wadje, BNP Paribas

Why 95% of AI Projects Fail: Model Risk & AI Governance | Sandip Wadje, BNP Paribas

Why do 95% of enterprise AI implementations fail? According to Sandip Wadje, Managing Director at BNP Paribas, many organizations attempt complex reasoning tasks on day one rather than building a matu...

27 Aug 45min

Why I Dont Trust Your AI Agent | Kane Narraway, Canva

Why I Dont Trust Your AI Agent | Kane Narraway, Canva

With over 200 AI security vendors in the market, how does an enterprise CISO decide whether to build a custom solution, buy an off-the-shelf product, or just wait out the hype?In this episode of the A...

20 Aug 52min

Baiting the Bot: How to Use Deception to Stop Autonomous AI Agents

Baiting the Bot: How to Use Deception to Stop Autonomous AI Agents

When AI agents start swarming your enterprise, they won't care about stealth. They will land a beachhead and instantly spawn 500 agents to crawl, probe, and exfiltrate data at machine speed. Is your d...

23 Juli 51min

Why AI Agents Are Forcing a Redesign of Application Security?

Why AI Agents Are Forcing a Redesign of Application Security?

When the CEO of Anthropic declares that human coding will disappear within six months, followed quickly by the death of software engineering itself, what does that mean for the future of cybersecurity...

26 Juni 51min

Why Asset Intelligence is Replacing the CMDB & Static Dashboards

Why Asset Intelligence is Replacing the CMDB & Static Dashboards

Why do CISOs still struggle with asset intelligence in 2026? Despite decades of security tooling, most organizations still have a massive 40% "dark matter" blind spot in their environment and the expl...

11 Juni 42min

The AI AuthZ Problem: Why Human Least Privilege Fails for Autonomous Agents

The AI AuthZ Problem: Why Human Least Privilege Fails for Autonomous Agents

Why are security leaders terrified of connecting AI agents to production data? Because unlike humans, AI agents don't apply judgment, and they operate at machine speed, meaning they can relentlessly h...

4 Juni 47min

Populärt inom Teknik

uppgang-och-fall
natets-morka-sida
elbilsveckan
bilar-med-sladd
skogsforum-podcast
rss-en-ai-till-kaffet
rss-elektrikerpodden
rss-laddstationen-med-elbilen-i-sverige
market-makers
rss-technokratin
rss-veckans-ai
rss-sakerhetspodcasten
developers-mer-an-bara-kod
rss-ai-med-jonas-benjamin
bli-saker-podden
rss-uppgang-och-fall
rss-fabriken-2
rss-powerboat-sverige-podcast
rss-en-liten-podd-om-it
rss-snacka-om-ai