Light

Episode #1011 Quiz

Jailbreaking AI
Date: 2025-02-04 | Length: 2.75 hrs | Episode page at twit.tv

About this episode

Security Now! Episode 1011 discusses the rise of AI jailbreaking, focusing on vulnerabilities in DeepSeek, a Chinese AI model rivaling GPT-4 in capability but cheaper to run. Jailbreak techniques like Bad Likert Judge, Crescendo, and Deceptive Delight bypass DeepSeek’s guardrails, producing detailed malicious code, phishing templates, and guides for explosives, posing serious security risks.

Your name and email are stored only in your browser local storage for convenience. They are not retained server-side.

Question 1: What fundamental architectural innovation allows DeepSeek's AI models to achieve comparable performance to GPT-4 while significantly reducing computational resources?
Question 2: According to Palo Alto Networks Unit 42's research on DeepSeek, which jailbreaking technique leverages a Likert scale to coax the AI into producing malicious content?
Question 3: What is the main security implication of DeepSeek's exposed ClickHouse database as reported by Wiz Research?
Question 4: According to the episode, how does DeepSeek's use of assembler programming contribute to its AI model training efficiency?
Question 5: Why is AI safety considered more challenging than traditional security issues like buffer overflows, based on the episode's discussion on AI jailbreaking?
Cancel