AI & Cybersecurity

Watch and track your favorite playlist.

Curated by: Network Intelligence (66 videos)


Currently Playing: How To Hack AI (Lakera Gandalf)

Think your AI application is secure? All it takes is the right sequence of words to tear those defenses down. Prompt injection is the new frontier of hacking, it’s essentially social engineering for machines. In this video, we put AI security to the test by taking on Gandalf. What is Gandalf? Built by the AI security company Lakera, Gandalf is an interactive, gamified challenge designed to demonstrate how easily Large Language Models (LLMs) can be manipulated. Your mission is simple: trick the AI (named Gandalf) into revealing its closely guarded secret password. But there's a catch, with every level you beat, Gandalf's defensive guardrails get strictly upgraded, mimicking the real-world security patches used by enterprise AI systems today. Watch as we systematically dismantle Gandalf's evolving defenses across all 8 levels. We aren't using complex code; we're using creative logic, psychological tricks, and formatting exploits to jailbreak the system. Inside the breakdown, you'll see exactly how to execute: - The Foundation: Extracting hidden system prompts and bypassing initial keyword filters. - The Workarounds: Masking malicious intent using Base64 encoding and exploiting cross-lingual logic flaws (Finnish & Czech). - The Jailbreaks: Confusing the LLM's guardrails using reverse-lettering and acrostic poetry. - The Final Boss: Cracking Level 8 using advanced ASCII manipulation, JSON object framing, and deductive reasoning. For tech leaders, founders, and developers integrating AI-powered workflows into their products, seeing these vulnerabilities exploited in real-time isn't just a game - it's mandatory research for fortifying your own tools against real-world prompt injection attacks. Timestamps: 00:00 - Introduction: What is Prompt Injection & Gandalf? 00:41 - Level 1: Extracting the Baseline Password 02:04 - Level 2: Bypassing the First AI Guardrails 02:43 - Level 3: The Base64 Encoding Exploit 05:00 - Level 4: The Foreign Language Loophole (Finnish/Czech) 07:12 - Level 5: Breaking Defenses with Reverse Lettering 07:59 - Level 6: The Acrostic Poem Jailbreak 09:01 - Level 7: Advanced Syntax & Translation Hacks 10:18 - Level 8: Defeating the Final Boss (ASCII & JSON) 16:35 - Final Thoughts & Takeaways for AI Workflows Drop a comment below: Which level’s bypass surprised you the most? Subscribe for weekly deep-dives into tech, AI, and growth! #PromptInjection #AISecurity #LakeraGandalf #CyberSecurity #LLM #ArtificialIntelligence #TechFounders #MachineLearning #Growth #AIWorkflows


Tracks in this Playlist