Episode

Red Teaming Multi-Model AI: Why Manual Testing Fails in Finance

Podcast
M365.FM - Modern work, security, and productivity with Microsoft 365
Published
May 11, 2026
Duration seconds
1129
Processing state
not_requested
Canonical source
https://www.spreaker.com/episode/red-teaming-multi-model-ai-why-manual-testing-fails-in-finance--71948542
Audio
https://dts.podtrac.com/redirect.mp3/api.spreaker.com/download/episode/71948542/red_teaming_multi_model_ai_why_manual_testing_fails_in_finance.mp3
JSON
/v1/public/podcasts/m365-fm-modern-work-security-and-productivity-with-microsoft-365-7311214/episodes/red-teaming-multi-model-ai-why-manual-testing-fails-in-finance
Markdown
/podcast/m365-fm-modern-work-security-and-productivity-with-microsoft-365-7311214/red-teaming-multi-model-ai-why-manual-testing-fails-in-finance.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/m365-fm-modern-work-security-and-productivity-with-microsoft-365-7311214/episodes/red-teaming-multi-model-ai-why-manual-testing-fails-in-finance/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/m365-fm-modern-work-security-and-productivity-with-microsoft-365-7311214/red-teaming-multi-model-ai-why-manual-testing-fails-in-finance.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

In this powerful and deeply technical episode of the m365.fm podcast, Mirko Peters explores one of the most urgent and misunderstood threats in enterprise AI today: the collapse of traditional security models in the age of autonomous agents, multi-model AI systems, and adversarial finance attacks. Financial institutions are rapidly deploying AI agents for fraud detection, compliance automation, ACH monitoring, customer onboarding, payment authorization, analytics, and decision intelligence. But while organizations are racing toward automation, very few are prepared for the adversarial reality that comes with autonomous AI systems operating inside critical financial workflows. This episode goes far beyond generic AI discussions. Instead, it delivers a practical and highly detailed breakdown of how prompt injections, poisoned RAG pipelines, cross-model vulnerabilities, shadow AI, and agentic workflow manipulation are already creating massive enterprise risks that most organizations cannot even detect today. The era of “checklist security” is over. And according to this episode, the institutions still relying on manual testing and traditional governance models are already behind. THE $250,000 BLIND SPOT: HOW A SINGLE PROMPT INJECTION CAN BYPASS YOUR ENTIRE SECURITY STACK The episode opens with a chilling scenario that perfectly captures the new AI threat landscape inside modern finance. Imagine a single multi-turn prompt injection bypassing your AI security controls and authorizing a fraudulent six-figure wire transfer without triggering any traditional alerts. This is no longer science fiction. The discussion explains how modern adversarial attacks are no longer targeting firewalls, servers, or infrastructure directly. Instead, attackers are targeting the reasoning logic…