Module 4 of 4 · 60 min
Adversarial Security Capstone: Audit, Attack & Fortify
Audit an enterprise AI agent, execute multi-vector adversarial attacks, implement multi-layered defenses, and verify zero exploit pass rate.
Core concept
By the end
You will be able to
- Perform a full black-box and white-box security audit of an AI agent.
- Demonstrate successful exploitation across prompt injection and insecure output channels.
- Fortify the architecture and prove 100% defense validation across regression suites.
01
Security Capstone Audit Standard
In this capstone, you will audit a vulnerable enterprise agent, uncover injection and data exfiltration flaws, build multi-layered defense guardrails, and generate an executive penetration testing report.
Practice activity
Deliver Fortified Agent Security Package
- Audit vulnerable agent codebase.
- Execute 20 attack vectors and document proof-of-concept exploits.
- Implement defense guardrails and verify zero successful exploit rate.
What to produce
- Security audit report, exploit PoCs, fortification patch diff, and verification logs.
Reflect before continuing
How do you balance aggressive security filtering with user experience and false-positive rejections?
Evidence
Sources and verification
- NIST AI Risk Management Framework (AI RMF 1.0)NIST · verified 2026-08-22
Knowledge check
Make it stick.
Choose the strongest answer for each question. Your attempts become part of your account transcript.