Red Teaming
Agentic AI systems are powerful but remain vulnerable to adversarial manipulation and jailbreaking. We advance their robustness by continuously probing and debugging them through red teaming. This raises the following question.
How do we train red agents that continuously debug target AI systems for hardening?
Keywords
- Agentic AI
- Reinforcement Learning
6 Related Publications
The Illusion of Rust Safety: Detecting Modular Unsafe Functions with LLMs
The ACM Conference on Computer and Communications Security (CCS), 2026
CCS
International Conference on Machine Learning (ICML), 2026
ICML
International Conference on Dependable Systems and Networks (DSN), 2026
2025
🏆 DARPA AIxCC Winner - $(4+2)M Award