Design and execute comprehensive adversarial testing campaigns against AI models, including large language models, multimodal systems, and autonomous agents.
Research attack vectors and prompt injection techniques to identify vulnerabilities, jailbreaks, and unintended behaviors.
Conduct red team exercises simulating real-world deployment scenarios and edge cases, systematically probing for bias, toxicity, misinformation, and other harmful outputs across diverse contexts and demographics.
Create adversarial datasets and benchmarks to evaluate model robustness under various attack conditions.
Collaborate with security teams to perform red teaming activities, document findings, vulnerabilities, and provide remediation recommendations to mitigate risks.
What's Needed?
Minimum 5 years of experience in penetration testing, red team operations, or AI Red Team activities.
Degree in computer science, information technology, programming, development, artificial intelligence, or cybersecurity from an accredited institution or equivalent work experience.
Background in AI red teaming, offensive cyber operations, or related fields.
Understanding of AI, Machine Learning, cloud computing, software applications, and networking.
Proficiency with testing tools such as SPLX, Garak, TextAttack, PyRIT, Burp Suite, Metasploit, Nessus, Cobalt Strike/Mythic/C2, NMAP, sqlmap, or similar