Middle AI Red Team – GenAI Adversarial Evaluation
We are looking for an AI Red Team Gym Creator to design and scale adversarial Evaluation for LLM and AI systems.
The role will focus on creating and reviewing indirect prompt injection attacks, jailbreaks, agent manipulation, tool-use attacks, and other adversarial techniques in controlled and authorized environments.
Responsibilities
- Create novel adversarial attacks.
- Help build scalable systems to generate and test attacks at large scale.
- Review attack quality, effectiveness, novelty, and reproducibility.
- Identify new attack patterns and model vulnerabilities.
- Develop adversarial datasets, benchmarks, and regression tests.
- Share findings, techniques, and structured attack data with the Data Science, Security, and Engineering teams.
- Help improve model robustness and defensive capabilities.
- 3+ years of professional experience in cybersecurity, security research, red teaming, or a related security field.
- Bachelor’s degree in Data Science, Computer Engineering, Computer Science, or a closely related technical field.
- Strong understanding of LLMs, prompt injection, and AI security.
- Experience with red teaming, security research, or adversarial testing.
- Strong Python and automation skills.
- Creative problem-solving and ability to develop new attack techniques.
- Ability to analyze results and communicate findings clearly.
Preferred Skills That Set You Apart
- Certifications in offensive cybersecurity (e.g., OSWA, OSWE, OSCE3, SEC542, SEC522)
Alice is a trust, safety, and security company built for the AI era. We safeguard the communicative technologies people use to create, collaborate, and interact- whether with each other or with machines.
In a world where AI has fundamentally changed the nature of risk, Alice provides end-to-end coverage across the entire AI lifecycle. We support frontier model labs, enterprises, and UGC platforms with a comprehensive suite of solutions: from model hardening evaluations and pre-deployment red-teaming to runtime guardrails and ongoing drift detection.
Alice is widely considered a global leader in online safety and AI security. We have some of the most forward-thinking and passionate minds in the world working to safeguard over 3 billion users across the largest AI and tech platforms.
If you're creative and driven to secure the future of AI, we want to hear from you!
Required Skills
Required Languages
🇬🇧 English