← Gen AI Evaluation & SafetyAll roles
Eval & safety · ContractRemote
GenAI Red Teamer
SeniorApply by 2026-12-31
Red TeamingJailbreak TestingOWASP LLMGuardrailsPythonPolicy Regression
Apply nowThe work
Run structured adversarial testing for LLM applications — jailbreaks, policy bypass attempts, and tool-abuse scenarios aligned to OWASP LLM Top 10-style risks. Outcomes are findings and regression tests, not marketing claims about eliminating risk.
You will
- Design adversarial prompt and tool-use scenarios per application threat model
- Convert repeatable findings into automated guardrail and policy regression tests
- Coordinate with eval engineers so safety suites share goldens and CI infrastructure
- Report residual risk in plain language for non-security stakeholders
You likely have
- 4+ years in application security, ML safety, or adversarial testing
- Familiarity with LLM-specific failure modes (prompt injection, tool abuse, data leakage paths)
- Disciplined documentation — no sensationalism in client-facing artefacts
- Ability to work without overstating coverage or guarantees
Candidate apply
Apply for GenAI Red Teamer
Six fields to start. Your profile is saved locally for faster re-apply across roles.
GenAI Red Teamer
Apply