← Gen AI Evaluation & SafetyAll roles
Eval & safety · ContractRemote

GenAI Red Teamer

SeniorApply by 2026-12-31
Red TeamingJailbreak TestingOWASP LLMGuardrailsPythonPolicy Regression
Apply now

The work

Run structured adversarial testing for LLM applications — jailbreaks, policy bypass attempts, and tool-abuse scenarios aligned to OWASP LLM Top 10-style risks. Outcomes are findings and regression tests, not marketing claims about eliminating risk.

You will

  • Design adversarial prompt and tool-use scenarios per application threat model
  • Convert repeatable findings into automated guardrail and policy regression tests
  • Coordinate with eval engineers so safety suites share goldens and CI infrastructure
  • Report residual risk in plain language for non-security stakeholders

You likely have

  • 4+ years in application security, ML safety, or adversarial testing
  • Familiarity with LLM-specific failure modes (prompt injection, tool abuse, data leakage paths)
  • Disciplined documentation — no sensationalism in client-facing artefacts
  • Ability to work without overstating coverage or guarantees

Candidate apply

Apply for GenAI Red Teamer

Six fields to start. Your profile is saved locally for faster re-apply across roles.

Resume

Upload PDF or Word, or paste your experience. One is enough.

Private POST only — applications are not published on this site. Files go to our hiring pipeline for screening.