← Back to Talent Bench
Ref #HL-QA-3308Lead Level · 9 Years Experience
Deployment Ready

AI QA Engineer (LLM Applications)

Specialization: Gen AI Evaluation & Safety · Industries: B2B SaaS, Healthcare, Financial Services · Vetting Rating: Interview-ready · quality engineering review

Contract EngagementContractor / Deployable
AvailabilityImmediate (Within 48 hrs)
Work ModelRemote worldwide
Benchmark Rate~$90 / hr
Verified Technical Stack:
PlaywrightTypeScriptAPI Contract TestingCI PipelinesPerformance Testingk6GitHub ActionsAccessibility TestingTest Data Strategy

Professional Summary

AI QA engineer for LLM-backed products — eval harnesses, golden datasets, and CI gates alongside Playwright where UI paths still matter.

Verified Impact & Project Highlights

  • Introduced LLM eval gates in pull-request CI alongside contract tests for API surfaces
  • Maintains golden datasets versioned with prompt and retrieval config changes
  • Partners with eval engineers on judge calibration instead of duplicating rubrics

Request an interview with HL-QA-3308

Hiring teams only. We reply within one business day with slots. Candidate names stay off this site.

Engineer looking at this profile? Do not fill the interview form.Apply to join the bench.