← Back to Talent Bench
Ref #HL-QA-3308Lead Level · 9 Years Experience
Deployment ReadyAI QA Engineer (LLM Applications)
Specialization: Gen AI Evaluation & Safety · Industries: B2B SaaS, Healthcare, Financial Services · Vetting Rating: Interview-ready · quality engineering review
Contract EngagementContractor / Deployable
AvailabilityImmediate (Within 48 hrs)
Work ModelRemote worldwide
Benchmark Rate~$90 / hr
Verified Technical Stack:
PlaywrightTypeScriptAPI Contract TestingCI PipelinesPerformance Testingk6GitHub ActionsAccessibility TestingTest Data Strategy
Professional Summary
AI QA engineer for LLM-backed products — eval harnesses, golden datasets, and CI gates alongside Playwright where UI paths still matter.
Verified Impact & Project Highlights
- Introduced LLM eval gates in pull-request CI alongside contract tests for API surfaces
- Maintains golden datasets versioned with prompt and retrieval config changes
- Partners with eval engineers on judge calibration instead of duplicating rubrics
Request an interview with HL-QA-3308
Hiring teams only. We reply within one business day with slots. Candidate names stay off this site.
Engineer looking at this profile? Do not fill the interview form.Apply to join the bench.