Role category

RAG Evaluation & Regression

Faithfulness, context precision/recall, citation checks, and regression packs so retrieval quality does not drift silently.

3 open roles · Apply on the listing · No login

Full-time · RAG evalRemote

Lead Generative AI & Agentic Systems Architect

You will own production Gen AI architecture: retrieval that stays grounded, agents that call tools over MCP, and evaluations that catch regressions before users...

Generative AIAgentic WorkflowsModel Context Protocol (MCP)Vector DatabasesPython
Closes 2026-12-31Apply
Contract · RAG evalRemote

GenAI Evaluation Engineer

Join a Halcer Gen AI delivery pod as the engineer who makes model behaviour measurable. You will build evaluation suites, prompt regression, and safety guardrai...

LLM EvaluationGuardrailsPythonPrompt RegressionRAG Quality
Closes 2026-12-31Apply
Contract · RAG evalRemote

RAG Evaluation Specialist

Build and maintain RAG regression packs: faithfulness, context precision/recall, and citation checks wired into CI. You work with client-chosen observability st...

RAG EvaluationFaithfulness MetricsCitation ChecksPythonGolden Datasets
Closes 2026-12-31Apply

Other categories