Frontier AI Safety Evaluator

3 giorni fa

Rome, Lazio, Italia Mercor Tempo pieno

Mercor is seeking experienced AI Safety Practitioners to evaluate safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback.

The role involves reviewing content on misinformation, political persuasion, self-harm, violence, biosecurity, and other sensitive domains; applying and refining evaluation rubrics for RLHF, SFT, and