Frontier AI Safety Evaluator

Il y a 2 jours

Brussel Hoofdstad, Belgique Obsidian Temps plein 70 000 € - 110 000 € Contrat

Obsidian is seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models. You will assess AI-generated responses, apply safety policies, and help improve model behavior through structured evaluations and feedback.

You will review content spanning misinformation, political persuasion, self-harm, violence, cyber, biosecurity, and other sensitive domains, and refine evaluation rubrics for RLHF, SFT, and AI safety benchmarking.