Obsidian is seeking experienced AI Safety Practitioners to evaluate the safety, quality, and alignment of frontier AI models across complex, policy-sensitive topics. You will assess AI-generated responses, apply safety policies, and provide structured feedback to improve model behavior through evaluations and collaboration with AI researchers and safety teams.
Responsibilities include iterating on rubrics for RLHF, SFT, and safety benchmarking, identifying unsafe outputs and misalignment, and
#J-18808-LjbffrFrontier AI Safety Evaluator in san francisco at Unknown Company
This position is listed as full time and onsite.