Mercor is seeking a highly skilled AI-evaluation specialist to join a leading GenAI team. This full-time, remote role within the United States focuses on red-teaming frontier models, designing robust benchmark tasks, and documenting findings for reproducibility.
You will work closely with researchers in a collaborative, results-driven environment. Ideal candidates have advanced degrees in STEM, 1+ years in AI evaluation or related roles, and strong Python/Git skills.
#J-18808-Ljbffr