Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code). You will author original, executable research problems that today's frontier models cannot solve.
In this six-week, part-time role (20+ hours/week), you will source material, write prompts, build grading criteria, and calibrate against frontier models to ensure the tasks are challenging and robust. Start date is immediate, with a collaborative, research-oriented environment.
#J-18808-LjbffrQuantum Computing Research Engineer for AI Benchmarks in san francisco at Unknown Company
This position is listed as part time and onsite.