Mercor is hiring PhD and Master's scientists to author AI evaluation tasks (Sci Code) for a new benchmark in scientific computing. You will author original, executable research problems that today’s frontier models cannot solve, sourced from published papers, datasets, or open-source repositories.
The role focuses on biology-related domains with coding, including constructing prompts, defining grading criteria, and calibrating model performance.
#J-18808-LjbffrAI Benchmark Scientist for Biochemistry & Genomics in san francisco at Unknown Company
This position is listed as full time and onsite.