Cincinnatus LLC is recruiting researchers to design and author multi-step evaluation tasks for frontier AI benchmarks. The role emphasizes translating scientific method into practical tasks, with a focus on Python-based analysis, rigorous evaluation, and clear written conclusions.
You will work remotely in the United States for approximately 35 hours per week as part of a full-time W-2 engagement, integrated with leading AI lab teams through Cincinnatus’ extended workforce model.
#J-18808-Ljbffr