Cincinnatus LLC is hiring a QA/Test Engineer to help design and review complex AI evaluation tasks. You will run tasks, probe edge cases, and debug environments in a remote US role around 35 hours per week. Collaboration with researchers and task authors is a core part of the job.
The ideal candidate has MSc/PhD or equivalent research/engineering experience, strong Python and Git skills, and a keen eye for detail. Prior AI training or model evaluation experience is a plus.
#J-18808-Ljbffr