Terac is conducting a remote study in the United States to assess the quality and realism of coding environments used for AI agent testing. As a software engineer, you will review tasks and evaluation harnesses, discuss their technical soundness, and identify potential flaws in the code structures.
The compensation is $75 per hour, and participants will provide feedback on difficulty, realism, and alignment with industry practices.
#J-18808-LjbffrSoftware Engineer – Evaluation Harness Reviewer (Remote) in Location not specified at Unknown Company
This position is listed as full time and able to be worked remotely.