Dorado seeks a contributor in the United States to help build a dataset for evaluating AI coding agents. You will design challenging tasks and evaluation criteria within realistic simulated environments—creating developer environments, tickets, docs, and conversations that form a believable history.
You will write tests that accept diverse valid solutions, iterate on tasks based on QA feedback, and participate end-to-end from applying to getting paid.
#J-18808-Ljbffr