Project Overview
- Evaluate and refine AI-generated code to ensure that it is efficient, scalable, and reliable.
- Collaborate with cross-functional teams to enhance AI-driven coding solutions against industry performance benchmarks.
- Build agents that can verify the quality of the code and identify error patterns.
- Hypothesize on steps in the software engineering cycle (prototyping, architecture design, API design, production implementation, launch, experiments, monitoring, operational maintenance) and evaluate model capabilities on them.
- Design verification mechanisms that can automatically verify a solution to a software engineering task.
Required Skills
- Several years of software engineering experience (+5 years), including 2+ years of continuous full-time experience at a top-tier product company (e.g., Google, Stripe, Amazon, Apple, Meta, Netflix, Microsoft, Datadog, Dropbox, Shopify, PayPal, IBM Research).
- Strong expertise in building full-stack applications and deploying scalable, production-grade software using modern languages and tools.
- Deep understanding of software architecture, design, development, debugging, and code quality/review assessment.
- Excellent oral and written communication skills for clear, structured evaluation rationales.
Commitment
- Flexible engagement, minimum 10 hrs/week, up to 40 hrs/week (partial PST overlap required).
Type
- Contractor (no medical/paid leave).
Duration
- 1 month (starting next week; potential extensions based on performance and fit).
Location
- Candidates must be based out of US, Canada or WEU countries (UK, Netherlands, Italy, Germany, …).
Notes
Note: As part of assessments you will go through an AI video interview.
After applying, you will receive an email with a login link. Please use that link to access the portal and complete your profile.
Refer talent at turing.com/referrals and earn money from your network.
#J-18808-Ljbffr