San Francisco AI Lab is hiring to build and own evaluation ecosystems for on-device agents and model releases. You will define readiness criteria, design human-scored rubrics, and ship dashboards and tooling to speed research and leadership decisions.
Ideal candidates excel at engineering instrumentation, rapid experiment cycles, and communicating what constitutes meaningful improvement under deadline pressure.
#J-18808-LjbffrOn-Device AI Evaluation Architect in san francisco at Unknown Company
This position is listed as full time and onsite.