Unknown Company

Software Engineer, Infrastructure - Core Experimentation

seattle, wa • Posted 4 days ago
Onsite Full Time General

About the RoleAs a Senior Infrastructure Engineer on the Statsig team, you will build and scale the foundational systems behind OpenAI's experimentation and rollout platform. You will work on the distributed control plane, SDK and server evaluation paths, ingestion pipelines, analytics foundations, and operational tooling that make launches safe and measurable at OpenAI scale.This role is deeply technical and centered on performance, scalability, reliability, and correctness. You will design systems that serve low-latency configuration decisions, handle high-throughput event ingestion, preserve data availability for experimentation and analytics workflows, and keep critical launch infrastructure dependable as usage grows.The work matters because OpenAI's next phase depends on learning quickly without compromising safety or reliability.

Every major product surface needs a trusted way to evaluate changes, progressively roll them out, understand impact, and recover cleanly. Statsig is one of the core infrastructure layers that makes that possible.In This Role, You WillDesign and operate low-latency configuration delivery systems powering feature flags, dynamic configs, and progressive rollouts across OpenAI product suites.Scale SDK, server-side evaluation, and control-plane systems so high-volume services can depend on Statsig without adding user-visible latency or operational fragility.Build high-throughput data ingestion and analytics infrastructure for experimentation, product analytics, feature performance monitoring, and model or product measurement workflows.Improve performance, efficiency, reliability, and observability of core Statsig infrastructure as OpenAI products scale globally.Optimize query performance, data freshness, and data availability for teams making launch decisions from experimentation and analytics workflows.Strengthen operational excellence through better SLOs, alerting, debugging tools, incident response, capacity planning, and failure-mode design.Partner with teams across ChatGPT, Codex, model measurement, consumer ads, business subscriptions, developer products, and infrastructure to turn recurring launch and measurement needs into durable platform capabilities.Lead large technical initiatives and shape the architecture of experimentation and rollout infrastructure used across the company.You Might Thrive In This Role If YouHave experience building large-scale distributed systems with strict performance, availability, and correctness requirements.Enjoy low-latency systems work, including real-time configuration delivery, SDK/runtime performance, caching, concurrency, and high-throughput service design.Have built or operated large-scale data platforms, event pipelines, analytics systems, or query infrastructure where freshness and correctness matter.Care deeply about reliability, observability, incident response, capacity planning, and the practical craft of keeping critical production systems boring in the best way.Are excited by the post-acquisition chapter of a high-performing infrastructure team: preserving Statsig's strengths while integrating deeply into OpenAI's product and platform stack.Want to work on infrastructure that directly affects how ChatGPT, Codex, model measurement, ads, business subscriptions, developer products, and future OpenAI surfaces ship.Take ownership of complex technical problems end to end and enjoy building infrastructure that lets other teams move faster with more confidence.

Back to Job Search