Skybitra is seeking a Senior Reliability Engineer to operate production systems and drive resilience across platforms. You will own SLOs, improve on‑call health, and design observable, failure‑tolerant architectures in collaboration with product teams.
Responsibilities include leading incident reviews, building dashboards, and automating runbooks. Strong Python/Go automation and Kubernetes familiarity are preferred, with a focus on scalable, secure operations.
#J-18808-LjbffrSite Reliability Engineer - Reliability & Incident Response in new york at Unknown Company
This position is listed as full time and onsite.