iCapital’s Site Reliability Engineering team is looking for an experienced SRE to design scalable, reliable systems and observability solutions. You will implement SLOs/SLIs, standardize monitoring, and drive automation across Kubernetes-based services in a hybrid work setup in Utah.
The role emphasizes incident response, postmortems, and cross-team collaboration to improve platform reliability with a focus on reducing alert fatigue.
#J-18808-Ljbffr