- The Platform & Infrastructure Engineering team sits at the core of our Data Infrastructure organization, responsible for the availability, reliability, scalability, and security of the company’s data platform
- We build and operate the foundational systems that underpin data ingestion, transformation, analytics, and internal AI workloads at scale — delivering with production-grade discipline across mission-critical services
- Our work is guided by a commitment to automation, deep observability, and resilience, supporting stringent uptime requirements that the broader organization depends on
- As a Senior Software Engineer, you will take end-to-end ownership of the reliability and performance of our Kubernetes-based data platform
- You will architect and operate highly available, multi-region systems designed to meet rigorous uptime and latency standards
- Your day-to-day work will span scaling infrastructure, optimizing deployment pipelines, and strengthening our security posture — while playing a central role in building, testing, and managing scalable systems that are production-ready from day one
You love building highly reliable systems that operate at scale.
Must-have experience
- Data Platform Experience: Familiarity operating data platforms or data-intensive workloads, including distributed processing and streaming frameworks such as Spark, Airflow, Kafka, or Flink.
- CI/CD Systems: Proven track record building and operating robust CI/CD pipelines using tools such as Argo CD and GitHub Actions.
- Distributed Systems: Hands‑on experience developing large-scale distributed systems, databases, and backend APIs.
- Cloud‑Native Security: Hands‑on experience applying security best practices in cloud-native settings, including secrets management, network policies, and vulnerability scanning.
We believe in investing in our people, and value candidates who can bring their own diversified experiences to our teams – even if you aren’t a 100% skill or experience match. Here are a few qualities we’ve found compatible with our team. If some of this describes you, we’d love to talk.
Desired qualities
- You bring 7+ years of hands‑on experience in Platform Engineering, Infrastructure Engineering, or building highly scalable distributed systems — with a strong foundation in software design, development, and algorithmic problem‑solving.
- Observability: Strong experience building and owning full‑stack observability solutions, including metrics, logging, and distributed tracing using tools such as Prometheus, Grafana, and OpenTelemetry.
- You’re an expert in diagnosing and solving complex distributed systems problems.
- Production Ownership: Demonstrated experience owning mission‑critical systems with high availability requirements (≥99.99% uptime), including incident response, SLI/SLO/SLA definition, error budget management, and blameless postmortems.
- Infrastructure as Code: Proficiency with IaC tooling such as Helm, Terraform, or Pulumi, and experience with automated environment provisioning.
- Multi‑Region Architecture: Practical experience designing and operating geo‑replicated, active‑active, multi‑region systems — with a solid grasp of traffic routing, failover strategies, and data consistency tradeoffs.
- Performance & Capacity: Strong command of system performance tuning, capacity planning, and resource optimization in distributed environments.
- You’re curious about how to continuously improve system resilience, security, and operations.
- Kubernetes & Containerization: Deep expertise in Kubernetes cluster design, day‑to‑day operations, and production troubleshooting across containerized service environments.
- Regulated Environments: Experience working within compliance‑driven environments and a working knowledge of regulatory frameworks such as GDPR, SOC 2, HIPAA, or SOX.
- Internal Developer Platforms: Background in designing and building internal developer platforms or self‑service infrastructure tooling that empowers engineering teams to move faster and operate independently.
Wondering if you’re a good fit?
#J-18808-Ljbffr