Sr Software Engineer (Cloud Native) – Bellevue, WA
Responsibilities
- Build reliable, distributed systems that deliver high performance, observability, and operational visibility at scale.
- Design and develop cloud‑native infrastructure and APIs using Go and Python, building custom Kubernetes controllers and operators to automate platform operations.
- Develop microservices and APIs in Golang, improving developer productivity with Swagger and WSDL.
- Architect multi‑cloud infrastructure across AWS and Azure, utilizing EC2/VMs, VPC/VNet, load balancers, RDS/Azure SQL, S3/Blob Storage, KMS/Key Vault, and DNS/IAM services.
- Implement observability platforms using Prometheus, Thanos, Grafana, OpenTelemetry, Splunk, and Outcold.
- Automate infrastructure provisioning and delivery with Terraform, Ansible, GitOps (Flux), and GitLab CI pipelines.
- Ensure security and compliance with OPA/Gatekeeper, Kubernetes network policies, and Wiz.
- Manage Kubernetes workloads with Containerd, Istio, NGINX Ingress; implement autoscaling (KEDA), progressive delivery (Flagger), disaster recovery (Velero, ETCD), and stateful storage (Portworx).
- Architect and manage GPU‑optimized Kubernetes clusters for AI/ML workloads using NVIDIA GPU nodes; streamline multi‑cluster operations with Spectro Cloud.
- Troubleshoot and debug Linux environments using Bash, systemd, and journalctl for stability and issue resolution.
- Telecommuting permitted, but must work at the worksite at least 3‑4 days per week.
Skill Requirements
- Experience with multiple programming languages (Python, Go, Java, JavaScript) in production environments and designing RESTful APIs, gRPC services, and event‑driven applications.
- Experience deploying and managing Kubernetes clusters on AWS EKS, Azure AKS, and on‑premises (Kubeadm, Spectro Cloud); configuring ContainerD & Docker, RBAC, network policies, Istio & Cilium, and designing scalable micro‑services architectures.
- Experience designing and deploying cloud‑native systems with 12‑Factor principles, building stateless microservices and connecting to external services.
- Experience with Ubuntu, RHEL, and Flatcar Linux OS, Bash scripting, troubleshooting kernel, network, performance issues, and root cause analysis.
- Experience with DevOps tools: GitLab CI/CD, Terraform, Ansible, Vault, KMS; and
- Experience with Prometheus, Grafana, OpenTelemetry, Splunk for monitoring, tracing, logging, and using observability data to improve performance.
Experience and Education Requirements
- Primary: Bachelor’s degree in Computer Engineering or related field plus 5 years of related work experience.
- Alternative: Master’s degree in Computer Engineering or related field plus 3 years of related work experience.
Additional Information
- Location: Bellevue, WA.
- Work hours: 40 hours per week.
- Washington pay range: $212,202.00 – $217,202.00 per year.
- Travel required: No.
T-Mobile USA, Inc. is an Equal Opportunity Employer. All decisions concerning the employment relationship will be made without regard to age, race, ethnicity, color, religion, creed, sex, sexual orientation, gender identity or expression, national origin, religious affiliation, marital status, citizenship status, veteran status, the presence of any physical or mental disability, or any other status or characteristic protected by federal, state, or local law. Discrimination, retaliation or harassment based upon any of these factors is wholly inconsistent with how we do business and will not be tolerated. Talent comes in all forms at the Un‑carrier. If you are an individual with a disability and need reasonable accommodation at any point in the application or interview process, please let us know by emailing or calling 1‑844‑873‑9500.
#J-18808-Ljbffr