Unknown Company

SRE - PERM REMOTE

Remote • Posted 4 days ago
Onsite Full Time Architecture and Engineering Occupations

Job Description

n

We are seeking a Site Reliability Engineer (SRE) / DevOps Engineer to join a growing team responsible for supporting and scaling a Microsoft Azure environment. This role is ideal for an engineer who enjoys blending infrastructure, automation, deployment support, and reliability engineering to ensure highly available, scalable systems.

n

As the organization continues to grow its customer base, this individual will play a key role in improving observability, monitoring system health, diagnosing issues, and proactively preventing outages before they occur. While the position currently leans more heavily toward DevOps and infrastructure support, it will evolve into a balanced SRE/DevOps role with increased ownership of reliability initiatives, capacity planning, root cause analysis, and system performance optimization.

n

Responsibilities

n

Support and maintain cloud infrastructure within Microsoft Azure

n

Build and manage CI/CD pipelines and deployment automation

n

Monitor application and system health using Azure observability tools

n

Analyze logs, diagnose incidents, and troubleshoot production issues

n

Improve monitoring, alerting, and overall system observability

n

Partner with engineering teams to improve application reliability and performance

n

Implement Infrastructure-as-Code (IaC) solutions

n

Participate in preventative maintenance, system optimization, and capacity planning efforts

n

Contribute to a culture of proactive reliability engineering and operational excellence

n

We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to learn more about how we collect, keep, and process your private information, please review Insight Global's Workforce Privacy Policy:

n

Skills and Requirements

n

Experience

n

2-5 years of experience in a DevOps Engineer, Site Reliability Engineer (SRE), Cloud Engineer, or related role

n

Experience working within Microsoft Azure environments

n

Strong troubleshooting, problem-solving, and systems-thinking abilities

n

Experience supporting applications in a .NET ecosystem

n

Azure & Cloud Technologies

n

Azure DevOps

n

Azure Monitor

n

Application Insights

n

Log Analytics

n

Azure CLI

n

Azure Container Apps

n

Infrastructure & Automation

n

YAML-based pipeline deployments

n

Bicep

n

Git repositories and source control best practices

n

Infrastructure-as-Code (IaC) experience

n

Reliability & Operations

n

Log analysis and troubleshooting

n

Incident diagnosis and resolution

n

System health monitoring

n

Strong observability mindset Kusto Query Language (KQL)

n

ARM Templates

n

Additional Azure platform expertise

n

Experience implementing or managing observability solutions

n

Capacity planning experience

n

Root cause analysis and post-incident review experience

n

Previous ownership of SRE initiatives or reliability programs

n

Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field

SRE - PERM REMOTE in Remote at Unknown Company

This position is listed as full time and onsite.

Back to Job Search