Unknown Company

Site Reliability Engineer

chicago, il • Posted 1 weeks ago
Onsite Full Time IT & Technology

Itasca, IL and moving to Chicago, IL in 2027 (must be ok with 4 days onsite in both locations)


US Citizens and Green Card Holders encouraged to apply (this opportunity does not offer sponsorship now or in the future)


Responsibilities



  • Develop and maintain the organization's monitoring and observability strategy, standards, and best practices.

  • Design, deploy, and manage monitoring platforms and related tools.

  • Collect, analyze, and visualize metrics, logs, traces, and events to deliver end-to-end observability across infrastructure and applications.

  • Build and maintain dashboards, alerts, and reports for infrastructure, operations, development, and security teams.

  • Improve performance, reliability, scalability, and operational maturity of systems and applications.

  • Own capacity planning — anticipating future growth and ensuring systems can handle increased loads without performance degradation.

  • Partner with development teams to improve services through rigorous testing and release procedures.

  • Troubleshoot and resolve issues involving application performance, monitoring, and observability tools.

  • Provide guidance and support to IT teams on effective use of monitoring and observability solutions and help standardize their usage across the organization.

  • Research and assess emerging technologies and trends in application performance, monitoring, and observability.


Required Skills



  • Proven experience in application performance and IT monitoring/observability roles.

  • Bachelor's degree in computer science, information systems, or a related field, or equivalent work experience.

  • At least 3 years of experience in IT monitoring and observability administration, engineering, or support.

  • Strong knowledge of monitoring and observability tools and platforms.

  • Experience with IT infrastructure and application technologies, including Linux, Windows, Azure, networking, databases, and web platforms.

  • Strong experience with scripting and automation tools such as Python, Bash, and Ansible.

  • Strong analytical and problem-solving skills.

  • Strong communication and collaboration skills.

  • Relevant certifications in monitoring and observability tools and platforms preferred.


Preferred



  • Credentials demonstrating a specific level of proficiency in IT monitoring and observability.

  • Working knowledge of the New Relic platform.

  • Application performance testing experience.

#J-18808-Ljbffr

Site Reliability Engineer in chicago at Unknown Company

This position is listed as full time and onsite.

Back to Job Search