‹ Back
WWAYPOINT
UnitedHealth Group·Schaumburg, IL

Senior Site Reliability Engineer

Support Staff
iPosting details
EducationBachelor's degree
Experience7+ years
SourceUnitedHealth Group · posted Oct 8, 2026
✓Requirements
Certifications
Industry certifications such as Certified Kubernetes Administrator (CKA), AWS Solutions Architect, Azure Solutions Architect Expert, HashiCorp Terraform Associate, or equivalent cloud certifications preferred
Education
✓Bachelor's degree in Computer Science, Engineering, Information Technology, or related field
Qualifications
✓7+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Cloud Engineering, or Software Engineering
✓4+ years of hands-on experience supporting cloud infrastructure in AWS, Azure, or GCP environments
✓3+ years of hands-on experience managing Kubernetes platforms including EKS, AKS, or GKE in production environments
✓3+ years of Infrastructure as Code (IaC) experience using Terraform or equivalent automation technologies
✓3+ years of hands-on experience with observability and monitoring platforms such as Datadog, Splunk, Dynatrace, Grafana, Prometheus, OpenTelemetry, or similar solutions
✓3+ years of experience implementing and supporting monitoring, logging, distributed tracing, alerting, SLIs, SLOs, and Error Budget frameworks
✓3+ years of experience building and supporting CI/CD pipelines using GitHub Actions, Azure DevOps, Jenkins, ArgoCD, or equivalent technologies
✓3+ years of experience with scripting and automation skills using Python, Bash, PowerShell, or similar languages
✓Ability to participate in rotating on-call support schedules
✓All employees working remotely will be required to adhere to UnitedHealth Group's Telecommuter Policy
Experience supporting mission-critical production systems and leading incident response and root cause analysis activities preferred
Strong understanding of distributed systems, cloud-native architectures, networking, security, IAM, encryption, and reliability engineering principles preferred
Proven ability to collaborate effectively across engineering, platform, architecture, and security teams preferred
Experience implementing enterprise observability solutions using Datadog APM, Splunk Observability Cloud, Dynatrace, Grafana, or OpenTelemetry preferred
Experience with AIOps, intelligent alerting, anomaly detection, operational automation, and predictive analytics platforms preferred
Experience supporting AI/ML, Generative AI, Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), or data-intensive workloads in production environments preferred
Experience with GitOps frameworks such as ArgoCD or Flux preferred
Experience supporting multi-region and multi-cluster cloud deployments preferred
Experience mentoring engineers and leading reliability improvements across multiple teams preferred
Experience working within regulated environments such as Healthcare, HIPAA, SOC2, NIST, or FedRAMP preferred
We comply with all minimum wage laws as applicable. preferred
+Benefits
✓401(k)
+UnitedHealth Group
8support roles open
SchaumburgIL · 24 mi to Chicago
Pay for this position
Employer-posted
$92k – $164k/yr
Apply to UnitedHealth GroupContact Recruiter about this role
✓You’ll need
Certifications
Industry certifications such as Certified Kubernetes Administrator (CKA), AWS Solutions Architect, Azure Solutions Architect Expert, HashiCorp Terraform Associate, or equivalent cloud certifications preferred
Education
✓Bachelor's degree in Computer Science, Engineering, Information Technology, or related field
Qualifications
✓7+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, Cloud Engineering, or Software Engineering
✓4+ years of hands-on experience supporting cloud infrastructure in AWS, Azure, or GCP environments
✓3+ years of hands-on experience managing Kubernetes platforms including EKS, AKS, or GKE in production environments
✓3+ years of Infrastructure as Code (IaC) experience using Terraform or equivalent automation technologies
✓3+ years of hands-on experience with observability and monitoring platforms such as Datadog, Splunk, Dynatrace, Grafana, Prometheus, OpenTelemetry, or similar solutions
✓3+ years of experience implementing and supporting monitoring, logging, distributed tracing, alerting, SLIs, SLOs, and Error Budget frameworks
✓3+ years of experience building and supporting CI/CD pipelines using GitHub Actions, Azure DevOps, Jenkins, ArgoCD, or equivalent technologies
✓3+ years of experience with scripting and automation skills using Python, Bash, PowerShell, or similar languages
✓Ability to participate in rotating on-call support schedules
✓All employees working remotely will be required to adhere to UnitedHealth Group's Telecommuter Policy
Experience supporting mission-critical production systems and leading incident response and root cause analysis activities preferred
Strong understanding of distributed systems, cloud-native architectures, networking, security, IAM, encryption, and reliability engineering principles preferred
Proven ability to collaborate effectively across engineering, platform, architecture, and security teams preferred
Experience implementing enterprise observability solutions using Datadog APM, Splunk Observability Cloud, Dynatrace, Grafana, or OpenTelemetry preferred
Experience with AIOps, intelligent alerting, anomaly detection, operational automation, and predictive analytics platforms preferred
Experience supporting AI/ML, Generative AI, Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), or data-intensive workloads in production environments preferred
Experience with GitOps frameworks such as ArgoCD or Flux preferred
Experience supporting multi-region and multi-cluster cloud deployments preferred
Experience mentoring engineers and leading reliability improvements across multiple teams preferred
Experience working within regulated environments such as Healthcare, HIPAA, SOC2, NIST, or FedRAMP preferred
We comply with all minimum wage laws as applicable. preferred
+Benefits
401(k)

More support jobs near Schaumburg, ILMore support jobs nearby

All 768 in Illinois →All 768 →

Highest-paying support roles in Illinois

Employer-posted ranges only
All IL support jobs →
Want the next support job in Illinois by email?
Weekly, free, unsubscribe with one click.

About the role

Lead the implementation and continuous improvement of Site Reliability Engineering (SRE) practices including Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets to improve system reliability and operational excellence