Demo

Site Reliability Engineer

Evlo AI
Phoenix, AZ Full Time
POSTED ON 9/25/2026
AVAILABLE BEFORE 10/23/2026
About The Role

The Site Reliability Engineer owns the reliability, availability, and operational maturity of production services running across AWS. The role spans Kubernetes, Terraform, CI/CD, observability, incident response, and the automation required to operate distributed systems at scale.

The team is building dependable platform capabilities for engineering teams that ship frequently and serve demanding workloads. This role matters because it turns operational risk into measurable engineering improvements through resilient architecture, clear service-level objectives, and disciplined automation.

Key Responsibilities

  • Design and operate highly available AWS infrastructure using Kubernetes, Terraform, Helm, and infrastructure-as-code best practices
  • Define and enforce service-level objectives, error budgets, and reliability standards across production services
  • Build and maintain observability systems using Prometheus, Grafana, OpenTelemetry, and centralized logging platforms
  • Automate deployment, scaling, backup, recovery, and routine operational workflows through Python, Go, or Bash
  • Lead incident response, coordinate remediation during high-severity events, and produce blameless post-incident reviews with actionable follow-up work
  • Improve CI/CD pipelines with deployment safety controls, automated testing, progressive delivery, and reliable rollback procedures
  • Partner with software engineers to identify performance bottlenecks, eliminate recurring operational toil, and strengthen failure-mode testing

What We Are Looking For

  • 3–8 years of experience in site reliability engineering, DevOps, platform engineering, or production infrastructure roles
  • Hands-on experience operating Kubernetes workloads in AWS, including networking, IAM, autoscaling, storage, and production troubleshooting
  • Strong proficiency with Terraform or an equivalent infrastructure-as-code framework, plus practical experience managing cloud resources through version control
  • Experience building observability and alerting systems with tools such as Prometheus, Grafana, OpenTelemetry, Datadog, or equivalent platforms
  • Proficiency in Python, Go, or Bash for operational automation, tooling, and service integration; familiarity with Linux internals and networking fundamentals
  • Bachelor’s degree in computer science, engineering, or a related technical field, or equivalent professional experience
  • Bonus: Experience with AWS certifications, service mesh technologies, GitOps, chaos engineering, distributed databases, or compliance requirements for production infrastructure

Salary.com Estimation for Site Reliability Engineer in Phoenix, AZ
$87,792 to $109,268
If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Site Reliability Engineer?

Sign up to receive alerts about other jobs on the Site Reliability Engineer career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$92,877 - $110,401
Income Estimation: 
$120,933 - $155,034
Income Estimation: 
$114,618 - $136,401
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Evlo AI

  • Evlo AI Washington, DC
  • About The Role The Business Analyst translates business needs into reliable data, system, and process improvements across revenue and operational workflows... more
  • 1 Day Ago

  • Evlo AI Phoenix, AZ
  • About The Role The Software Engineer - Full Stack role builds and operates customer-facing web products across the application stack, from responsive React... more
  • 1 Day Ago

  • Evlo AI Chicago, IL
  • About The Role The Operations Manager coordinates the processes, systems, and cross-functional programs that support a cloud-based technology business. The... more
  • 2 Days Ago

  • Evlo AI Chicago, IL
  • About The Role The Mobile Engineer will build and evolve iOS and Android applications that serve as primary interfaces for user-facing products. The role c... more
  • 2 Days Ago


Not the job you're looking for? Here are some other Site Reliability Engineer jobs in the Phoenix, AZ area that may be a better fit.

  • Precisely Phoenix, AZ
  • At Precisely, we're not just building software — we're shaping the future of data integrity. As a global leader in data quality, data enrichment, and locat... more
  • 9 Days Ago

  • Chainlink Labs Phoenix, AZ
  • About Chainlink Chainlink is the industry-standard oracle platform bringing the capital markets onchain and powering the majority of decentralized finance ... more
  • 2 Days Ago

AI Assistant is available now!

Feel free to start your new journey!