Demo

Senior Site Reliability Engineer (Cloud Platform)

Salve.Lab
Atlanta, GA Full Time
POSTED ON 7/18/2026
AVAILABLE BEFORE 8/15/2026
B2B Contract

Role Overview

We're looking for a Senior Site Reliability Engineer to help build, operate, and continuously improve a highly available cloud platform supporting mission-critical production services.

In this role, you'll work at the intersection of cloud infrastructure, software engineering, and operations, helping engineering teams build reliable, scalable systems while driving automation, observability, and operational excellence. You'll play a key role in strengthening platform reliability, improving incident response, and embedding SRE best practices throughout the software development lifecycle.

Key Responsibilities:

  • Maintain the reliability, availability, and performance of production and pre-production environments.
  • Monitor platform health and improve alerting, automation, and operational processes.
  • Respond to production incidents, participate in root cause analysis, and implement long-term improvements.
  • Design, build, and enhance observability solutions using metrics, logs, traces, and dashboards.
  • Partner with software engineers to improve application reliability throughout the development lifecycle.
  • Develop and maintain operational documentation, troubleshooting guides, and runbooks.
  • Automate repetitive operational tasks to improve efficiency and reduce manual intervention.
  • Participate in on-call rotations while continuously improving incident response processes.
  • Promote reliability engineering principles, operational excellence, and continuous improvement across engineering teams.

Requirements:

  • Bachelor's or Master's degree in Engineering, Computer Science, or a related field.
  • Strong experience operating Kubernetes or other container orchestration platforms.
  • Experience supporting large-scale production services.
  • Hands-on experience with AWS.
  • Experience with Prometheus, Grafana, and ELK.
  • Strong scripting skills (Bash, Python, or Go).
  • Experience administering Linux-based production environments.
  • Experience with Infrastructure as Code or configuration management tools such as Terraform or Ansible.
  • Solid understanding of networking fundamentals (TCP/IP, DNS, load balancing, routing).
  • Excellent troubleshooting, communication, and collaboration skills.
  • A proactive mindset with a passion for automation and reliability.

Nice to Have:

  • Experience with SIP or VoIP technologies.
  • Familiarity with MySQL or PostgreSQL.
  • Experience with Redis or other NoSQL databases.

What's on Offer:

  • Long-term, full-time collaboration.
  • Flexible remote working environment.
  • Professional development opportunities, including training and technical learning.
  • The opportunity to work on innovative cloud technologies used by customers worldwide.
  • Collaborative engineering culture focused on knowledge sharing and continuous improvement.
  • Modern Apple equipment provided.

Diversity and Inclusion Commitment

We are dedicated to creating and sustaining an inclusive, respectful workplace for all -regardless of gender, ethnicity, or background. We actively encourage applicants from all identities and experience levels to apply and bring your authentic self to our fast-paced, supportive team.

Salary.com Estimation for Senior Site Reliability Engineer (Cloud Platform) in Atlanta, GA
$112,135 to $131,530
If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Senior Site Reliability Engineer (Cloud Platform)?

Sign up to receive alerts about other jobs on the Senior Site Reliability Engineer (Cloud Platform) career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$114,618 - $136,401
Income Estimation: 
$144,264 - $191,312
Income Estimation: 
$140,435 - $166,410
Income Estimation: 
$114,618 - $136,401
Income Estimation: 
$144,264 - $191,312
Income Estimation: 
$140,435 - $166,410
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Salve.Lab

  • Salve.Lab Denver, CO
  • Role Overview A fast-growing AI SaaS company is expanding its enterprise sales team. The platform helps financial services firms, insurance companies, and ... more
  • 13 Days Ago

  • Salve.Lab Cincinnati, OH
  • Role Overview A fast-scaling AI SaaS company is looking for an enterprise Account Executive. The platform helps large enterprises replace legacy CRM and BP... more
  • 13 Days Ago

  • Salve.Lab San Diego, CA
  • Role Overview A rapidly growing native-AI SaaS company is expanding its enterprise sales team. The platform replaces outdated CRM and workflow tools with a... more
  • 13 Days Ago

  • Salve.Lab Boston, MA
  • About The Role Join our client’s Revenue team as a Sales Development Representative, where you’ll play a key role in driving business growth within the hea... more
  • 3 Days Ago


Not the job you're looking for? Here are some other Senior Site Reliability Engineer (Cloud Platform) jobs in the Atlanta, GA area that may be a better fit.

  • BioSpace Atlanta, GA
  • Company Description About AbbVie AbbVie's mission is to discover and deliver innovative medicines and solutions that solve serious health issues today and ... more
  • 14 Days Ago

  • Priority Technology Holdings, LLC Alpharetta, GA
  • Job title: Senior Site Reliability Engineer Reports to: Director, Site Reliability Engineering Department: Cloud Platforms Location: Remote Grade: 20 About... more
  • 1 Month Ago

AI Assistant is available now!

Feel free to start your new journey!