Demo

Cloud Engineer/Site Reliability Architect

Slalom
Portland, OR Full Time
POSTED ON 9/18/2026
AVAILABLE BEFORE 10/16/2026

This is a Kubernetes-heavy platform engineering role supporting large-scale, multi-cloud GPU infrastructure. Success requires deep operational judgment, disciplined debugging, strong automation skills, and the ability to protect platform stability while capacity grows rapidly.

What You’ll Do

    Operate Kubernetes platforms at significant scale across providers, including Amazon Elastic Kubernetes Service (EKS), CoreWeave Kubernetes Service (CKS), and Google Kubernetes Engine (GKE).

    Own the Kubernetes cluster lifecycle, including node-pool design and management, scheduler troubleshooting, container network interface (CNI) and network-policy troubleshooting, capacity planning, and safe rolling upgrades across large fleets.

    Develop, review, and maintain infrastructure as code with Terraform, along with Python tooling and automation that improve reliability, repeatability, and operational efficiency.

    Define and maintain service-level indicators (SLIs) and service-level objectives (SLOs); build monitoring and alerting that surface meaningful risks before they affect workloads.

    Debug complex distributed-system failures by forming hypotheses, testing them methodically, and separating temporal correlation from causation.

    Provision high-performance computing (HPC) and GPU infrastructure through the Conveyor CI/CD system across AWS, CoreWeave, Google Cloud Platform (GCP), and Oracle Cloud Infrastructure (OCI), with additional providers as the platform expands.

    Coordinate daily with Networking, Storage, Security, and AI/ML platform teams to resolve cross-system dependencies and improve the end-to-end developer and researcher experience.

    Bring forward well-reasoned proposals, solutions, and informed opinions; use available tools and resources to self-unblock and drive issues to resolution.

What You’ll Bring

  • Hands-on Kubernetes experience including node pool sizing, scheduler debugging, CNI troubleshooting, and rolling upgrades across fleets.
  • Strong Terraform proficiency and experience writing and reviewing production infrastructure as code on a daily basis.
  • Practical Python skills for tooling and automation, including the ability to write code without relying on artificial intelligence assistance.
  • Strong critical-thinking and debugging skills, with a disciplined approach to forming hypotheses, validating evidence, and isolating root causes.
  • Ability to read unfamiliar code, understand its behavior, and diagnose issues effectively.
  • Resourcefulness and a consistent ability to self-unblock using documentation, telemetry, code, peers, and available tooling.
  • A proactive working style: you arrive with proposals, formulate solutions, and communicate informed technical opinions.
  • Clear communication and effective collaboration across engineering teams and technical stakeholders.
  • Working knowledge of AWS or similar Cloud services relevant to compute and data-intensive platforms, including Amazon EC2, Amazon S3, Amazon EFS, and Amazon FSx for Lustre.
  • Experience designing or operating CI/CD pipelines and automated infrastructure provisioning workflows.
  • Knowledge of cloud networking, storage, security, observability, reliability engineering, and platform governance.

Nice-to-have:

  • Experience using AI coding tools responsibly: you remain accountable for the solution, validate generated code, identify edge cases, and reject unnecessary or incorrect output.


About Us

Slalom is a fiercely human business and technology consulting company that leads with outcomes to bring more value, in all ways, always. From strategy through delivery, our agile teams across 52 offices in 12 countries partner with clients to co-create powerful customer experiences, modern ways of working, and meaningful impact. 

What sets us apart? We believe work should be challenging and fulfilling, not perfect, but possible. That’s why we prioritize purpose, flexibility, connection, and recognition, so our people can thrive and love what they do, most days. 

Compensation and Benefits

Slalom prides itself on helping team members thrive in their work and life. As a result, Slalom is proud to invest in benefits that include meaningful time off and paid holidays, parental leave, 401(k) with a match, a range of choices for highly subsidized health, dental, & vision coverage, adoption and fertility assistance, and short/long-term disability. We also offer yearly $350 reimbursement account for any well-being-related expenses, as well as discounted home, auto, and pet insurance.

Slalom is committed to fair and equitable compensation practices. For this role, we are hiring at the following levels and targeted base pay salary ranges: The target base salary range for Senior Consultant in East Bay, Silicon Valley, and San Francisco is $149,000-$185,000. The target base salary range for Senior Consultant in is $149,000-$185,000. The target base salary range in San Diego, Los Angeles, Orange County, Seattle, Houston, New Jersey, New York City, Westchester, Boston, Washington DC for Senior Consultant is $136,500-$169,500. The target base salary range for Senior Consultant in all other US based Slalom locations is $125,000-$155,500. In addition, individuals may be eligible for an annual discretionary bonus. Actual compensation will depend upon an individual’s skills, experience, qualifications, location, and other relevant factors. The salary pay range is subject to change and may be modified at any time.    

We will accept applicants until 10/5/2026, or until the position is filled.

We are committed to pay transparency and compliance with applicable laws. If you have questions or concerns about the pay range or other compensation information in this posting, please contact us at: peopleone@slalom.com. Please note, this recipient is not able to support recruitment inquiries beyond this purpose.

EEO and Accommodations

Slalom is an equal opportunity employer and is committed to attracting, developing and retaining highly qualified talent who empower our innovative teams through unique perspectives and experiences. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, disability status, protected veterans’ status, or any other characteristic protected by federal, state, or local laws. Slalom will also consider qualified applications with criminal histories, consistent with legal requirements. Slalom welcomes and encourages applications from individuals with disabilities. Reasonable accommodations are available for candidates during all aspects of the selection process. Please advise the talent acquisition team or contact accomodationrequest@slalom.com if you require accommodations during the interview process. 

Salary : $136,500 - $169,500

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Cloud Engineer/Site Reliability Architect?

Sign up to receive alerts about other jobs on the Cloud Engineer/Site Reliability Architect career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$92,369 - $122,605
Income Estimation: 
$117,024 - $149,811
Income Estimation: 
$103,114 - $138,258
Income Estimation: 
$118,163 - $145,996
Income Estimation: 
$120,777 - $151,022
Income Estimation: 
$129,363 - $167,316
Income Estimation: 
$86,891 - $130,303
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Slalom

  • Slalom Portland, ME
  • Who You'll Work With: As a modern technology company, our Slalom Technologists are disrupting the market and bringing to life the art of the possible for o... more
  • 11 Days Ago

  • Slalom Portland, ME
  • Who You'll Work With: As a modern technology company, our Slalom Technologists are disrupting the market and bringing to life the art of the possible for o... more
  • 11 Days Ago

  • Slalom Las Vegas, NV
  • Who You'll Work With: As a modern technology company, our Slalom Technologists are disrupting the market and bringing to life the art of the possible for o... more
  • 11 Days Ago

  • Slalom Las Vegas, NV
  • Who You'll Work With: As a modern technology company, our Slalom Technologists are disrupting the market and bringing to life the art of the possible for o... more
  • 11 Days Ago


Not the job you're looking for? Here are some other Cloud Engineer/Site Reliability Architect jobs in the Portland, OR area that may be a better fit.

  • Electrical Reliability Services, Inc. Portland, OR
  • Job Description The primary function will be to plan and perform sales and marketing efforts for the Company. These territory/account based sales and marke... more
  • 3 Days Ago

  • Electrical Reliability Services, Inc. Portland, OR
  • Job Description Electrical Reliability Services, a rapidly growing third-party, NETA-accredited, electrical testing company is currently seeking a Power Sy... more
  • 30 Days Ago

AI Assistant is available now!

Feel free to start your new journey!