Demo

Senior Systems Engineer, Infrastructure & Platform Reliability

Lambda
San Francisco, CA Full Time
POSTED ON 11/18/2025
AVAILABLE BEFORE 12/3/2025
Lambda, The Superintelligence Cloud, builds Gigawatt-scale AI Factories for Training and Inference. Lambda’s mission is to make compute as ubiquitous as electricity and give every person access to artificial intelligence. One person, one GPU.

If you'd like to build the world's best deep learning cloud, join us.

  • Note: This position requires presence in our San Francisco or San Jose office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.

Information Systems at Lambda is responsible for building and scaling the internal systems that power our business. We partner across the company—Finance, GTM, Engineering, and People—to implement tools, automate workflows, and ensure data flows securely and accurately. Our scope includes enterprise applications, integrations, data platform and analytics, compliance automation, and all things IT.

What You’ll Do

  • Design, write, and deliver software and services to improve the availability, scalability, reliability, and efficiency of Lambda’s internal IT systems and platforms.
  • Solve problems relating to mission critical services and build automation to prevent problem recurrence with the goal of automating response to all non-exceptional events.
  • Work with Lambda Engineering and internal teams to Influence and create new designs, architectures, standards, and methods for large-scale distributed systems.
  • Engage in service capacity planning and demand forecasting, software performance analysis, and system tuning.
  • Be an excellent communicator, producing documentation and related artifacts for the systems you are responsible for.

You

  • Have a keen interest in system design, architecting for performance, scalability, and experience with multiple cloud infrastructure platforms (AWS, GCP, Azure, etc.).
  • Think carefully about systems: edge cases, failure modes, behaviors, and specific implementations.
  • Know and prefer configuration management systems and toolchains (Chef, Ansible, Terraform, GitHub Actions, etc.)
  • Have solid programming skills: Python, Go, etc.
  • Have an urge to collaborate and communicate asynchronously, combined with a desire to record and document issues and solutions.
  • Have an enthusiastic, go-for-it attitude. When you see something broken, you can’t help but fix it.
  • Have an urge for delivering quickly and effectively, and iterating fast.

Nice to Have

  • Experience and interest in ML/AI workloads and compute
  • Practical experience implementing and managing paging, alerting, and on-call scheduling flows
  • A positive attitude, combined with a desire to learn and collaborate

Salary Range Information

The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.

About Lambda

  • Founded in 2012, ~400 employees (2025) and growing fast
  • We offer generous cash & equity compensation
  • Our investors include Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, US Innovative Technology, Gradient Ventures, Mercato Partners, SVB, 1517, Crescent Cove.
  • We are experiencing extremely high demand for our systems, with quarter over quarter, year over year profitability
  • Our research papers have been accepted into top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG
  • Health, dental, and vision coverage for you and your dependents
  • Wellness and Commuter stipends for select roles
  • 401k Plan with 2% company match (USA employees)
  • Flexible Paid Time Off Plan that we all actually use

A Final Note

You do not need to match all of the listed expectations to apply for this position. We are committed to building a team with a variety of backgrounds, experiences, and skills.

Equal Opportunity Employer

Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.

Compensation Range: $206K - $310K

Salary : $206,000 - $310,000

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Senior Systems Engineer, Infrastructure & Platform Reliability?

Sign up to receive alerts about other jobs on the Senior Systems Engineer, Infrastructure & Platform Reliability career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$110,730 - $135,754
Income Estimation: 
$128,617 - $162,576
Income Estimation: 
$117,033 - $148,289
Income Estimation: 
$69,994 - $90,028
Income Estimation: 
$91,370 - $117,201
Income Estimation: 
$84,222 - $112,497
Income Estimation: 
$83,184 - $105,164
Income Estimation: 
$83,184 - $105,164
Income Estimation: 
$115,390 - $147,559
Income Estimation: 
$106,780 - $140,358
Income Estimation: 
$104,963 - $131,876
Income Estimation: 
$104,963 - $131,876
Income Estimation: 
$136,671 - $177,110
Income Estimation: 
$128,093 - $158,900
Income Estimation: 
$128,093 - $158,900
Income Estimation: 
$148,304 - $196,737
Income Estimation: 
$163,759 - $202,445
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Lambda

Lambda
Hired Organization Address Seattle, WA Full Time
Lambda, The Superintelligence Cloud, builds Gigawatt-scale AI Factories for Training and Inference. Lambda’s mission is ...
Lambda
Hired Organization Address Columbus, OH Full Time
Lambda, The Superintelligence Cloud, builds Gigawatt-scale AI Factories for Training and Inference. Lambda’s mission is ...
Lambda
Hired Organization Address San Francisco, CA Full Time
Lambda, The Superintelligence Cloud, builds Gigawatt-scale AI Factories for Training and Inference. Lambda’s mission is ...
Lambda
Hired Organization Address San Francisco, CA Full Time
Lambda, The Superintelligence Cloud, builds Gigawatt-scale AI Factories for Training and Inference. Lambda’s mission is ...

Not the job you're looking for? Here are some other Senior Systems Engineer, Infrastructure & Platform Reliability jobs in the San Francisco, CA area that may be a better fit.

Infrastructure Engineer — Systems & Platform

Sixtyfour, San Francisco, CA

AI Assistant is available now!

Feel free to start your new journey!