Demo

Software Engineer - Model Serving Infrastructure

Anyscale
San Francisco, CA Full Time
POSTED ON 1/4/2026
AVAILABLE BEFORE 2/10/2026
About Anyscale

At Anyscale, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray, a popular open-source project that's creating an ecosystem of libraries for scalable machine learning. Companies like OpenAI, Uber, Spotify, Instacart, and many more, have Ray in their tech stacks to accelerate the progress of AI applications out into the real world.

With Anyscale, we’re building the best place to run Ray, so that any developer or data scientist can scale an ML application from their laptop to the cluster without needing to be a distributed systems expert.

Proud to be backed by Andreessen Horowitz, NEA, and Addition with $250 million raised to date.

Anyscale is actively seeking talented engineers to join our team and contribute to the development of next-generation, high-performance machine learning serving systems. We value diversity and inclusion, and we encourage individuals from underrepresented groups to apply.

Many existing ML serving tools are inherited from previous infrastructure generations, but emerging ML applications present new requirements, such as high compute demands, specialized hardware needs, and the integration of multiple models and business logic within a single request. At Anyscale, our mission is to provide a powerful yet simple set of tools that enable the seamless deployment of complex ML applications in production.

About The Serving Infrastructure Team

The Serving Infrastructure team is dedicated to creating world-class systems for serving ML models in production. This includes building and maintaining the open-source Ray Serve library and contributing directly to the Anyscale platform used by our customers for critical applications. Our work is often user-facing, allowing you to collaborate with open-source users and customers ranging from lean ML engineering teams at small startups to industry-leading companies such as Uber, Shopify, and ByteDance.

As Part Of This Role, You Will

  • Develop a highly available service for ML model serving.
  • Enhance Ray Serve and our other libraries to simplify the development of next-generation ML applications in production.
  • Improve our autoscaling capabilities to drive performance enhancements and cost savings.
  • Optimize latency and throughput for both single- and multi-model serving scenarios.

We'd Love To Hear From You If You Have

  • A solid background in algorithms, data structures, and system design.
  • Experience working with modern machine learning tooling, including PyTorch, TensorFlow, and JAX.
  • At least 2 year of relevant work experience.

Bonus points!

  • Experience in building and maintaining open-source projects.
  • Experience in building and operating machine learning infrastructure in production.
  • Experience in building highly available serving systems.

Compensation

  • At Anyscale, we take a market-based approach to compensation. We are data-driven, transparent, and consistent. The target salary for this role is $170,112 ~ $237,000. As the market data changes over time, the target salary for this role may be adjusted.
  • This role is also eligible to participate in Anyscale's Equity and Benefits offerings, including the following:
    • Stock Options
    • Healthcare plans, with premiums covered by Anyscale at 99%
    • 401k Retirement Plan
    • Education & Wellbeing Stipend
    • Paid Parental Leave
    • Fertility Benefits
    • Flexible Time Off
    • Commute reimbursement
    • 100% of in office meals covered
Anyscale Inc. is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law.

Anyscale Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish

Salary : $170,112 - $237,000

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Software Engineer - Model Serving Infrastructure?

Sign up to receive alerts about other jobs on the Software Engineer - Model Serving Infrastructure career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$97,257 - $120,701
Income Estimation: 
$123,167 - $152,295
Income Estimation: 
$119,030 - $151,900
Income Estimation: 
$149,493 - $192,976
Income Estimation: 
$184,796 - $233,226
Income Estimation: 
$179,606 - $233,815
Income Estimation: 
$101,387 - $124,118
Income Estimation: 
$119,030 - $151,900
Income Estimation: 
$77,900 - $95,589
Income Estimation: 
$101,387 - $124,118

Sign up to receive alerts about other jobs with skills like those required for the Software Engineer - Model Serving Infrastructure.

Click the checkbox next to the jobs that you are interested in.

  • Bug/Defect Analysis Skill

    • Income Estimation: $72,620 - $96,681
    • Income Estimation: $74,092 - $105,774
  • Debugging Skill

    • Income Estimation: $72,620 - $96,681
    • Income Estimation: $74,206 - $95,716
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Anyscale

  • Anyscale San Francisco, CA
  • About Anyscale: At Anyscale , we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We... more
  • 14 Days Ago

  • Anyscale San Francisco, CA
  • At Anyscale, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializi... more
  • 14 Days Ago

  • Anyscale San Francisco, CA
  • About Anyscale At Anyscale, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’r... more
  • 14 Days Ago

  • Anyscale San Francisco, CA
  • About Anyscale At Anyscale, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’r... more
  • 4 Days Ago


Not the job you're looking for? Here are some other Software Engineer - Model Serving Infrastructure jobs in the San Francisco, CA area that may be a better fit.

  • Databricks San Francisco, CA
  • At Databricks, we are passionate about enabling data teams to solve the world's toughest problems — from making the next mode of transportation a reality t... more
  • 4 Days Ago

  • Scale AI, Inc. San Francisco, CA
  • As a Software Engineer on the ML Infrastructure team, you will design and build platforms for scalable, reliable, and efficient serving of LLMs. Our platfo... more
  • 17 Days Ago

AI Assistant is available now!

Feel free to start your new journey!