Demo

Research Scientist, Efficient ML Systems

Goaly AI
Sunnyvale, CA Full Time
POSTED ON 6/4/2026
AVAILABLE BEFORE 11/30/2026

About Goaly At Goaly, our mission is to make custom AI affordable for every business. Our founding team comes from the front lines of top AI labs and tech giants (Meta MSL, TikTok AI, Google DeepMind, xAI, Microsoft Research, etc.), where we built large-scale training infrastructure powering trillion-parameter models and scaled GenAI models to a global user base. Now, we are building something we wish we had before: a platform that makes training and adapting custom AI affordable for all modern companies, not just Big Tech. Our north star is ambitious: for a domain-specific task, reach 90% of SOTA performance at less than 10% of the cost. To get a taste of what we are doing, see our first tech blog.


Role Description As an AI Research Scientist (Efficient ML Systems) at Goaly, you will research and build the systems that make frontier-scale models practical. This role sits at the intersection of algorithms, systems, and hardware efficiency.You will design and evaluate new training and inference techniques, prototype them in real systems, and push them to production-scale workloads.


Your work will either ship directly into our core platform or lead to publications at top venues such as NeurIPS, ICML, ICLR, or CVPR.This is not a paper-only role. You will write real systems code, run large-scale experiments, and directly shape how modern LLMs and RL systems are trained and deployed.


Core Responsibilities

  • Research efficient AI/ML systems: Invent and evaluate algorithms and system techniques that improve LLM and agentic RL training and inference efficiency (memory, compute, communication, and stability).
  • Scale agentic RL: Design and optimize large-scale agentic RL pipelines, including asynchronous training, experience management, reward modeling, and long-horizon stability.
  • End-to-end experimentation: Design large-scale experiments spanning model architecture, training algorithms, distributed systems, and hardware-aware optimization.
  • System-aware research: Prototype research ideas directly in training and inference stacks (e.g., parallelism strategies, attention kernels, RL training pipelines) and validate them at scale.
  • Production & publication: Translate successful ideas into production-ready systems and/or publish them at top-tier conferences with full internal support.


Qualifications

  • Ph.D. or Master's degree in CS, AI, Systems, or related fields (Exceptional undergraduates with strong research capabilities may be considered).
  • Strong foundation in LLM or large-scale ML training, including Transformers, attention mechanisms, distributed training, and optimization methods.
  • Experience or strong interest in agentic RL or large-scale reinforcement learning systems, including stability, scalability, or long-horizon training challenges.
  • Demonstrated interest in efficiency-focused research, such as training acceleration, memory optimization, parallelism, kernels, or RL system robustness.
  • Proficient in PyTorch or JAX. Clean coding style and strong command of Python.
  • Adaptability: A fast learner with a strong sense of responsibility, capable of wearing multiple hats and handling cross-stack challenges.


Why join us?

  • Expert Mentorship: Partner with AI veterans who have trained trillion-parameter models at scale and applied it to solve real-worldproblems in the billon-user products
  • Compute Freedom: Access to abundant GPU cluster resources—don't let your creativity be limited by compute.
  • Flat Culture: Flat management structure that rejects office politics and values only technology and results.
  • Competitive Compensation: Competitive full-time offers, huge upside, and extra equity incentives when hitting key milestones.


Salary.com Estimation for Research Scientist, Efficient ML Systems in Sunnyvale, CA
$139,384 to $174,325
If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Research Scientist, Efficient ML Systems?

Sign up to receive alerts about other jobs on the Research Scientist, Efficient ML Systems career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$108,245 - $136,486
Income Estimation: 
$136,683 - $171,343
Income Estimation: 
$108,245 - $136,486
Income Estimation: 
$136,683 - $171,343
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Goaly AI

  • Goaly AI Sunnyvale, CA
  • About Goaly At Goaly, our mission is to make custom AI affordable for every business. Our founding team comes from the front lines of top AI labs and tech ... more
  • 7 Days Ago


Not the job you're looking for? Here are some other Research Scientist, Efficient ML Systems jobs in the Sunnyvale, CA area that may be a better fit.

  • TikTok San Jose, CA
  • Responsibilities About the Team The Vision-Applied Research team focuses on applied research in Generative AI and CV/Multimodal Understanding, and deliveri... more
  • 10 Days Ago

  • TikTok San Jose, CA
  • Responsibilities TEAM INTRODUCTION The Vision Engineering Team at TikTok is at the forefront of delivering GenAI technologies directly into the TikTok prod... more
  • 1 Month Ago

AI Assistant is available now!

Feel free to start your new journey!