Demo

Post-Training Research Scientist

Two Sigma
York, NY Full Time
POSTED ON 9/27/2026
AVAILABLE BEFORE 11/18/2026
Position Summary

Two Sigma is a leading quantitative investment management and trading firm. The company applies a scientific approach to investing, combining cutting-edge technology, artificial intelligence, data science, and quantitative research with rigorous human inquiry to capitalize on market opportunities and deliver alpha for investors.

Our team of engineers, quantitative researchers and data scientists looks beyond the traditional to test hypotheses and develop creative solutions to some of the world’s most complex economic problems.

We are applying large language models and transformer-based architectures to problems where ground truth is delayed, noisy, and non-stationary. Our systems generate code, run experiments, and iterate autonomously, and we are looking to go beyond supervised fine-tuning.

We are hiring a Post-Training Research Scientist to build RLHF, DPO, and reward modeling capabilities from the ground up. This is a greenfield role: you will define the infrastructure, research agenda, and evaluation frameworks for aligning LLMs to sophisticated, multi-step workflows in a domain where the reward signal is fundamentally different from existing research on human preference or deterministic task completion.

This hire will help own methodology across training, fine-tuning, context management, and model evaluation. You will shape not only the post-training capability but the broader research direction of the team.

You Will Take On The Following Responsibilities

  • Lead post-training efforts for LLMs applied to financial time series and quantitative reasoning
  • Design and execute RLHF, DPO, and related alignment methods at scale, including deployment of substantial compute budgets (O($100mm))
  • Build infrastructure for preference data collection, reward modeling, and policy optimization on financial datasets
  • Drive research agenda connecting post-training methods to quantitative finance applications
  • Collaborate with quant researchers to define task distributions and evaluation frameworks
  • Unblock production systems dependent on post-training capabilities

You Should Possess The Following Qualifications

  • BS or equivalent work experience in Science, Technology, Engineering or Math (an MS is a plus).
  • Minimum 1 year of experience required; 1-10 years of experience preferred (ideally 1-5 years) at a frontier AI lab (OpenAI, Anthropic, DeepMind, Meta FAIR, or equivalent)
  • Shipped post-training systems in production: RLHF, DPO, RLAIF, or related methods
  • Deep understanding of distributed training infrastructure: multi-node GPU clusters, training stability, checkpointing
  • Track record managing large-scale compute: budgeting, experiment design, ablations
  • Publications or demonstrated expertise in alignment, preference learning, or reward modeling
  • Hands-on implementation skills: PyTorch/JAX, distributed frameworks (DeepSpeed, FSDP, etc.)

You Will Enjoy The Following Benefits

  • Core Benefits: Fully paid medical and dental insurance premiums for employees and dependents, competitive 401k match, employer-paid life & disability insurance
  • Perks: Onsite gyms with laundry service, wellness activities, casual dress, snacks, game rooms
  • Learning: Tuition reimbursement, conference and training sponsorship
  • Time Off: Generous vacation and unlimited sick days, competitive paid caregiver leaves
  • Hybrid Work Policy: Flexible in-office days with budget for home office setup

The base pay for this role will be between $165,000 and $300,000. This role may also be eligible for other forms of compensation and benefits, such as a discretionary bonus, health, dental and other wellness plans and 401(k) contributions. Discretionary bonus can be a significant portion of total compensation. Actual compensation for successful candidates will be carefully determined based on a number of factors, including their skills, qualifications and experience.

We are proud to be an equal opportunity workplace. We do not discriminate based upon race, religion, color, national origin, sex, sexual orientation, gender identity/expression, age, status as a protected veteran, status as an individual with a disability, or any other applicable legally protected characteristics.

Two Sigma is committed to providing reasonable accommodations to qualified individuals in accordance with applicable federal, state, and local laws.

If you believe you need an accommodation, please visit our website for additional information.

Salary : $165,000 - $300,000

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Post-Training Research Scientist?

Sign up to receive alerts about other jobs on the Post-Training Research Scientist career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$100,407 - $125,193
Income Estimation: 
$120,989 - $162,093
Income Estimation: 
$74,806 - $91,633
Income Estimation: 
$71,928 - $87,026
Income Estimation: 
$145,337 - $174,569
Income Estimation: 
$100,407 - $125,193
Income Estimation: 
$120,989 - $162,093
Income Estimation: 
$74,806 - $91,633
Income Estimation: 
$71,928 - $87,026
Income Estimation: 
$145,337 - $174,569
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Two Sigma

  • Two Sigma York, NY
  • Position Summary Two Sigma is a leading quantitative investment management and trading firm. The company applies a scientific approach to investing, combin... more
  • 1 Day Ago

  • Two Sigma York, NY
  • Position Summary Two Sigma is a leading quantitative investment management and trading firm. The company applies a scientific approach to investing, combin... more
  • 1 Day Ago

  • Two Sigma York, NY
  • Position Summary Two Sigma is a leading quantitative investment management and trading firm. The company applies a scientific approach to investing, combin... more
  • 4 Days Ago

  • Two Sigma York, NY
  • Position Summary Two Sigma is a leading quantitative investment management and trading firm. The company applies a scientific approach to investing, combin... more
  • 4 Days Ago


Not the job you're looking for? Here are some other Post-Training Research Scientist jobs in the York, NY area that may be a better fit.

  • Lightning AI York, NY
  • Who We Are Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying A... more
  • 1 Day Ago

  • lightningai York, NY
  • Who We Are Lightning AI is the company behind PyTorch Lightning. Founded in 2019, we build an end-to-end platform for developing, training, and deploying A... more
  • 1 Month Ago

AI Assistant is available now!

Feel free to start your new journey!