Demo

Staff Software Engineer, Data Curation

altoslabs
San Francisco, CA Full Time
POSTED ON 12/4/2025
AVAILABLE BEFORE 2/4/2026

Our Mission

Our mission is to restore cell health and resilience through cell rejuvenation to reverse disease, injury, and the disabilities that can occur throughout life.

For more information, see our website at altoslabs.com.

Our Value

Our Single Altos Value: Everyone Owns Achieving Our Inspiring Mission.

Diversity at Altos

We believe that diverse perspectives are foundational to scientific innovation and inquiry. At Altos, exceptional scientists and industry leaders from around the world work together to advance a shared mission. Our intentional focus is on Belonging, so that all employees know that they are valued for their unique perspectives. We are all accountable for sustaining a diverse and inclusive environment.

What You Will Contribute To Altos

Use AI agents to make complex research data FAIR—Findable, Accessible, Interoperable, Reusable—so scientists and product teams can ask richer questions, move faster, and advance discovery. Be part of a team using knowledge and data engineering to enable the transition from manual to LLM‑enabled, agentic data ingestion and curation.  You’ll sit at the intersection of data curation, data and knowledge engineering. Your job is to automate the ingestion and standardization of multi‑source datasets into governed, searchable, analytics-ready assets, and to model the domain knowledge that ties them together. 

Responsibilities

  • Curate and harmonize data. Ingest, profile, clean, normalize, and annotate multi‑modal research datasets (e.g., genomics/transcriptomics, proteomics, imaging/microscopy, CRISPR screens, assay/instrument metadata). Map to controlled vocabularies and standards; manage identifiers, synonyms, and crosswalks.
  • Deliver insights from curated data. Focus on the substance—entities, relationships, and annotations that answer real research and product questions using public domain assets from Ensembl, GEO, PubMed, OMIM, OLS, amongst others. Use pipelines and existing data sources storage pragmatically as tools to deliver content and outcomes.
  • Model knowledge to serve decisions. Capture the concepts and links researchers actually use; keep schemas lightweight and purpose‑built. Leverage OBO Foundry ontologies; define with LinkML; align to the BioLink/Biolink Model; and integrate/serve with platforms such as BioCypher.
  • Quality, governance & AI enablement. Instrument automated checks (tests/expectations), process development to improvement data FAIRification, and LLM‑assisted validations; capture provenance/lineage; codify SOPs; and work to facilitate the migration of processes from manual → automation → agentic (MCP‑integrated) workflows.
  • Serve as a key technical liaison between scientific, data science, and engineering teams, translating complex research needs into scalable and maintainable data solutions.
  • Define and evangelize best practices for data and knowledge engineering across the organization, mentoring junior team members and building reusable, AI-enhanced, enterprise-level components.

Who You Are

Minimum Qualifications

  • PhD, Biological Sciences, Computer Science, Software Engineering, or related quantitative field, or equivalent technical experience
  • Candidates should have 8 years of relevant experience in data curation, ontology/knowledge engineering, or data engineering (or equivalent experience) at a biotechnology company.
  • Mindset: You prioritize data and business objectives over tools; technology is a means to an end.
  • Demonstrably strong Python expertise, particularly in the context of data modeling and processing, with strong skills in both relational (SQL) and graph data stores, and the ability to choose pragmatically between them (e.g., Postgres/Redshift vs. Neo4j/Neptune).
  • Comfortable building pragmatic ETL/ELT workflows in a major cloud (preferably AWS), using orchestration frameworks or AWS-native tools.
  • Active user of AI coding editors such as Cursor, with an active interest in designing and building Model Context Protocol (MCP) applications; motivated to migrate processes from manual → automation → agentic.
  • Mature understanding of data quality, provenance, versioning, and “curation as code,” including hands-on use of testing/validation frameworks.

Preferred Qualifications

  • Experience in basic/exploratory life‑science research across multiple modalities (genomics/transcriptomics, proteomics, imaging/microscopy, screening, model organisms); a user of curated content to achieve research/business outcomes.
  • Experience with a data platform such as lamin.ai.
  • Experience with vector databases and search (e.g., Weaviate, FAISS, pgvector) and AI/LLM frameworks (e.g., LiteLLM, LangChain, LlamaIndex) for retrieval-augmented generation and agent workflows.
  • Experience with OBO Foundry ontologies and modern frameworks such as LinkML, BioLink, and BioCypher, familiarity with graph database technologies (e.g., Neo4j, AWS Neptune) and semantic standards (OWL, RDF, SPARQL).
  • Experience creating lightweight semantic layers and AI/LLM‑assisted curation workflows (LiteLLM, FastMCP).

The salary range for Redwood City, CA:

  • Staff Software Engineer: $221,850 - $300,150

Exact compensation may vary based on skills, experience, and location.

 

For UK applicants, before submitting your application:

- Please click here to read the Altos Labs EU and UK Applicant Privacy Notice (bit.ly/eu_uk_privacy_notice)
- This Privacy Notice is not a contract, express or implied and it does not set terms or conditions of employment.

Equal Opportunity Employment

We value collaboration and scientific excellence.

We believe that diverse perspectives and a culture of belonging are foundational to scientific innovation and inquiry. At Altos Labs, exceptional scientists and industry leaders from around the world work together to advance a shared mission. Our intentional focus is on Belonging, so that all employees know that they are valued for their unique perspectives. We are all accountable for sustaining an inclusive environment.

Altos Labs provides equal employment opportunities to all employees and applicants for employment, without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state or local laws. Altos prohibits unlawful discrimination and harassment. This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation and training.

Thank you for your interest in Altos Labs where we strive for a culture of scientific excellence, learning, and belonging.

Note: Altos Labs will not ask you to download a messaging app for an interview or outlay your own money to get started as an employee. If this sounds like your interaction with people claiming to be with Altos, it is not legitimate and has nothing to do with Altos. Learn more about a common job scam at https://www.linkedin.com/pulse/how-spot-avoid-online-job-scams-biron-clark/

Salary : $222 - $300

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Staff Software Engineer, Data Curation?

Sign up to receive alerts about other jobs on the Staff Software Engineer, Data Curation career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$122,257 - $154,284
Income Estimation: 
$143,391 - $179,890
Income Estimation: 
$176,149 - $220,529
Income Estimation: 
$156,679 - $196,968
Income Estimation: 
$77,657 - $95,021
Income Estimation: 
$97,257 - $120,701
Income Estimation: 
$97,257 - $120,701
Income Estimation: 
$123,167 - $152,295
Income Estimation: 
$146,673 - $180,130
Income Estimation: 
$176,149 - $220,529
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at altoslabs

  • altoslabs San Diego, CA
  • Our Mission Our mission is to restore cell health and resilience through cell rejuvenation to reverse disease, injury, and the disabilities that can occur ... more
  • 13 Days Ago

  • altoslabs San Francisco, CA
  • Our Mission Our mission is to restore cell health and resilience through cell rejuvenation to reverse disease, injury, and the disabilities that can occur ... more
  • 4 Days Ago


Not the job you're looking for? Here are some other Staff Software Engineer, Data Curation jobs in the San Francisco, CA area that may be a better fit.

  • Pinterest San Francisco, CA
  • About Pinterest Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that w... more
  • 14 Days Ago

  • RAPIDFORT San Francisco, CA
  • We are looking for a skilled DevOps Software Engineer to join our team and play a key role in building, maintaining, and optimizing curated container image... more
  • 18 Days Ago

AI Assistant is available now!

Feel free to start your new journey!