Demo

Software Engineer – Presto / Spark Development

IBM
IBM Salary
San Jose, CA Full Time
POSTED ON 7/17/2026
AVAILABLE BEFORE 8/24/2026
Introduction

At IBM Software, we transform client challenges into solutions, building the world's leading AI-powered, cloud-native products that shape the future of business and society. We are building the next generation of watsonx.data—a GPU-accelerated, open data lakehouse engineered to deliver category-leading price-performance for analytics and AI workloads. Working in Software means joining a team fueled by curiosity and collaboration, where you'll work deep inside Presto and Spark to build connectors, custom UDFs/UDAFs, optimizer rules, and performance-critical execution paths that shape query throughput and latency for customers at petabyte scale. With a culture that values innovation, growth, and continuous learning, IBM Software places you at the heart of IBM's product and technology landscape. Here, you'll have the tools and opportunities to advance your career while creating software that changes the world.

Your Role And Responsibilities

As a Software Engineer with deep Presto and/or Spark internals expertise, you will design, develop, test, and deliver connectors, optimizer extensions, and engine-level performance improvements that power the watsonx.data query layer. You will work in an Agile, collaborative environment to understand stakeholder requirements and directly impact query throughput, latency, and reliability at petabyte scale. Your primary responsibilities will include:

  • Develop Engine Internals: Build and maintain Presto and/or Spark connectors, operator implementations, optimizer rules, and custom UDFs/UDAFs for open table formats (Iceberg, Delta, Hudi).
  • Optimize Performance at Scale: Diagnose and resolve data skew, broadcast-join sizing, shuffle bottlenecks, and memory pressure; tune operator memory, spill thresholds, and off-heap usage using async-profiler and flamegraphs.
  • Contribute to CI/CD & Benchmarks: Contribute to the automated CI/CD pipeline and maintain CI benchmarks that guard against regressions in query latency, throughput, and resource consumption.
  • Support Production & Debug: Support Presto/Spark deployments on Kubernetes and bare metal, unit-test fixes for engine-related and customer-reported issues, and participate in on-call.
  • Collaborate in Agile Environment: Partner with query optimization, storage, GPU acceleration, and AI/ML teams, conduct reviews with measurable acceptance criteria, and document connector interfaces and engine internals.

Preferred Education

Bachelor's Degree

Required Technical And Professional Expertise

  • Engine Development Experience: 6 years of professional software engineering, including at least 2 years developing against Presto/Trino or Apache Spark internals.
  • Language & Codebase Proficiency: Strong Java or Scala skills with comfort navigating and modifying a large, complex open-source codebase.
  • Connectors & Optimizer Work: Hands-on experience building connectors, UDFs/UDAFs, or optimizer extensions, plus working knowledge of Spark/Presto query planning, physical execution, and the operator/stage model.
  • Performance & Formats: Experience resolving data skew, shuffle bottlenecks, broadcast-join sizing, and memory pressure at scale; familiarity with an open table format (Iceberg, Delta, or Hudi); JVM tuning in production.
  • Communication & Education: Clear written communication—able to file actionable bugs, write design docs, and explain engine trade-offs; Bachelor's degree in Computer Science, Engineering, or equivalent practical experience.

Preferred Technical And Professional Experience

  • OSS & GPU Acceleration: Committer or significant contributor to Apache Spark, Trino, or Presto, and experience integrating GPU-accelerated execution (RAPIDS Accelerator, cuDF) into query paths.
  • Vectorization & Multi-Tenancy: Familiarity with vectorized execution and columnar formats (Arrow, ORC, Parquet), ML feature and inference pipelines on Spark/Presto, FinOps cost modeling, and multi-tenant deployments with fairness scheduling and workload isolation.

Salary : $131,000 - $245,000

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Software Engineer – Presto / Spark Development?

Sign up to receive alerts about other jobs on the Software Engineer – Presto / Spark Development career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$77,900 - $95,589
Income Estimation: 
$101,387 - $124,118
Income Estimation: 
$176,149 - $220,529
Income Estimation: 
$156,679 - $196,968
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at IBM

  • IBM Junction, VT
  • Introduction A career in IBM Consulting is built on long-term client relationships and close collaboration worldwide. You’ll work with leading companies ac... more
  • 12 Days Ago

  • IBM Washington, DC
  • Introduction A career in IBM Consulting is built on long-term client relationships and close collaboration worldwide. You’ll work with leading companies ac... more
  • 12 Days Ago

  • IBM Annapolis, MD
  • Introduction At IBM, the work is about building things that matter. You will be working on real problems in a mission environment where the output of your ... more
  • 12 Days Ago

  • IBM Bellevue, WA
  • Entry Level Technical Support Engineer - Cloudability Bellevue, Washington, United States Software Product Development Entry Level Introduction We are seek... more
  • 12 Days Ago


Not the job you're looking for? Here are some other Software Engineer – Presto / Spark Development jobs in the San Jose, CA area that may be a better fit.

  • Cloudera San Jose, CA
  • Business Area: Engineering Seniority Level: Director Job Description: At Cloudera, we empower people to transform complex data into clear and actionable in... more
  • 13 Days Ago

  • Amazon Palo Alto, CA
  • Description The Open Data analytics team is looking for an experienced engineer to join the Spark engines team. EMR, Glue ETL and Athena are services that ... more
  • 15 Days Ago

AI Assistant is available now!

Feel free to start your new journey!