Demo

Performance Engineer

Aziro Technologies LLC
Santa Clara, CA Full Time
POSTED ON 8/4/2026
AVAILABLE BEFORE 9/4/2026

What you'll do

Set the performance architecture agenda. Bring deep, current expertise across file, block, and object storage protocols and translate it into concrete performance requirements for //EXA's data path (NFSv3 direct-to-DN, pNFS layouts, S3 via MDN) informed by how HPC and hyperscale environments actually push storage systems (checkpointing, small-file metadata storms, GPU-starved read patterns, mixed-tenant burst I/O).

Track and act on the NeoCloud / sovereign-cloud shift. Maintain a living view of where //EXA's highest-value deployments are heading GPU-cloud and sovereign-cloud operators (CoreWeave, Crusoe, Nscale, and similar) and make sure //EXA's performance roadmap, reference architectures, and sizing guidance map to how these operators actually buy and operate infrastructure (multi-tenant GPU clusters, bursty training/inference mixes, strict SLAs to end customers).

Own competitive performance positioning. Build and maintain deep, technically substantiated comparisons against VAST Data, DDN, and WEKA not marketing bullet points, but real architectural analysis (metadata scaling model, erasure coding/durability tradeoffs, protocol support, GPU-direct paths, cost/performance at scale) that engineering and field teams can use to win technical evaluations and POCs.

Drive performance tuning for multi-tenant HPC/AI workloads. Lead tuning and validation work spanning the full stack a GPU cluster touches storage (MDN/DN geometry, pack groups, erasure coding layout), networking (RDMA, RoCE/InfiniBand fabric behavior, NIC/queue tuning), and compute (GPU-side I/O patterns, checkpoint/restore, data loader behavior) with particular focus on how these interact when multiple tenants/workloads share the same //EXA fleet.

Build and run the benchmark suite. Own //EXA's benchmark framework and result credibility: MLPerf Storage (v2/v3), elbencho, IO500, fio/vdbench-class synthetic tests, and workload-representative benchmarks for AI training/inference and traditional HPC. Ensure results are reproducible, defensible in public disclosure, and directly comparable to published competitor numbers.

Define QoS, limits, and workload segmentation. Drive the technical requirements and validation for quality-of-service guarantees, per-tenant/per-workload throughput and IOPS limits, and workload isolation the mechanisms that let //EXA make hard SLA commitments in shared, multi-tenant NeoCloud deployments rather than best-effort performance.

What makes you competitive for this role

Deep, hands-on background in storage performance engineering across file, block, and object protocols, ideally with direct HPC or hyperscale exposure (parallel filesystems, pNFS/NFS at scale, S3-scale object stores).

Working knowledge of GPU cluster architecture RDMA fabrics, GPUDirect Storage, checkpoint/restore patterns for large model training and how storage bottlenecks manifest in mixed compute/network/storage systems.

Fluency with industry benchmark standards (MLPerf Storage, IO500) and load-generation tooling (elbencho, fio, vdbench), plus the judgment to design workload-representative tests beyond canned benchmarks.

Demonstrated ability to build rigorous, technically credible competitive analysis (not slideware) against systems like VAST, DDN, and WEKA architecture-level understanding, not just spec-sheet comparison.

Experience with multi-tenant resource management concepts (QoS, rate limiting, workload isolation) in a distributed systems context.

Comfortable operating across the stack and across audiences deep enough to debug an RDMA queue-pair stall or a metadata hot-partition, articulate enough to brief a NeoCloud customer's technical evaluation team.

Why this role matters

//EXA's disaggregated design (independent MDN/DN scaling, direct-to-data-node pNFS) is a real architectural bet against how VAST, DDN, and WEKA scale metadata and data. Whether that bet wins in the market depends on whether performance claims hold up under the exact multi-tenant, GPU-bound, bursty workloads that NeoCloud and sovereign-cloud operators run and whether we can prove it with numbers that stand up to public scrutiny. This role is the person who makes sure both are true.

Salary : $100,000 - $140,000

If your compensation planning software is too rigid to deploy winning incentive strategies, it’s time to find an adaptable solution. Compensation Planning
Enhance your organization's compensation strategy with salary data sets that HR and team managers can use to pay your staff right. Surveys & Data Sets

What is the career path for a Performance Engineer?

Sign up to receive alerts about other jobs on the Performance Engineer career path by checking the boxes next to the positions that interest you.
Income Estimation: 
$97,257 - $120,701
Income Estimation: 
$123,167 - $152,295
Income Estimation: 
$104,606 - $124,147
Income Estimation: 
$111,859 - $131,446
Income Estimation: 
$110,457 - $133,106
Income Estimation: 
$105,809 - $128,724
Income Estimation: 
$122,763 - $145,698
Employees: Get a Salary Increase
View Core, Job Family, and Industry Job Skills and Competency Data for more than 15,000 Job Titles Skills Library

Job openings at Aziro Technologies LLC

  • Aziro Technologies LLC Lansing, MI
  • References are required for this position. Please include a separate attachment with your submission that includes 2-3 professional references. Each refere... more
  • 6 Days Ago

  • Aziro Technologies LLC Santa Clara, CA
  • Job Title : Storage Benchmarking Engineer Location : Santa Clara, CA Full-Time On-Site The Role The Storage Benchmarking Engineer will design, execute, and... more
  • 6 Days Ago

  • Aziro Technologies LLC Santa Clara, CA
  • Job Title: Burn-in Tool & Support Tools Preferred Location: Highly desired in Santa Clara , CA , or Costa Rica (Lower Cost Center Preferred) Duration: Full... more
  • 6 Days Ago

  • Aziro Technologies LLC Santa Clara, CA
  • Job Title: System Test Engineer Preferred Location: Highly desired in Santa Clara , CA , or Costa Rica (Lower Cost Center Preferred) Duration: Long term Co... more
  • 6 Days Ago


Not the job you're looking for? Here are some other Performance Engineer jobs in the Santa Clara, CA area that may be a better fit.

  • Hippocratic AI Menlo Park, CA
  • About Us Hippocratic AI is the leading generative AI company in healthcare. We have the only system that can have safe, autonomous, clinical conversations ... more
  • 1 Month Ago

  • asteralabs San Jose, CA
  • Astera Labs (NASDAQ: ALAB) provides rack-scale AI infrastructure through purpose-built connectivity solutions. By collaborating with hyperscalers and ecosy... more
  • 16 Days Ago

AI Assistant is available now!

Feel free to start your new journey!