Staff HPC Software Engineer

Sdstealthco
Location
San Diego, California
Job Type
Full-time
Posted
September 2, 2026
Views
1
Salary Range
$174k - $185k USD

Job Description

Location: San Diego, CA

Job Type: Full-Time

Salary: $174,000 - $185,000

Position Overview

We are looking for an experienced systems engineer to build and improve a measured, reliable compute architecture. The work spans high-throughput data ingestion, processing, storage, and delivery.

Role mission

Build and improve a high-throughput compute stack, so it is fast, observable, recoverable, and practical to operate in production environments.

This person will work with algorithms, platforms, and infrastructure engineers. They will not be expected to own every algorithm or infrastructure service. Their core responsibility is making the data and compute path reliable under real throughput, storage, network, and latency constraints.

Early work

In the first three to six months, this person should help:

    Establish reproducible hardware benchmarks for accelerated compute, CPU workloads, memory transfers, storage throughput, and network streaming.
    Build or harden stateful, multi-threaded pipelines that move data from ingestion through compute and output.
    Productize machine learning models, including neural networks, tree-based models, and unsupervised models, so they meet production requirements for performance, reliability, observability, and quality.
    Define backpressure, checkpointing, retry, and recovery behavior for disk pressure, slow consumers, and network outages.
    Compare alternative processing designs using wall-clock time, memory, storage, and quality measurements.
    Make the production interfaces and performance tests durable enough that later algorithm changes do not quietly break throughput or recovery behavior.

Required experience

    This role requires a PhD in Computer Science, Life Sciences, or a related discipline with 3+ years of relevant experience; a master's degree with 6+ years of relevant experience; or a bachelor's degree with 8+ years of relevant experience.
    Has contributed to a complex production software system with state machines, concurrency, and real compute or I/O bottlenecks. They do not need to have been the technical lead but must understand how these systems fail and how to debug them.
    Has shipped production-quality software in at least one of C++, Rust, CUDA, C, or C#. Comfortable with the normal engineering tools: profiling, tracing, debugging, testing, code review, builds, and CI.
    Can reason concretely about throughput, latency, buffering, memory, storage, network behavior, scheduling, contention, and failure recovery.
    Uses measurements to guide performance work: can identify a bottleneck, make a targeted change, quantify the gain, and add a regression guard.
    Has experience productizing machine learning models, including neural networks, tree-based models, or unsupervised models. Can make these models reliable, measurable, and efficient in a production system. This is not a model-research role.&...

Frequently Asked Questions

Where is the job located, and is it remote/hybrid/on-site?
The job is located in San Diego, California. The posting indicates it is a Full-Time position, but it does not specify a remote, hybrid, or on-site work-mode policy.
What is the salary range for this position?
The salary range for this Staff HPC Software Engineer role is $174,000 - $185,000.
What are the required educational qualifications and years of experience?
You need a PhD in Computer Science, Life Sciences, or a related discipline with 3+ years of experience; a master's degree with 6+ years of experience; or a bachelor's degree with 8+ years of experience.
What technical experience and programming languages are required?
You must have shipped production-quality software in C++, Rust, CUDA, C, or C#. You should have experience with complex production software systems featuring state machines, concurrency, and compute/IO bottlenecks, and be comfortable with profiling, tracing, debugging, testing, code review, builds, and CI.
What are the key responsibilities of this role?
You will build and improve a fast, observable, and reliable high-throughput compute stack. Responsibilities include establishing hardware benchmarks, building multi-threaded pipelines, productizing machine learning models, defining recovery behaviors, and comparing processing designs.

Ready to Apply?

Apply for this Position

You'll be redirected to the company's application page

Share this job:

Job Information

Source: greenhouse
AI Relevance: 75/100 (Relevant)
Remote Type: onsite
Allowed Locations: San Diego, California
Skills & Tags:
Primary Analysis

Get Similar Jobs by Email

Weekly digest of Sdstealthco and similar companies. Free.

Related Jobs

Apply for this Position

Get weekly job alerts