374 hpc jobs at 146 companies in California

2mo
Save
Mark Applied
Hide
HPC Engineer
Sunnyvale, California, United States
$150k-$300k/yr OnsiteFull Time
Institute of Foundation Models
Institute of Foundation Models: Develops open-source frontier-class AI foundation models and research.
2+ YOE2+ years in Linux systems administration, SRE, DevOps, cloud operations, HPC or infrastructure operations; strong Linux troubleshooting; scripting in Python or Bash; Bachelor’s in a related field.
Slurm, GPU infrastructure, AWS, Azure, GCP, Grafana, Prometheus, Datadog, Containers, Kubernetes, Python, Bash, Linux
2mo
Save
Mark Applied
Hide
HPC Systems Engineer
San Francisco, California, United States
$120k-$196k/yr HybridFull Time
University of California, San Francisco
University of California, San Francisco: Public research university dedicated to health sciences and education.
6+ YOEBachelor's in CS/engineering plus 6+ years HPC experience (or 10+ yrs related); expert HPC infrastructure design; experience with parallel filesystems (GPFS, Lustre, Vast, DDN); security to NIST/HIPAA standards; SLURM/PBS, Warewulf/InfiniBand, automated testing, advanced scripting, and technical documentation skills.
GPFS, Lustre, Vast, DDN, SLURM, PBS, Warewulf, InfiniBand
2mo
Save
Mark Applied
Hide
Staff HPC Engineer
Milpitas, California, United States
$163k-$285k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
8+ YOEExtensive Linux systems engineering in large-scale compute; HPC schedulers (Slurm), MPI, GPUs; scripting and automation; troubleshooting; collaboration.
Slurm, PBS, LSF, Lustre, GPFS, BeeGFS, Ansible, Terraform, Kubernetes, Prometheus, Grafana, ELK, Docker, Singularity
2mo
Save
Mark Applied
Hide
HPC Hardware Engineer
Milpitas, California, United States
$95k-$162k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management solutions for semiconductors.
0+ YOEMaster's degree (0 years) or Bachelor's + 2 years; foundational computer architecture and Linux knowledge; experience or exposure to HPC, distributed systems, or server platforms; Python/Bash scripting; strong problem-solving and teamwork skills.
Linux, Python, Bash
2w
Save
Mark Applied
Hide
Senior HPC Systems Architect
San Jose or San Francisco
$255k-$340k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
8+ YOE8+ years designing large-scale HPC infrastructures, expertise with GPU clusters, high-speed networking, liquid cooling, benchmarking, capacity planning, and architecture documentation.
InfiniBand, Ethernet, Ansible, Terraform, Kubernetes
2mo
Save
Mark Applied
Hide
Senior HPC Storage Engineer
Santa Clara or Austin
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOE8+ years designing or operating large-scale storage infrastructure; Linux, Python and bash proficiency; experience with containers, distributed filesystems, storage performance tuning, and HPC/AI workloads.
Centos/RHEL, Ubuntu Linux, Python, bash, Docker, Enroot, Ceph, Weka.io, Vast, Lustre, GPFS, NVIDIA GPUs, CUDA, NCCL, MLPerf, Network Appliance, PyTorch, TensorFlow
2mo
Save
Mark Applied
Hide
HPC Scientific Support Engineer
Berkeley, California, United States
$139k-$268k/yr HybridFull Time
Lawrence Berkeley National Laboratory
Lawrence Berkeley National Laboratory: Conducts multidisciplinary scientific research for the U.S. Department of Energy.
8+ YOE8+ years related experience with a Bachelor's (CSE3) or 12+ years (CSE4); 2+ years using HPC systems; deep HPC expertise (Linux/Unix, Fortran, C/C++, MPI, OpenMP, CUDA, Python, containers, debuggers, performance tools); strong communication and teaching skills.
Linux/Unix, Fortran, C/C++, MPI, OpenMP, OpenACC, CUDA, shell scripting, Python, parallel algorithms, programming models, AI models, debuggers, performance tools, containers, REST, JavaScript, SQL
1mo
Save
Mark Applied
Hide
Staff HPC Network Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building specialized industrial robots and physical AI systems.
Experience designing and scaling high-performance networks for HPC/GPU compute, hands-on switch management (Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC), packet analysis, fiber optics knowledge, cloud networking, config management, automation scripting, and monitoring tools.
Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC, tcpdump, Wireshark, AWS, GCP, Azure, Ansible, Salt, Python, Prometheus, Grafana, ELK, GitHub
1mo
Save
Mark Applied
Hide
Staff HPC Network Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building robotics and software to automate physical world industries.
Experience designing and scaling networks for HPC/GPU compute; hands-on with Arista EOS, Cisco NX-OS, Nvidia Cumulus, and/or SONiC; packet analysis (tcpdump, Wireshark); fiber optics to 800 Gbps; familiarity with AWS/GCP/Azure, Ansible/Salt, Python, Prometheus, Grafana, ELK, and GitHub.
Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC, tcpdump, Wireshark, AWS, GCP, Azure, Ansible, Salt, Python, Prometheus, Grafana, ELK, GitHub
1mo
Save
Mark Applied
Hide
HPC/AI Performance Specialist
Berkeley or California or United States
$139k-$268k/yr HybridFull Time
Lawrence Berkeley National Laboratory
Lawrence Berkeley National Laboratory: Conducts scientific research to address global challenges.
8+ YOEAdvanced experience in HPC/AI performance engineering, data management, storage and I/O, scientific software development, and collaboration with domain scientists; typically 8+ years (CSE3) or 12+ years (CSE4) of related experience with relevant degree.
NVIDIA Vera-Rubin, VAST file system, Doudna
2mo
Save
Mark Applied
Hide
HPC Systems Engineer, Modeling & Simulation
Costa Mesa, California, United States
$132k-$198k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
5+ YOE5+ years administering Linux/Unix in HPC or scientific computing; hands-on cluster, storage, and scientific code support; Bash/Python scripting; experience with MPI/OpenMP, CMake, compilers, filesystems, and secure computing; bachelor's degree in a technical field; ability to obtain US security clearance.
Linux, Unix, Windows, Bash, Python, CTH, ALE3D, Sierra, ParaView, Cubit, VisIt, CMake, compilers, MPI, OpenMPI, OpenMP, NFS, NAS, TCP/IP, Slurm
1mo
Save
Mark Applied
Hide
HPC Systems Administrator (Hardware & Infrastructure Operations)
Stanford, California, United States
$150k-$172k/yr OnsiteFull Time
Stanford University
Stanford University: A private research university providing higher education and research.
3+ YOEManage and maintain large-scale HPC hardware and infrastructure, perform diagnostics and root-cause analysis, collaborate with data center teams, automate provisioning and monitoring; bachelor's degree plus experience required.
Linux, Slurm, Lustre, InfiniBand, DCIM, NVIDIA H200, x86
1w
Save
Mark Applied
Hide
Director, HPC Systems Software Engineering
Houston or New York City or San Francisco or Seattle
$230k-$343k/yr OnsiteFull Time
Nscale
Nscale: Vertically integrated AI infrastructure provider for high-performance computing.
Requires leadership of HPC or infrastructure engineering organizations, Linux, distributed systems, production operations, GPU infrastructure, networking, budgeting, and people management; Slurm experience strongly preferred.
Slurm, Kueue, Linux
4w
Save
Mark Applied
Hide
HPC Performance Engineer
Oregon or Texas or California or Massachusetts
$152k-$242k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years programming experience; BS/MS or equivalent in CS or related field; strong Fortran/C/C++ and parallel programming knowledge; experience with OpenACC, OpenMP, MPI, CUDA; strong performance analysis and communication skills.
Fortran, C, C++, OpenACC, OpenMP, MPI, CUDA, NVHPC, assembly
3w
Save
Mark Applied
Hide
Senior HPC Storage Engineer
Los Angeles, California, United States
HybridFull Time
University of California, Los Angeles
University of California, Los Angeles: A public research university providing higher education and research.
7+ YOE7+ years managing large-scale production storage for research or hyperscale environments; deep knowledge of scale-out, parallel, distributed, object, and federated storage; hands-on experience with Lustre, VAST Data, GPFS/Spectrum Scale, Ceph, BeeGFS, WekaFS, or MinIO; advanced Linux administration; Bash/Python scripting; Ansible and Git; identity ...
Lustre, VAST Data, GPFS/Spectrum Scale, Ceph, BeeGFS, WekaFS, MinIO, InfiniBand, RoCE, Bash, Python, Ansible, Git, Active Directory, LDAP, CILogon, OIDC, SAML, Globus Auth
2mo
Save
Mark Applied
Hide
Flight Sciences Tools and HPC engineer
San Jose, California, United States
$163k-$218k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
5+ YOE5+ years in scientific/engineering software for flight sciences; strong Python; HPC design/administration; CI/CD; Linux; MPI; Slurm/PBS; OpenHPC; containerization; IT security.
Python (NumPy SciPy Pandas Scikit-learn TensorFlow PyTorch), VTK, Slurm, PBS, Torque, OpenHPC, Bright, Warewulf, XCat, Spack, EasyBuild, Lmod, Linux, MPI, AWS, Docker, Singularity, Shifter, Chef, Puppet, Ansible, Jenkins, TeamCity, Artifactory, CI/CD tooling, SQL, NoSQL, time-series databases, MATLAB, Simulink, C/C++, Fortran, Rust, CAD, CFD, GNC, aerodynamics, loads, structural analysis, mass properties
1mo
Save
Mark Applied
Hide
Sr. High Performance Computing (HPC) Systems Engineer
Brownsville or Hawthorne
OnsiteFull Time
SpaceX
SpaceX: Designs and launches advanced rockets and satellite internet constellations.
5+ YOE5+ years systems engineering experience, hands-on Linux and HPC cluster administration, Kubernetes, scripting (Bash/Python), networking, and security; eligible for TS/SCI with polygraph.
Kubernetes, Linux, Bash, Python, Slurm, PBS, LSF, Prometheus, Grafana, Nagios, CUDA, Docker, Podman, Singularity, Puppet, Ansible
2mo
Save
Mark Applied
Hide
HPC Systems Engineer (Linux / Infrastructure)
San Diego, California, United States
$116k-$209k/yr OnsiteFull Time
General Atomics
General Atomics: Designs and manufactures unmanned aircraft and nuclear technology systems.
15+ YOEBachelor's degree or equivalent experience, 15+ years systems administration experience, deep Linux stack knowledge (NFS, ZFS, BTRFS, mdadm, LVM), scripting with Python/Ansible/Bash, troubleshooting distributed systems; SLURM and parallel file system experience desirable.
Lustre, BeeGFS, WEKA, Ceph, SLURM, Ansible, Python, Bash, Docker, Podman, Apptainer, NFS, ZFS, BTRFS, mdadm, LVM
2d
Save
Mark Applied
Hide
AI & HPC Infrastructure Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irvine or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
5+ YOERequires 5+ years designing AI infrastructure, accelerated computing, clusters, orchestration, and automation; bachelor's degree or equivalent experience. Strong Kubernetes, Python, Terraform, and cloud expertise required.
Slurm, Run:ai, Kubernetes, NVIDIA Base Command Manager (BCM), NVIDIA NGC, NCCL, NVLink, CUDA, TensorRT-LLM, vLLM, SGLang, Triton Inference Server, NVIDIA Dynamo, llm-d, MLPerf, NCCL tests, fio, iperf, InfiniBand, Ethernet, SONiC, NVMe, NVMe-oF, VAST, Weka, DDN, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, Ansible, REST, OpenAPI, JSON, YAML, TensorFlow, PyTorch, JAX, Jupyter notebooks, Google Colab, MCP
1mo
Save
Mark Applied
Hide
Senior Engineer – AI & HPC Observability
Austin or San Diego
$155k-$206k/yr OnsiteFull Time
Cirrascale
Cirrascale: Provides specialized GPU-based cloud infrastructure for AI workloads.
5+ YOEBachelors in CS/CE or equivalent; 5+ years building distributed systems and modern observability tooling (Open Telemetry, Prometheus, Grafana, Datadog, ELK, etc.); 1+ year HPE OpsRamp; Bash and Python; cloud (AWS/GCP/OpenStack) and k8s; on-call participation.
HPE OpsRamp, Open Telemetry, Prometheus, Grafana, Nagios, Datadog, ELK, Thousand Eyes, Bash, Python, AWS, GCP, OpenStack, Proxmox, k8s, Netbox, MaaS, Redfish