235 hpc engineer jobs at 81 companies in Mountain View, CA

1mo
Save
Mark Applied
Hide
HPC Engineer
Sunnyvale, California, United States
$150k-$300k/yr OnsiteFull Time
Institute of Foundation Models
Institute of Foundation Models: Develops open-source frontier-class AI foundation models and research.
2+ YOE2+ years in Linux systems administration, SRE, DevOps, cloud operations, HPC or infrastructure operations; strong Linux troubleshooting; scripting in Python or Bash; Bachelor’s in a related field.
Slurm, GPU infrastructure, AWS, Azure, GCP, Grafana, Prometheus, Datadog, Containers, Kubernetes, Python, Bash, Linux
1mo
Save
Mark Applied
Hide
Staff HPC Engineer
Milpitas, California, United States
$163k-$285k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
8+ YOEExtensive Linux systems engineering in large-scale compute; HPC schedulers (Slurm), MPI, GPUs; scripting and automation; troubleshooting; collaboration.
Slurm, PBS, LSF, Lustre, GPFS, BeeGFS, Ansible, Terraform, Kubernetes, Prometheus, Grafana, ELK, Docker, Singularity
2mo
Save
Mark Applied
Hide
Staff HPC Engineer
San Francisco, California, United States
$214k-$300k/yr HybridFull Time
Chan Zuckerberg Biohub
Chan Zuckerberg Biohub: Builds AI-powered technologies for biomedical research and disease study.
10+ YOE10+ years HPC infra, hybrid on-prem/cloud, GPU AI workloads; Slurm, Kubernetes; PyTorch/TensorFlow/JAX; MLOps; strong collaboration and leadership.
Slurm, Kubernetes, SUNK, Docker, Singularity, Terraform, Ansible, Python, Bash, Git, Horovod, DeepSpeed, Ray, PyTorch, TensorFlow, JAX, RAPIDS, Coreweave, AWS, GCP
1mo
Save
Mark Applied
Hide
HPC Hardware Engineer
Milpitas, California, United States
$95k-$162k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management solutions for semiconductors.
0+ YOEMaster's degree (0 years) or Bachelor's + 2 years; foundational computer architecture and Linux knowledge; experience or exposure to HPC, distributed systems, or server platforms; Python/Bash scripting; strong problem-solving and teamwork skills.
Linux, Python, Bash
3mo
Save
Mark Applied
Hide
HPC Operations Engineer
Santa Clara, California, United States
$124k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
2+ YOEAdminister CentOS/RHEL, Docker, Python and bash; experience with cluster config management (Ansible), troubleshooting large-scale HPC systems, BS in Computer Science or equivalent with 2+ years post-degree experience, strong problem-solving and communication.
CentOS, RHEL, Docker, Python, bash, Ansible, IBM Spectrum LSF, SLURM, FlexLM, Perl, InfiniBand, RDMA, RoCE, Lustre, GPFS
1mo
Save
Mark Applied
Hide
HPC Systems Engineer
San Francisco, California, United States
$120k-$196k/yr HybridFull Time
University of California, San Francisco
University of California, San Francisco: Public research university dedicated to health sciences and education.
6+ YOEBachelor's in CS/engineering plus 6+ years HPC experience (or 10+ yrs related); expert HPC infrastructure design; experience with parallel filesystems (GPFS, Lustre, Vast, DDN); security to NIST/HIPAA standards; SLURM/PBS, Warewulf/InfiniBand, automated testing, advanced scripting, and technical documentation skills.
GPFS, Lustre, Vast, DDN, SLURM, PBS, Warewulf, InfiniBand
4w
Save
Mark Applied
Hide
Staff HPC Network Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building specialized industrial robots and physical AI systems.
Experience designing and scaling high-performance networks for HPC/GPU compute, hands-on switch management (Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC), packet analysis, fiber optics knowledge, cloud networking, config management, automation scripting, and monitoring tools.
Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC, tcpdump, Wireshark, AWS, GCP, Azure, Ansible, Salt, Python, Prometheus, Grafana, ELK, GitHub
4w
Save
Mark Applied
Hide
Staff HPC Network Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building robotics and software to automate physical world industries.
Experience designing and scaling networks for HPC/GPU compute; hands-on with Arista EOS, Cisco NX-OS, Nvidia Cumulus, and/or SONiC; packet analysis (tcpdump, Wireshark); fiber optics to 800 Gbps; familiarity with AWS/GCP/Azure, Ansible/Salt, Python, Prometheus, Grafana, ELK, and GitHub.
Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC, tcpdump, Wireshark, AWS, GCP, Azure, Ansible, Salt, Python, Prometheus, Grafana, ELK, GitHub
1mo
Save
Mark Applied
Hide
HPC Systems Engineer
San Francisco, California, United States
HybridFull Time
UCSF Health
UCSF Health: Academic medical center providing advanced patient care and research.
6+ YOEBachelor's in a related field and 6+ years experience with large-scale/HPC systems (or 10+ years related). Expert HPC infrastructure, parallel filesystems, security controls (NIST/HIPPA), SLURM/PBS, Warewulf/InfiniBand, EM automation and advanced scripting.
GPFS, Lustre, Vast, DDN, NIST 800-171, NIST 800-223, HIPPA, SLURM, PBS, Warewulf, InfiniBand
2mo
Save
Mark Applied
Hide
HPC & AI Senior Performance Engineer
Bloomington or Spring or San Jose
$136k-$275k/yr RemoteFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Provides global edge-to-cloud technology solutions and IT infrastructure services.
7+ YOEMasters or PhD in CS/engineering/math; 7+ years HPC performance analysis; strong analytical skills; excellent communication; U.S. citizenship.
OpenMP, MPI, CUDA, HIP, NCCL/RCCL, HPL, HPCG, OSU MPI, CPU/GPU benchmarking
3mo
Save
Mark Applied
Hide
HPC Operations Engineer
Santa Clara, California, United States
$124k-$242k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
2+ YOEBS in Computer Science or equivalent; 2+ years experience; CentOS/RHEL, Docker, Python, Bash; strong problem solving & communication; Ansible; HPC/cluster familiarity.
Docker, Ansible, Python, Bash, CentOS/RHEL, Linux, LSF, SLURM, NFS, LDAP, DNS, TCP/IP, Perl, InfiniBand, Lustre, GPFS
3d
Save
Mark Applied
Hide
Senior HPC Platform Hardware Engineer
San Jose, California, United States
$255k-$340k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
5+ YOE5+ years technical lead experience in hardware NPI for HPC/data center, hands-on lab experience, PLM/BOM familiarity, expertise in compute/storage/network hardware, and cross-functional collaboration.
PLM, BMC, BIOS
1mo
Save
Mark Applied
Hide
Flight Sciences Tools and HPC engineer
San Jose, California, United States
$163k-$218k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
5+ YOE5+ years in scientific/engineering software for flight sciences; strong Python; HPC design/administration; CI/CD; Linux; MPI; Slurm/PBS; OpenHPC; containerization; IT security.
Python (NumPy SciPy Pandas Scikit-learn TensorFlow PyTorch), VTK, Slurm, PBS, Torque, OpenHPC, Bright, Warewulf, XCat, Spack, EasyBuild, Lmod, Linux, MPI, AWS, Docker, Singularity, Shifter, Chef, Puppet, Ansible, Jenkins, TeamCity, Artifactory, CI/CD tooling, SQL, NoSQL, time-series databases, MATLAB, Simulink, C/C++, Fortran, Rust, CAD, CFD, GNC, aerodynamics, loads, structural analysis, mass properties
1mo
Save
Mark Applied
Hide
HPC Scientific Support Engineer
Berkeley, California, United States
$139k-$268k/yr HybridFull Time
Lawrence Berkeley National Laboratory
Lawrence Berkeley National Laboratory: Conducts multidisciplinary scientific research for the U.S. Department of Energy.
8+ YOE8+ years related experience with a Bachelor's (CSE3) or 12+ years (CSE4); 2+ years using HPC systems; deep HPC expertise (Linux/Unix, Fortran, C/C++, MPI, OpenMP, CUDA, Python, containers, debuggers, performance tools); strong communication and teaching skills.
Linux/Unix, Fortran, C/C++, MPI, OpenMP, OpenACC, CUDA, shell scripting, Python, parallel algorithms, programming models, AI models, debuggers, performance tools, containers, REST, JavaScript, SQL
4w
Save
Mark Applied
Hide
HPC Scientific Support Engineer
Berkeley, California, United States
$139k-$268k/yr HybridFull Time
Lawrence Berkeley National Laboratory
Lawrence Berkeley National Laboratory: Conducts scientific research to address global challenges.
8+ YOE8+ years related experience with a Bachelor's (CSE3) or 12+ years (CSE4); 2+ years using HPC systems; expertise in high-performance scientific computing including at least four of Linux/Unix, Fortran, C/C++, MPI, OpenMP, OpenACC, CUDA, shell scripting, Python, parallel algorithms, debuggers, performance tools, containers; strong communication and t...
Linux/Unix, Fortran, C/C++, MPI, OpenMP, OpenACC, CUDA, shell scripting, Python, debuggers, performance tools, containers, REST, JavaScript, SQL
1mo
Save
Mark Applied
Hide
HPC/ML Infrastructure Engineer
San Francisco or Tokyo
OnsiteFull Time
Spellbrush
Spellbrush: Develops anime-themed video games using proprietary generative AI technology.
Experienced HPC/ML infrastructure engineer with Linux sysadmin skills, cluster bring-up and operations experience, familiarity with SLURM and parallel filesystems, networking and datacenter hardware handling.
SLURM, Slinky, K8s, Warewulf, MAAS, Ansible, WEKA, VAST, Ceph, Tailscale, Grafana, Prometheus, LDAP, dmesg, HGX, VLAN
2mo
Save
Mark Applied
Hide
HPC & AI Senior Performance Engineer
Bloomington or Spring or San Jose
$120k-$275k/yr RemoteFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Providing global edge-to-cloud infrastructure and IT solutions for businesses.
7+ YOEMaster’s or PhD in CS/engineering/math; 7+ years HPC performance analysis; strong communication; US citizenship.
OpenMP, MPI, CUDA, HIP, NCCL/RCCL, HPL, HPCG, OSU-MPI, Performance profiling tools, GPU architectures, Multiprocessor, CPU/GPU
2mo
Save
Mark Applied
Hide
HPC & AI Senior Performance Engineer
San Jose or Bloomington
$120k-$275k/yr RemoteFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Provides edge-to-cloud IT infrastructure and platform services.
7+ YOEMaster’s or PhD in CS/engineering/math; 7+ years HPC performance experience; strong analytical, communication skills; US citizenship.
OpenMP, MPI, CUDA, HIP, HPL, HPCG, NCCL, RCCL, Performance profiling tools
3d
Save
Mark Applied
Hide
AI/HPC Network Performance Engineer
Menlo Park, California, United States
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
8+ YOEBachelor's or equivalent,8+ years in system or network performance engineering for large-scale distributed/HPC environments; experience with datacenter networks, network automation, and coding in Python,C++,Go.
Python, C++, Go, IB, RDMA, RoCE
3w
Save
Mark Applied
Hide
GPU Systems Engineer – HPC / Parallel Computing
San Francisco or Los Angeles
$160k-$320k/yr OnsiteFull Time
Vast.ai
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Experience with HPC/parallel programming, GPU systems optimization, CUDA/C++, Python, and Linux; familiarity with parallel frameworks (HIP, SYCL, OpenCL, OpenACC) and HPC performance tooling.
CUDA, C++, GPGPU, Python, Linux, C++17, C++20, HIP, SYCL, OpenCL, OpenACC