261 hpc jobs at 113 companies in San Francisco, CA

2mo
Save
Mark Applied
Hide
HPC Engineer
Sunnyvale, California, United States
$150k-$300k/yr OnsiteFull Time
Institute of Foundation Models
Institute of Foundation Models: Develops open-source frontier-class AI foundation models and research.
2+ YOE2+ years in Linux systems administration, SRE, DevOps, cloud operations, HPC or infrastructure operations; strong Linux troubleshooting; scripting in Python or Bash; Bachelor’s in a related field.
Slurm, GPU infrastructure, AWS, Azure, GCP, Grafana, Prometheus, Datadog, Containers, Kubernetes, Python, Bash, Linux
2mo
Save
Mark Applied
Hide
HPC Systems Engineer
San Francisco, California, United States
HybridFull Time
UCSF Health
UCSF Health: Academic medical center providing advanced patient care and research.
6+ YOEBachelor's in a related field and 6+ years experience with large-scale/HPC systems (or 10+ years related). Expert HPC infrastructure, parallel filesystems, security controls (NIST/HIPPA), SLURM/PBS, Warewulf/InfiniBand, EM automation and advanced scripting.
GPFS, Lustre, Vast, DDN, NIST 800-171, NIST 800-223, HIPPA, SLURM, PBS, Warewulf, InfiniBand
2mo
Save
Mark Applied
Hide
HPC Systems Engineer
San Francisco, California, United States
$120k-$196k/yr HybridFull Time
University of California, San Francisco
University of California, San Francisco: Public research university dedicated to health sciences and education.
6+ YOEBachelor's in CS/engineering plus 6+ years HPC experience (or 10+ yrs related); expert HPC infrastructure design; experience with parallel filesystems (GPFS, Lustre, Vast, DDN); security to NIST/HIPAA standards; SLURM/PBS, Warewulf/InfiniBand, automated testing, advanced scripting, and technical documentation skills.
GPFS, Lustre, Vast, DDN, SLURM, PBS, Warewulf, InfiniBand
4w
Save
Mark Applied
Hide
HPC Software Engineer
Milpitas, California, United States
$136k-$232k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management for semiconductor manufacturing.
Design and develop multi-threaded, distributed HPC systems using C++, Java, Linux, Shell, and Python; diagnose complex system issues and communicate technical solutions.
C++, Java, Shell, Python, Linux
2mo
Save
Mark Applied
Hide
HPC Hardware Engineer
Milpitas, California, United States
$95k-$162k/yr OnsiteFull Time
KLA
KLANASDAQ: KLAC: Provides process control and yield management solutions for semiconductors.
0+ YOEMaster's degree (0 years) or Bachelor's + 2 years; foundational computer architecture and Linux knowledge; experience or exposure to HPC, distributed systems, or server platforms; Python/Bash scripting; strong problem-solving and teamwork skills.
Linux, Python, Bash
3w
Save
Mark Applied
Hide
Senior HPC Systems Architect
San Jose or San Francisco
$255k-$340k/yr HybridFull Time
Lambda
Lambda: Provides high-performance GPU cloud infrastructure for AI development.
8+ YOE8+ years designing large-scale HPC infrastructures, expertise with GPU clusters, high-speed networking, liquid cooling, benchmarking, capacity planning, and architecture documentation.
InfiniBand, Ethernet, Ansible, Terraform, Kubernetes
2mo
Save
Mark Applied
Hide
HPC Scientific Support Engineer
Berkeley, California, United States
$139k-$268k/yr HybridFull Time
Lawrence Berkeley National Laboratory
Lawrence Berkeley National Laboratory: Conducts multidisciplinary scientific research for the U.S. Department of Energy.
8+ YOE8+ years related experience with a Bachelor's (CSE3) or 12+ years (CSE4); 2+ years using HPC systems; deep HPC expertise (Linux/Unix, Fortran, C/C++, MPI, OpenMP, CUDA, Python, containers, debuggers, performance tools); strong communication and teaching skills.
Linux/Unix, Fortran, C/C++, MPI, OpenMP, OpenACC, CUDA, shell scripting, Python, parallel algorithms, programming models, AI models, debuggers, performance tools, containers, REST, JavaScript, SQL
1mo
Save
Mark Applied
Hide
Staff HPC Network Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building specialized industrial robots and physical AI systems.
Experience designing and scaling high-performance networks for HPC/GPU compute, hands-on switch management (Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC), packet analysis, fiber optics knowledge, cloud networking, config management, automation scripting, and monitoring tools.
Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC, tcpdump, Wireshark, AWS, GCP, Azure, Ansible, Salt, Python, Prometheus, Grafana, ELK, GitHub
1mo
Save
Mark Applied
Hide
Staff HPC Network Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building robotics and software to automate physical world industries.
Experience designing and scaling networks for HPC/GPU compute; hands-on with Arista EOS, Cisco NX-OS, Nvidia Cumulus, and/or SONiC; packet analysis (tcpdump, Wireshark); fiber optics to 800 Gbps; familiarity with AWS/GCP/Azure, Ansible/Salt, Python, Prometheus, Grafana, ELK, and GitHub.
Arista EOS, Cisco NX-OS, Nvidia Cumulus, SONiC, tcpdump, Wireshark, AWS, GCP, Azure, Ansible, Salt, Python, Prometheus, Grafana, ELK, GitHub
1mo
Save
Mark Applied
Hide
HPC/AI Performance Specialist
Berkeley or California or United States
$139k-$268k/yr HybridFull Time
Lawrence Berkeley National Laboratory
Lawrence Berkeley National Laboratory: Conducts scientific research to address global challenges.
8+ YOEAdvanced experience in HPC/AI performance engineering, data management, storage and I/O, scientific software development, and collaboration with domain scientists; typically 8+ years (CSE3) or 12+ years (CSE4) of related experience with relevant degree.
NVIDIA Vera-Rubin, VAST file system, Doudna
1mo
Save
Mark Applied
Hide
HPC Systems Administrator (Hardware & Infrastructure Operations)
Stanford, California, United States
$150k-$172k/yr OnsiteFull Time
Stanford University
Stanford University: A private research university providing higher education and research.
3+ YOEManage and maintain large-scale HPC hardware and infrastructure, perform diagnostics and root-cause analysis, collaborate with data center teams, automate provisioning and monitoring; bachelor's degree plus experience required.
Linux, Slurm, Lustre, InfiniBand, DCIM, NVIDIA H200, x86
1w
Save
Mark Applied
Hide
Director, HPC Systems Software Engineering
Houston or New York City or San Francisco or Seattle
$230k-$343k/yr OnsiteFull Time
Nscale
Nscale: Vertically integrated AI infrastructure provider for high-performance computing.
Requires leadership of HPC or infrastructure engineering organizations, Linux, distributed systems, production operations, GPU infrastructure, networking, budgeting, and people management; Slurm experience strongly preferred.
Slurm, Kueue, Linux
1mo
Save
Mark Applied
Hide
HPC Middleware Developer
Santa Clara or Illinois or Colorado or Boulder or California or Holmdel
$152k-$288k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years programming in C/C++, 3+ years Linux experience, deep knowledge of InfiniBand and Ethernet, computer architecture, OS, and performance optimization; MSc or equivalent preferred.
C, C++, Linux, InfiniBand, Ethernet, MPI, RDMA
2mo
Save
Mark Applied
Hide
Flight Sciences Tools and HPC engineer
San Jose, California, United States
$163k-$218k/yr OnsiteFull Time
Archer
ArcherNYSE: ACHR: Develops electric vertical takeoff and landing aircraft for urban mobility.
5+ YOE5+ years in scientific/engineering software for flight sciences; strong Python; HPC design/administration; CI/CD; Linux; MPI; Slurm/PBS; OpenHPC; containerization; IT security.
Python (NumPy SciPy Pandas Scikit-learn TensorFlow PyTorch), VTK, Slurm, PBS, Torque, OpenHPC, Bright, Warewulf, XCat, Spack, EasyBuild, Lmod, Linux, MPI, AWS, Docker, Singularity, Shifter, Chef, Puppet, Ansible, Jenkins, TeamCity, Artifactory, CI/CD tooling, SQL, NoSQL, time-series databases, MATLAB, Simulink, C/C++, Fortran, Rust, CAD, CFD, GNC, aerodynamics, loads, structural analysis, mass properties
3d
Save
Mark Applied
Hide
AI & HPC Infrastructure Engineer
Albany or Arlington or Atlanta or Austin or Beaverton or Bentonville or Boston or Carmel or Charlotte or Chicago or Cincinnati or Cleveland or Columbus or Culver City or Denver or Des Moines or Detroit or Hartford or Houston or Irvine or Irving or Kirkland or Miami or Milwaukee or Minneapolis or Morristown or Mountain View or Nashville or New York City or Oklahoma City or Overland Park or Philadelphia or Pittsburgh or Raleigh or Redmond or Sacramento or San Diego or San Francisco or Scottsdale or Seattle or St. Louis or St. Petersburg or Walnut Creek or United States
$80k-$266k/yr HybridFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
5+ YOERequires 5+ years designing AI infrastructure, accelerated computing, clusters, orchestration, and automation; bachelor's degree or equivalent experience. Strong Kubernetes, Python, Terraform, and cloud expertise required.
Slurm, Run:ai, Kubernetes, NVIDIA Base Command Manager (BCM), NVIDIA NGC, NCCL, NVLink, CUDA, TensorRT-LLM, vLLM, SGLang, Triton Inference Server, NVIDIA Dynamo, llm-d, MLPerf, NCCL tests, fio, iperf, InfiniBand, Ethernet, SONiC, NVMe, NVMe-oF, VAST, Weka, DDN, AWS, Azure, GCP, VMware, Nutanix, Python, Terraform, Ansible, REST, OpenAPI, JSON, YAML, TensorFlow, PyTorch, JAX, Jupyter notebooks, Google Colab, MCP
1mo
Save
Mark Applied
Hide
HPC Platform Engineer, Software, Center for Quantum Computing
San Francisco, California, United States
OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
2+ YOEExperience automating and supporting large-scale infrastructure, programming in at least one modern language, Linux/Unix, CI/CD and infrastructure-as-code, and 2+ years designing or architecting systems.
Python, Ruby, Golang, Java, C++, C#, Rust, Linux, MPI, Docker, Kubernetes, AWS CDK, CloudFormation, EC2, S3, EBS, SQS, Lambda, VPC, DNS, DHCP, TCP/IP, HTTP, Palace
1mo
Save
Mark Applied
Hide
Staff Technical Product Manager – Electronics AI & HPC Simulation
Sunnyvale, California, United States
$117k-$175k/yr OnsiteFull Time
Synopsys
SynopsysNasdaq: SNPS: Provides software and IP for semiconductor design and manufacturing.
Bachelor's in engineering or CS preferred, strong Python scripting, experience with simulation/HPC/AI workflows and APIs, familiarity with HFSS/Icepak/Maxwell/AEDT, and experience with GPU/cloud HPC deployment.
Python, HFSS, Icepak, Maxwell, AEDT, NVIDIA Omniverse, APIs, HPC, GPU
2mo
Save
Mark Applied
Hide
HPC/ML Infrastructure Engineer
San Francisco or Tokyo
OnsiteFull Time
Spellbrush
Spellbrush: Develops anime-themed video games using proprietary generative AI technology.
Experienced HPC/ML infrastructure engineer with Linux sysadmin skills, cluster bring-up and operations experience, familiarity with SLURM and parallel filesystems, networking and datacenter hardware handling.
SLURM, Slinky, K8s, Warewulf, MAAS, Ansible, WEKA, VAST, Ceph, Tailscale, Grafana, Prometheus, LDAP, dmesg, HGX, VLAN
21h
Save
Mark Applied
Hide
AI Systems Engineer - HPC
San Jose, California, United States
$152k-$217k/yr OnsiteFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Experience with HPC infrastructure, GPU clusters, AI workload schedulers, SLURM, Kubernetes, Python, Linux tools, automation, monitoring, and distributed computing; bachelor's or master's degree preferred.
Python, SLURM, Kubernetes, RoCEv2, K8s, KVM, Ubuntu, Shell, Ansible, Saltstack, Terraform, Prometheus, Grafana
1mo
Save
Mark Applied
Hide
GPU Systems Engineer – HPC / Parallel Computing
San Francisco or Los Angeles
$160k-$320k/yr OnsiteFull Time
Vast.ai
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Experience with HPC/parallel programming, GPU systems optimization, CUDA/C++, Python, and Linux; familiarity with parallel frameworks (HIP, SYCL, OpenCL, OpenACC) and HPC performance tooling.
CUDA, C++, GPGPU, Python, Linux, C++17, C++20, HIP, SYCL, OpenCL, OpenACC