338 gpu jobs at 163 companies in Rohnert Park, CA

1mo
Save
Mark Applied
Hide
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yr OnsiteFull Time
TypeSafe AI
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
CUDA, CuTe DSL
1mo
Save
Mark Applied
Hide
Systems/GPU Research Engineer
San Francisco or Los Angeles
$160k-$320k/yr OnsiteFull Time
Vast.ai
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Expertise in systems and GPU engineering, GPU architectures, neural network performance, C++/CUDA/Python proficiency, and strong research background with publications preferred.
C++, CUDA, GPGPU, Python, Linux
1mo
Save
Mark Applied
Hide
Senior GPU Capacity Planner
San Francisco or Sunnyvale or Bellevue
$160k-$195k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
3+ YOE3+ years capacity planning or systems engineering experience; hyperscaler cloud experience; GPU topology knowledge (NVIDIA H100/B200); Bachelor’s or Master’s in quantitative field; strong cross-functional communication and modeling skills.
AWS, GCP, Azure, Oracle Cloud, NVIDIA H100, NVIDIA B200
5d
Save
Mark Applied
Hide
Senior Systems GPU Engineer – AI & Robotics
San Francisco or Sunnyvale
$160k-$271k/yr OnsiteFull Time
Intuitive
IntuitiveNASDAQ: ISRG: Robotic-assisted systems for minimally invasive surgery.
6+ YOEMaster’s degree in a technical field and 6+ years of embedded systems software experience required. Requires CUDA/OpenCL, C/C++/Python/Bash, Linux, real-time systems, machine learning, and robotics expertise.
NVIDIA Nsight Systems, NVIDIA Nsight Compute, Linux, Docker, NVIDIA Container Toolkit, Kubernetes, CUDA, OpenCL, C, C++, Python, Bash, QNX, Yocto, PyTorch, TensorFlow
2w
Save
Mark Applied
Hide
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
Linux, SLURM, Kubernetes, Ray, Hadoop
1mo
Save
Mark Applied
Hide
GPU/CPU Systems Engineer
Seattle or San Francisco
$135k-$306k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
10+ YOEExpertise in GPU/CPU hardware and platform engineering, firmware and diagnostics (BMC, UEFI/BIOS, Linux), board-level tools, FPGA and server architectures (x86/ARM); 10+ years experience preferred; strong debugging and communication skills.
BMC firmware, UEFI, BIOS, Linux, FPGA, PCIe, DDR, Ethernet, USB, SPI, GPU
3w
Save
Mark Applied
Hide
Staff Software Engineer, GPU Infrastructure Lifecycle Management
San Francisco, California, United States
$240k-$280k/yr OnsiteFull Time
Together AI
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Go, Python, Rust, Temporal, Cadence, Kubernetes, Kafka, NATS, SQS, PXE, iPXE, Redfish, IPMI, BMC, NCCL, CUDA, InfiniBand, RoCE
3mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
1d
Save
Mark Applied
Hide
Senior ML Accelerator Engineer - GPU
Sunnyvale or Washington or Austin or San Francisco or Warren
$170k-$258k/yr HybridFull Time
General Motors
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
2+ YOERequires 2+ years of relevant experience and a CS or related technical degree. Strong CUDA and C++ skills, GPU architecture knowledge, performance optimization experience, and analytical, collaborative communication skills.
CUDA, C++, NSight, CUTLASS, CuTe
3w
Save
Mark Applied
Hide
Engineering Manager, GPU Infrastructure
Toronto or San Francisco or New York City or London or Paris or Montreal
HybridFull Time
Cohere
Cohere: Provides enterprise-grade large language models and AI software platforms.
Experience managing engineering teams focused on GPU/ML infrastructure, Kubernetes, IaC, observability, and collaboration with AI researchers; strong communication and mentorship skills.
JAX, PyTorch, TensorFlow, Kubernetes, Prometheus, Grafana, Terraform, ArgoCD
2d
Save
Mark Applied
Hide
Photoshop Developer, GPU/Imaging
San Francisco or San Jose or Seattle or New York City
$139k-$258k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Bachelor's or master's in computer science or equivalent experience; graphics and GPU programming fundamentals; modern C++, graphics APIs, asynchronous production code, and strong analytical and debugging skills.
C++, WebGPU, WGSL, Metal, DirectX, Vulkan, WebAssembly, Adobe Photoshop, AI tools
1mo
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
1w
Save
Mark Applied
Hide
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yr OnsiteFull Time
Perplexity
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Kubernetes, kubectl, NVIDIA, CUDA, InfiniBand, RoCE, CoreWeave, AWS, GCP, Go, Rust, C++, vLLM, SGLang, TensorRT-LLM, Slurm, Triton, RDMA, Prometheus, Grafana, Weights & Biases
1mo
Save
Mark Applied
Hide
Member of Technical Staff, AMD GPU Performance Engineering
San Francisco, California, United States
$200k-$400k/yr OnsiteFull Time
Inferact
Inferact: A building an AI inference engine and GPU-optimized software to accelerate model inference performance.
Bachelor's or equivalent experience; hands-on AMD GPU optimization using ROCm/HIP/Triton/CK/AITER; deep understanding of AMD GPU execution, memory, toolchains; experience optimizing ML kernels and strong profiling/benchmarking skills.
ROCm, HIP, Triton, CK, AITER, vLLM, SGLang, TensorRT-LLM, MLIR, LLVM, PyTorch
6d
Save
Mark Applied
Hide
GPU Cluster Engineer, Networking
San Francisco, California, United States
$150k-$180k/yr OnsiteFull Time
Sciforium
Sciforium: Building multimodal AI models and high-performance model serving infrastructure.
7+ YOERequires 7+ years designing production data center networks, large-scale HPC/AI fabric experience, expert routing and switching, RDMA, network security, automation with Python and Ansible, and cloud networking experience.
InfiniBand, RoCE v2, NVLink, NVSwitch, BMC, IPMI, ConnectX, BlueField, OSFP, QSFP-DD, BGP, VLAN, VRF, EVPN-VXLAN, PFC, ECN/DCQCN, QoS, DSCP, UFM, OpenSM, SHARP, perftest, ib_write_bw, nccl-tests, rccl-tests, GPUDirect RDMA, OSPF, ECMP, Arista EOS, NVIDIA Cumulus, Cisco NX-OS, Junos, SONiC, WireGuard, IPsec, NAT, ACL, NetBox, Ansible, Nornir, NAPALM, Git, gNMI, sFlow, Prometheus, Grafana, DWDM, AWS Direct Connect, Azure ExpressRoute, GCP Interconnect, VPC, Transit Gateway, Python, MOFED, DOCA, NCCL, RCCL, Kubernetes, CNI, SR-IOV, Multus, CCIE, JNCIE
1mo
Save
Mark Applied
Hide
Software Engineer- GPU Fabric Observability
San Francisco, California, United States
$200k-$380k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Kubernetes
1d
Save
Mark Applied
Hide
Cloud Platforms and Infrastructure Engineer, TPU/GPU
Austin or San Francisco or Atlanta or Boulder or Addison or Miami or Sunnyvale or Chicago
$152k-$221k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
6+ YOEBachelor's degree or equivalent; 6 years in infrastructure automation, DevOps, CI/CD, Kubernetes, and Linux; 3 years in project management and technical delivery; coding experience and cloud provider experience required.
Kubernetes, Linux, Python, Java, Go, C, C++, Google Cloud Platform (GCP), PANW, Fortinet, VMWare, TCP/IP, Hypertext Transfer Protocol, Border Gateway Protocol (BGP), IAM, PyTorch, JAX, TensorFlow, Slurm, Google Kubernetes Engine (GKE)
1mo
Save
Mark Applied
Hide
Member of Technical Staff - Research Engineer
Freiburg im Breisgau or San Francisco
$180k-$290k/yr HybridFull Time
Black Forest Labs
Black Forest Labs: Developing frontier generative AI models for visual intelligence.
Deep experience with large-scale training systems, PyTorch, distributed training, GPU profiling and low-precision training; ability to debug training failures and implement GPU-level optimizations.
PyTorch, FSDP, NCCL, Nsight Systems, Nsight Compute, torch profiler, CUDA, Triton, CuTe, CUTLASS
1mo
Save
Mark Applied
Hide
Staff Cluster Infrastructure Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building specialized industrial robots and physical AI systems.
6+ YOE6+ years operating GPU compute on Kubernetes, strong Python/Go programming, experience with Terraform or CloudFormation, bare-metal Linux and GPU hardware familiarity, and strong automation and reliability focus.
Kubernetes, Python, Go, Terraform, CloudFormation, Linux, GPU
2w
Save
Mark Applied
Hide
Staff Engineer, Inference Optimizations
San Francisco, California, United States
$191k-$239k/yr RemoteFull Time
DigitalOcean
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, deep GPU and low-level optimization expertise, experience with CUDA/Triton/ROCm, distributed GPU parallelization, and system design for inference workloads.
CUDA, ROCm, TensorRT, OpenAI Triton, AITER, FlashAttention