658 gpu jobs at 257 companies in San Rafael, CA

3mo
Save
Mark Applied
Hide
GPU Compiler Lead
Sunnyvale, California, United States
$175k-$250k/yr OnsiteFull Time
Bolt Graphics
Bolt Graphics: Designing high-efficiency graphics processors for professional rendering and simulation.
8+ YOE8+ years leading a high-performance GPU compiler stack; strong C++/assembly; LLVM/GCC internals; GPU architectures; parallel languages.
LLVM, GCC, C++, assembly, MLIR, ISPC, Vulkan, DirectX, CUDA, Metal, OpenCL, OpenGL
2w
Save
Mark Applied
Hide
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yr OnsiteFull Time
TypeSafe AI
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
CUDA, CuTe DSL
3w
Save
Mark Applied
Hide
Systems/GPU Research Engineer
San Francisco or Los Angeles
$160k-$320k/yr OnsiteFull Time
Vast.ai
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Expertise in systems and GPU engineering, GPU architectures, neural network performance, C++/CUDA/Python proficiency, and strong research background with publications preferred.
C++, CUDA, GPGPU, Python, Linux
1mo
Save
Mark Applied
Hide
Senior GPU Capacity Planner
San Francisco or Sunnyvale or Bellevue
$160k-$195k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
3+ YOE3+ years capacity planning or systems engineering experience; hyperscaler cloud experience; GPU topology knowledge (NVIDIA H100/B200); Bachelor’s or Master’s in quantitative field; strong cross-functional communication and modeling skills.
AWS, GCP, Azure, Oracle Cloud, NVIDIA H100, NVIDIA B200
3mo
Save
Mark Applied
Hide
GPU Performance Engineer, Platform Architecture
Austin or Boston or San Francisco or San Diego
HybridFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
3+ YOEBachelor’s degree; 3+ years in GPU/CPU modeling; strong C++; Python or Ruby; experience with data analysis tools.
C++, Python, Ruby, Tableau, Pandas, Excel, Matplotlib
3mo
Save
Mark Applied
Hide
Hyperbolic Labs - Senior GPU Infrastructure Engineer
San Francisco, California, United States
RemoteFull Time
YieldNest
YieldNest: Liquid restaking protocol for risk-adjusted DeFi yields.
Senior infrastructure/DevOps engineer with expertise in bare-metal provisioning, GPU scheduling, Terraform/Pulumi, CI/CD for infrastructure, storage for AI/ML workloads, and cloud-init provisioning.
Terraform, Pulumi, CI/CD, infrastructure as code, secrets management, configuration management, observability stack, object storage, block storage, distributed file systems, cloud-init, CUDA, GPU topology, GPU orchestration
3mo
Save
Mark Applied
Hide
Senior Software Engineer - C++ GPU Performance
Foster City or Seattle or Boston or San Diego
$217k-$307k/yr HybridFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
7+ YOEBS in CS or related field; 7+ years; strong CUDA, C++, Linux; GPU performance, instrumentation, debugging, profiling.
CUDA, Nsight, C++, Linux, TensorRT, XLA, OpenGL, RocM
1mo
Save
Mark Applied
Hide
GPU Driver & Memory Management Architect
Spring or Palo Alto or Houston
$147k-$231k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Manufactures personal computers, printers, and 3D printing hardware.
10+ YOEDesign GPU/NPU driver architectures and memory management for shared CPU/GPU/NPU use; 10+ years experience recommended; degree in electrical or related engineering preferred; expertise in memory/GPU design, hardware architecture, debugging, and laboratory tools.
DVMT, WDDM, Windows, Linux, MATLAB, Oscilloscope, FPGA
1mo
Save
Mark Applied
Hide
GPU Driver & Memory Management Architect
Spring or Palo Alto or Houston
$147k-$231k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Produces personal computers, printers, and related digital imaging products.
10+ YOEDevelop GPU/NPU driver architectures owning memory management to enable dynamic shared memory across CPU/GPU/NPU, optimize iGPU for AI+graphics, diagnose performance bottlenecks, and lead cross-organization hardware architecture efforts.
MATLAB, Oscilloscope, Field-Programmable Gate Array (FPGA), Schematic Capture, WDDM, DVMT, Windows, Linux
1mo
Save
Mark Applied
Hide
GPU Performance Engineer
Sunnyvale or Bellevue
$120k-$160k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
2+ YOE2+ years software engineering experience; proficiency in Go and/or Python; production Kubernetes experience; develop performance tests, automation, and platform tooling; participate in on-call rotation.
Go, Python, Kubernetes
1mo
Save
Mark Applied
Hide
GPU Driver & Memory Management Architect
Spring or Palo Alto
$147k-$231k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Manufacturer of personal computers, printers, and imaging devices.
10+ YOE10+ years experience in electrical/hardware or driver architecture; expertise in GPU/memory design, OS-driver contracts (Windows/Linux), debugging, FPGA and hardware architecture; degree in electrical engineering or related preferred.
MATLAB, Oscilloscope, Schematic Capture, Simulations, Field-Programmable Gate Array (FPGA), Printed Circuit Board, WDDM, DVMT, Windows, Linux
1w
Save
Mark Applied
Hide
Member of Technical Staff, GPU / ML Systems
San Mateo, California, United States
OnsiteFull Time
SkyPilot
SkyPilot: Unified compute platform for orchestrating AI workloads across clouds.
Hands-on experience with GPU/accelerator systems and ML training or inference infrastructure; strong Python and systems-level skills; experience operating large-scale training or high-throughput inference.
vLLM, PyTorch, CUDA, Slime, Kueue, KAI, KServe, Python, Kubernetes, Spark, Databricks
1mo
Save
Mark Applied
Hide
Principal Software Engineer, GPU Compute
San Mateo, California, United States
$345k-$399k/yr HybridFull Time
Roblox
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
Go, CUDA, Kubernetes, NVLink, InfiniBand, RoCE, BMC, IPMI, Redfish
11h
Save
Mark Applied
Hide
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
Linux, SLURM, Kubernetes, Ray, Hadoop
1mo
Save
Mark Applied
Hide
GPU/CPU Systems Engineer
Seattle or San Francisco
$135k-$306k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
10+ YOEExpertise in GPU/CPU hardware and platform engineering, firmware and diagnostics (BMC, UEFI/BIOS, Linux), board-level tools, FPGA and server architectures (x86/ARM); 10+ years experience preferred; strong debugging and communication skills.
BMC firmware, UEFI, BIOS, Linux, FPGA, PCIe, DDR, Ethernet, USB, SPI, GPU
1w
Save
Mark Applied
Hide
Staff Software Engineer, GPU Infrastructure Lifecycle Management
San Francisco, California, United States
$240k-$280k/yr OnsiteFull Time
Together AI
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Go, Python, Rust, Temporal, Cadence, Kubernetes, Kafka, NATS, SQS, PXE, iPXE, Redfish, IPMI, BMC, NCCL, CUDA, InfiniBand, RoCE
2mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
19h
Save
Mark Applied
Hide
Staff Software Engineer, GPU Inference
Toronto or Sunnyvale
HybridFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
8+ YOE8+ years software engineering experience, strong C++ and Python skills, GPU inference experience, Linux, containers and Kubernetes, benchmarking and production optimization for latency-sensitive services.
C++, Python, vLLM, PyTorch, ROCm, HIP, RCCL, rocprofiler, AMD SMI, AITER, hipBLASLt, Composable Kernel, CUDA, SGLang, TensorRT-LLM, Triton Inference Server, Kubernetes, Linux, RDMA
2mo
Save
Mark Applied
Hide
System Software Engineer, Robot Platform — GPU & Accelerated Compute
Redwood City, California, United States
OnsiteFull Time
Sunday
Sunday: Developing autonomous robots to perform household chores.
2+ YOE2+ years in GPU systems software; proficient in CUDA and a systems language (C++, C, or Rust); strong understanding of GPU architecture and time-slicing; experience with CUDA ecosystem and GPU sharing; solid Linux fundamentals.
CUDA, CUDA Graphs, CUDA IPC, Nsight Systems, Nsight Compute, NVDEC, NVENC, MPS, MIG, Linux
1w
Save
Mark Applied
Hide
Engineering Manager, GPU Infrastructure
Toronto or San Francisco or New York City or London or Paris or Montreal
HybridFull Time
Cohere
Cohere: Provides enterprise-grade large language models and AI software platforms.
Experience managing engineering teams focused on GPU/ML infrastructure, Kubernetes, IaC, observability, and collaboration with AI researchers; strong communication and mentorship skills.
JAX, PyTorch, TensorFlow, Kubernetes, Prometheus, Grafana, Terraform, ArgoCD