437 gpu jobs at 196 companies in Novato, CA

2w
Save
Mark Applied
Hide
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yr OnsiteFull Time
TypeSafe AI
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
CUDA, CuTe DSL
3w
Save
Mark Applied
Hide
Systems/GPU Research Engineer
San Francisco or Los Angeles
$160k-$320k/yr OnsiteFull Time
Vast.ai
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Expertise in systems and GPU engineering, GPU architectures, neural network performance, C++/CUDA/Python proficiency, and strong research background with publications preferred.
C++, CUDA, GPGPU, Python, Linux
1mo
Save
Mark Applied
Hide
Senior GPU Capacity Planner
San Francisco or Sunnyvale or Bellevue
$160k-$195k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
3+ YOE3+ years capacity planning or systems engineering experience; hyperscaler cloud experience; GPU topology knowledge (NVIDIA H100/B200); Bachelor’s or Master’s in quantitative field; strong cross-functional communication and modeling skills.
AWS, GCP, Azure, Oracle Cloud, NVIDIA H100, NVIDIA B200
3mo
Save
Mark Applied
Hide
GPU Performance Engineer, Platform Architecture
Austin or Boston or San Francisco or San Diego
HybridFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
3+ YOEBachelor’s degree; 3+ years in GPU/CPU modeling; strong C++; Python or Ruby; experience with data analysis tools.
C++, Python, Ruby, Tableau, Pandas, Excel, Matplotlib
3mo
Save
Mark Applied
Hide
Hyperbolic Labs - Senior GPU Infrastructure Engineer
San Francisco, California, United States
RemoteFull Time
YieldNest
YieldNest: Liquid restaking protocol for risk-adjusted DeFi yields.
Senior infrastructure/DevOps engineer with expertise in bare-metal provisioning, GPU scheduling, Terraform/Pulumi, CI/CD for infrastructure, storage for AI/ML workloads, and cloud-init provisioning.
Terraform, Pulumi, CI/CD, infrastructure as code, secrets management, configuration management, observability stack, object storage, block storage, distributed file systems, cloud-init, CUDA, GPU topology, GPU orchestration
3mo
Save
Mark Applied
Hide
Senior Software Engineer - C++ GPU Performance
Foster City or Seattle or Boston or San Diego
$217k-$307k/yr HybridFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
7+ YOEBS in CS or related field; 7+ years; strong CUDA, C++, Linux; GPU performance, instrumentation, debugging, profiling.
CUDA, Nsight, C++, Linux, TensorRT, XLA, OpenGL, RocM
1w
Save
Mark Applied
Hide
Member of Technical Staff, GPU / ML Systems
San Mateo, California, United States
OnsiteFull Time
SkyPilot
SkyPilot: Unified compute platform for orchestrating AI workloads across clouds.
Hands-on experience with GPU/accelerator systems and ML training or inference infrastructure; strong Python and systems-level skills; experience operating large-scale training or high-throughput inference.
vLLM, PyTorch, CUDA, Slime, Kueue, KAI, KServe, Python, Kubernetes, Spark, Databricks
1mo
Save
Mark Applied
Hide
Principal Software Engineer, GPU Compute
San Mateo, California, United States
$345k-$399k/yr HybridFull Time
Roblox
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
Go, CUDA, Kubernetes, NVLink, InfiniBand, RoCE, BMC, IPMI, Redfish
19h
Save
Mark Applied
Hide
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
Linux, SLURM, Kubernetes, Ray, Hadoop
1mo
Save
Mark Applied
Hide
GPU/CPU Systems Engineer
Seattle or San Francisco
$135k-$306k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
10+ YOEExpertise in GPU/CPU hardware and platform engineering, firmware and diagnostics (BMC, UEFI/BIOS, Linux), board-level tools, FPGA and server architectures (x86/ARM); 10+ years experience preferred; strong debugging and communication skills.
BMC firmware, UEFI, BIOS, Linux, FPGA, PCIe, DDR, Ethernet, USB, SPI, GPU
1w
Save
Mark Applied
Hide
Staff Software Engineer, GPU Infrastructure Lifecycle Management
San Francisco, California, United States
$240k-$280k/yr OnsiteFull Time
Together AI
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Go, Python, Rust, Temporal, Cadence, Kubernetes, Kafka, NATS, SQS, PXE, iPXE, Redfish, IPMI, BMC, NCCL, CUDA, InfiniBand, RoCE
2mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
2mo
Save
Mark Applied
Hide
System Software Engineer, Robot Platform — GPU & Accelerated Compute
Redwood City, California, United States
OnsiteFull Time
Sunday
Sunday: Developing autonomous robots to perform household chores.
2+ YOE2+ years in GPU systems software; proficient in CUDA and a systems language (C++, C, or Rust); strong understanding of GPU architecture and time-slicing; experience with CUDA ecosystem and GPU sharing; solid Linux fundamentals.
CUDA, CUDA Graphs, CUDA IPC, Nsight Systems, Nsight Compute, NVDEC, NVENC, MPS, MIG, Linux
1w
Save
Mark Applied
Hide
Engineering Manager, GPU Infrastructure
Toronto or San Francisco or New York City or London or Paris or Montreal
HybridFull Time
Cohere
Cohere: Provides enterprise-grade large language models and AI software platforms.
Experience managing engineering teams focused on GPU/ML infrastructure, Kubernetes, IaC, observability, and collaboration with AI researchers; strong communication and mentorship skills.
JAX, PyTorch, TensorFlow, Kubernetes, Prometheus, Grafana, Terraform, ArgoCD
3w
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
1mo
Save
Mark Applied
Hide
Member of Technical Staff, AMD GPU Performance Engineering
San Francisco, California, United States
$200k-$400k/yr OnsiteFull Time
Inferact
Inferact: A building an AI inference engine and GPU-optimized software to accelerate model inference performance.
Bachelor's or equivalent experience; hands-on AMD GPU optimization using ROCm/HIP/Triton/CK/AITER; deep understanding of AMD GPU execution, memory, toolchains; experience optimizing ML kernels and strong profiling/benchmarking skills.
ROCm, HIP, Triton, CK, AITER, vLLM, SGLang, TensorRT-LLM, MLIR, LLVM, PyTorch
3w
Save
Mark Applied
Hide
Software Engineer- GPU Fabric Observability
San Francisco, California, United States
$200k-$380k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Kubernetes
1mo
Save
Mark Applied
Hide
Senior Power Design Engineer – GPU Server Boards
San Carlos, California, United States
$150k-$200k/yr OnsiteFull Time
Cowboy Space
Cowboy Space: Building satellite-based orbital infrastructure for AI data centers.
Expertise in board-level DC-DC power conversion, schematic capture, PCB layout, circuit simulation, system-level power delivery design, and lab validation; strong communication and cross-functional collaboration skills.
schematic capture, PCB layout, circuit simulation, CAD, oscilloscope, electronic load, frequency response analyzer
3w
Save
Mark Applied
Hide
Sr. Software Development Engineer - Video Rendering
Seattle or San Francisco or San Jose
$174k-$331k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years software engineering with deep GPU/graphics or rendering systems experience, strong modern C++, and knowledge of GPU APIs, shading languages, and performance debugging.
C++, DirectX 12, Metal, Vulkan, CUDA, OpenCL, HLSL, Slang, SPIR-V, DXIL, DXC, Nsight, PIX, TensorRT, ONNX Runtime
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - AI Infrastructure
San Francisco or United States
RemoteFull Time
Andromeda Cluster
Andromeda Cluster: AI compute orchestration platform for GPU clusters.
Senior SRE with GPU infra, distributed training, and networking expertise.
NVIDIA GPUs, InfiniBand, RoCE, NVLink, NCCL, CUDA, PyTorch, DeepSpeed, Megatron, FSDP, Linux, Kubernetes, Slurm, Terraform, Helm, Ansible, DCGM, nvidia-smi