399 gpu jobs at 178 companies in Fairfield, CA

2w
Save
Mark Applied
Hide
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yr OnsiteFull Time
TypeSafe AI
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
CUDA, CuTe DSL
3w
Save
Mark Applied
Hide
Systems/GPU Research Engineer
San Francisco or Los Angeles
$160k-$320k/yr OnsiteFull Time
Vast.ai
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Expertise in systems and GPU engineering, GPU architectures, neural network performance, C++/CUDA/Python proficiency, and strong research background with publications preferred.
C++, CUDA, GPGPU, Python, Linux
1mo
Save
Mark Applied
Hide
Senior GPU Capacity Planner
San Francisco or Sunnyvale or Bellevue
$160k-$195k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
3+ YOE3+ years capacity planning or systems engineering experience; hyperscaler cloud experience; GPU topology knowledge (NVIDIA H100/B200); Bachelor’s or Master’s in quantitative field; strong cross-functional communication and modeling skills.
AWS, GCP, Azure, Oracle Cloud, NVIDIA H100, NVIDIA B200
3mo
Save
Mark Applied
Hide
GPU Performance Engineer, Platform Architecture
Austin or Boston or San Francisco or San Diego
HybridFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
3+ YOEBachelor’s degree; 3+ years in GPU/CPU modeling; strong C++; Python or Ruby; experience with data analysis tools.
C++, Python, Ruby, Tableau, Pandas, Excel, Matplotlib
3mo
Save
Mark Applied
Hide
Hyperbolic Labs - Senior GPU Infrastructure Engineer
San Francisco, California, United States
RemoteFull Time
YieldNest
YieldNest: Liquid restaking protocol for risk-adjusted DeFi yields.
Senior infrastructure/DevOps engineer with expertise in bare-metal provisioning, GPU scheduling, Terraform/Pulumi, CI/CD for infrastructure, storage for AI/ML workloads, and cloud-init provisioning.
Terraform, Pulumi, CI/CD, infrastructure as code, secrets management, configuration management, observability stack, object storage, block storage, distributed file systems, cloud-init, CUDA, GPU topology, GPU orchestration
3mo
Save
Mark Applied
Hide
Senior Software Engineer - C++ GPU Performance
Foster City or Seattle or Boston or San Diego
$217k-$307k/yr HybridFull Time
Zoox
ZooxNASDAQ: AMZN: Developing autonomous robotaxis for urban ride-hailing services.
7+ YOEBS in CS or related field; 7+ years; strong CUDA, C++, Linux; GPU performance, instrumentation, debugging, profiling.
CUDA, Nsight, C++, Linux, TensorRT, XLA, OpenGL, RocM
1w
Save
Mark Applied
Hide
Member of Technical Staff, GPU / ML Systems
San Mateo, California, United States
OnsiteFull Time
SkyPilot
SkyPilot: Unified compute platform for orchestrating AI workloads across clouds.
Hands-on experience with GPU/accelerator systems and ML training or inference infrastructure; strong Python and systems-level skills; experience operating large-scale training or high-throughput inference.
vLLM, PyTorch, CUDA, Slime, Kueue, KAI, KServe, Python, Kubernetes, Spark, Databricks
1mo
Save
Mark Applied
Hide
Principal Software Engineer, GPU Compute
San Mateo, California, United States
$345k-$399k/yr HybridFull Time
Roblox
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
Go, CUDA, Kubernetes, NVLink, InfiniBand, RoCE, BMC, IPMI, Redfish
1mo
Save
Mark Applied
Hide
GPU/CPU Systems Engineer
Seattle or San Francisco
$135k-$306k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
10+ YOEExpertise in GPU/CPU hardware and platform engineering, firmware and diagnostics (BMC, UEFI/BIOS, Linux), board-level tools, FPGA and server architectures (x86/ARM); 10+ years experience preferred; strong debugging and communication skills.
BMC firmware, UEFI, BIOS, Linux, FPGA, PCIe, DDR, Ethernet, USB, SPI, GPU
1w
Save
Mark Applied
Hide
Staff Software Engineer, GPU Infrastructure Lifecycle Management
San Francisco, California, United States
$240k-$280k/yr OnsiteFull Time
Together AI
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Go, Python, Rust, Temporal, Cadence, Kubernetes, Kafka, NATS, SQS, PXE, iPXE, Redfish, IPMI, BMC, NCCL, CUDA, InfiniBand, RoCE
2mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
1w
Save
Mark Applied
Hide
Engineering Manager, GPU Infrastructure
Toronto or San Francisco or New York City or London or Paris or Montreal
HybridFull Time
Cohere
Cohere: Provides enterprise-grade large language models and AI software platforms.
Experience managing engineering teams focused on GPU/ML infrastructure, Kubernetes, IaC, observability, and collaboration with AI researchers; strong communication and mentorship skills.
JAX, PyTorch, TensorFlow, Kubernetes, Prometheus, Grafana, Terraform, ArgoCD
3w
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
1mo
Save
Mark Applied
Hide
Member of Technical Staff, AMD GPU Performance Engineering
San Francisco, California, United States
$200k-$400k/yr OnsiteFull Time
Inferact
Inferact: A building an AI inference engine and GPU-optimized software to accelerate model inference performance.
Bachelor's or equivalent experience; hands-on AMD GPU optimization using ROCm/HIP/Triton/CK/AITER; deep understanding of AMD GPU execution, memory, toolchains; experience optimizing ML kernels and strong profiling/benchmarking skills.
ROCm, HIP, Triton, CK, AITER, vLLM, SGLang, TensorRT-LLM, MLIR, LLVM, PyTorch
3w
Save
Mark Applied
Hide
Software Engineer- GPU Fabric Observability
San Francisco, California, United States
$200k-$380k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Kubernetes
2w
Save
Mark Applied
Hide
Sr. Software Development Engineer - Video Rendering
Seattle or San Francisco or San Jose
$174k-$331k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years software engineering with deep GPU/graphics or rendering systems experience, strong modern C++, and knowledge of GPU APIs, shading languages, and performance debugging.
C++, DirectX 12, Metal, Vulkan, CUDA, OpenCL, HLSL, Slang, SPIR-V, DXIL, DXC, Nsight, PIX, TensorRT, ONNX Runtime
3mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - AI Infrastructure
San Francisco or United States
RemoteFull Time
Andromeda Cluster
Andromeda Cluster: AI compute orchestration platform for GPU clusters.
Senior SRE with GPU infra, distributed training, and networking expertise.
NVIDIA GPUs, InfiniBand, RoCE, NVLink, NCCL, CUDA, PyTorch, DeepSpeed, Megatron, FSDP, Linux, Kubernetes, Slurm, Terraform, Helm, Ansible, DCGM, nvidia-smi
1mo
Save
Mark Applied
Hide
Staff Cluster Infrastructure Engineer
San Francisco, California, United States
$224k-$284k/yr OnsiteFull Time
Atoms
Atoms: Building specialized industrial robots and physical AI systems.
6+ YOE6+ years operating GPU compute on Kubernetes, strong Python/Go programming, experience with Terraform or CloudFormation, bare-metal Linux and GPU hardware familiarity, and strong automation and reliability focus.
Kubernetes, Python, Go, Terraform, CloudFormation, Linux, GPU
5d
Save
Mark Applied
Hide
Staff Engineer, Inference Optimizations
San Francisco, California, United States
$191k-$239k/yr RemoteFull Time
DigitalOcean
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, deep GPU and low-level optimization expertise, experience with CUDA/Triton/ROCm, distributed GPU parallelization, and system design for inference workloads.
CUDA, ROCm, TensorRT, OpenAI Triton, AITER, FlashAttention
2mo
Save
Mark Applied
Hide
Machine Learning Infra Engineer
San Francisco, California, United States
$150k-$300k/yr OnsiteFull Time
Reducto
Reducto: AI platform extracting structured data from unstructured documents
Strong Python, Kubernetes, distributed training, multi-node GPUs, and production observability.
Python, Kubernetes, Distributed training, GPU clusters, Machine learning frameworks