428 gpu jobs at 201 companies in American Canyon, CA
1mo
Save
Mark Applied
Hide
1mo
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yrOnsiteFull Time
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
Vast.ai: Decentralized marketplace for GPU cloud computing resources.
Expertise in systems and GPU engineering, GPU architectures, neural network performance, C++/CUDA/Python proficiency, and strong research background with publications preferred.
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
3+ YOE3+ years capacity planning or systems engineering experience; hyperscaler cloud experience; GPU topology knowledge (NVIDIA H100/B200); Bachelor’s or Master’s in quantitative field; strong cross-functional communication and modeling skills.
IntuitiveNASDAQ: ISRG: Robotic-assisted systems for minimally invasive surgery.
6+ YOEMaster’s degree in a technical field and 6+ years of embedded systems software experience required. Requires CUDA/OpenCL, C/C++/Python/Bash, Linux, real-time systems, machine learning, and robotics expertise.
SkyPilot: Unified compute platform for orchestrating AI workloads across clouds.
Hands-on experience with GPU/accelerator systems and ML training or inference infrastructure; strong Python and systems-level skills; experience operating large-scale training or high-throughput inference.
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
10+ YOEExpertise in GPU/CPU hardware and platform engineering, firmware and diagnostics (BMC, UEFI/BIOS, Linux), board-level tools, FPGA and server architectures (x86/ARM); 10+ years experience preferred; strong debugging and communication skills.
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
Sunnyvale or Washington or Austin or San Francisco or Warren
$170k-$258k/yrHybridFull Time
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
2+ YOERequires 2+ years of relevant experience and a CS or related technical degree. Strong CUDA and C++ skills, GPU architecture knowledge, performance optimization experience, and analytical, collaborative communication skills.
System Software Engineer, Robot Platform — GPU & Accelerated Compute
Redwood City, California, United States
OnsiteFull Time
Sunday: Developing autonomous robots to perform household chores.
2+ YOE2+ years in GPU systems software; proficient in CUDA and a systems language (C++, C, or Rust); strong understanding of GPU architecture and time-slicing; experience with CUDA ecosystem and GPU sharing; solid Linux fundamentals.
CUDA, CUDA Graphs, CUDA IPC, Nsight Systems, Nsight Compute, NVDEC, NVENC, MPS, MIG, Linux
Toronto or San Francisco or New York City or London or Paris or Montreal
HybridFull Time
Cohere: Provides enterprise-grade large language models and AI software platforms.
Experience managing engineering teams focused on GPU/ML infrastructure, Kubernetes, IaC, observability, and collaboration with AI researchers; strong communication and mentorship skills.
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
San Francisco or San Jose or Seattle or New York City
$139k-$258k/yrOnsiteFull Time
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
Bachelor's or master's in computer science or equivalent experience; graphics and GPU programming fundamentals; modern C++, graphics APIs, asynchronous production code, and strong analytical and debugging skills.
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yrOnsiteFull Time
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Member of Technical Staff, AMD GPU Performance Engineering
San Francisco, California, United States
$200k-$400k/yrOnsiteFull Time
Inferact: A building an AI inference engine and GPU-optimized software to accelerate model inference performance.
Bachelor's or equivalent experience; hands-on AMD GPU optimization using ROCm/HIP/Triton/CK/AITER; deep understanding of AMD GPU execution, memory, toolchains; experience optimizing ML kernels and strong profiling/benchmarking skills.
Sciforium: Building multimodal AI models and high-performance model serving infrastructure.
7+ YOERequires 7+ years designing production data center networks, large-scale HPC/AI fabric experience, expert routing and switching, RDMA, network security, automation with Python and Ansible, and cloud networking experience.
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
6+ YOEBachelor's degree or equivalent; 6 years in infrastructure automation, DevOps, CI/CD, Kubernetes, and Linux; 3 years in project management and technical delivery; coding experience and cloud provider experience required.