242 software engineer gpu jobs at 126 companies in San Rafael, CA
1w
Save
Mark Applied
Hide
1w
Staff Software Engineer, GPU Inference
Toronto or Sunnyvale
HybridFull Time
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
8+ YOE8+ years software engineering experience, strong C++ and Python skills, GPU inference experience, Linux, containers and Kubernetes, benchmarking and production optimization for latency-sensitive services.
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE5+ years building HPC/GPU software, hands-on CUDA kernel authoring and optimization, C++/Python coding, GPU profiling, and experience delivering performance at scale.
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
System Software Engineer, Robot Platform — GPU & Accelerated Compute
Redwood City, California, United States
OnsiteFull Time
Sunday: Developing autonomous robots to perform household chores.
2+ YOE2+ years in GPU systems software; proficient in CUDA and a systems language (C++, C, or Rust); strong understanding of GPU architecture and time-slicing; experience with CUDA ecosystem and GPU sharing; solid Linux fundamentals.
CUDA, CUDA Graphs, CUDA IPC, Nsight Systems, Nsight Compute, NVDEC, NVENC, MPS, MIG, Linux
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yrOnsiteFull Time
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
MatX: Developing custom silicon chips optimized for large language models.
7+ YOE2+ Mgmt2+ years management, 7+ years engineering, BS in Computer Science or equivalent, experience optimizing software for specialized hardware (SIMD, parallelism, assembly, GPU/CUDA), strong communication and stakeholder alignment.
Ollama: Software platform for running open-source large language models.
Experience with systems programming (Go,C,C++), GPU or low-level performance work, profiling and optimizing real workloads, and shipping software across macOS, Linux, and Windows.
Sr. Software Development Engineer - Video Rendering
Seattle or San Francisco or San Jose
$174k-$331k/yrOnsiteFull Time
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years software engineering with deep GPU/graphics or rendering systems experience, strong modern C++, and knowledge of GPU APIs, shading languages, and performance debugging.
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
6+ YOEBachelor's in CS (or equivalent) with 6+ years engineering experience; strong coding in C/C++/C#/Java/JavaScript/Python; GPU and performance optimization experience; experience with DL frameworks and GPU tooling preferred.
Anyscale: Cloud platform for scaling distributed machine learning applications.
5+ YOE5+ years in distributed systems or software engineering; strong C/C++ and low-level OS experience; experience building scalable, fault-tolerant distributed systems; knowledge of distributed model training and inference; GPU programming preferred.
Software Engineer, ML Systems & Training Architecture
San Francisco, California, United States
$295k-$380k/yrOnsiteFull Time
OpenAI: Develops artificial intelligence models and generative AI software services.
Senior Software Engineer with ML systems, training frameworks, GPUs, and distributed systems experience; strong code review, debugging, and infrastructure skills; hands-on IC.
GPUs, Distributed systems, Training frameworks, Python, C++, Linux
8+ YOE5+ MgmtBachelor's or equivalent,8+ years software development,7+ years leading technical projects and ML infrastructure,experience with distributed systems,Linux and GPU technologies,EM and leadership experience preferred.
Waymo: Autonomous driving technology for ride-hailing and logistics.
Experience developing GPU kernels and runtimes for machine learning systems; strong ML systems background and ability to work on autonomous driving software.