115 software engineer gpu jobs at 77 companies in Petaluma, CA
3w
Save
Mark Applied
Hide
3w
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yrOnsiteFull Time
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yrOnsiteFull Time
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
Sr. Software Development Engineer - Video Rendering
Seattle or San Francisco or San Jose
$174k-$331k/yrOnsiteFull Time
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years software engineering with deep GPU/graphics or rendering systems experience, strong modern C++, and knowledge of GPU APIs, shading languages, and performance debugging.
Anyscale: Cloud platform for scaling distributed machine learning applications.
5+ YOE5+ years in distributed systems or software engineering; strong C/C++ and low-level OS experience; experience building scalable, fault-tolerant distributed systems; knowledge of distributed model training and inference; GPU programming preferred.
Software Engineer, ML Systems & Training Architecture
San Francisco, California, United States
$295k-$380k/yrOnsiteFull Time
OpenAI: Develops artificial intelligence models and generative AI software services.
Senior Software Engineer with ML systems, training frameworks, GPUs, and distributed systems experience; strong code review, debugging, and infrastructure skills; hands-on IC.
GPUs, Distributed systems, Training frameworks, Python, C++, Linux
Sony Interactive EntertainmentNYSE: SONY: Provides PlayStation gaming hardware, software, and digital network services.
Strong systems programming in C/C++; multithreading and low-level programming; GPU/graphics experience (HLSL, Vulkan, CUDA); debugging/profiling; experience with IDEs, git/SVN, and bug tracking; bachelor’s degree or equivalent experience.
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
Hands-on experience diagnosing and repairing rack-mounted GPU systems, proficiency coding in Golang, strong Linux skills, familiarity with NVIDIA/AMD GPU platforms and high-speed interconnects, and ability to operate in fast-paced data center environments.
Zendar: High-resolution radar perception systems for autonomous vehicles and robotics
5+ YOE5+ years software engineering experience, proficiency in C++ and Python, experience leading features to production, strong problem decomposition and communication skills, ability to work onsite in Berkeley at least Tue–Thu.
Tavus: Develops AI models for creating emotionally intelligent digital humans.
Fluent in Python with concurrent programming experience, built real-time systems, senior-level ownership and communication skills; familiarity with GPUs, video streaming, and LLMs preferred.
Edinburgh or New York City or Atlanta or San Francisco or Seattle
HybridFull Time
BlackRockNYSE: BLK: Provides investment management and financial technology services globally.
3+ YOE3+ years professional software engineering experience; strong Python and SQL skills; experience building, testing, deploying, and operating cloud-native applications, services, APIs, or data pipelines; strong engineering fundamentals and communication.
Chef Robotics: AI-powered robotic systems for automated food manufacturing.
Experience building Linux-based systems in C/C++/Rust, deploying software to networked edge devices, ROS 2/DDS familiarity, strong engineering fundamentals, and debugging low-level systems.
Gradient Robotics: Developing autonomous humanoid robots for data center industrial environments.
5+ YOE5+ years building production software close to hardware (drivers,kernels,embedded,robotics); experience with Linux kernel, multithreading, timing and concurrency debugging; able to work on-site in San Francisco.