168 software engineer gpu jobs at 107 companies in Vallejo, CA

3w
Save
Mark Applied
Hide
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yr OnsiteFull Time
TypeSafe AI
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
CUDA, CuTe DSL
2mo
Save
Mark Applied
Hide
Principal Software Engineer, GPU Compute
San Mateo, California, United States
$345k-$399k/yr HybridFull Time
Roblox
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
Go, CUDA, Kubernetes, NVLink, InfiniBand, RoCE, BMC, IPMI, Redfish
1mo
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
2w
Save
Mark Applied
Hide
Staff Software Engineer, GPU Infrastructure Lifecycle Management
San Francisco, California, United States
$240k-$280k/yr OnsiteFull Time
Together AI
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Go, Python, Rust, Temporal, Cadence, Kubernetes, Kafka, NATS, SQS, PXE, iPXE, Redfish, IPMI, BMC, NCCL, CUDA, InfiniBand, RoCE
2mo
Save
Mark Applied
Hide
System Software Engineer, Robot Platform — GPU & Accelerated Compute
Redwood City, California, United States
OnsiteFull Time
Sunday
Sunday: Developing autonomous robots to perform household chores.
2+ YOE2+ years in GPU systems software; proficient in CUDA and a systems language (C++, C, or Rust); strong understanding of GPU architecture and time-slicing; experience with CUDA ecosystem and GPU sharing; solid Linux fundamentals.
CUDA, CUDA Graphs, CUDA IPC, Nsight Systems, Nsight Compute, NVDEC, NVENC, MPS, MIG, Linux
4w
Save
Mark Applied
Hide
Software Engineer- GPU Fabric Observability
San Francisco, California, United States
$200k-$380k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Kubernetes
5h
Save
Mark Applied
Hide
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yr OnsiteFull Time
Perplexity
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Kubernetes, kubectl, NVIDIA, CUDA, InfiniBand, RoCE, CoreWeave, AWS, GCP, Go, Rust, C++, vLLM, SGLang, TensorRT-LLM, Slurm, Triton, RDMA, Prometheus, Grafana, Weights & Biases
3mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
1w
Save
Mark Applied
Hide
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
Linux, SLURM, Kubernetes, Ray, Hadoop
2mo
Save
Mark Applied
Hide
Sr Software Engineer
San Mateo, California, United States
$193k-$290k/yr HybridFull Time
Sony Interactive Entertainment
Sony Interactive EntertainmentNYSE: SONY: Provides PlayStation gaming hardware, software, and digital network services.
Experience with video codecs, C/C++, GPU programming (HLSL/Vulkan/CUDA), ML/AI, and software development tools.
C, C++, HLSL, Vulkan, CUDA, Git, SVN, JIRA, IDEs
3w
Save
Mark Applied
Hide
Software Engineer, Runtime
Palo Alto, California, United States
OnsiteFull Time
Ollama
Ollama: Software platform for running open-source large language models.
Experience with systems programming (Go,C,C++), GPU or low-level performance work, profiling and optimizing real workloads, and shipping software across macOS, Linux, and Windows.
Go, C, C++, CUDA, Metal, SYCL, MLX
4w
Save
Mark Applied
Hide
Sr. Software Development Engineer - Video Rendering
Seattle or San Francisco or San Jose
$174k-$331k/yr OnsiteFull Time
Adobe
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
10+ YOE10+ years software engineering with deep GPU/graphics or rendering systems experience, strong modern C++, and knowledge of GPU APIs, shading languages, and performance debugging.
C++, DirectX 12, Metal, Vulkan, CUDA, OpenCL, HLSL, Slang, SPIR-V, DXIL, DXC, Nsight, PIX, TensorRT, ONNX Runtime
2mo
Save
Mark Applied
Hide
Software Engineer (Ray Core)
San Francisco or Palo Alto
HybridFull Time
Anyscale
Anyscale: Cloud platform for scaling distributed machine learning applications.
5+ YOE5+ years in distributed systems or software engineering; strong C/C++ and low-level OS experience; experience building scalable, fault-tolerant distributed systems; knowledge of distributed model training and inference; GPU programming preferred.
C/C++, GPU programming, Ray
2mo
Save
Mark Applied
Hide
Software Engineer, ML Systems & Training Architecture
San Francisco, California, United States
$295k-$380k/yr OnsiteFull Time
OpenAI
OpenAI: Develops artificial intelligence models and generative AI software services.
Senior Software Engineer with ML systems, training frameworks, GPUs, and distributed systems experience; strong code review, debugging, and infrastructure skills; hands-on IC.
GPUs, Distributed systems, Training frameworks, Python, C++, Linux
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Inference
Palo Alto, California, United States
$185k-$250k/yr HybridFull Time
Pika
Pika: AI-powered platform for generating and editing professional videos
5+ YOE5+ years engineering experience in inference acceleration, GPU programming (CUDA, NCCL), model deployment, quantization, attention optimization, and parallelism for production-scale AI systems.
CUDA, NCCL
1w
Save
Mark Applied
Hide
Staff Software Engineer (Cloud Infrastructure)
San Francisco, California, United States
$215k-$260k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
Hands-on experience diagnosing and repairing rack-mounted GPU systems, proficiency coding in Golang, strong Linux skills, familiarity with NVIDIA/AMD GPU platforms and high-speed interconnects, and ability to operate in fast-paced data center environments.
Golang, NVIDIA DCGM, NVIDIA NCCL, InfiniBand, NVLink, RDMA over Converged Ethernet (RoCE), Ubuntu, Rocky Linux, CentOS
1mo
Save
Mark Applied
Hide
Software Engineer – Nonlinear Solid Mechanics & High-Performance Computing
Palo Alto, California, United States
$190k-$230k/yr OnsiteFull Time
Vinci
Vinci: AI software for rapid hardware design and physics simulation.
3+ YOEExperience in computational solid mechanics, nonlinear FEM, iterative solvers, GPU programming (CUDA/HIP/SYCL), proficiency in C++ or Python, strong software engineering and CI/CD practices, and 3+ years of relevant experience.
CUDA, HIP, SYCL, C++, Python, Git, Nsight, VTune, Roofline analysis, CI/CD
3w
Save
Mark Applied
Hide
Senior Software Engineer (CVI)
San Francisco, California, United States
OnsiteFull Time
Tavus
Tavus: Develops AI models for creating emotionally intelligent digital humans.
Fluent in Python with concurrent programming experience, built real-time systems, senior-level ownership and communication skills; familiarity with GPUs, video streaming, and LLMs preferred.
Python, asyncio, multiprocessing, WebRTC, GPUs, LLMs
1mo
Save
Mark Applied
Hide
Senior Software Engineer
Berkeley, California, United States
$160k-$180k/yr HybridFull Time
Zendar
Zendar: High-resolution radar perception systems for autonomous vehicles and robotics
5+ YOE5+ years software engineering experience, proficiency in C++ and Python, experience leading features to production, strong problem decomposition and communication skills, ability to work onsite in Berkeley at least Tue–Thu.
C++, Python, JavaScript, Linux, CANbus, GPU, CI/CD, SDK, radar, lidar, camera
1mo
Save
Mark Applied
Hide
Software Engineer, AI Labs
Edinburgh or New York City or Atlanta or San Francisco or Seattle
HybridFull Time
BlackRock
BlackRockNYSE: BLK: Provides investment management and financial technology services globally.
3+ YOE3+ years professional software engineering experience; strong Python and SQL skills; experience building, testing, deploying, and operating cloud-native applications, services, APIs, or data pipelines; strong engineering fundamentals and communication.
Python, SQL, Spark, Airflow, Dagster, Flyte, GPUs, TPUs, AWS Inferentia, CI/CD, AI coding assistants