612 software engineer gpu jobs at 159 companies in Berkeley, CA

3mo
Save
Mark Applied
Hide
Senior GPU Software Performance Engineer — Post‑Training
San Jose, California, United States
$179k-$306k/yr HybridFull Time
AMD
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Senior GPU software performance engineer with strong cross-stack problem-solving skills in training workloads; experience across data loaders, kernels, distributed training, and compilers.
ROCm, HIP, Triton, PyTorch, Python, C++, AMD Instinct, Kernels, Distributed training
2mo
Save
Mark Applied
Hide
System Software Engineer - GPU
Santa Clara, California, United States
$152k-$242k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOEBS or MS in Electrical/Computer Engineering or Computer Science; 5+ years in hardware/software; strong C/C++; knowledge of PC architecture; GPU/driver/app exposure; experience with AI tooling and modern AI-assisted workflows.
C/C++, CUDA, Vulkan, PCIe, Nvlink, Infiniband, Ethernet, AI tooling
1w
Save
Mark Applied
Hide
Staff Software Engineer, GPU Inference
Toronto or Sunnyvale
HybridFull Time
Cerebras Systems
Cerebras SystemsNasdaq: CBRS: Manufactures specialized computer chips designed for AI.
8+ YOE8+ years software engineering experience, strong C++ and Python skills, GPU inference experience, Linux, containers and Kubernetes, benchmarking and production optimization for latency-sensitive services.
C++, Python, vLLM, PyTorch, ROCm, HIP, RCCL, rocprofiler, AMD SMI, AITER, hipBLASLt, Composable Kernel, CUDA, SGLang, TensorRT-LLM, Triton Inference Server, Kubernetes, Linux, RDMA
3w
Save
Mark Applied
Hide
Senior Systems Software Engineer – GPU Software
Santa Clara, California, United States
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
10+ YOE10+ years software development experience, strong C and low-level driver/kernel experience, BS/MS in Computer Engineering or Computer Science (or equivalent), Linux/Android/Windows kernel experience, system-level debugging and performance optimization.
C, Linux, Android, Chrome, Windows, XenServer, KVM, Hyper-V, RTOS
2w
Save
Mark Applied
Hide
Senior GPU Engineer
Santa Clara, California, United States
$177k-$233k/yr OnsiteFull Time
Qualcomm
QualcommNASDAQ: QCOM: Designs and manufactures semiconductors and wireless telecommunications products.
Architect, design, implement, verify, and optimize GPU cores; build functional models, develop software/tools/tests, run benchmarks, and perform pre/post-silicon verification. Master's in EE/CE/CS or related.
3w
Save
Mark Applied
Hide
GPU Kernel Engineer
San Francisco, California, United States
$180k-$280k/yr OnsiteFull Time
TypeSafe AI
TypeSafe AI: Building reliable, general frontier AI models for automation.
Deep CUDA/GPU kernel expertise, experience building and optimizing training and inference kernels, LLM training experience, profiling and eliminating performance bottlenecks.
CUDA, CuTe DSL
2mo
Save
Mark Applied
Hide
Principal Software Engineer, GPU Compute
San Mateo, California, United States
$345k-$399k/yr HybridFull Time
Roblox
RobloxNYSE: RBLX: Platform for creating and playing user-generated 3D digital experiences.
10+ YOE10+ years building large-scale distributed systems; deep GPU and accelerator expertise; experience with driver/firmware lifecycle, CUDA, GPU scheduling, Kubernetes; strong Go proficiency and technical leadership.
Go, CUDA, Kubernetes, NVLink, InfiniBand, RoCE, BMC, IPMI, Redfish
3w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Kernel Authoring & Optimization
Sunnyvale or Bellevue
$182k-$242k/yr OnsiteFull Time
CoreWeave
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE5+ years building HPC/GPU software, hands-on CUDA kernel authoring and optimization, C++/Python coding, GPU profiling, and experience delivering performance at scale.
CUDA, Nsight Compute, Nsight Systems, C++, Python, Triton, Mojo, CuTe DSL, JAX, HIP, ROCm, NCCL, Kubernetes, SUNK, Slurm, MLPerf, vLLM, TensorRT-LLM, llm-d, SGLang, KNYFE, Pallas, CUTLASS
1mo
Save
Mark Applied
Hide
Sr. Staff/Principal Engineer — GPU Driver & Systems Software
San Diego or San Jose
$179k-$286k/yr OnsiteFull Time
MediaTek
MediaTekTaiwan Stock Exchange: 2454: Designs and develops system-on-chip solutions for electronic devices.
10+ YOE10+ years building GPU drivers or low-level systems software; fluent C/C++; deep knowledge of Vulkan/DirectX/OpenCL/OpenGL ES and GPU internals; experience with kernel-mode drivers, firmware, performance optimization.
C, C++, Vulkan, DirectX, OpenGL ES, OpenCL
1mo
Save
Mark Applied
Hide
Software Engineer, Compute (GPU)
San Francisco or New York or Austin or Seattle
$175k-$300k/yr OnsiteFull Time
Fluidstack
Fluidstack: Provides high-performance cloud GPU infrastructure for AI development.
Experience building automation and observability for large GPU fleets; familiarity with firmware/BMC/Redfish/IPMI, Kubernetes, metrics and alerting, and production automation; comfortable with incident response and on-call.
Kubernetes, Redfish, BMC, IPMI, Temporal, Cadence, Prometheus, Grafana, Go, Python, LLM APIs, Claude Code, Cursor
2w
Save
Mark Applied
Hide
Staff Software Engineer, GPU Infrastructure Lifecycle Management
San Francisco, California, United States
$240k-$280k/yr OnsiteFull Time
Together AI
Together AI: Cloud platform for training and deploying artificial intelligence models.
Strong software engineering experience with Go, Python, or Rust; durable workflow orchestration (Temporal/Cadence); control-plane/orchestration and event-driven system experience; product mindset building internal platforms.
Go, Python, Rust, Temporal, Cadence, Kubernetes, Kafka, NATS, SQS, PXE, iPXE, Redfish, IPMI, BMC, NCCL, CUDA, InfiniBand, RoCE
2mo
Save
Mark Applied
Hide
System Software Engineer, Robot Platform — GPU & Accelerated Compute
Redwood City, California, United States
OnsiteFull Time
Sunday
Sunday: Developing autonomous robots to perform household chores.
2+ YOE2+ years in GPU systems software; proficient in CUDA and a systems language (C++, C, or Rust); strong understanding of GPU architecture and time-slicing; experience with CUDA ecosystem and GPU sharing; solid Linux fundamentals.
CUDA, CUDA Graphs, CUDA IPC, Nsight Systems, Nsight Compute, NVDEC, NVENC, MPS, MIG, Linux
1w
Save
Mark Applied
Hide
GPU Image Processing Framework Software Engineer
Cupertino, California, United States
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience developing GPU-centric image processing frameworks for iOS and macOS and collaborating across teams to deliver imaging features.
Core Image, iOS, macOS
4w
Save
Mark Applied
Hide
Software Engineer- GPU Fabric Observability
San Francisco, California, United States
$200k-$380k/yr HybridFull Time
Baseten
Baseten: Scalable infrastructure platform for deploying and serving AI models.
Staff-level experience building production infrastructure software, strong distributed systems and telemetry pipeline background, networking and high-performance network knowledge, experience processing high-volume operational data.
Kubernetes
8h
Save
Mark Applied
Hide
Member of Technical Staff (Software Engineer, GPU Cluster Infrastructure)
San Francisco or Seattle or New York City or United States
$250k-$485k/yr OnsiteFull Time
Perplexity
Perplexity: AI-powered search engine providing conversational answers with citations.
Deep Kubernetes and GPU cluster experience, multi-cloud orchestration, strong distributed systems fundamentals, systems-level coding in Go/Rust/C++, and experience with training and inference workloads.
Kubernetes, kubectl, NVIDIA, CUDA, InfiniBand, RoCE, CoreWeave, AWS, GCP, Go, Rust, C++, vLLM, SGLang, TensorRT-LLM, Slurm, Triton, RDMA, Prometheus, Grafana, Weights & Biases
3mo
Save
Mark Applied
Hide
Infrastructure Engineer (GPU & Compute)
New York or San Francisco or Seattle
$180k-$200k/yr RemoteFull Time
Lightning AI
Lightning AI: Unified platform to build, train, and deploy AI models.
5+ YOE5+ years in infrastructure or systems engineering; strong Linux in production; GPU hardware and software experience; bare-metal provisioning; Python automation; debugging across hardware/OS/GPU.
NVIDIA DCGM, PXE, IPMI, Redfish, iDRAC, LiveCD, InfiniBand, NVLink, Linux, Python
1w
Save
Mark Applied
Hide
Member of Technical Staff - GPU Infrastructure Engineer
San Francisco, California, United States
OnsiteFull Time
Liquid AI
Liquid AI: Develops efficient general-purpose artificial intelligence foundation models.
Strong software engineering with production infrastructure tooling, deep distributed systems/Linux/networking/storage knowledge, experience operating shared compute clusters and supporting production users.
Linux, SLURM, Kubernetes, Ray, Hadoop
2mo
Save
Mark Applied
Hide
Sr Software Engineer
San Mateo, California, United States
$193k-$290k/yr HybridFull Time
Sony Interactive Entertainment
Sony Interactive EntertainmentNYSE: SONY: Provides PlayStation gaming hardware, software, and digital network services.
Experience with video codecs, C/C++, GPU programming (HLSL/Vulkan/CUDA), ML/AI, and software development tools.
C, C++, HLSL, Vulkan, CUDA, Git, SVN, JIRA, IDEs
3w
Save
Mark Applied
Hide
Lead Software Engineer - Kernels
Mountain View, California, United States
$120k-$600k/yr HybridFull Time
MatX
MatX: Developing custom silicon chips optimized for large language models.
7+ YOE2+ Mgmt2+ years management, 7+ years engineering, BS in Computer Science or equivalent, experience optimizing software for specialized hardware (SIMD, parallelism, assembly, GPU/CUDA), strong communication and stakeholder alignment.
assembly, C++, C, Zig, Rust, SIMD, GPU, CUDA, AllReduce, AllToAll
3w
Save
Mark Applied
Hide
Software Engineer, Runtime
Palo Alto, California, United States
OnsiteFull Time
Ollama
Ollama: Software platform for running open-source large language models.
Experience with systems programming (Go,C,C++), GPU or low-level performance work, profiling and optimizing real workloads, and shipping software across macOS, Linux, and Windows.
Go, C, C++, CUDA, Metal, SYCL, MLX