8 gpu kernel development engineer jobs at 4 companies in Seattle, WA

2d
Save
Mark Applied
Hide
GPU Kernel Engineer
Redmond, Washington, United States
$60k-$149k/yr OnsiteFull Time
Wipro
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
5+ YOERequires 5–8 years of experience, strong Python and C/C++ skills, PyTorch and Triton experience, GPU kernel development, GPU architecture, performance optimization, compiler optimization, custom SDKs, and hardware abstraction layers.
GPU, Python, C++, PyTorch, Triton, CUDA, SDK
1d
Save
Mark Applied
Hide
GPU Kernel Deployment Engineer
Redmond, Washington, United States
$60k-$149k/yr OnsiteFull Time
Wipro
WiproNYSE: WIT: Global technology services and consulting for digital transformation.
5+ YOERequires 5–8 years of experience, strong Python and C/C++ skills, PyTorch and Triton experience, GPU kernel development, GPU architecture, performance optimization, compiler optimization, and runtime execution knowledge.
Python, C, C++, PyTorch, Triton
1mo
Save
Mark Applied
Hide
Senior Linux Kernel Systems Software Engineer – CSP Engagements
Santa Clara or Austin or Redmond or Seattle
$184k-$357k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
10+ YOE10+ years system software experience; expert in Linux kernel internals, device drivers, PCIe/USB/Ethernet, ARM (aarch64) and x86, kernel debugging (GDB, kdump, eBPF), C/C++, Python, virtualization, Kubernetes, NUMA and performance optimization.
GDB, kdump, eBPF, ARM (aarch64), x86, C, C++, Python, Kubernetes, CUDA, PCIe, USB, Ethernet, IOMMU, NUMA, CXL
2w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Local AI Platforms
Santa Clara or Austin or Westford or Durham or Seattle
$224k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOE12+ years software engineering experience with GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Python, C++, CUDA, Triton, Docker, OCI, NVIDIA Container Toolkit, NCCL, RCCL
2w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Local AI Platforms
Santa Clara or Westford or Austin or Durham or Seattle
$224k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years software engineering experience in GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Python, C++, CUDA, Triton, Docker, OCI, NCCL, RCCL, NVIDIA Container Toolkit
1mo
Save
Mark Applied
Hide
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin or Seattle or Cupertino
$151k-$235k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
6+ YOE6+ years systems/software development and systems design experience; strong programming in C++, C#, Java, Python, Golang, PowerShell, or Ruby; Linux/Unix experience; experience building reliable, scalable automation, diagnostics, and CI/CD for server fleets.
C++, C#, Java, Python, Golang, PowerShell, Ruby, Linux, Linux kernel, CI/CD, BMC/IPMI, PCIe, NVMe, GPU, ARM, x86
1mo
Save
Mark Applied
Hide
Embedded Linux Engineer
Seattle or Costa Mesa
$166k-$220k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
U.S. Person required. Experience with Linux kernel development, board bring-up (device trees, bootloaders, kernel drivers), uboot/EDK2, triaging and patching vulnerabilities, C or Rust, and interest in Nix/NixOS instead of Yocto/buildroot.
Linux kernel, uboot, EDK2, Nix, NixOS, Yocto, buildroot, C, Rust, CUDA, C++, Python, Go, Haskell
2mo
Save
Mark Applied
Hide
Senior Software Engineer, CUTLASS Kernels
Santa Clara or Austin or Hillsboro or Durham or Redmond
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
3+ YOEMasters/PhD (or equivalent experience), 3+ years industry experience, strong C++ and CUDA skills, experience with parallel-accelerator programming (e.g., OpenCL/HIP/SYCL), assembly-level understanding, and performance optimization experience.
CUTLASS, CUDA C++, C++, Python, CUDA, OpenCL, HIP, SYCL, Mojo, Pallas, Triton, Mosaic, Halide, PTX, NVVM, cuTile