11 gpu kernel development engineer jobs at 4 companies in Texas
2d
Save
Mark Applied
Hide
2d
Research Kernel Engineer
Singapore or Austin
OnsiteFull Time
BitdeerNASDAQ: BTDR: Operates cryptocurrency mining and high-performance computing data centers.
Requires a degree in computer science, electrical engineering, or related field; CUDA or Triton, Python, C++, GPU architecture, profiling, inference optimization, and high-performance computing experience.
Senior Linux Kernel Systems Software Engineer – CSP Engagements
Santa Clara or Austin or Redmond or Seattle
$184k-$357k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
10+ YOE10+ years system software experience; expert in Linux kernel internals, device drivers, PCIe/USB/Ethernet, ARM (aarch64) and x86, kernel debugging (GDB, kdump, eBPF), C/C++, Python, virtualization, Kubernetes, NUMA and performance optimization.
Santa Clara or Westford or Austin or Durham or Seattle
$224k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years software engineering experience in GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Santa Clara or Austin or Westford or Durham or Seattle
$224k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOE12+ years software engineering experience with GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Senior Software Engineer, AI and DL Kernel Libraries
Santa Clara or Georgia or Texas or Colorado or Washington or California or Oregon or Massachusetts
$184k-$288k/yrRemoteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOEMasters (or equivalent experience) in CS/EE, 6+ years ML/DL systems experience, strong Python and C/C++ skills, GPU kernel development experience (CUDA, Triton, cuTile), familiarity with deep learning frameworks and inference runtimes.
Sr. System Development Engineer, Edge & High Performance Accelerator Servers for AI/ML
Austin or Seattle or Cupertino
$151k-$235k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
6+ YOE6+ years systems/software development and systems design experience; strong programming in C++, C#, Java, Python, Golang, PowerShell, or Ruby; Linux/Unix experience; experience building reliable, scalable automation, diagnostics, and CI/CD for server fleets.
C++, C#, Java, Python, Golang, PowerShell, Ruby, Linux, Linux kernel, CI/CD, BMC/IPMI, PCIe, NVMe, GPU, ARM, x86
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Strong C/C++ development for embedded Linux and RTOS, computer vision and AI pipeline integration, experience with real-time and low-latency systems, familiarity with OpenCV and debugging/optimization across CPU/GPU/FPGA accelerators.
Senior Software Engineer - CUDA and Unified Memory
Santa Clara or Austin
$184k-$357k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
10+ YOEBS/MS in CS, EE or equivalent experience; 10+ years development experience; strong C skills; experience with OS interfaces, multithreading, large codebases; kernel/driver experience preferred.
Senior Software Engineer, CUDA Deep Learning Systems
Santa Clara or Austin or Texas
$184k-$357k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
8+ YOE8+ years industry experience; BS/MS/PhD in CS/CE/EE or equivalent; strong C++ and Python skills; deep learning (transformers) and distributed computing expertise; CUDA, kernel optimization, and profiling experience.
Santa Clara or Austin or Hillsboro or Durham or Redmond
$152k-$288k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
3+ YOEMasters/PhD (or equivalent experience), 3+ years industry experience, strong C++ and CUDA skills, experience with parallel-accelerator programming (e.g., OpenCL/HIP/SYCL), assembly-level understanding, and performance optimization experience.