5 gpu performance profiling engineer jobs at 2 companies in Oregon

1mo
Save
Mark Applied
Hide
Senior Performance Engineer - DGX Cloud
Santa Clara or Austin or Redmond or Washington or Oregon
$224k-$431k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
12+ YOE12+ years experience, strong C++ and Python programming, foundations in OS/architecture/distributed systems, performance engineering and profiling experience, BS in CS/CE or equivalent.
C++, Python, CUDA, PyTorch, JAX, XLA
1mo
Save
Mark Applied
Hide
Senior Performance Engineer - DGX Cloud
Santa Clara or Austin or Redmond or Oregon or Washington
$224k-$431k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
12+ YOE12+ years experience; BS or higher in CS/CE or equivalent; strong C++ and Python skills; foundation in OS, computer architecture, distributed systems; performance engineering and profiling experience.
C++, Python, CUDA, PyTorch, JAX, XLA
2w
Save
Mark Applied
Hide
Principal Software Engineer, E2E Performance and Goodput — CSP Engagements
Santa Clara or Austin or California or Oregon
$272k-$431k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
15+ YOE15+ years in systems performance engineering; BS or MS in computer science, computer engineering, or related field; GPU profiling, distributed training, statistical analysis, Python, data visualization, and cross-team technical influence.
STREAM, GPU Burn, GPU BLAST, CUDA, NCCL, nsight systems, nsight compute, DCGM, Python, pandas, Megatron-LM, DeepSpeed, FSDP, DGX, HGX, NVLink, vLLM, TensorRT-LLM, SGLang
2mo
Save
Mark Applied
Hide
Principal Software Engineer, E2E Performance and Goodput — CSP Engagements
Santa Clara or Austin or Oregon or California
$272k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
15+ YOE15+ years systems performance engineering experience, GPU workload profiling, distributed training performance expertise, statistical analysis, Python/pandas, strong communication and cross-team influence.
Nsight Systems, Nsight Compute, DCGM, CUDA, NCCL, STREAM, GPU Burn, GPU BLAST, Python, pandas, Megatron-LM, DeepSpeed, FSDP, DGX, HGX, NVLink, TensorRT-LLM, vLLM, SGLang
1mo
Save
Mark Applied
Hide
Sr. Inference Optimization Engineer (local / edge runtime)
Santa Clara or Hillsboro or Folsom or Phoenix
$195k-$361k/yr HybridFull Time
Intel
IntelNASDAQ: INTC: Design and manufacture of semiconductors and computing technology.
8+ YOE8+ years software development; strong C++ and/or Python; experience with LLM inference, profiling and optimizing CPU/GPU performance; Linux and low-level debugging expertise.
C++, Python, llama.cpp, vLLM, ggml, Vulkan, SYCL, oneAPI, CUDA, Metal, SIMD, Linux, GGUF, AWQ, GPTQ