9 gpu performance profiling engineer jobs at 5 companies in Washington
1mo
Save
Mark Applied
Hide
1mo
Senior Performance Engineer - DGX Cloud
Santa Clara or Austin or Redmond or Washington or Oregon
$224k-$431k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
12+ YOE12+ years experience, strong C++ and Python programming, foundations in OS/architecture/distributed systems, performance engineering and profiling experience, BS in CS/CE or equivalent.
Santa Clara or Austin or Redmond or Oregon or Washington
$224k-$431k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
12+ YOE12+ years experience; BS or higher in CS/CE or equivalent; strong C++ and Python skills; foundation in OS, computer architecture, distributed systems; performance engineering and profiling experience.
CoreWeaveNasdaq: CRWV: Specialized cloud provider for large-scale AI and machine learning.
5+ YOE5+ years building HPC/GPU software, hands-on CUDA kernel authoring and optimization, C++/Python coding, GPU profiling, and experience delivering performance at scale.
Fellow Software Engineer — AI Performance & Reliability
San Jose or Bellevue
$235k-$402k/yrHybridFull Time
AMDNASDAQ: AMD: Leader in high-performance computing, graphics, and visualization technologies.
PhD or equivalent in AI/ML/CS, strong software engineering, experience profiling and optimizing ML models and AI workloads, proficiency in Python/C++, ML frameworks, customer-facing troubleshooting and performance analysis.
Santa Clara or Utah or Remote or Remote or Seattle or Redmond or Salt Lake City
$152k-$242k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
5+ YOE5+ years software engineering; Python (C++ a plus); CI/CD and automation; GPU/accelerator performance analysis; experience with DL frameworks (PyTorch, TensorFlow, JAX, TensorRT); data analysis and profiling; able to debug complex systems.
Senior Deep Learning Systems Engineer, Datacenters
Santa Clara or Redmond
$184k-$357k/yrHybridFull Time
NVIDIANASDAQ: NVDA: Computing platform for AI and accelerated graphics.
8+ YOEBachelor's in EE/CS or equivalent, 8+ years experience, strong system architecture and performance analysis, programming in C++ and Python, familiarity with Linux, CUDA, DL frameworks, and profiling tools.
Veeda AI: Private Canadian AI startup building multimodal world models and simulated environments for robotics and Physical AI.
Bachelor's degree or equivalent experience in a related technical field; deep PyTorch and multi-node parallelism experience; Python and C++/CUDA fluency; training-run profiling experience; expertise in low-precision numerics, kernels, or fault diagnosis.
Anthropic: AI research developing safe and steerable AI systems.
Senior IC with deep systems or ML infrastructure experience, hands-on performance profiling and optimization, accelerator ecosystem expertise (CUDA/TPU/Trainium), strong software engineering and cross-org alignment skills, and a relevant bachelor’s degree or equivalent.