7 inference engineer jobs at 3 companies in Danville, VA

18h
Save
Mark Applied
Hide
Software Engineer 3 - Inference
Vancouver or San Jose or Durham or Mexico City or Bangalore or Pune or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
$132k-$198k/yr HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
4+ YOERequires 4–7 years developing resilient software, strong algorithms, Docker, Kubernetes, Go, Python, CI/CD, distributed systems, datacenter architecture, and a computer science bachelor's or master's degree.
Kubernetes, Docker, Go, Python, TensorFlow, PyTorch, CI/CD, GPU
3w
Save
Mark Applied
Hide
Software Engineer 4 - LLM Inference
San Jose or Durham or Mexico or Canada or India or Netherlands or Serbia or Spain or Singapore or Australia or Japan
$148k-$222k/yr HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
6+ YOE6+ years building distributed, high-performance cloud-native systems; strong Go/Python, Docker, Kubernetes, CI/CD, systems and networking knowledge; bachelor’s or master’s in CS or equivalent.
Docker, Kubernetes, Go, Python, CI/CD, TensorFlow, PyTorch
3w
Save
Mark Applied
Hide
Senior Software Engineer - LLM Inference
San Jose or Durham or Mexico City or Vancouver or Bengaluru or Pune or Hoofddorp or Belgrade or Barcelona or Singapore or Sydney or Tokyo
$171k-$257k/yr HybridFull Time
Nutanix
NutanixNASDAQ: NTNX: Sells cloud software and hyperconverged infrastructure for enterprises.
8+ YOE8+ years building distributed, high-performance systems; strong Go/Python, Docker, Kubernetes, CI/CD; knowledge of datacenter, OS internals, virtualization, and ML frameworks.
Docker, Kubernetes, Go, Python, CI/CD, TensorFlow, PyTorch, GPUs
1w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Local AI Platforms
Santa Clara or Austin or Westford or Durham or Seattle
$224k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOE12+ years software engineering experience with GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Python, C++, CUDA, Triton, Docker, OCI, NVIDIA Container Toolkit, NCCL, RCCL
1w
Save
Mark Applied
Hide
Senior Software Engineer - GPU Local AI Platforms
Santa Clara or Westford or Austin or Durham or Seattle
$224k-$431k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOE12+ years software engineering experience in GPU computing or ML systems, strong Python or C++ skills, GPU kernel optimization (CUDA/Triton), container engineering, and LLM inference knowledge.
Python, C++, CUDA, Triton, Docker, OCI, NCCL, RCCL, NVIDIA Container Toolkit
1mo
Save
Mark Applied
Hide
Senior Deep Learning Framework Communications Engineer
Santa Clara or Austin or Westford or Durham or United States
$152k-$288k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOE5+ years software engineering experience in HPC/AI, experience with PyTorch/JAX and inference engines, Python/C++/CUDA development, performance benchmarking and profilers, understanding of multi-GPU communication and compilers.
PyTorch, TRT-LLM, vLLM, SGLang, JAX, NCCL, NVSHMEM, GPUDirect, torch.compile, PyTorch profiler, NVIDIA Nsight Systems, Python, C++, CUDA, Triton, cuTe, MPI
1mo
Save
Mark Applied
Hide
Director of Product Management – AI Essentials
Spring or San Jose or Durham or Fort Collins or Andover
$170k-$413k/yr HybridFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Provides global edge-to-cloud technology solutions and IT infrastructure services.
15+ YOE5+ MgmtBachelor's in CS/engineering required; 15+ years product management experience with 5+ years AI/ML product leadership; experience building AI/ML platforms, inference/model serving, GPU ecosystem, executive communication, and GTM strategy.
GreenLake, GenAI, GPU, agent frameworks, AI tools, MLOps, DevOps, SaaS