4 inference engineer jobs at 3 companies in Colorado

1w
Save
Mark Applied
Hide
Senior Engineer, Inference Data Plane
Denver or Seattle
$139k-$174k/yr RemoteFull Time
DigitalOcean
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
Design and deliver high-scale, resilient data-plane services for inference; experience with LLM serving, distributed systems, and AI hardware; expert in Go or Python and gRPC; familiarity with inference frameworks and observability.
llm-d, vLLM, SGLang, TensorRT, NVIDIA Dynamo, Ray Serve, KServe, TensorRT-LLM, TGI, Modular MAX, gRPC, GoLang, Python, Kubernetes
1w
Save
Mark Applied
Hide
Staff Engineer, Inference Optimizations
Denver, Colorado, United States
$191k-$239k/yr RemoteFull Time
DigitalOcean
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
5+ YOE5+ years in high-performance computing or AI infrastructure, GPU architecture expertise, CUDA/Triton experience, attention-layer and kernel optimization, system design and low-level GPU programming.
AITER, CUDA, ROCm, TensorRT, OpenAI Triton, Triton
2mo
Save
Mark Applied
Hide
Senior Software Engineer, AI and DL Kernel Libraries
Santa Clara or Georgia or Texas or Colorado or Washington or California or Oregon or Massachusetts
$184k-$288k/yr RemoteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOEMasters (or equivalent experience) in CS/EE, 6+ years ML/DL systems experience, strong Python and C/C++ skills, GPU kernel development experience (CUDA, Triton, cuTile), familiarity with deep learning frameworks and inference runtimes.
PyTorch, JAX, TensorFlow, ONNX, vLLM, SGLang, MLC, Python, C/C++, CUDA C/C++, cuTile, Triton, FlashInfer, Flash Attention, Apache TVM, MLIR
1mo
Save
Mark Applied
Hide
Director of Product Management – AI Essentials
Spring or San Jose or Durham or Fort Collins or Andover
$170k-$413k/yr HybridFull Time
Hewlett Packard Enterprise
Hewlett Packard EnterpriseNYSE: HPE: Provides global edge-to-cloud technology solutions and IT infrastructure services.
15+ YOE5+ MgmtBachelor's in CS/engineering required; 15+ years product management experience with 5+ years AI/ML product leadership; experience building AI/ML platforms, inference/model serving, GPU ecosystem, executive communication, and GTM strategy.
GreenLake, GenAI, GPU, agent frameworks, AI tools, MLOps, DevOps, SaaS