2 inference optimization engineer jobs at 1 company in Albany, NY

5d
Save
Mark Applied
Hide
NVIDIA GPU AI SME
Albany, New York, United States
$150k-$164k/yr RemoteFull Time
LTIMindtree
LTIMindtreeNational Stock Exchange of India: LTIM: Global technology consulting and digital solutions.
8+ YOERequires 8+ years in infrastructure or ML engineering, hands-on NVIDIA GPU operations, Kubernetes GPU workloads, model serving, GPU scheduling and partitioning, autoscaling, and inference optimization.
AWS EKS, AWS CloudFormation, NVIDIA AI Enterprise, NVIDIA GPU Operator, CUDA, DCGM, NIM microservices, NVIDIA Dynamo, RunAI, KEDA, OpenAI-compatible API, Triton, TensorRT-LLM, vLLM, Kubernetes, MIG, AWS Outposts
5d
Save
Mark Applied
Hide
NVIDIA GPU AI SME
Albany, New York, United States
$150k-$164k/yr RemoteFull Time
LTIMindtree
LTIMindtreeNational Stock Exchange of India: LTIM: Global technology consulting and digital solutions.
8+ YOERequires 8+ years in infrastructure or ML engineering, hands-on NVIDIA GPU operations, Kubernetes GPU workloads, model serving, GPU scheduling and partitioning, KEDA autoscaling, and inference performance optimization.
Amazon Web Services (AWS), Amazon Elastic Kubernetes Service (EKS), AWS CloudFormation, NVIDIA AI Enterprise (NVAIE), NVIDIA GPU Operator, CUDA, Data Center GPU Manager (DCGM), NIM, NVIDIA Dynamo, OpenAI-compatible API, RunAI, Kubernetes, KEDA, NVIDIA Triton Inference Server, TensorRT-LLM, vLLM, NVIDIA Multi-Instance GPU (MIG), Amazon Outposts