2 inference optimization engineer jobs at 1 company in Schenectady, NY
5d
Save
Mark Applied
Hide
5d
NVIDIA GPU AI SME
Albany, New York, United States
$150k-$164k/yrRemoteFull Time
LTIMindtreeNational Stock Exchange of India: LTIM: Global technology consulting and digital solutions.
8+ YOERequires 8+ years in infrastructure or ML engineering, hands-on NVIDIA GPU operations, Kubernetes GPU workloads, model serving, GPU scheduling and partitioning, autoscaling, and inference optimization.
LTIMindtreeNational Stock Exchange of India: LTIM: Global technology consulting and digital solutions.
8+ YOERequires 8+ years in infrastructure or ML engineering, hands-on NVIDIA GPU operations, Kubernetes GPU workloads, model serving, GPU scheduling and partitioning, KEDA autoscaling, and inference performance optimization.
Amazon Web Services (AWS), Amazon Elastic Kubernetes Service (EKS), AWS CloudFormation, NVIDIA AI Enterprise (NVAIE), NVIDIA GPU Operator, CUDA, Data Center GPU Manager (DCGM), NIM, NVIDIA Dynamo, OpenAI-compatible API, RunAI, Kubernetes, KEDA, NVIDIA Triton Inference Server, TensorRT-LLM, vLLM, NVIDIA Multi-Instance GPU (MIG), Amazon Outposts