435 ai ml infrastructure engineer jobs at 228 companies in United States
🚀PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
AccentureNYSE: ACN: Global provider of management consulting and technology services.
7.5+ YOE7.5+ years experience building and operating ML/AI platform infrastructure; must-have skills: Machine Learning Operations, DevOps, LLMs; experience with CI/CD, GenAI, cloud AI services; team leadership and platform observability experience.
Machine Learning Operations, DevOps, Large Language Models (LLMs), CI/CD, GenAI, Cloud AI services
Vultr: Global cloud infrastructure and bare metal server hosting provider.
5+ YOE5+ years in bare metal infrastructure, GPU platforms, Linux, and automation with Python/Bash; hardware, BIOS/firmware, and driver experience; leadership experience.
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
5+ YOE5+ years in DevOps/platform/infrastructure engineering; deep Kubernetes experience; experience building developer-facing platforms, Helm and GitOps workflows, storage/networking for GPU workloads; Terraform, monitoring, and ML framework exposure preferred.
Booz Allen HamiltonNYSE: BAH: Provides technology and management consulting services to diverse organizations.
5+ YOE5+ years DevOps/DevSecOps experience automating CI/CD and infrastructure, scripting with Python or Linux Shell, containerization (Kubernetes/Docker), troubleshooting pipelines, AI/ML ops experience, bachelor's degree, ability to obtain Secret clearance.
Python, Linux Shell Script, Kubernetes, Docker, Jenkins, CI/CD, AIOps, ML Ops, model ops, data ops, AWS, Azure, Google Cloud Platform, ML flow
San Francisco or Minneapolis or Washington, D.C. or United States
$120k-$215k/yrRemoteFull Time
UnitedHealth GroupNYSE: UNH: Provides health insurance and technology-enabled health care services.
4+ YOE2+ MgmtBachelor's degree or 4+ years equivalent, 4+ years Python, 4+ years cloud infrastructure (AWS/Azure/GCP), 4+ years AI/ML infrastructure experience, 2+ years team lead, 1+ year LLM experience.
Python, AWS, Azure, GCP, Large Language Models (LLMs), GitHub, GitHub Actions, Docker, Terraform, CI/CD
Utilidata: Provides edge AI for power grid and data center optimization.
5+ YOE5+ years software engineering with a focus on AI infrastructure, ML model serving, distributed systems, and GPU-based deployments; Python experience; Kubernetes/Docker; travel up to 10%.
Delta Air LinesNYSE: DAL: Major US-based airline providing passenger and cargo air transport.
Lead AI/ML initiatives including identifying business problems for AI, defining frameworks and infrastructure strategy, building POCs/prototypes and reusable models, and promoting an AI-first mindset.
IPT Associates: Provides IT and professional services to federal government agencies.
10+ YOEBachelor's in CS/Engineering/Data Science/Math,10+ years enterprise IT experience, experience with AI/ML applications, familiarity with LLM patterns, strong communication, and active security clearance.
Seekr: Transparent AI platform for enterprise and government decision-making.
5+ YOE5–8 years building distributed systems or cloud/platform services; production Kubernetes and ML infra experience; strong software engineering in Python and Go/Rust/C++; familiarity with GPU inference, model serving frameworks, and cloud platforms.
GameStopNYSE: GME: Video game, consumer electronics, and collectibles retail.
9+ YOE9+ years software development; dynamic pricing and real-time data pipelines; web scraping; AWS, Kubernetes, CI/CD; Java/Spring; e-commerce/retail tech.
GuidewireNYSE: GWRE: Provides a software platform for property and casualty insurers.
10+ YOE10+ years software engineering; 5+ years ML platforms/infrastructure; distributed systems; Python/Go/Java; Docker/Kubernetes; MLOps tools; cloud experience; knowledge of ML models.
Artificial Analysis: Independent AI benchmarking and performance analysis platform.
3+ YOE3+ years software engineering experience; strong Python, pandas and OpenAI API proficiency; experience with data-intensive backend systems, data visualization, cloud infrastructure, and leading projects end-to-end.
Docker: Provides a platform for building, sharing, and running containerized applications.
5+ YOE5+ years applied ML/AI experience, 4+ years software engineering, experience with LLM-based systems, model lifecycle and ML infrastructure, bachelor's in CS/Engineering or equivalent, strong communication and mentoring skills.
Senior Applied AI and AI Infrastructure Engineer - Chip Design and DFX
Santa Clara, California, United States
$200k-$380k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOEAdvanced degree or equivalent experience in engineering with 6+–12+ yrs experience in AI infrastructure, applied ML/GenAI; experience with agents, SQL/ETL/data modeling, cloud (AWS/Azure/GCP), Python and C++; strong communication.
8+ YOEBachelor's in CS or equivalent, 8+ years software development, 5+ years building large-scale infrastructure, 3+ years software design, experience with ML infrastructure; leadership and deep learning framework experience preferred.
Senior Applied AI and AI Infrastructure Engineer - Chip Design and DFX
Santa Clara, California, United States
$200k-$380k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOEBSEE/MSEE/PhD with 12+/10+/6+ years experience in AI infrastructure, applied ML and Gen AI. Experience with SQL, ETL, data modeling, cloud (AWS/Azure/GCP), Python, C++, distributed systems, and mentoring.
Senior AI Infrastructure Engineer - Model Training
Mountain View, California, United States
$190k-$260k/yrOnsiteFull Time
Kodiak RoboticsNASDAQ: KDK: Develops autonomous driving technology for commercial trucking and defense.
2+ YOEDegree in CS or related field,2+ years ML systems experience,expertise in distributed training,high-performance data pipelines,GPU performance and profiling,Python and PyTorch skills.
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
5+ YOE5+ years in high-scale infrastructure or ML systems; BS in Computer Science or related field; strong Python and PyTorch; Kubernetes; cloud experience; GPU/ML infra focus.