1,996 ai infrastructure engineer jobs at 931 companies in United States
2d
Save
Mark Applied
Hide
2d
AI Infrastructure Engineer
New York City, New York, United States
$215k-$350k/yrHybridFull Time
FortinetNASDAQ: FTNT: Cybersecurity providing integrated network security and protection solutions.
Linux infrastructure administration, KVM or containerization, DevOps toolchains, Python or Bash automation, GPU-based compute or AI/ML infrastructure; bachelor's degree preferred.
TruistNYSE: TFC: US-based banking and financial services institution.
5+ YOEBachelor’s degree in computer science, engineering, information systems, or related field; 5+ years of infrastructure engineering experience; strong enterprise infrastructure technology knowledge.
Sciforium: AI infrastructure building multimodal models and high-efficiency serving software for developers and teams.
5+ YOE5+ years in systems or infrastructure engineering with GPU, HPC, or ML infrastructure experience; technical bachelor's or master's degree; Linux, Kubernetes, schedulers, configuration management, Python, Bash, containers, GPUs, and RDMA expertise.
Together AI: Research-driven AI cloud infrastructure provider offering inference, fine-tuning, GPU clusters, and model training to developers and enterprises.
5+ YOE5+ years in AI infrastructure or related roles; BS in CS or equivalent; knowledge of Ansible, Terraform, Kubernetes; programming/scripting; monitoring/observability; cloud services; collaborative work
Seekr Technologies: Private American enterprise AI providing explainable, secure AI software and hardware to government and critical-infrastructure customers.
8+ YOE8+ years building distributed systems and cloud-native AI infrastructure; strong Python and systems programming skills; Kubernetes, GPU inference, and platform engineering experience; leadership and architecture experience.
Palona AI: AI platform helping restaurants capture demand, convert revenue, and manage operations through voice, text, and visual agents.
3+ YOE3+ years in a relevant technical domain, distributed systems and software engineering experience, cloud platform expertise, infrastructure automation, production debugging, and Python or another modern programming language.
VideoAmp: VideoAmp is a privately held media measurement software helping advertisers, agencies, and publishers plan, optimize, and measure campaigns.
6+ YOE6+ years software engineering with 1+ years in AI/ML infrastructure or LLM platform work; strong Go, Python, SQL; experience with multi-tenant APIs, on-call production operations, LLM APIs, and agentic systems.
ZoomNasdaq Global Select Market: ZM: American publicly traded communications platform serving businesses and individuals with AI-assisted video, voice, chat, and phone services.
5+ YOEBachelor's in CS/Engineering/AI, 5+ years software engineering experience in infrastructure and distributed systems, GPU/CUDA expertise, container and cloud experience, Python/C++ programming, PyTorch and Transformers experience.
International Materials: Privately owned bulk raw-materials trader and logistics provider serving global cement, construction, steel, wallboard, and energy industries.
5+ YOERequires 5+ years in software, cloud infrastructure, platform engineering, MLOps, or DevOps; production application experience; cloud platform expertise; AI/LLM experience; API integration and security knowledge.
AWS, Microsoft Azure, GCP, CI/CD, APIs, Docker, Kubernetes, Terraform, Model Context Protocol (MCP), Anthropic Claude, OpenAI, Google Gemini, Hugging Face, LangChain, LangGraph, SOC 2, LLMs, RAG
Scout Motors: Reviving the iconic Scout brand with American-made electric vehicles.
8+ YOE8+ years in LLM/MLOps and MLOps platforms; experience with Terraform, Docker, cloud (AWS), Kubernetes, CI/CD, Python/PySpark/Bash; production AI/platform deployment and security best practices.
AccentureNYSE: ACN: Global professional services delivering 360° value.
5+ YOE5+ years designing, deploying, and managing AI/HPC infrastructure; experience with accelerated computing, cluster management, orchestration, networking, storage, and automation; bachelor’s degree or equivalent.
Thinking Machines Lab: Private AI research and product building customizable multimodal systems for researchers and the wider public.
4+ YOE4+ years operating large-scale distributed systems in production; expertise in Linux, networking, complex failure debugging, and Python, Go, or C++; on-call experience required.
NIONYSE: NIO: Global smart electric vehicle and battery technology provider.
5+ YOERequires 5+ years building AI inference systems, LLM/VLM internals, performance engineering, GPU/NPU programming, C/C++, systems programming, and a BS/MS in computer science, computer engineering, or a related field.
Large Language Models (LLMs), Vision-Language Models (VLMs), GPU, NPU, DSP, CUDA, PyTorch, TensorFlow, C, C++, AIOS
Interactive Process Technology LLC: Service-disabled veteran-owned IT contractor delivering process automation, software, cloud, cybersecurity, and enterprise solutions to federal organizations.
10+ YOEBachelor's in CS/Engineering/Data Science/Math,10+ years enterprise IT experience, experience with AI/ML applications, familiarity with LLM patterns, strong communication, and active security clearance.
HR Certification Institute: Independent nonprofit credentialing and learning organization serving human resource professionals and their employers worldwide.
3+ YOE3+ years software engineering and Azure experience, CI/CD and DevOps with Azure DevOps/GitHub, containerization, Python scripting, production AI coding agent experience, security-focused instincts, and ability to communicate across teams.
Escalent Group: AI-enabled market research and advisory firm helping brands understand human and market behavior and navigate disruption.
5+ YOE5+ years infrastructure/cloud/DevOps experience with AWS and Azure, Terraform, Docker, CI/CD, monitoring, Linux, scripting, and cloud security/IAM.
Echelon: AI business-operations platform connecting enterprise HR, finance, and operations data for decision-makers.
5+ YOERequires 5+ years building production infrastructure, distributed systems, developer platforms, or execution runtimes; multi-tenant workloads; strong Linux and programming skills; Kubernetes or comparable scheduler; AWS or Azure experience.