119 ml infrastructure engineer jobs at 50 companies in Washington

PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
Python, PyTorch, vLLM, SGLang, TensorRT, LLMs
PromotedHiringCafe
Founding Machine Learning / AI Search Engineer
Cupertino, CA, US
$160k-$310k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
Python, PyTorch, Elasticsearch, LLMs
PromotedHiringCafe
Founding Backend / Infra Engineer
Cupertino, CA, US
$160k-$300k/yr On-SiteFull Time
HiringCafe
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
Node.js, Python, Elasticsearch, Redis
3mo
Save
Mark Applied
Hide
Senior ML Infrastructure Engineer, Proactive
Cupertino or Seattle
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience designing large-scale ML/AI platforms; proficiency in ML systems, LLMs, and distributed data; strong multi-threaded programming skills.
Python, Distributed systems, MacOS/iOS development, LLMs, Machine learning platforms
1mo
Save
Mark Applied
Hide
ML Engineer
Palo Alto or Seattle or Paris
$139k-$226k/yr RemoteFull Time
Docker
Docker: Provides a platform for building, sharing, and running containerized applications.
5+ YOE5+ years applied ML/AI experience, 4+ years software engineering, experience with LLM-based systems, model lifecycle and ML infrastructure, bachelor's in CS/Engineering or equivalent, strong communication and mentoring skills.
Docker Desktop, Docker Hub, Docker Scout, LLM, MCP, Agentic Platform
1w
Save
Mark Applied
Hide
ML Infrastructure engineer - Early Career
Mountain View or Washington
$143k-$186k/yr OnsiteFull Time
Unity
UnityNYSE: U: Provides software for creating real-time 3D interactive content.
Bachelor's degree in CS or related, strong foundation in ML systems or distributed systems, experience with Python, familiarity with ML frameworks and distributed tools, and ability to build scalable ML infrastructure.
Python, PyTorch, Ray, Airflow, Flyte, TensorFlow, Spark
2w
Save
Mark Applied
Hide
Sr. Mgr., ML Infrastructure, PV Personalization and Discovery
Sunnyvale or Seattle or New York
$242k-$328k/yr OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
10+ YOE5+ Mgmt10+ years engineering experience, 5+ years managing engineering teams, expertise in ML infrastructure, retrieval/recommendation systems, LLMs/generative AI, partnering with applied scientists, experience delivering large-scale consumer software.
AWS
1mo
Save
Mark Applied
Hide
Staff ML/LLM Ops Engineer
Seattle, Washington, United States
$213k-$272k/yr OnsiteFull Time
LiveView Technologies
LiveView Technologies: Provides mobile, solar-powered security units with AI-driven surveillance software.
8+ YOE8+ years in ML-infrastructure/MLOps building and operating model deployment, serving, CI/CD, monitoring, and LLM/VLM ops; strong API design and technical leadership; BS/MS in CS/Engineering or equivalent experience.
Kubernetes, Argo, LangGraph, MCP, NVIDIA Jetson, vector databases
6d
Save
Mark Applied
Hide
Principal Core Infrastructure Engineer
Seattle, Washington, United States
$85k-$210k/yr OnsiteFull Time
Oracle
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
6+ YOEDesign, deploy, and manage AI/ML and HPC infrastructure; scripting and automation (Ansible,Terraform,Python); containerization (Docker,Kubernetes); strong Linux, networking, security, and troubleshooting skills.
Ansible, Terraform, Python, Kubernetes, Docker, Slurm, PBS, TensorFlow, PyTorch, scikit-learn, Jenkins, GitLab CI/CD, Prometheus, GitHub, Scala, Oracle Linux, RHEL, CentOS, Ubuntu, Debian
2w
Save
Mark Applied
Hide
Machine Learning Infrastructure Engineer
San Francisco or New York or Los Angeles or Seattle
$200k-$345k/yr HybridFull Time
Whatnot
Whatnot: Social marketplace for buying and selling via live streams
4+ YOE4+ years building ML systems,3+ years software engineering,1+ year Python,experience with databases,monitoring,cloud services and production ML deployments.
Python, PostgreSQL, DynamoDB, Elasticsearch, Redis, DataDog, Grafana, AWS Sagemaker, Lambda, Kinesis, S3, EC2, EKS, ECS, Apache Kafka, Flink
2mo
Save
Mark Applied
Hide
SWE - Backend Infrastructure Engineer
San Francisco or Bellevue or New York
$175k-$280k/yr OnsiteFull Time
Sesame
Sesame: Designing wearable computers with lifelike voice-driven AI agents.
3+ YOEStrong systems thinker with reliability engineering experience; 3+ years in infrastructure, platform, or ML systems; Kubernetes production experience; strong communication.
Kubernetes, Terraform, CloudFormation, Pulumi, TorchServe, Seldon, KServe, Ray Serve, PyTorch, APIs, Database design
1mo
Save
Mark Applied
Hide
Software Engineer, Systems ML
Bellevue or Menlo Park or New York
$184k-$257k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Develops social networking platforms and virtual reality technologies.
8+ YOEBachelor's or equivalent, 8+ years in systems engineering/ML infrastructure, experience with distributed ML training/inference, PyTorch/JAX/TensorFlow, low-level C++/CUDA optimizations, and leading cross-functional ML systems projects.
PyTorch, JAX, TensorFlow, C++, CUDA, MLIR, XLA, TVM, Triton
3mo
Save
Mark Applied
Hide
Senior AI/ML Capacity and Performance Engineer
Sunnyvale or Seattle
$145k-$261k/yr HybridFull Time
General Motors
General MotorsNYSE: GM: Manufactures and sells automobiles and automotive parts globally.
5+ YOE5+ years in high-scale infrastructure or ML systems; BS in Computer Science or related field; strong Python and PyTorch; Kubernetes; cloud experience; GPU/ML infra focus.
Python, PyTorch, Kubernetes, Nvidia DCGM, nvidia-smi, Grafana, AWS, GCP, Azure, BigQuery, Hugging Face, Nvidia Nsight, Nsight Compute
1w
Save
Mark Applied
Hide
Member of Technical Staff - Machine Learning Infrastructure Engineer
San Francisco or Toronto or Seattle
$180k-$300k/yr OnsiteFull Time
Preference Model
Preference Model: Building reinforcement learning environments to train frontier AI models.
Experienced software engineer with production ML/data infrastructure skills, proficiency with PyTorch or JAX, distributed systems, AWS/GCP, Kubernetes, data pipelines, and familiarity with transformers and inference libraries like vLLM.
PyTorch, JAX, AWS, GCP, Kubernetes, transformers, vLLM, SGLang
1mo
Save
Mark Applied
Hide
Staff Software Engineer - ML Platform
San Francisco or Seattle or New York City
$217k-$289k/yr HybridFull Time
Grow Therapy
Grow Therapy: Platform connecting mental health providers with patients and insurance.
Proven experience designing and building real-time ML systems and infrastructure (feature stores, real-time serving, deployment safety, monitoring). Strong backend engineering, Terraform experience, technical leadership, and cross-team partnership.
Terraform, Gem, One Medical, Headspace, Talkspace
1mo
Save
Mark Applied
Hide
Principal Infrastructure Architect
Seattle, Washington, United States
$184k-$231k/yr HybridFull Time
DigitalOcean
DigitalOceanNew York Stock Exchange: DOCN: Simplifies cloud infrastructure for developers, startups, and SMBs.
10+ YOE10+ years in server/storage/infrastructure architecture; deep expertise with CPU/GPU/xPU platforms, AI/ML hardware roadmaps, CXL/DPUs/SmartNICs and storage topologies; strong communication and leadership.
CXL memory pooling, DPUs/SmartNICs, disaggregated storage, computational storage, CPU, GPU, xPU, AI/ML, 800VDC rack scale systems
1d
Save
Mark Applied
Hide
AI and ML Infra Software Engineer, GPU Clusters - New College Grad 2026
Santa Clara or Redmond
$124k-$242k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
Recent MS/PhD graduate or equivalent with experience in AI/ML and HPC infrastructure, knowledge of GPUs, storage, schedulers, networking, containers, distributed training, and proficiency in Python/Go/Bash.
Lustre, GPFS, BeeGFS, Slurm, Kubernetes, LSF, Infiniband, RoCE, Amazon EFA, Docker, Enroot, PyTorch, NeMo, JAX, Python, Go, Bash, AWS, GCP, Azure
1d
Save
Mark Applied
Hide
AI and ML Infra Software Engineer, GPU Clusters - New College Grad 2026
Santa Clara or Redmond
$124k-$242k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
MS/PhD or equivalent experience in CS or related field; experience with AI/ML and HPC workloads, GPU infrastructure, storage, orchestration, networking, containers, distributed training; proficiency in Python/Go/Bash and cloud platforms.
Lustre, GPFS, BeeGFS, Slurm, Kubernetes, LSF, Infiniband, RoCE, Amazon EFA, Docker, Enroot, PyTorch, NeMo, JAX, Python, Go, Bash, AWS, GCP, Azure
1w
Save
Mark Applied
Hide
Software Engineer - AI Compute Infrastructure
Seattle, Washington, United States
OnsiteFull Time
ByteDance
ByteDance: Developing AI-driven content platforms and mobile applications.
2+ YOE2+ years building cloud/ML infrastructure; strong knowledge of large-model inference, distributed systems, scheduling, and container orchestration; proficiency in Go, Rust, Python, or C++; experience with Kubernetes and GPU orchestration.
AIBrix, Kubernetes, vLLM, SGLang, TensorRT-LLM, Docker, Ray, Go, Rust, Python, C++, CUDA, AWS, Azure, GCP, SageMaker, Azure ML, Vertex AI, DeepSpeed, PyTorch
1mo
Save
Mark Applied
Hide
Software Engineer, ML Platform
Seattle, Washington, United States
$140k-$215k/yr OnsiteFull Time
Xaira Therapeutics
Xaira Therapeutics: Develops AI models for protein and antibody drug discovery.
5+ YOE5+ years building and deploying ML systems, strong Python skills, experience with ML training infrastructure, Terraform/Ansible, PyTorch/JAX, and leading technical projects; degree in CS/ML/Computational Biology preferred.
Python, Terraform, Ansible, PyTorch, JAX, Slurm, Kubernetes
1mo
Save
Mark Applied
Hide
Staff Software Engineer, ML Infrastructure, Level 6
Bellevue or Seattle or Palo Alto
$229k-$343k/yr OnsiteFull Time
Snap
SnapNYSE: SNAP: Develops social media applications and augmented reality technology.
9+ YOE9+ years software development (or equivalent with advanced degree), experience building large-scale ML, distributed systems or big data processing, strong programming in Python/Java/Scala/C++, and system performance focus.
Python, Java, Scala, C++, Spark, Flink, Ray, PyTorch, TensorFlow
1mo
Save
Mark Applied
Hide
Staff Software Engineer, ML Infrastructure, Level 6
Bellevue or Seattle or Palo Alto
$229k-$343k/yr OnsiteFull Time
Snap
SnapNYSE: SNAP: Provides visual messaging software and augmented reality wearable devices.
9+ YOE9+ years post-Bachelor's software development experience (or advanced degree with reduced experience), experience building large-scale ML, distributed systems or big data processing, proficiency in Python/Java/Scala/C++, strong systems and scalability skills.
Python, Java, Scala, C++, Spark, Flink, Ray, PyTorch, TensorFlow
4d
Save
Mark Applied
Hide
Senior Software Engineer - AI / ML Platforms and Infrastructure
Seattle, Washington, United States
$168k-$210k/yr HybridFull Time
Allen Institute
Allen Institute: Non-profit organization conducting large-scale open science and bioscience research.
5+ YOEBachelor's or equivalent, 5+ years building AI/ML platforms, proficiency with Python, shell scripting, Git, cloud platforms, and project tools; experience with federated learning preferred.
Jira, Python, shell scripting, Git, AWS, Google Cloud, Azure