182 ai ml infrastructure engineer jobs at 81 companies in Ross, CA
1w
Save
Mark Applied
Hide
1w
AI Infrastructure Engineer
San Francisco, California, United States
$150k-$220k/yrOnsiteFull Time
Sciforium: Building multimodal AI models and high-performance model serving infrastructure.
5+ YOE5+ years in systems or infrastructure engineering with GPU, HPC, or ML infrastructure experience; technical bachelor's or master's degree; Linux, Kubernetes, schedulers, configuration management, Python, Bash, containers, GPUs, and RDMA expertise.
San Francisco or Minneapolis or Washington, D.C. or United States
$120k-$215k/yrRemoteFull Time
UnitedHealth GroupNYSE: UNH: Provides health insurance and technology-enabled health care services.
4+ YOE2+ MgmtBachelor's degree or 4+ years equivalent, 4+ years Python, 4+ years cloud infrastructure (AWS/Azure/GCP), 4+ years AI/ML infrastructure experience, 2+ years team lead, 1+ year LLM experience.
Python, AWS, Azure, GCP, Large Language Models (LLMs), GitHub, GitHub Actions, Docker, Terraform, CI/CD
GuidewireNYSE: GWRE: Provides a software platform for property and casualty insurers.
10+ YOE10+ years software engineering; 5+ years ML platforms/infrastructure; distributed systems; Python/Go/Java; Docker/Kubernetes; MLOps tools; cloud experience; knowledge of ML models.
Artificial Analysis: Independent AI benchmarking and performance analysis platform.
3+ YOE3+ years software engineering experience, Python and pandas proficiency, OpenAI API experience, cloud infrastructure familiarity, data visualization and communication skills; Bachelor’s or Master’s preferred.
Docker: Provides a platform for building, sharing, and running containerized applications.
5+ YOE5+ years applied ML/AI experience, 4+ years software engineering, experience with LLM-based systems, model lifecycle and ML infrastructure, bachelor's in CS/Engineering or equivalent, strong communication and mentoring skills.
Senior AI Infrastructure Engineer - Model Training
Mountain View, California, United States
$190k-$260k/yrOnsiteFull Time
Kodiak RoboticsNASDAQ: KDK: Develops autonomous driving technology for commercial trucking and defense.
2+ YOEDegree in CS or related field,2+ years ML systems experience,expertise in distributed training,high-performance data pipelines,GPU performance and profiling,Python and PyTorch skills.
2+ YOEBachelor's degree or equivalent practical experience, 2+ years programming in Python or C++, and experience with ML infrastructure and a specialized ML area.
Phenix Space: Builds robotic systems for on-orbit satellite assembly and upgrades.
5+ YOE5+ years building or operating ML infrastructure; deep GPU and distributed training knowledge; experience with PyTorch/DeepSpeed/Megatron/Ray, inference stacks (vLLM, TGI, Triton), Python and C++/Rust/Go, Kubernetes and IaC, and observability tooling.
Senior Lead AI Engineer (GenAI Platform, Agentic Infrastructure)
New York City or San Francisco or San Jose or Cambridge or McLean or Plano
$209k-$286k/yrOnsiteFull Time
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
6+ YOEBachelor's degree plus 6 years or master's degree plus 4 years developing AI/ML technologies; 6 years programming with Python, Go, Scala, or Java. Leadership and cloud AI experience preferred.
Cargomatic: Digital marketplace connecting shippers with local trucking capacity.
3+ YOE3+ years software engineering with focus on AI/ML or automation; hands-on AI-powered applications; experience with agentic AI concepts, LLMs, and modern AI frameworks; cloud infrastructure (AWS); backend (Node.js); frontend (React).
Denver or San Francisco or New York City or San Jose or Scottsdale
$160k-$240k/yrHybridFull Time
Gusto: Cloud-based payroll and HR software for small businesses.
4+ YOERequires 4+ years of software engineering experience in Python, Ruby, or Java; ML infrastructure and platform-service experience; cloud-platform experience; and familiarity with AI frameworks and AI-assisted development tools.
TargetNYSE: TGT: Operates a chain of general merchandise stores and supermarkets.
MS or equivalent preferred; extensive experience designing and operating large-scale cloud-native ML platforms, Kubernetes-based infrastructure, MLOps, model governance, observability, and platform automation.
Zensors: AI platform that turns existing cameras into intelligent sensors.
BS/MS/PhD in CS or EE; strong C/C++ and Python; model optimization, quantization, pruning; GPU performance tuning; profiling tools; cross-functional collaboration.
Harell Data: A managed platform that enables secure sharing and high-performance use of scientific datasets for model training and deployment.
5+ YOE5+ years building and operating production infrastructure for ML workloads; hands-on Kubernetes on AWS or GCP (GPU workloads); strong CS fundamentals, system design, incident response, and customer-facing debugging experience.
Thumbtack: Online marketplace connecting homeowners with local service professionals.
3+ YOE1-3 years software engineering; strong data structures/algorithms; Go and Python; Postgres or DynamoDB; AI tooling; adaptable in fast-paced environment.
Seattle or Bellevue or United States or San Francisco Bay Area
HybridFull Time
Stealth AI Startup: Building infrastructure to deploy and manage enterprise AI applications.
5+ YOE5+ years engineering experience, AI/ML infrastructure or MLOps expertise, cloud-native systems knowledge, Python and a systems or backend language, debugging ability, and strong customer-facing communication.
Senior Software Engineer, AI Infrastructure - LVM Inference & Evaluation
Redwood City, California, United States
$168k-$205k/yrHybridFull Time
Ambient.ai: AI-powered physical security platform for proactive threat detection.
4+ YOE4+ years building infrastructure or production AI systems; strong Python; experience with ML infrastructure, LLM/LVM inference, inference optimization, evaluation frameworks, cloud and GPU workloads; BS/MS or equivalent.