4 ml systems engineer jobs at 4 companies in Davis, CA
🚀PromotedHiringCafe
ML Engineer - Inference & Model Deployment
Cupertino, CA, US
$250k-$310k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Turn powerful AI and ML models into fast, reliable production systems. Own inference latency, throughput, model-serving architecture, multi-GPU systems, and production deployment for millions of users.
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Own the crawlers, pipelines, and infrastructure powering a real-time job search engine. Strong Node.js and Python fundamentals; bonus points for security and reverse-engineering chops.
Richmond or Indianapolis or Walnut Creek or Seattle or Grand Prairie or Winston-Salem or Chicago or Atlanta or Miami or Nashville or Tampa or Woburn or Norfolk
$231k-$417k/yrHybridFull Time
Elevance HealthNYSE: ELV: Provides health insurance plans and integrated healthcare services.
15+ YOEBachelor's in CS/IT (or equivalent) and 15+ years software engineering experience; production AI/ML, LLM and agentic systems experience; strong Python/Java skills; APIs, microservices, cloud, CI/CD and DevOps experience required.
Python, Java, APIs, microservices, event-driven architectures, cloud platforms, CI/CD, DevOps, MLOps, LLMOps, Large Language Models (LLMs), agentic systems, distributed systems
San Francisco or Atlanta or Austin or Boston or Chicago or Dallas or Denver or Houston or Jacksonville or Los Angeles or Miami or New York City or Phoenix or Portland or Sacramento or Salt Lake City or San Diego or Seattle or Washington, D.C. or Ottawa or Toronto or Vancouver or Mexico City
$128k-$235k/yrHybridFull Time
Scribd: Subscription-based digital library for e-books, audiobooks, and documents.
5+ YOE5+ years data engineering; AI/ML infra or LLM apps; Python and SQL; Databricks; NLP/LLM systems; governance and security focus; autonomous in ambiguous problems; strong communication.
AMDNASDAQ: AMD: Designs and manufactures computer processors and graphics technology.
Develop performance models and prototypes for GPU/SoC systems; analyze ML/HPC workloads; write and optimize GPU kernels; strong programming in C/C++/Python; knowledge of CUDA/OpenCL and ML frameworks.
TensorFlow, PyTorch, CUDA, OpenCL, C, C++, Python, Hip, Triton, MLIR, LLVM, RTL, System C
New York City or Milwaukee or Dallas or Columbus or Kirkland or Cincinnati or Cleveland or Oklahoma City or Austin or Albany or Chicago or St. Petersburg or Hartford or Pittsburgh or St. Louis or Miami or Sacramento or Raleigh or Minneapolis or Mountain View or Scottsdale or San Francisco or Morristown or Denver or Boston or Philadelphia or Des Moines or Overland Park or Los Angeles or Charlotte or Walnut Creek or Carmel or Seattle or Houston or Arlington or Atlanta or Redmond or Bentonville or Beaverton or Nashville or Detroit or San Diego
$80k-$225k/yrHybridFull Time
AccentureNYSE: ACN: Global provider of management consulting and technology services.
5+ YOE5+ years in AI/ML/automation or software engineering, 5+ years Python and APIs/automation frameworks experience, 2+ years deploying AI/ML to production, bachelor’s degree or equivalent, experience with cloud/distributed systems and GenAI preferred.
Python, APIs, automation frameworks, data pipelines, GenAI, LLMs