Principal Machine Learning Engineer, Accelerated Apache Spark
Santa Clara, California, United States
$272k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
12+ YOESenior ML/DS engineer with 12+ years of experience; 5+ years leading ML model development; strong Python and ML tooling; experience with GPU-accelerated Spark; expertise in ML methods and feature engineering.
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
6+ YOE6+ years software engineering experience with ≥2 years on Presto/Trino or Apache Spark internals; strong Java or Scala; connectors, UDFs/UDAFs, optimizer work; JVM tuning, open table formats, and clear written communication.
San Francisco or Sunnyvale or Seattle or New York City
$131k-$285k/yrHybridFull Time
DoorDashNASDAQ: DASH: On-demand delivery platform connecting consumers with local merchants.
24+ YOEExperience operating production distributed systems and Apache Spark at scale, Kubernetes production experience, familiarity with schedulers and observability stacks, cloud (AWS) experience, programming in Python/Go/Scala/Java, and SQL fluency. BS/MS/PhD in CS or equivalent.
Principal Machine Learning Engineer, Accelerated Apache Spark
Santa Clara, California, United States
$272k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOEBS/MS/PhD or equivalent; 12+ years ML/DL experience; 5+ years as technical lead; 2+ years with Apache Spark; strong Python and data-science libraries experience; expertise in LLM/GenAI, RL, XGBoost; leadership and deployment experience.
Databricks: A unified platform for data analytics and artificial intelligence.
8+ YOE8+ years software engineering with 2+ years building production LLM/agentic applications; deep Python fluency; experience with agentic frameworks, REST/GraphQL integrations, Databricks/Spark, and technical leadership.
Wayve: Develops end-to-end artificial intelligence for autonomous driving systems.
10+ YOE10+ years building large-scale distributed systems or ML infrastructure, 3+ years at staff/principal level, experience with Spark, Ray, Kubernetes, Airflow, MLflow, web frameworks, reliability engineering, and mentoring engineers.
5+ YOEBachelor's degree or equivalent experience,5+ years software development,3+ years building large-scale distributed infrastructure,EMR/Excel not required,experience with distributed computing and API design.
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
5+ YOE5+ years applied machine learning experience; strong statistical, optimization, and experimental methodology knowledge; proficient in Spark, Python or Java; experience with real-time, low-latency systems and distributed ML frameworks.
Eightfold: An AI-native enterprise talent platform that builds scalable machine learning solutions and agentic AI for talent workflows.
6+ YOEDeep expertise in ML/DL/NLP and LLMs; technical leadership and team mentoring; proficiency with Python, TensorFlow, PyTorch, big data (Hadoop, Spark); experience deploying ML at scale. MS/PhD and 6+ years preferred.
Wayve: Develops AI software for autonomous vehicle navigation.
10+ YOE10+ years building large-scale distributed systems or ML infrastructure, 3+ years at staff/principal level, experience with Spark, Ray, Kubernetes, Airflow, MLflow, reliability/observability, mentoring, and optimization or scheduling systems.
TikTok USDS Joint Venture: Operates and secures TikTok services for U.S. users.
5+ YOEBachelor's in CS or related plus 5+ years deploying/administering large-scale distributed systems; strong Unix/Linux and networking knowledge; experience with programming, debugging, and systems like Nginx, Kubernetes, Docker, Hadoop, Spark, Flink, Kafka.
Senior Principal Machine Learning Engineer - Optimization
Redwood City or United States
$260k-$330k/yrHybridFull Time
PubMaticNASDAQ: PUBM: Sell-side platform for digital advertising and programmatic media buying.
10+ YOE10+ years building production ML, ranking, or optimization systems; strong ML fundamentals, experience with prediction/CTR/CVR/ calibration/experimentation; proficiency in Python, Java, SQL, Spark, TensorFlow, PyTorch, XGBoost.
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
2+ YOEBachelor's or equivalent and 2+ years engineering experience; coding in C,C++,C#,Java,JavaScript or Python; experience with Spark and data engineering preferred; must pass Microsoft security screening.
Senior Site Reliability Engineer, Apple Data Platform SRE / Apple Services Engineering
Cupertino, California, United States
OnsiteFull Time
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Apply SRE principles to mentor teams, ensure reliability for large-scale analytics infrastructure across Hadoop, HBase, Spark, Data Lakes, and Airflow; participate in production on-call.
Livingston or New York City or Sunnyvale or San Francisco or Bellevue
$207k-$275k/yrOnsiteFull Time
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
10+ YOE10+ years in platform or infrastructure engineering with Kubernetes, CI/CD, IaC, observability, multi-region systems, and production ownership for high-availability services.