153 spark data engineer jobs at 45 companies in Marina, CA
🚀PromotedHiringCafe
Founding Machine Learning / AI Search Engineer
Cupertino, CA, US
$160k-$310k/yrOn-SiteFull Time
HiringCafe: Building a 100× better job search engine to take on Indeed and LinkedIn.
Build the ML and AI search behind HiringCafe — ranking, recommenders, retrieval, and LLM agents that surface jobs people would never find on their own.
Qcells: Provider of solar modules, energy storage, and EPC services.
10+ YOE10+ years in data engineering or architecture, deep SQL and Python, Azure data services experience, distributed processing (Spark), data modeling, governance, and leadership experience.
Azure Fabric, Data Lake, Data Factory, Synapse, NetSuite, SAP, Salesforce, APIs, Delta Lake, Delta Tables, Snowflake, Kafka, Event Hub, Spark, Python, SQL
PlusAI: AI-based virtual driver software for factory-built autonomous trucks.
1+ YOEMS in CS/Electrical or related, 1+ years software engineering, expertise in Python and SQL, experience with large-scale data processing (Spark, MapReduce, Kafka), AWS (S3, EC2, RDS), and building/maintaining data pipelines.
WalmartNYSE: WMT: Multinational retail operating discount stores and supermarkets.
2+ YOEDesign and operate large-scale data pipelines on GCP using Spark, Airflow, BigQuery; optimize cloud cost; build automation and AI-driven engineering tools; 2+ years data engineering experience.
RokuNASDAQ: ROKU: Operates a TV streaming platform and sells streaming hardware.
8+ YOE8+ years data engineering experience, strong SQL, Python preferred, expertise with big-data tech (Hadoop, Spark, Kafka, Hive, Airflow), data modeling, and cloud platforms (AWS/GCP).
PayPalNASDAQ: PYPL: Global digital payments platform for consumers and merchants.
8+ YOEMaster's in CS/Engineering plus 8 years (or Bachelor's plus 10 years). Requires data engineering, Python, Shell, SQL, ETL, BigQuery, GCP, Kafka, Airflow, Spark/Hadoop, BI tools, data modeling experience.
Python, Shell, SQL, Tableau, ThoughtSpot, Google BigQuery, Google Cloud Platform, Erwin, Apache Kafka, Apache Airflow, Automic UC4, Spark, Hadoop
TikTok: Global short-form video hosting and social media platform.
BS/MS in CS or equivalent; experience with Hadoop, Hive, Spark, Presto, Kafka, ClickHouse, Flink; ETL, data ingestion, schema design and SQL; big data system architecture experience.
Boston ScientificNYSE: BSX: Manufacturer of interventional medical devices and technologies.
13+ YOEBachelor's in CS or related, 13+ years data engineering with 6+ years building AWS cloud-native data platforms; expertise with AWS data services, Snowflake, Python, Terraform, Spark/Flink, and governance automation.
3+ YOEBachelor's in CS/Math/related or equivalent; 3+ years with data processing (Hadoop, Spark, Pig, Hive), DB administration or data engineering, software experience in Java/C++/Python/Go/JavaScript, and client-facing project experience.
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
5+ YOE5+ years building analytical data models and production data systems; expertise in Python/Scala/Rust, advanced SQL, Spark/Flink, lakehouse technologies, and dimensional modeling for OLAP workloads.
Boston ScientificNYSE: BSX: Developer and manufacturer of innovative medical devices and therapies.
13+ YOEBachelor's degree required; 13+ years data engineering experience including 6+ years designing cloud-native AWS data platforms and 4+ years building scalable data platform solutions. Expertise with AWS data services, Snowflake, Python, Terraform, Bash, Spark/Flink, and governance automation.
AdobeNASDAQ: ADBE: Provides software for digital media creation and marketing analytics
8+ YOESenior data engineer with 8+ years experience, distributed data and cloud technologies (Hadoop, Spark, Databricks, S3, EMR), streaming (Kafka/Kinesis), strong SQL and Python/PySpark/Scala skills, CI/CD and orchestration experience, and strong communication.
Gridmatic: AI-powered platform for optimizing energy trading and battery storage.
Experience building large-scale production data pipelines, designing storage and schema for timeseries and warehouse data, proficiency with DBT and data processing tools, strong software engineering skills, and startup experience.
AWS Principal Data Engineer (Valencia, CA, US, 91355)
Valencia or Santa Clara or Arden Hills
$107k-$203k/yrHybridFull Time, Contract
Boston ScientificNYSE: BSX: Developing and manufacturing innovative medical devices for less-invasive treatments.
13+ YOEBachelor's degree required; 13+ years data engineering experience with 6+ years building cloud-native AWS data platforms and 4+ years building scalable data platform solutions. Expertise in AWS data services, Snowflake, Python, Terraform, Spark/Flink, and governance automation.
Hewlett Packard EnterpriseNYSE: HPE: Providing global edge-to-cloud infrastructure and IT solutions for businesses.
4+ YOEDegree in Computer Science or Information Systems; Master’s with 2 years data engineering experience; at least 4 years in relevant technologies; Python and Java programming; big data (Spark/Kafka); AI/Generative AI experience; containerization (Kubernetes, Airflow); cloud development.
Principal Machine Learning Engineer, Accelerated Apache Spark
Santa Clara, California, United States
$272k-$431k/yrOnsiteFull Time
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
12+ YOEBS/MS/PhD or equivalent; 12+ years ML/DL experience; 5+ years as technical lead; 2+ years with Apache Spark; strong Python and data-science libraries experience; expertise in LLM/GenAI, RL, XGBoost; leadership and deployment experience.
BILLNYSE: BILL: Automated financial operations software for small and midsize businesses.
8+ YOELead architecture and delivery of data platform capabilities (ingest, lake, streaming, feature store, query, graph, search). Requires distributed systems, streaming and batch experience, SQL and Python, and technical leadership.