27 spark data engineer jobs at 22 companies in Isleton, CA
2mo
Save
Mark Applied
Hide
2mo
Data Engineer III
San Ramon, California, United States
$104k-$153k/yrRemoteFull Time
Robert HalfNYSE: RHI: Provides specialized staffing and global business consulting services.
5+ YOE5+ years Python/SQL data engineering; 3+ years cloud AWS; ETL/ELT pipelines; Docker/Kubernetes; Terraform/CloudFormation; Spark; CI/CD; collaboration with data science and MLOps; mentoring.
AVEVA: Industrial software for engineering and operational performance management.
3+ YOEBachelor's in a STEM field,minimum 3 years data engineering experience,proficiency with data pipelines,ETL/ELT,SQL,and cloud data platforms (Azure preferred).
Microsoft Azure, Azure Synapse Analytics, Azure Data Factory, Azure Data Lake, Python, Spark, SQL
Houston or Golden or Reno or Oakland or Salt Lake City
HybridFull Time
Fervo Energy: Generates clean electricity using advanced geothermal drilling technology.
2+ YOE2+ years building and operating production data pipelines with Apache Spark, Python/SQL, Databricks, Azure, and Snowflake; experience with streaming, data modeling, data quality, governance, and CI/CD.
Databricks, Delta Lake, Delta Live Tables, Unity Catalog, Databricks Workflows, Databricks SQL, Apache Spark, PySpark, Spark SQL, Structured Streaming, Kafka, Microsoft Event Hubs, Microsoft Azure Data Factory, Microsoft Azure Data Lake Storage (ADLS), Snowflake, Microsoft Power BI, Spotfire, Python, SQL, Git, Azure DevOps, GitHub Actions, dbt, Terraform, Docker, Airflow, Microsoft Entra ID, Microsoft Key Vault, MQTT, OPC UA, SparkplugB, Canary, Ignition, Snowflake Semantic Views, Databricks Metric Views
Sakata Seed AmericaTokyo Stock Exchange: 1377: Breeding and distributing high-quality vegetable and flower seeds.
5+ YOEDesign and maintain scalable ETL/ELT pipelines, data models, and Microsoft Fabric solutions; strong SQL and Python skills; 5+ years data engineering experience; knowledge of data governance and monitoring.
Microsoft Fabric, Data Factory, Lakehouse, Warehouse, Dataflow Gen2, notebooks, semantic models, SQL Server, PostgreSQL, MySQL, SQL, Python, REST APIs, Spark
PRISM: Provides risk management and insurance for public entities.
Proven data engineering experience designing pipelines and models; strong Python, Spark SQL, T-SQL, and MS SQL Server skills; experience with Microsoft Fabric, Power BI, Azure DevOps, CI/CD, and data governance preferred.
Microsoft Fabric, Fabric Lakehouse, Fabric Data Factory, Fabric Data Warehouse, Fabric Notebooks, Power BI, Python, Spark SQL, T-SQL, MS SQL Server, OneLake, Azure DevOps, Qlik, CI/CD, Sparkhire
IntuitNASDAQ: INTU: Provides financial software for accounting, tax, and personal finance.
10+ YOEBachelor's in CS/Engineering,10+ years in software/data engineering,proficiency in Python/Java/Scala and SQL,experience building large-scale data pipelines and data models with modern data technologies.
Python, Java, Scala, SQL, Google Dataflow, BigQuery, Airflow/Composer, Spark, Flink
DocuSignNASDAQ: DOCU: Provides electronic signature and agreement management software solutions.
8+ YOEBachelor's degree required; 8+ years experience in dimensional/relational data modeling, strong SQL and data pipeline skills in Python or Java, experience with cloud data platforms, ETL tools, and data architecture.
Peterson Holding: Sells Caterpillar equipment and construction technology solutions.
3+ YOEBachelor's in a related field or equivalent, 3+ years data engineering or analytics experience, strong SQL, cloud data platform experience (Snowflake, Microsoft Fabric), Python, Microsoft Power BI, SSIS/Azure Data Factory, and experience building ETL/ELT pipelines and analytics-ready datasets.
Snowflake, Microsoft Fabric, SQL Server Integration Services (SSIS), Microsoft Azure Data Factory, SQL, Python, Microsoft Power BI, Microsoft Dynamics, Spark, PySpark, Microsoft 365
San Francisco or Ottawa or Phoenix or Toronto or Los Angeles or Denver or Salt Lake City or Atlanta or Chicago or Houston or Portland or New York City or Vancouver or San Diego or Washington, D.C. or Sacramento or Jacksonville or Seattle or Mexico City or Austin or Miami or Boston or Dallas
$155k-$265k/yrHybridFull Time
Scribd: Subscription-based digital library for e-books, audiobooks, and documents.
10+ YOE3+ Mgmt10+ years in data engineering, 3+ years managing engineering teams; advanced SQL and experience with Python/Scala, Spark, Databricks/Delta Lake/Snowflake/BigQuery; strong data architecture, dimensional modeling, and cross-functional leadership.
SQL, Python, Scala, Spark, Databricks, Delta Lake, Snowflake, BigQuery
Global Business Intelligence Solutions Data Architect
Fremont, California, United States
$137k-$287k/yrHybridFull Time
Lam ResearchNASDAQ: LRCX: Designs and manufactures wafer fabrication equipment for the semiconductor industry.
12+ YOEBachelor's in CS/IT/MIS/Data Science/Engineering required; 12+ years modern data engineering with cloud (Azure, AWS, Databricks), 10+ years SQL, 3+ years Spark and Python, experience with medallion architecture, Workday Adaptive integrations, and cloud analytics stacks.
Workday Adaptive, Microsoft Fabric, Azure Synapse, Microsoft Power BI, Databricks, Azure, AWS, Redshift, Anaplan, SAP BPC, Spark, Python, REST, SOAP, SQL, Data Fabric
United States or Colorado or Hawaii or Illinois or Maryland or Massachusetts or Minnesota or New Jersey or New York or Vermont or District of Columbia or Washington or California or San Francisco or Oakland or San Jose or Washington
$139k-$204k/yrRemoteFull Time
TwilioNYSE: TWLO: Cloud communications platform for building customer engagement applications.
5+ YOEBachelor's or Master's in CS/Engineering, 5+ years software development, proficiency in Python/Java/Scala, experience with big data tools (Kafka, Spark, Hive, Hudi, Presto, Airflow) and AWS services, strong problem-solving and communication skills.
AbbottNYSE: ABT: Manufactures medical devices, diagnostics, and nutritional health products.
8+ YOE8+ years in data architecture/engineering, 5+ years building enterprise cloud data platforms; deep Databricks, AWS, Azure, SQL and Python experience; data governance, security, and healthcare data knowledge preferred.
Databricks, Delta Lake, Unity Catalog, Databricks Workflows, Databricks SQL, Spark, PySpark, AWS, S3, Glue, IAM, Redshift, Lambda, Kinesis, Lake Formation, Microsoft Azure, Azure Data Lake Storage (ADLS), Azure Data Factory, Synapse, Azure Key Vault, Microsoft Purview, SQL, Python, Kafka, Spark Structured Streaming, EMR/EHR, FHIR, HL7, MLOps
Advanced quantitative degree, deep probability/statistics knowledge, expertise in risk and catastrophe modeling, strong programming and data engineering skills, and ability to communicate insights to underwriting teams.
AbbottNYSE: ABT: Provides medical devices, diagnostics, and science-based nutritional products.
8+ YOEBachelor's degree required; 8+ years in data/solution architecture; 5+ years designing enterprise cloud data platforms; deep Databricks, AWS, Azure, SQL and Python experience; governance and healthcare data experience preferred.
Databricks, Delta Lake, Unity Catalog, Databricks Workflows, Databricks SQL, Spark, PySpark, Amazon S3, AWS Glue, AWS IAM, Amazon Redshift, AWS Lambda, Amazon Kinesis, AWS Lake Formation, Azure Data Lake Storage (ADLS), Azure Data Factory, Azure Synapse, Azure Key Vault, Microsoft Purview, SQL, Python, Kafka, Spark Structured Streaming
Distributed Systems Engineer, Data & Inference Platform
San Francisco or Fremont or Palo Alto or Berkeley or Sunnyvale or Mountain View or San Jose or Oakland or Redwood City
HybridFull Time
Adaption: Develops efficient AI systems that adapt in real-time.
5+ YOE5+ years building and operating distributed systems; experience with Ray/Spark/Flink/Beam/Dask; Python and at least one systems language; GPU/accelerator stack; Kubernetes; incident ownership.
Lead Principal Application Software Engineer- Forward Deployed Engineer
Pleasanton or United States or California
$105k-$264k/yrOnsiteFull Time
OracleNYSE: ORCL: Provides cloud infrastructure and enterprise software for global businesses.
10+ YOERequires 10–12+ years in enterprise data platforms, expertise in cloud data engineering and Apache Spark, migration experience, and proficiency with Databricks, Snowflake, Azure, AWS, Kafka, Python, SQL, Scala, or Java.
Databricks, Snowflake, Azure, AWS, Oracle AI Data Platform, Apache Spark, SPARK, Delta Lake, Kafka, Python, SQL, Scala, Java, DataOps, CI/CD, Infrastructure as Code, Lakehouse, Medallion Architecture, OCI, LLMs, Retrieval-Augmented Generation (RAG), MLOps, OCI - DevOps, OCI - Generative AI Agent Hub, OCI Generative AI
Ridgeline: Cloud-native platform for investment management operations.
8+ YOE8+ years engineering experience building shared platforms; strong architecture judgment; experience with data models, APIs, distributed systems, observability, and mentoring.
New York City or Milwaukee or Dallas or Columbus or Kirkland or Cincinnati or Cleveland or Oklahoma City or Austin or Albany or Chicago or St. Petersburg or Hartford or Pittsburgh or St. Louis or Miami or Sacramento or Raleigh or Minneapolis or Mountain View or Scottsdale or San Francisco or Morristown or Denver or Boston or Philadelphia or Des Moines or Overland Park or Los Angeles or Charlotte or Walnut Creek or Carmel or Seattle or Houston or Arlington or Atlanta or Redmond or Bentonville or Beaverton or Nashville or Detroit or San Diego
$80k-$294k/yrHybridFull Time
AccentureNYSE: ACN: Global provider of management consulting and technology services.
6+ YOEBachelor's degree or equivalent experience, 6+ years developing and deploying AI/ML solutions, programming proficiency, experience with ML frameworks, big data, databases, and cloud platforms.
Python, R, Java, TensorFlow, PyTorch, Hadoop, Spark, AWS, Microsoft Azure, Google Cloud, D3.js, ggplot
Workday: A provider of enterprise cloud applications and platforms for managing finance and human resources.
5+ YOEBachelor's in CS or related and 5+ years experience; strong Java/Scala and Python; experience with ML/NLP, deep learning, data lake, AWS, S3, Spark, Kafka and ML workflow orchestration.