303 spark data engineer jobs at 159 companies in Fairfax, CA
1mo
Save
Mark Applied
Hide
1mo
Data Engineer
New York City or San Francisco
OnsiteFull Time
Bedrock Robotics: Automates heavy construction machinery with AI retrofit kits.
5+ YOE5+ years data engineering experience in large-scale data lake/warehouse environments; strong SQL and distributed query engine experience; pipeline orchestration (Airflow/Prefect) and Spark/Databricks experience; data quality and observability expertise.
Stitch FixNASDAQ: SFIX: Provides personalized apparel styling services through algorithms and stylists.
2+ YOEBachelor's in engineering or computer science, 2+ years data engineering experience, familiarity with Spark, dbt, Fivetran, Airflow, S3; production Python and SQL; experience with AI coding agents; strong communication.
Spark, dbt, Fivetran, Airflow, S3, Claude Code, Codex, Python, SQL
LendingClubNYSE: LC: Digital marketplace bank providing personal loans and banking services.
8+ YOE2+ Mgmt8+ years data engineering experience with 2+ years technical lead experience; strong SQL, data modeling, ETL, orchestration, Spark, AWS, modern data platforms, and experience applying AI tools to engineering workflows.
United States or San Francisco or New York City or Washington or Austin
$118k-$179k/yrRemoteFull Time
SamsaraNYSE: IOT: Connected Operations Cloud platform for physical operations and IoT.
8+ YOEBachelor's in CS or equivalent, 8+ years as a software/data engineer, 5+ years building production data pipelines and Spark/PySpark experience, strong Python and SQL, cloud data warehouse and ETL tooling experience.
United States or Foster City or Orange County or Morristown
RemoteFull Time
Bridgepointe Technologies: Vendor-agnostic IT strategy and technology procurement services.
5+ YOERequires 5+ years in data engineering, advanced SQL, Python, Spark/PySpark, Microsoft Fabric, Azure data services, APIs, data modeling, CI/CD, Git, and Azure DevOps; bachelor's degree or equivalent experience.
Microsoft Fabric, Lakehouse, Warehouse, Data Pipelines, Notebooks, OneLake, Git, Azure DevOps, Spark, PySpark, SQL, Python, Azure, Delta Lake, Parquet, REST APIs, JSON, Power BI, DP-600, DP-700
Robert HalfNYSE: RHI: Provides specialized staffing and global business consulting services.
5+ YOE5+ years Python/SQL data engineering; 3+ years cloud AWS; ETL/ELT pipelines; Docker/Kubernetes; Terraform/CloudFormation; Spark; CI/CD; collaboration with data science and MLOps; mentoring.
CoupangNYSE: CPNG: Provides an end-to-end e-commerce and logistics network.
8+ YOE3+ Mgmt8+ years data engineering experience, 3+ years people management; proficiency in data modeling, ETL, Python, SQL; experience with Spark, HDFS, S3 and ETL schedulers (Airflow, Dagster, DBT); experimentation and ML familiarity.
Windfall: Provides consumer financial data and AI-driven go-to-market insights.
4+ YOE4-8 years in data engineering; experience with Apache Beam/Spark/Flink or MapReduce; JVM language proficiency; distributed data processing; experience at a small company; strong communication and ownership.
United States or San Francisco or New York City or Chicago
$170k-$230k/yrHybridFull Time
Komodo Health: Provides AI-driven healthcare data and patient journey analytics software.
Experience building production data pipelines at scale with advanced Python, SQL, Airflow, Spark, and AWS; healthcare data expertise and strong data quality, reliability, troubleshooting, and collaboration skills required.
The Walt Disney CompanyNYSE: DIS: Produces media content and operates global theme parks.
5+ YOEBachelor's or equivalent experience, 5+ years big data engineering, strong Python/Scala/SQL, Spark/Presto/Hive, Databricks and cloud MPP databases, Airflow, AWS S3, CI/CD, on-call and production operations experience.
OpenAI: Develops artificial intelligence models and generative AI software services.
3+ YOE3+ years data engineering; 8+ years software engineering; Python/Scala/Java; Databricks, Snowflake; ETL schedulers; Spark/Hadoop/Flink; S3/HDFS; strong data pipelines and collaboration.
Sapiom: Financial infrastructure for autonomous AI agents.
5+ YOE5+ years building production data pipelines; hands-on SQL, Python, Spark, AWS Glue, EMR, DBT, Airflow; 3+ years with MPP databases (Snowflake/Redshift/Teradata); on-call experience and strong cross-team communication.
Checkr: AI-powered platform for background checks and identity verification.
10+ YOE10+ years building scalable data platforms; expert PySpark, Python, SQL; experience with Kafka, Spark, Iceberg, data lakes, AWS; strong data modeling and security awareness.
Tatari: A platform for buying and measuring TV advertising campaigns.
5+ YOE5+ years building and operating production ETL pipelines; strong Python and SQL skills; Spark/PySpark, Databricks/Delta Lake, Airflow experience; strong data modeling and data quality practices.
Python, SQL, Spark, PySpark, Databricks, Delta Lake, Airflow, ClickHouse
Foundry Robotics: An AI-native robotics manufacturer focused on deploying advanced assembly and production capability for robotics companies and national-security-critical hardware.
5+ YOE5-7 years data engineering experience; TB–PB-scale pipelines; Spark/Flink/Kafka/Airflow/dbt; Python/SQL; Go/Java/TypeScript a plus; cloud data services; architecture and governance; independent, end-to-end ownership.
White Plains or Greensboro or Charleston or Greenville or Maitland or Harrisonburg or Clearwater or Louisville or Birmingham or Huntsville or Alpharetta or Clayton or Honolulu or Lake Oswego or Gainesville or Fayetteville or Naples or Hibbing or Miami or Myrtle Beach or Houston or Denver or Piscataway or Irvine or Key Largo or Columbia or Little Rock or Austin or Pensacola or Lexington or Bloomington or Fort Myers or Panama City or San Antonio or Rochester or Mobile or Salt Lake City or Bend or Charlotte or Glen Allen or Raleigh or Abilene or Tallahassee or San Ramon or Dallas or Pittsburgh or Boston or Missoula or Cincinnati or Boise or Eugene or McMinnville or Knoxville or Dakota Dunes or Kissimmee or Fort Worth or Grand Forks or Canonsburg or Buffalo or Eastland or Philadelphia or Chicago or Phoenix or Dublin or Jacksonville or St. Louis or Los Angeles or Atlanta or Norwalk or Colchester or Grand Rapids or Richmond or Hoboken or Indianapolis or Kansas City or Chagrin Falls or Orland Park or Anchorage or High Point or Bonita Springs or Rancho Cordova or Las Vegas or Irving or Brandenburg or Minneapolis or Fort Lauderdale or Shreveport or Fargo or Sarasota or Kennesaw or Dayton or Sacramento or Princeton or Durham or Conshohocken or Baton Rouge or Asheville or Chattanooga or Memphis or Nashville or United States
$96k-$168k/yrRemoteFull Time
Marsh McLennanNYSE: MRSH: Global professional services firm providing risk, strategy, and people solutions.
10+ YOERequires 10+ years of enterprise data engineering experience and advanced Azure, Databricks, Spark, Python, SQL, ETL, data modeling, streaming, and cloud architecture expertise.
Microsoft Azure, Azure Databricks, Azure Functions, Azure Data Factory, Databricks, Apache Spark, PySpark, Python, SQL, Scala, Unity Catalog, Microsoft SQL Server, Kafka, Event Hub, Azure DevOps, Git, Terraform, ARM templates, Databricks Genie
VisaNYSE: V: Global payment technology facilitating electronic funds transfers.
2+ YOE2+ years experience with a Bachelor's (or Advanced) degree; strong programming in Java/Python/Go/Scala; experience building scalable data pipelines, Spark/Kafka/Hadoop ecosystem, cloud platforms, GenAI and data modeling.
Chartmetric: Provides a global music data analytics platform for professionals.
6+ YOE6+ years building production data systems; expert in Python, PostgreSQL, Airflow, Clickhouse, Snowflake, Elasticsearch, Spark, and AWS; Bachelor's in CS/Data Engineering or equivalent experience.
Scottsdale or Chicago or San Francisco or New York City
$97k-$149k/yrHybridFull Time
Early Warning Services: Operates payment and risk solutions for the financial industry.
2+ YOEBachelor's degree typical; 2+ years experience building data engineering tools with Spark and Python, experience with ML toolkits, Hadoop or AWS/S3, and relational databases; strong communication and background/drug screen required.
AVEVA: Industrial software for engineering and operational performance management.
3+ YOEBachelor's in a STEM field,minimum 3 years data engineering experience,proficiency with data pipelines,ETL/ELT,SQL,and cloud data platforms (Azure preferred).
Microsoft Azure, Azure Synapse Analytics, Azure Data Factory, Azure Data Lake, Python, Spark, SQL