123 spark data engineer jobs at 51 companies in Shadow Hills, CA
1mo
Save
Mark Applied
Hide
1mo
Senior Data Engineer
North Hollywood or New York
$140k-$160k/yrHybridFull Time
VersantNasdaq: VSNT: Operates cable television networks and digital media entertainment platforms.
5+ YOE5+ years building production data pipelines with Spark-based platforms, expertise in PySpark and SQL, Databricks experience preferred, bachelor's in CS/data engineering or equivalent, strong CI/CD, orchestration, and data governance skills.
Delta Lake, Databricks Auto Loader, Databricks, Structured Streaming, Change Data Capture (CDC), Unity Catalog, Spark, Photon, Databricks Workflows, Delta Live Tables, Airflow, MLflow, Presto, Flink, Git, Dagster, EMR, Dataproc, Lakeflow Spark Declarative Pipelines (SDP), ChatGPT
Anduril Industries: Defense technology building autonomous military hardware and software.
3+ YOE3+ years in data engineering, strong Python and SQL skills, experience with Spark/PySpark, dbt, SQLMesh, Palantir Foundry, cloud platforms (AWS/Azure/GCP), data orchestration (Flyte), and data formats (Apache Iceberg).
The Walt Disney CompanyNYSE: DIS: Produces media content and operates global theme parks.
5+ YOE5+ years data engineering experience building large data pipelines; strong SQL, Python/PySpark; experience with Snowflake/Redshift, Databricks, Spark, Airflow, AWS; data modeling and performance tuning; Bachelor's degree or equivalent.
Hadrian: Building autonomous factories for aerospace and defense manufacturing.
Production data-model ownership, expert SQL, Spark, dbt, Dagster or equivalent, Python, data modeling, lake and warehouse internals, semantic layers, pipelines, testing, documentation, and CI/CD.
Publicis GroupeEuronext Paris: PUB: Global advertising and digital transformation agency holding.
3+ YOE3–5 years data engineering experience; strong SQL, Python, Spark/PySpark; experience with Databricks, AWS, Snowflake or Google Cloud; familiarity with AI tools and modern lakehouse architectures.
Python, PySpark, SQL, Spark, Claude Code, Codex, GitHub Copilot, Databricks Genie, Databricks, AWS, Snowflake, Google Cloud, SFTP, APIs
Tatari: A platform for buying and measuring TV advertising campaigns.
5+ YOE5+ years building and operating production ETL pipelines, strong Python and SQL skills, Spark/PySpark, Databricks/Delta Lake, Airflow, data modeling and data quality expertise.
Python, SQL, Spark, PySpark, Databricks, Delta Lake, Airflow, ClickHouse
San Francisco or Los Angeles or Denver or Austin or Chicago or New York City or Seattle or Toronto or Santa Barbara or San Diego
$127k-$190k/yrRemoteFull Time
Invoca: AI platform for conversation intelligence and revenue execution.
5+ YOE5+ years in data or software engineering; advanced Python, SQL, Databricks, Spark, and data modeling; experience with orchestration, streaming, cloud infrastructure, data quality, ML datasets, and privacy; bachelor's degree or equivalent.
Austin or Boston or Charleston or Charlotte or Chicago or Dallas or Durham or Harrisburg or Houston or Irvine or Kansas City or Los Angeles or Miami or Nashville or New York or Newark or Palo Alto or Pittsburgh or Portland or Raleigh or San Francisco or Seattle or Washington or Wilmington
$128k-$249k/yrHybridFull Time
K&L Gates: Global law firm providing comprehensive legal and regulatory counsel.
5+ YOE5+ years designing enterprise Microsoft Fabric data platforms; strong SQL and Python skills; experience with Spark, CI/CD, Azure DevOps/GitHub Actions; knowledge of data governance and Azure AI Foundry; Bachelor's degree or equivalent.
Microsoft Fabric, OneLake, Fabric Data Factory, Dataflow Gen2, Fabric Notebooks, Azure AI Foundry, Microsoft Copilot, Claude, SQL, Python, Spark, Azure DevOps, GitHub Actions, Microsoft Purview
ParamountNASDAQ: PSKY: Produces and distributes media content across global entertainment platforms.
2+ YOE2+ years building ETL/ELT pipelines, strong SQL and Python, experience with Airflow, Spark, Kafka/Pub-Sub, cloud-native (GCP preferred), data modeling, and observability for analytics/ML.
Boston or Chicago or Dallas or Los Angeles or Minneapolis or New York City or San Francisco or Seattle or Washington, D.C.
$120k-$162k/yrHybridFull Time
West Monroe: Business and technology consultancy specializing in digital transformation.
4+ YOE4+ years building enterprise data platforms with Databricks, Snowflake or Microsoft Fabric; Spark/SQL/Python; cloud (AWS/Azure/GCP); Git/CI-CD; experience enabling analytics and AI; travel 30–50%; US work authorization required.
Databricks, Snowflake, Microsoft Fabric, AWS, Azure, GCP, Spark, SQL, Python, Azure Data Factory, Azure Storage, Azure Synapse, Git, CI/CD, ChatGPT
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
6+ YOEHands-on experience with Databricks, dbt, Python, PySpark, Spark, and SQL; 6+ years experience; proven track record in data platform migration to Databricks Lakehouse; leadership and mentoring experience.
Astrana HealthNASDAQ: ASTH: Operates technology-powered platforms to coordinate and deliver medical care.
2+ YOEBachelor's degree required; experience with relational databases, Python, Spark, SQL, Databricks, BI tools, cloud services (AWS/GCP/Azure), version control, and 2+ years in data/analytics landscape.
Python, Spark, SQL, Databricks, Tableau, Microsoft Power BI, Microsoft Excel, AWS, GCP, Azure
San Jose or Los Angeles or Singapore or New York City or London or Dublin or Paris or Berlin or Dubai or Jakarta or Seoul or Tokyo
$162k-$388k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
1+ YOERequires 1+ year in data warehouse design, production ETL with Hive, Spark, Hadoop, or SQL Server, production SQL, Python, or Java, distributed infrastructure debugging, and scalable data services.
Data Engineer, Prime Video - GSS Planning & Strategy
Culver City or Seattle
$132k-$179k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOEBachelor's degree, 3+ years data engineering, experience with Hadoop/Hive/Spark/EMR, SQL and Python, data modeling, ETL/ELT pipelines, BI tools (Tableau/QuickSight), and strong cross-team communication.
Publicis GroupeEuronext Paris: PUB: Global communications, advertising, and digital transformation holding.
3+ YOE3–5 years in data/analytics engineering, strong SQL and Python, experience with Databricks/AWS/Snowflake/Google Cloud, Spark/PySpark, ETL/ELT pipelines, and modern lakehouse architectures.
Databricks, AWS, Snowflake, Google Cloud, GitHub Copilot, Claude Code, Codex, Databricks Genie, Python, PySpark, SQL
Principal Data Engineer, Personalization - Central Product Insights
Los Angeles, California, United States
$210k-$293k/yrOnsiteFull Time
Riot Games: Develops and publishes video games and esports content.
8+ YOEBachelor's or master's degree in a related field, 8+ years in data engineering and tech-focused publishing or marketing, expertise in Python, GoLang, Spark, Scala, SQL, Airflow, dbt, cloud infrastructure, and Databricks.
Principal Data Engineer, Personalization - Central Product Insights
Los Angeles, California, United States
$210k-$293k/yrOnsiteFull Time
Riot Games: Developing and publishing competitive multiplayer video games.
8+ YOEBachelor's degree in CS or related,8+ years data engineering experience in publishing/marketing,expertise with Python,GoLang,Spark,Scala,SQL,Airflow,dbt,cloud (AWS/GCP),Databricks,mentoring and cross-functional collaboration.
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
6+ YOE2+ MgmtLead data engineering role requiring 6-8 years of experience, Databricks stack, SQL, Python, PySpark, data warehousing, Airflow, cloud platforms, and leadership responsibilities.
Austin or San Francisco or New York City or Seattle or Los Angeles or Chicago or Gurugram or Bengaluru
$170k-$230k/yrRemoteFull Time
SentiLink: Provides identity verification and fraud prevention for financial institutions.
5+ YOE5+ years engineering experience; strong Python or Golang skills; building ETL/ELT pipelines at scale with Spark/Hadoop/Kafka; cloud (AWS/Azure/GCP) and database expertise; containerization and IaC experience.