87 spark data engineer jobs at 39 companies in Wrightwood, CA
2w
Save
Mark Applied
Hide
2w
Senior Data Engineer
Los Angeles, California, United States
$110k-$244k/yrHybridFull Time
UCLA Health: Provides comprehensive hospital services and academic medical research.
5+ YOE5+ years data engineering experience; strong Python and SQL; experience with Databricks, Spark, data lakehouse, ETL/ELT, cloud platforms; bachelor's degree in a technical field.
Python, SQL Server, Oracle, PostgreSQL, Databricks, Spark, Delta Lake, Snowflake, Microsoft Fabric, Synapse, Azure, AWS, GCP, Azure Data Factory, Databricks Workflows, Airflow, MLflow, Git
United States or Foster City or Orange County or Morristown
RemoteFull Time
Bridgepointe Technologies: Vendor-agnostic IT strategy and technology procurement services.
5+ YOERequires 5+ years in data engineering, advanced SQL, Python, Spark/PySpark, Microsoft Fabric, Azure data services, APIs, data modeling, CI/CD, Git, and Azure DevOps; bachelor's degree or equivalent experience.
Microsoft Fabric, Lakehouse, Warehouse, Data Pipelines, Notebooks, OneLake, Git, Azure DevOps, Spark, PySpark, SQL, Python, Azure, Delta Lake, Parquet, REST APIs, JSON, Power BI, DP-600, DP-700
VersantNasdaq: VSNT: Operates cable television networks and digital media entertainment platforms.
5+ YOE5+ years building production data pipelines with Spark-based platforms, expertise in PySpark and SQL, Databricks experience preferred, bachelor's in CS/data engineering or equivalent, strong CI/CD, orchestration, and data governance skills.
Delta Lake, Databricks Auto Loader, Databricks, Structured Streaming, Change Data Capture (CDC), Unity Catalog, Spark, Photon, Databricks Workflows, Delta Live Tables, Airflow, MLflow, Presto, Flink, Git, Dagster, EMR, Dataproc, Lakeflow Spark Declarative Pipelines (SDP), ChatGPT
Tatari: A platform for buying and measuring TV advertising campaigns.
5+ YOE5+ years building and operating production ETL pipelines, strong Python and SQL skills, Spark/PySpark, Databricks/Delta Lake, Airflow, data modeling and data quality expertise.
Python, SQL, Spark, PySpark, Databricks, Delta Lake, Airflow, ClickHouse
ParamountNASDAQ: PSKY: Produces and distributes media content across global entertainment platforms.
2+ YOE2+ years building ETL/ELT pipelines, strong SQL and Python, experience with Airflow, Spark, Kafka/Pub-Sub, cloud-native (GCP preferred), data modeling, and observability for analytics/ML.
Boston or Chicago or Dallas or Los Angeles or Minneapolis or New York City or San Francisco or Seattle or Washington, D.C.
$120k-$162k/yrHybridFull Time
West Monroe: Business and technology consultancy specializing in digital transformation.
4+ YOE4+ years building enterprise data platforms with Databricks, Snowflake or Microsoft Fabric; Spark/SQL/Python; cloud (AWS/Azure/GCP); Git/CI-CD; experience enabling analytics and AI; travel 30–50%; US work authorization required.
Databricks, Snowflake, Microsoft Fabric, AWS, Azure, GCP, Spark, SQL, Python, Azure Data Factory, Azure Storage, Azure Synapse, Git, CI/CD, ChatGPT
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
6+ YOEHands-on experience with Databricks, dbt, Python, PySpark, Spark, and SQL; 6+ years experience; proven track record in data platform migration to Databricks Lakehouse; leadership and mentoring experience.
Astrana HealthNASDAQ: ASTH: Operates technology-powered platforms to coordinate and deliver medical care.
2+ YOEBachelor's degree required; experience with relational databases, Python, Spark, SQL, Databricks, BI tools, cloud services (AWS/GCP/Azure), version control, and 2+ years in data/analytics landscape.
Python, Spark, SQL, Databricks, Tableau, Microsoft Power BI, Microsoft Excel, AWS, GCP, Azure
Cotality: Provides property data, analytics, and workflow intelligence solutions.
7+ YOE7+ years IT experience, bachelor's in CS/Engineering or equivalent, expertise in cloud data platforms, Python/PySpark, data modeling, big data technologies, graph DBs, orchestration tools, and enterprise data architecture.
AWS, Google Cloud Platform (GCP), Python, PySpark, Neo4j, Google Spanner, Hadoop, Spark, Elasticsearch, Google Cloud Dataflow, Apache Beam, Apache Airflow, BMAD, Spec Kit, SQL, NoSQL
San Jose or Los Angeles or Singapore or New York City or London or Dublin or Paris or Berlin or Dubai or Jakarta or Seoul or Tokyo
$162k-$388k/yrOnsiteFull Time
TikTok: Global short-form video hosting and social media platform.
1+ YOERequires 1+ year in data warehouse design, production ETL with Hive, Spark, Hadoop, or SQL Server, production SQL, Python, or Java, distributed infrastructure debugging, and scalable data services.
Adastra: Global provider of data, AI, and digital transformation consulting services.
5+ YOEBachelor's degree in a related field and 5+ years of data engineering experience. Requires Databricks, PySpark, Spark SQL, SQL, Python, cloud, data modeling, ETL/ELT, Git, CI/CD, and testing expertise.
Databricks, Databricks Lakehouse Platform, PySpark, Spark SQL, Delta Lake, Databricks Workflows, Unity Catalog, SQL, Python, Git, CI/CD, AWS, Azure, GCP, Power BI, Structured Streaming, Auto Loader, Kafka, Kinesis, Braze
Data Engineer, Prime Video - GSS Planning & Strategy
Culver City or Seattle
$132k-$179k/yrOnsiteFull Time
AmazonNASDAQ: AMZN: Global online retail and cloud computing technology provider.
3+ YOEBachelor's degree, 3+ years data engineering, experience with Hadoop/Hive/Spark/EMR, SQL and Python, data modeling, ETL/ELT pipelines, BI tools (Tableau/QuickSight), and strong cross-team communication.
Spokeo: People search engine that aggregates public records and data.
7+ YOE7+ years in production data engineering; 5+ years with AWS, EMR, Python, Spark, SQL, data modeling, and Airflow; 2+ years with non-relational databases; bachelor's degree required.
Principal Data Engineer, Personalization - Central Product Insights
Los Angeles, California, United States
$210k-$293k/yrOnsiteFull Time
Riot Games: Developing and publishing competitive multiplayer video games.
8+ YOEBachelor's degree in CS or related,8+ years data engineering experience in publishing/marketing,expertise with Python,GoLang,Spark,Scala,SQL,Airflow,dbt,cloud (AWS/GCP),Databricks,mentoring and cross-functional collaboration.
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
6+ YOE2+ MgmtLead data engineering role requiring 6-8 years of experience, Databricks stack, SQL, Python, PySpark, data warehousing, Airflow, cloud platforms, and leadership responsibilities.
Austin or San Francisco or New York City or Seattle or Los Angeles or Chicago or Gurugram or Bengaluru
$170k-$230k/yrRemoteFull Time
SentiLink: Provides identity verification and fraud prevention for financial institutions.
5+ YOE5+ years engineering experience; strong Python or Golang skills; building ETL/ELT pipelines at scale with Spark/Hadoop/Kafka; cloud (AWS/Azure/GCP) and database expertise; containerization and IaC experience.
San Francisco or New York City or Los Angeles or Seattle
$180k-$260k/yrRemoteFull Time
Whatnot: Social marketplace for buying and selling via live streams
5+ YOE5+ years building data warehouses or distributed/event-driven systems; skilled with data modeling, modern data tooling, cloud warehouses, Python/SQL, and cross-functional partnership.
Kafka, Debezium, dbt, Spark, Flink, Dagster, Airflow, Monte Carlo, Great Expectations, Snowflake, BigQuery, Redshift, Python, SQL, CI/CD
Cotality: Provider of property intelligence, data, and analytics solutions.
7+ YOE7+ years IT experience in data architecture; hands-on GCP, Dataflow, BigQuery, Python/PySpark, Airflow; experience with graph and SQL/NoSQL databases; strong communication and leadership skills.
Hyundai AutoEver AmericaKorea Exchange: 307950: Provides automotive software and IT services for mobility systems.
7+ YOEBachelor's degree required; 7+ years in data warehouse or MDM applications, strong SQL, Python and PL/SQL, relational databases, ETL/ELT pipelines, and production data workload experience.
Python, Apache Airflow, SQL, PL/SQL, Azure Data Lake, AWS, GCP, Spark, PySpark, Git, CI/CD, DevOps