408 spark data engineer jobs at 158 companies in Lakewood, NJ
2w
Save
Mark Applied
Hide
2w
Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI)
McLean or New York City or Richmond
$179k-$246k/yrOnsiteFull Time
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's degree, 4+ years application development, 2+ years big data, 1+ year cloud (AWS/Azure/GCP); experience with Python, SQL, Spark, data modeling, and cloud data platforms.
Bedrock Robotics: Automates heavy construction machinery with AI retrofit kits.
5+ YOE5+ years data engineering experience in large-scale data lake/warehouse environments; strong SQL and distributed query engine experience; pipeline orchestration (Airflow/Prefect) and Spark/Databricks experience; data quality and observability expertise.
FanDuelNYSE: FLUT: Offers online sports betting and daily fantasy sports services.
3+ YOE3+ years in data or software engineering with strong SQL, Python/Java/Scala, Databricks, Airflow, dbt, Spark, Kafka, and cloud (AWS/GCP/Azure) experience.
Sr Data Engineer, Python + Spark (Data Federation skillset - Data Lakehouse - Eg: Starburst) - New York
New York, New York, United States
$40k-$140k/yrOnsiteFull Time
Photon: Global technology services provider.
5+ YOESenior Data Engineer with 5+ years in Python, Spark; expertise in data federation, lakehouse architectures (Delta Lake/Iceberg/Hudi); Starburst/Trino/Dremio; cloud platforms (AWS/Azure/GCP); strong SQL and data modeling.
Bank of AmericaNYSE: BAC: Provides banking, investment, and financial risk management services.
5+ YOEBachelor's degree and 5+ years of data engineering experience required. Advanced SQL, Python, Spark, ETL, Airflow, data lake architecture, analytical, communication, and collaboration skills required.
Python, Spark, SQL, Apache Airflow, dbt, Git, Snowflake, Databricks, Delta Lake, HDFS, Hive, Oracle Exadata, SQL Server, Teradata, SSIS, Kafka, Tableau, Microsoft Power BI, SSRS, SSAS, CI/CD
United States or San Francisco or New York City or Washington or Austin
$118k-$179k/yrRemoteFull Time
SamsaraNYSE: IOT: Connected Operations Cloud platform for physical operations and IoT.
8+ YOEBachelor's in CS or equivalent, 8+ years as a software/data engineer, 5+ years building production data pipelines and Spark/PySpark experience, strong Python and SQL, cloud data warehouse and ETL tooling experience.
VersantNasdaq: VSNT: Operates cable television networks and digital media entertainment platforms.
5+ YOE5+ years building production data pipelines with Spark-based platforms, expertise in PySpark and SQL, Databricks experience preferred, bachelor's in CS/data engineering or equivalent, strong CI/CD, orchestration, and data governance skills.
Delta Lake, Databricks Auto Loader, Databricks, Structured Streaming, Change Data Capture (CDC), Unity Catalog, Spark, Photon, Databricks Workflows, Delta Live Tables, Airflow, MLflow, Presto, Flink, Git, Dagster, EMR, Dataproc, Lakeflow Spark Declarative Pipelines (SDP), ChatGPT
The Leading Hotels of the World: Provides marketing, sales, and reservation services for independent hotels.
7+ YOE7+ years in data engineering; Snowflake/BigQuery/Redshift/Databricks; ETL/ELT tools (Airflow, dbt, Spark); SQL and Python; large-scale data processing; data governance and security; strong cross-functional communication.
5+ YOE5+ years data engineering experience with Databricks, cloud-native data platforms, Python/SQL/Spark, GenAI/LLM experience, data pipeline and governance expertise for clinical/cross-study data.
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Currently pursuing a quantitative degree; familiarity with Python or scripting, data integration, statistical analysis, databases, SQL, Spark, cloud platforms, and data engineering; willingness to travel.
Python, SQL, Spark, IBM Cloud, Microsoft Azure, Amazon Web Services (AWS), Snowflake, Databricks, JavaScript
Lead Data Engineer (Python, AWS, Spark, Kafka, SQL, Snowflake, Databricks, GenAI)
Richmond or McLean or New York City
$179k-$246k/yrOnsiteFull Time
Capital OneNYSE: COF: A diversified financial services providing banking and credit products.
4+ YOEBachelor's degree, 4+ years of application development, 2+ years with big data technologies, and 1+ year with cloud computing. Preferred experience includes Python, SQL, Spark, AWS, streaming, data warehousing, and AI tools.
Point72: Global alternative investment firm managing capital and venture investments.
3+ YOERequires 3+ years in data engineering or a related field, strong Python development, experience with Spark or Scala, big data technology knowledge, and innovative problem-solving skills.
United States or Arlington or Tysons or Washington or New York City or Chicago or Austin or Atlanta or Boston or Boulder
$113k-$188k/yrRemoteFull Time
Guidehouse: Provides management and technology consulting services to diverse organizations.
3+ YOEBachelor's degree,3+ years data engineering experience,proficiency with Python/PySpark/SQL,Databricks/Spark/Delta Lake experience,knowledge of data pipelines,governance,and cloud platforms.
United States or San Francisco or New York City or Chicago
$170k-$230k/yrHybridFull Time
Komodo Health: Provides AI-driven healthcare data and patient journey analytics software.
Experience building production data pipelines at scale with advanced Python, SQL, Airflow, Spark, and AWS; healthcare data expertise and strong data quality, reliability, troubleshooting, and collaboration skills required.
Linero: AI-native, venture-backed building data infrastructure software.
5+ YOE5+ years data engineering; Airflow, dbt, Spark; streaming and batch processing; ontology design; multi-tenant, high-security data environments; based in or relocate to New York City.
Mizuho Financial GroupTokyo Stock Exchange: 8411: Global financial group providing banking and investment services.
0+ YOERequires 0–2 years in data engineering, analytics, or related technical work; SQL, Python, OOP, data concepts, and bachelor's degree in a related field. Experience with Databricks, Spark, cloud, ETL, and databases is valued.
Databricks, Microsoft Azure, Apache Spark, PySpark, SQL, Python, Git, Delta Lake, Amazon Web Services (AWS), Google Cloud Platform (GCP), Scala, SSIS, Azure Data Factory, Kafka
Balyasny Asset Management: Global multi-strategy investment firm managing diverse alternative asset classes.
3+ YOE3+ years data engineering experience; strong Python, SQL, Spark; Snowflake and cloud (AWS/Azure/GCP) experience; pipeline orchestration and data quality expertise.
Python, SQL, Spark, NoSQL, Snowflake, Hive, Hadoop, Airflow, Luigi, Oozie, NiFi, AWS, Azure, Google Cloud, Go
Bristol Myers SquibbNew York Stock Exchange: BMY: Develops and distributes innovative medicines for serious diseases.
5+ YOE5+ years data engineering experience with cloud platforms, Databricks (Delta Lake, Unity Catalog), Python, SQL, Spark/PySpark, GenAI/LLM (RAG, embeddings), and building production-grade ETL/ELT pipelines and data products.
Tatari: A platform for buying and measuring TV advertising campaigns.
5+ YOE5+ years building and operating production ETL pipelines; strong Python and SQL skills; Spark/PySpark, Databricks/Delta Lake, Airflow experience; strong data modeling and data quality practices.
Python, SQL, Spark, PySpark, Databricks, Delta Lake, Airflow, ClickHouse