71 data pipeline engineer jobs at 44 companies in Cornwall, NY
2w
Save
Mark Applied
Hide
2w
Lead Data Engineer - Data Engineering 4C
New York City, New York, United States
$60k-$75k/yrHybridFull Time
GenpactNYSE: G: Provides business process management and digital transformation services.
Lead and mentor data engineers; design scalable data pipelines and ETL; ensure data quality, integrity, and security; collaborate with stakeholders; perform data analysis and troubleshooting.
Databricks, Snowflake, Microsoft Azure, Oracle Database 12c, Riversand, Apache Spark
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
4+ YOERequires 4+ years in data engineering, a bachelor's or master's degree, SQL, Python, PySpark, cloud data platforms, ETL/ELT pipelines, data modeling, Airflow, cloud ecosystems, and engineering team leadership.
SQL, Python, PySpark, Snowflake, Databricks, Apache Airflow, AWS, Azure, GCP, Tableau, Microsoft Power BI, Looker, Spark, Hadoop, Hive, HBase, Kafka
MDCalc: Clinical decision support platform providing medical calculators for healthcare professionals.
5+ YOE5+ years data engineering experience, strong SQL and Python, experience with cloud data warehouses (Snowflake), ETL/ELT pipelines, and tools like dbt, Airflow, or Dagster; strong data modeling and data architecture skills.
VersantNasdaq: VSNT: Operates cable television networks and digital media entertainment platforms.
5+ YOE5+ years building production data pipelines with Spark-based platforms, expertise in PySpark and SQL, Databricks experience preferred, bachelor's in CS/data engineering or equivalent, strong CI/CD, orchestration, and data governance skills.
Delta Lake, Databricks Auto Loader, Databricks, Structured Streaming, Change Data Capture (CDC), Unity Catalog, Spark, Photon, Databricks Workflows, Delta Live Tables, Airflow, MLflow, Presto, Flink, Git, Dagster, EMR, Dataproc, Lakeflow Spark Declarative Pipelines (SDP), ChatGPT
NYSTEC: Provides technology consulting services to New York public agencies.
5+ YOEDesign and maintain enterprise data pipelines, data warehouses/lakehouses, ETL/ELT, SQL, Python/PySpark, data modeling, governance, and platform optimization; bachelor’s degree and 5 years' related experience.
Microsoft Fabric, Azure Data Factory, Azure Data Lake Storage, Azure SQL, SQL Server Integration Services (SSIS), SQL Server, SQL, Python, PySpark
New York Blood Center: Non-profit organization collecting and distributing blood and stem cells.
6+ YOE6+ years data engineering experience; bachelor’s in CS/Data Science/IT or related; expert SQL and Python (PySpark); deep Azure data platform experience (ADF, Databricks, Synapse, ADLS); data modeling, pipeline ownership, governance, and HIPAA compliance.
SQL, Python, PySpark, Azure Data Factory, Azure Databricks, Azure Synapse Analytics, Azure Data Lake Storage, Microsoft Purview, Microsoft SQL Server, Oracle, CI/CD
Reality Defender: Detects deepfakes and AI-generated media to identify fraud.
Hands-on experience with Kubernetes and AWS, distributed computing (Spark, Ray), workflow orchestration (Airflow), Python and SQL; experience building multi-terabyte and streaming data pipelines for audio/video is preferred.
CapgeminiEuronext Paris: CAP: Provides global IT consulting and digital transformation services.
Experience with market data feeds, EDM vendors, medallion architecture, and MS Fabric; designing/maintaining data pipelines, validation and data quality rules; experience in asset management or financial services preferred.
MS Fabric, Microsoft Purview, Gresham, Golden Source, RIMES, Bloomberg, EDM
BarclaysLondon Stock Exchange: BARC: Global bank providing retail, corporate, and investment financial services.
Experience building Java-based data pipelines, data warehouses and lakes; strong skills in Java, Spark, Kafka, Kubernetes, OAuth2/JWT, JUnit, Mockito, GitLab/Bitbucket; experience leading technical teams and ensuring data governance and observability.
Subway: Global restaurant chain specializing in made-to-order submarine sandwiches.
5+ YOE5+ years building data pipelines and integrations with 3+ years cloud experience; hands-on Databricks or Snowflake, PySpark/SQL, Python, Airflow, Git; bachelor’s degree preferred.
Research Triangle Park or Yorktown Heights or Austin or Chicago or New York
$116k-$200k/yrHybridFull Time
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Bachelor's degree required; Master's preferred. Strong data pipeline design, ETL/ELT, SQL (PostgreSQL/Presto), Python, Airflow, Kafka, and data lakehouse experience. Visa sponsorship not provided. Hybrid work in multiple U.S. locations.
CognizantNasdaq: CTSH: Provides global information technology and business process outsourcing services.
Deep hands-on expertise with Coalesce and SQL, strong Snowflake proficiency, experience with transformation pipelines, CI/CD using GitHub Actions, dimensional modeling and data quality practices.
Bachelor's degree in computer science or engineering, or related experience; applied data engineering experience with programming, databases, cloud, DevOps, testing, and data pipeline technologies.
CognizantNASDAQ: CTSH: Provides IT consulting and technology services to global enterprises.
Expert-level Coalesce and SQL experience to build Bronze→Silver→Gold transformation pipelines on Snowflake, create reusable node templates, and integrate deployments into CI/CD.
Interactive BrokersNASDAQ: IBKR: Automated global electronic brokerage and trading services provider.
5+ YOE5+ years building data pipelines with advanced Python (Pandas or Polars); experience with Oracle/SQL optimization, Tableau, Linux; bachelor's degree in computer science or related field required.
Python, Java, Pandas, Polars, Oracle, SQL, Tableau, Linux
Senior Data Engineer - PGIM Technology (Hybrid - Newark, NJ)
Newark, New Jersey, United States
$115k-$155k/yrHybridFull Time
PGIMNYSE: PRU: Provides global investment management services across various asset classes.
3+ YOE3–6 years data/AI engineering experience; strong Python, SQL, Spark/PySpark; experience with data pipelines (ETL/ELT) and tools like Microsoft Fabric or Azure Data Factory; ML deployment and basic MLOps; knowledge of generative AI (embeddings, RAG).
Microsoft Fabric, Azure Data Factory, PySpark, Spark, SQL, Python, AWS Glue, REST APIs, GitHub Actions, Azure DevOps, Git
KSL Capital Partners: Private equity firm investing in travel and leisure businesses.
5+ YOEBachelor’s degree in a technical or related field and 5+ years in data engineering, architecture, or related roles; requires Snowflake, SQL, dbt, pipeline, API, and AI/ML data infrastructure experience.
Senior Software Engineer, Server Fleet Infrastructure
Livingston or New York or Sunnyvale or San Francisco or Bellevue
$153k-$204k/yrOnsiteFull Time
CoreWeaveNASDAQ: CRWV: Cloud platform providing GPU-accelerated infrastructure for AI workloads.
4+ YOEBachelor's degree, 4+ years data engineering experience, strong SQL, Python/Java/Scala, experience with data lakes, ETL pipelines, orchestration and observability tools (Airflow, Spark), and cloud platforms.