43 pyspark data engineer jobs at 28 companies in New Castle, DE
2d
Save
Mark Applied
Hide
2d
Hadoop and PySpark Data Engineer
Charlotte or Jacksonville or Jersey City or Newark or Plano or Delaware or Florida or New Jersey or North Carolina or Texas or United States
$82k-$127k/yrOnsiteFull Time
InfosysNYSE: INFY: Global provider of digital services, consulting, and technology solutions.
Bachelor's degree or equivalent experience; Hadoop, Hive, Spark, PySpark, Python, Unix shell scripting, SQL, data warehousing, orchestration tools, Agile, and production Big Data experience required.
InfosysNYSE: INFY: Provides IT consulting, software development, and business outsourcing services.
Experience with Hadoop ecosystem (Cloudera), PySpark, Spark, Hive, Python, Unix shell scripting, Autosys/Airflow, SQL; Bachelor\u0002s or equivalent experience; ability to work across SDLC and travel/relocate as needed.
Atlanta or New York City or Philadelphia or Dallas
OnsiteFull Time
KPMG: Global professional services network providing audit, tax, and advisory.
5+ YOE5+ Mgmt5+ years leading data science/engineering teams, Master’s degree in CS/DS/Statistics/Engineering, hands-on Python, PySpark, SQL, Databricks, Snowflake, experience with cloud platforms, ability to travel, U.S. work authorization required.
Python, PySpark, SQL, Databricks, Snowflake, Linux, AWS, GCP, Azure, R
Asset Based Lending: A private lending firm providing bridge financing and lending solutions for real estate investors.
Extensive data engineering experience building cloud data platforms, strong Python/PySpark and SQL skills, dbt and ELT tool experience, data governance and observability knowledge, and bachelor's in a quantitative field.
Python, PySpark, SQL, dbt, Power BI, Snowflake, Databricks, Microsoft Fabric, Fivetran, Airbyte, Collibra, Microsoft Purview, AWS, Azure, Google Cloud Platform
ifm electronic: Manufactures sensors and control systems for industrial automation.
5+ YOE5+ years data engineering; bachelor's in Computer Science/Data Science/Engineering; Azure data platform; SQL/Python/PySpark; ETL/ELT; DevOps; AI/advanced analytics exposure.
Azure Synapse Analytics, Azure Data Factory, Microsoft Fabric, Azure Data Lake Storage Gen2, Azure SQL Server, SQL, Python, PySpark
Medical Guardian: Provides medical alert systems and emergency monitoring for seniors.
10+ YOE3+ Mgmt10+ years in data or software engineering, 7+ years building production data pipelines, strong Databricks/Spark/PySpark/SQL experience, Azure and streaming expertise, plus leadership and production operations experience.
Booz Allen HamiltonNYSE: BAH: Consulting and technology services for government and commercial clients
5+ YOE5+ years application development experience, 3+ years building ETL/ELT workflows, experience with PySpark/Spark SQL and Palantir Foundry, ability to build/operate scalable data platforms, Secret clearance, Bachelor's degree.
LMI: Provides management consulting and digital solutions to federal government agencies.
5+ YOE5–8 years data engineering experience with expert Python, SQL, PySpark; experience with Palantir Foundry (Vantage) and Advana; active DoD Secret clearance; familiarity with Army data sources preferred.
LMI: Provides management consulting and digital solutions to government agencies.
5+ YOEActive DoD Secret clearance,5+ years data engineering experience,expert Python,SQL,PySpark,experience with Palantir Foundry/Advana,ability to scale ETL pipelines and manage ontologies.
Atlanta or Baltimore or Boston or Charlotte or Columbus or Detroit or Hartford or Miami or New York or Philadelphia or Raleigh or Tampa or Washington
$91k-$122k/yrOnsiteFull Time
Slalom: Provides business and technology consulting and software engineering services.
4+ YOE4+ years building production data engineering solutions; experience with cloud data platforms, streaming, orchestration, Python/Java; strong communication and critical thinking.
Campbell'sNasdaq: CPB: Manufacturer of branded soups, snacks, and beverage products.
Pursuing a degree in CS/Data Engineering/Data Science, basic SQL, programming in Python or PySpark, familiarity with Databricks, Snowflake, Microsoft Azure, ADF/ADLS and Microsoft Power BI, and strong problem-solving and communication skills.
Databricks, Snowflake, Microsoft Azure, ADLS, ADF, Microsoft Power BI, SAP, Python, PySpark, SQL
Lutron Electronics: Designs and manufactures lighting and shading control systems.
8+ YOERequired bachelor’s or master’s in CS/Data Science/Computer Engineering/Software Engineering, 8+ years related experience, expertise in data modeling, real-time/batch processing, cloud architectures, and SQL.
NBME: Develops and manages medical licensing and healthcare professional assessments.
7+ YOEBachelor's degree, 7+ years application development experience, 4+ years in big data and AWS, proficiency in Python/SQL/Scala/Java, distributed data tools, PySpark, NoSQL, Redshift, UNIX/Linux, process orchestration and CI/CD; experience with ML/LLM pipelines and vector DBs preferred.
Algonomy: AI-powered algorithmic customer engagement platform for retailers.
8+ YOE8+ years software development; experience with Big Data technologies; Scala, PySpark, Python; data pipeline migration, automation, and monitoring; AWS services and Databricks.
CACINYSE: CACI: Provides information technology and professional services to government clients.
5+ YOE5+ years in data engineering/science with TS/SCI clearance; BS in data science/CS; experience building data pipelines and using Python data libraries; experience with git.
CACI InternationalNYSE: CACI: Provides information technology and engineering services for government agencies.
5+ YOEActive TS/SCI clearance, B.S. in Data Science/AI/CS or related, 5+ years experience, building data pipelines using Python (NumPy, Pandas, Polars), version control (git/GitLab/Bitbucket), and contributing to AI/ML solutions.
NFI Industries: A logistics and supply chain providing transportation, warehousing, and fulfillment services.
5+ YOE5+ years data engineering experience, advanced T-SQL, PySpark/Python proficiency, Azure Synapse/ADLS/DevOps experience, ETL/ELT pipeline development, and ability to refactor legacy SQL into modern notebook-based pipelines.
T-SQL, PySpark, Python, Azure Synapse Analytics, Azure Pipeline, Azure Data Lake Storage (ADLS Gen2), Microsoft Fabric, OneLake, Lakehouse, Warehouse, Git, CI/CD, GitHub Copilot, Google Gemini