1,920 pyspark data engineer jobs at 787 companies in United States

4w
Save
Mark Applied
Hide
Senior PySpark Data Developer
Irving, Texas, United States
$100k-$125k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesBSE: 532540: Global leader in IT services, consulting, and business solutions.
6+ YOE6+ years experience designing and building scalable PySpark/Spark data pipelines, cloud big-data tooling, Hive/SQL, optimized storage formats, and performance tuning.
Apache Spark, PySpark, Hive, Spark SQL, Spark UI, Parquet, ORC, Avro, Python, HiveQL, ANSI SQL, Amazon Web Services (AWS), AWS EMR, AWS Glue, Microsoft Azure, Azure Databricks, Azure Synapse, Google Cloud Platform (GCP), Apache Airflow, Git, Jenkins, Ansible, HBase, Cassandra, MongoDB
3w
Save
Mark Applied
Hide
Pyspark Developer
Hartford or Indianapolis or Phoenix or Raleigh or Richardson
OnsiteFull Time
Infosys
InfosysNYSE: INFY: Global leader in next-generation digital services and consulting.
Bachelor's degree or equivalent experience; experience with PySpark, AWS Redshift, AWS Glue and Gen AI; ability to work across SDLC, support production systems; US work authorization required; relocation/travel possible.
PySpark, AWS Redshift, Gen AI, AWS Glue
1mo
Save
Mark Applied
Hide
Data Engineer - PySpark
Tampa or Irving
$60k-$135k/yr HybridFull Time
Wipro
WiproNYSE: WIT: Global information technology, consulting, and business process services.
5+ YOE5+ years experience with PySpark in Big Data environments, Hadoop ecosystem, shell scripting, Autosys, complex SQL, streaming platforms, data modeling, and strong analytical and communication skills.
PySpark, Hadoop, Hive, HDFS, Sqoop, Spark, Impala, Scala, Autosys, SQL
3w
Save
Mark Applied
Hide
Data Engineer (PySpark)
Tampa, Florida, United States
$80k-$90k/yr OnsiteFull Time
Capgemini
CapgeminiEuronext Paris: CAP: Global leader in consulting, digital transformation, and engineering services.
5+ YOERequires a bachelor's degree or equivalent, 5–8 years of relevant experience, systems analysis and programming expertise, project implementation experience, and advanced HDFS, Spark, Python, PySpark, SQL, and data-platform skills.
HDFS, YARN, MapReduce, Spark, Flink, Zookeeper, Spark Core, Spark SQL, Spark Streaming, Spark GraphX, Python, PySpark, Pandas, NumPy, SQL, HBase, Cassandra, MongoDB, RDBMS
1mo
Save
Mark Applied
Hide
Big Data PySpark Lead Engineer - Vice President
Jersey City, New Jersey, United States
$142k-$213k/yr HybridFull Time
Citi
CitiNYSE: C: Global financial services organization enabling growth and economic progress.
6+ YOEExpertise in PySpark and Hadoop ecosystem, strong SQL, data modeling, shell/Autosys automation, 6+ years apps development/systems analysis experience, bachelor’s degree or equivalent.
PySpark, Hive, HDFS, Sqoop, Spark, Impala, Scala, SQL, Autosys, Apache Kafka, Shell scripting
2w
Save
Mark Applied
Hide
Hadoop and PySpark Data Engineer
Charlotte or Jacksonville or Jersey City or Newark or Plano or Delaware or Florida or New Jersey or North Carolina or Texas or United States
$82k-$127k/yr OnsiteFull Time
Infosys
InfosysNYSE: INFY: Global leader in next-generation digital services and consulting.
Bachelor's degree or equivalent experience; Hadoop, Hive, Spark, PySpark, Python, Unix shell scripting, SQL, data warehousing, orchestration tools, Agile, and production Big Data experience required.
Hadoop, PySpark, Cloudera, Hive, Spark, Python, Unix, Autosys, Airflow, SQL, Agile, Tableau, AI, cloud, containerization
1mo
Save
Mark Applied
Hide
Big Data PySpark Lead Engineer - Vice President
Jersey City, New Jersey, United States
$142k-$213k/yr HybridFull Time
Citi
CitiNYSE: C: Global financial services organization enabling growth and economic progress.
6+ YOEExpertise in PySpark and Hadoop ecosystem, complex SQL, data modeling, Autosys and shell scripting; 6+ years apps development or systems analysis experience; bachelor’s degree preferred.
PySpark, Hadoop, Hive, HDFS, Sqoop, Spark, Impala, Scala, SQL, Autosys, Apache Kafka, Shell scripting
1mo
Save
Mark Applied
Hide
Big Data PySpark Lead Engineer - Vice President
Jersey City, New Jersey, United States
$142k-$213k/yr HybridFull Time
Citi
CitiNYSE: C: Global financial services organization enabling growth and economic progress.
6+ YOERequires 6–10 years of applications development or systems analysis experience, PySpark and Hadoop expertise, complex SQL, distributed systems, data modeling, shell scripting, Autosys, and a bachelor's degree or equivalent.
PySpark, Hadoop, Hive, HDFS, Sqoop, Spark, Impala, Scala, SQL, shell scripting, Autosys, Apache Kafka
2w
Save
Mark Applied
Hide
Platform Data Engineer - (DataBricks, PySpark, AWS)
West Chester, Pennsylvania, United States
OnsiteFull Time
Comcast
ComcastNASDAQ: CMCSA: A global media and technology.
5+ YOEBachelor's degree or equivalent experience, 5+ years in data engineering, and hands-on AWS, PySpark, and Databricks experience. Requires Python, ETL/ELT, distributed systems, Airflow, Kubernetes, and production support expertise.
Amazon Web Services (AWS), PySpark, Databricks, Kafka, Kubernetes, Apache Airflow, Managed Workflows for Apache Airflow (MWAA), Python, Snowflake, Amazon EKS (Elastic Kubernetes Service)
1mo
Save
Mark Applied
Hide
Lead Data Engineer - Finance Technology (Hadoop, PySpark, Scala/Java)
Brooklyn Park, Minnesota, United States
$132k-$238k/yr HybridFull Time
Target
TargetNYSE: TGT: An American general merchandise retailer.
7+ YOE7+ years data engineering experience with Hadoop/Spark, PySpark, Scala/Java, cloud data services, API and streaming architectures; strong problem solving and leadership skills.
Hadoop, Spark, PySpark, Scala, Java, AWS, GCP, Azure, S3, BigQuery, Databricks, EMR, Snowflake, REST, GraphQL, gRPC, Kafka, Flink, Kinesis
1mo
Save
Mark Applied
Hide
Lead Data Engineer - Finance Technology (Hadoop, PySpark, Scala/Java)
Brooklyn Park, Minnesota, United States
$132k-$238k/yr HybridFull Time
Target
TargetNYSE: TGT: An American general merchandise retailer.
7+ YOE7+ years data engineering experience; BS/MS preferred; expertise with Hadoop/Spark, PySpark/Scala/Java, cloud platforms (AWS/GCP/Azure), APIs, streaming (Kafka/Flink/Kinesis), and data modeling.
Hadoop, Spark, Scala, Java, PySpark, AWS, GCP, Azure, S3, BigQuery, Databricks, EMR, Snowflake, REST, GraphQL, gRPC, Kafka, Flink, Kinesis
2w
Save
Mark Applied
Hide
Senior Data Engineer – Enterprise Data Hub (AWS + Snowflake + DBT + PySpark + CI/CD)
United States
OnsiteFull Time
Ness Digital Engineering
Ness Digital Engineering: Private global technology consulting firm delivering data, AI, cloud, and software engineering services to enterprise clients.
10+ YOERequires 10+ years of data engineering experience, strong AWS, Snowflake, DBT, PySpark, SQL, data modeling, CI/CD, deployment automation, and cloud data platform optimization expertise.
AWS, Snowflake, dbt, PySpark, SQL, CI/CD, Git
4d
Save
Mark Applied
Hide
Data Engineer III - Python/PySpark/Databricks/AI
Atlanta, Georgia, United States
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services and investment banking firm.
3+ YOEFormal data engineering training or certification and 3+ years of applied experience. Requires SQL, NoSQL, statistical data analysis, data lifecycle expertise, AI-output validation, and data sensitivity awareness.
Python, PySpark, Databricks, AI, SQL, NoSQL
1mo
Save
Mark Applied
Hide
Lead Data Engineer (Python, PySpark, AWS)
McLean or Richmond
$179k-$225k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: A technology-driven bank providing diverse financial services.
4+ YOEBachelor's degree, 4+ years of application development, 2+ years of big data experience, and 1+ year of cloud computing experience; preferred Python, SQL, Scala, Java, AWS, and distributed data tools.
Python, PySpark, AWS, Java, Scala, RDBMS, NoSQL, Redshift, Snowflake, machine learning, distributed microservices, SQL, Microsoft Azure, Google Cloud, MapReduce, Hadoop, Hive, EMR, Kafka, Spark, Gurobi, MySQL, Mongo, Cassandra, UNIX, Linux, shell scripting, Agile
1mo
Save
Mark Applied
Hide
Lead Data Engineer (Python, PySpark, AWS)
McLean or Richmond
$179k-$225k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: A technology-driven bank providing diverse financial services.
4+ YOEBachelor's degree, 4+ years application development experience, 2+ years big data experience, 1+ year cloud experience (AWS/Azure/GCP); preferred: extensive Python, SQL, Scala/Java, distributed data and cloud tooling.
Python, PySpark, Java, Scala, SQL, Redshift, Snowflake, MapReduce, Hadoop, Hive, EMR, Kafka, Spark, Gurobi, MySQL, Mongo, Cassandra, UNIX/Linux, AWS, Microsoft Azure, Google Cloud
2w
Save
Mark Applied
Hide
Data Engineer
United States
HybridFull Time
Varonis
VaronisNASDAQ: VRNS: Public cybersecurity software helping enterprises protect sensitive data across cloud, SaaS, and on-premises environments.
4+ YOERequires 4+ years in data engineering, cloud data solutions, large-scale data solutions, Python, PySpark, ETL/ELT, Databricks, Apache Spark, and end-to-end data solution leadership.
Databricks, Python, PySpark, Apache Spark
3mo
Save
Mark Applied
Hide
Data Engineer
Plano, Texas, United States
OnsiteFull Time
Ascentt
Ascentt: Enterprise AI and data solutions firm serving manufacturing and automotive organizations with analytics and intelligent platforms.
2+ YOE2-5 years data engineering; Databricks/Snowflake; SQL & Python; PySpark; ETL/ELT; cloud environments; Bachelor's degree.
Databricks, Snowflake, PySpark, SQL, Python, Airflow, dbt, Azure Data Factory, Git, Delta Lake, Kafka, Spark Streaming
3w
Save
Mark Applied
Hide
Data Engineer
Atlanta or Washington
$87k-$198k/yr OnsiteFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Global firm providing management, technology, and engineering consulting services.
3+ YOEBachelor’s degree and 3+ years with PySpark, SQL, and Azure Databricks; data engineering, analytics, technical collaboration, public health data, and Public Trust eligibility required.
PySpark, SQL, Azure Databricks, R, Scala, Java, AWS, Azure, GCP, Spark, Databricks, Hadoop, Hive, EMR, Kafka, Redshift, MySQL, Snowflake, FHIR, EHR, ETL, ELT
5d
Save
Mark Applied
Hide
Data Engineer
North America or Atlanta
RemoteContract
ProArch
ProArch: Technology consulting and IT services firm delivering cloud, data, cybersecurity, and custom software solutions to global clients.
8+ YOERequires a bachelor's or master's degree, 8+ years of data engineering experience, expertise in data modeling, governance, cloud platforms, Python, SQL, PySpark, Power BI, and scalable data architecture.
Erwin, SQL Data Modeler, AWS SageMaker, Glue ML, Databricks, AWS Glue Data Catalog, AWS Redshift, Amazon S3, Amazon EMR, AWS Lambda, Azure Synapse, Python, SQL, PySpark, Power BI, DAX, Power Query, Tableau, UDP framework
3w
Save
Mark Applied
Hide
Data Engineer
Illinois or Arkansas or Idaho or Georgia or Alabama or Iowa or Utah or California or Arizona or Kansas or Florida or Colorado or Connecticut or Indiana or Delaware or Kentucky
$65k-$173k/yr RemoteFull Time
CVS Health
CVS HealthNYSE: CVS: Diversified healthcare integrating retail, pharmacy, and insurance services.
4+ YOERequires 4+ years with SQL and relational databases, 2+ years in cloud and on-premises data engineering platforms, SQL/Python/PySpark, data modeling, warehousing, troubleshooting, and a bachelor's degree in computer science or engineering.
SQL, Amazon Web Services (AWS), Google Cloud Platform (GCP), Microsoft Azure, Databricks, Snowflake, Microsoft SQL Server, Oracle, Teradata, Python, PySpark