This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

Tata Consultancy Services
Posted 1mo ago

Pyspark/Python Data Engineer

Tata Consultancy Services
Irving, Texas, United States
$100k-$120k/yrOnsiteFull Time
Responsibilities
  • designing pipelines
  • developing pipelines
  • optimizing jobs
Requirements
  • Hands-on PySpark and Python experience for designing and optimizing ETL/data pipelines
  • Strong SQL skills
  • 6+ years data engineering experience preferred
Technical tools mentioned
PySparkPythonSpark SQLDataFrame APISQLAWSS3AWS GlueAmazon EMRAirflowSnowflakeGitCI/CD

Job description

Must Have Technical/Functional Skills

We are looking for a skilled PySpark Data Engineer with strong hands-on experience in PySpark and Python to design, build, and optimize scalable data processing pipelines. The ideal candidate will have practical experience working with distributed data processing and a solid foundation in writing efficient, production-grade Python code

Required Technical Skills

- Strong hands-on experience in PySpark (Spark SQL, DataFrame API)

- Advanced proficiency in Python (data processing, performance tuning, modular coding)

- Solid understanding of ETL design patterns and data pipeline architecture

- Good working knowledge of SQL for data transformation and analysis

- Experience with data processing in distributed environments

-  

Preferred Skills (Good to Have)

- Experience with cloud platforms (AWS preferred – S3, Glue, EMR or equivalent services)

- Familiarity with workflow orchestration tools such as Airflow or similar schedulers

- Exposure to data warehousing concepts (e.g., Snowflake or similar platforms)

- Knowledge of code versioning (Git) and CI/CD practices

Experience

• 3–8 years of experience in Data Engineering / PySpark development

• Proven hands-on project experience in PySpark + Python

Roles & Responsibilities

• Design, develop, and maintain ETL/ELT pipelines using PySpark

• Write optimized and scalable PySpark transformations using DataFrames and Spark SQL

• Develop reusable and efficient Python-based data processing components

• Ensure data quality, integrity, and performance across pipelines

• Perform debugging, performance tuning, and optimization of PySpark jobs

• Collaborate with cross-functional teams (Data Analysts, Architects, DevOps)

• Contribute to CI/CD pipelines and deployment workflows for data applications

• Monitor and troubleshoot data workloads in production environments





Salary Range: $100,000 to $120,000 per year

About Tata Consultancy Services

Global provider of IT services, consulting, and business solutions.

Similar jobs

Data Engineer roles near Irving, Texas
13h
Save
Mark Applied
Hide
Data Engineer
Plano or Teaneck
$65k/yr OnsiteFull Time
Cognizant
CognizantNASDAQ: CTSH: Provides IT consulting and technology services to global enterprises.
Bachelor's or master's degree in a related field; Python and SQL skills; data orchestration, cloud, data architecture, containerization, CI/CD, analytical, problem-solving, and communication skills.
Python, SQL, Apache Spark, AWS, Microsoft Azure, Google Cloud Platform (GCP), Apache Airflow, Prefect, Snowflake, Databricks, BigQuery, Docker, Kubernetes, GitHub
14h
Save
Mark Applied
Hide
Data Engineer
Plano or Teaneck
$65k/yr OnsiteFull Time
Cognizant
CognizantNasdaq: CTSH: Provides global information technology and business process outsourcing services.
Bachelor's or master's degree in a relevant field; Python and SQL skills; data pipelines, orchestration, cloud, data architecture, containerization, CI/CD, and ETL/ELT knowledge.
Python, SQL, Spark, AWS, Azure, GCP, Airflow, Prefect, Snowflake, Databricks, BigQuery, Docker, Kubernetes, GitHub, JSON, XML, ETL, ELT, CI/CD
20h
Save
Mark Applied
Hide
Data Engineer
Denton, Texas, United States
OnsiteFull Time
PACCAR
PACCARNasdaq: PCAR: Designs and manufactures heavy-duty commercial trucks and diesel engines.
5+ YOEBachelor's degree in a technical field and 5+ years of data engineering experience. Requires AWS, data pipelines, Python, SQL, databases, data warehousing, BI, and cross-functional communication skills.
AWS, Amazon EC2, Amazon ECS, Amazon S3, Amazon SNS, AWS Lambda, AWS IAM, Informatica, Attunity, Snowflake, Tableau, Kimball, SQL Server, PostgreSQL, MongoDB, Amazon DynamoDB, Python, SQL, Java, Scala, C#, Spark, dbt, Apache Airflow, Apache Kafka, Amazon Kinesis, Great Expectations, Soda, Monte Carlo, Terraform, AWS CloudFormation, Jenkins, AWS CodePipeline, Docker, Kubernetes, Flask, Plumber, Swagger, E-Verify
1d
Save
Mark Applied
Hide
Data Engineer
Mesa or Plano or Teaneck
$65k/yr HybridFull Time
Cognizant
CognizantNasdaq: CTSH: Provides IT consulting and digital business process services.
Bachelor’s or Master’s degree in a related field; Python and SQL skills; data pipeline, orchestration, cloud, data architecture, containerization, CI/CD, and analytics experience.
Python, SQL, Spark, Airflow, Prefect, Snowflake, Databricks, BigQuery, Docker, Kubernetes, GitHub, AWS, Azure, GCP, JSON, XML, ETL, ELT, CI/CD
1d
Save
Mark Applied
Hide
Lead Data Engineer
Plano, Texas, United States
$179k-$205k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Provides credit card, banking, and auto loan services.
4+ YOEBachelor's degree, 4+ years in application development, 2+ years in big data, and 1+ year in cloud computing. Preferred experience includes Java, Python, SQL, Scala, streaming, distributed data, and cloud platforms.
Flink, Spark Streaming, Java, Scala, Python, Spark, Kafka, Snowflake, AWS, Redshift, CI/CD, Microsoft Azure, Google Cloud, SQL, DynamoDB, OpenSearch, UNIX/Linux, Agile
1d
Save
Mark Applied
Hide
Lead Data Engineer
Plano, Texas, United States
$179k-$205k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: A diversified financial services providing banking and credit products.
4+ YOEBachelor's degree, 4+ years of application development, 2+ years of big data experience, and 1+ year of cloud computing experience; Java, Python, SQL, or Scala preferred.
Flink, Spark Streaming, Java, Scala, Python, Spark, Kafka, Snowflake, AWS Big Data Services, Redshift, CI/CD, machine learning, distributed microservices, AWS, Microsoft Azure, Google Cloud, SQL, DynamoDB, OpenSearch, UNIX/Linux, shell scripting
1d
Save
Mark Applied
Hide
Lead Data Engineer
Plano, Texas, United States
$179k-$205k/yr OnsiteFull Time
Capital One
Capital OneNYSE: COF: Financial services offering credit cards, banking, and loans.
4+ YOEBachelor's degree, 4+ years of application development, 2+ years of big data technologies, and 1+ year of cloud computing experience; preferred Java, Python, SQL, Scala, Flink, Kafka, Spark, and data warehousing.
Flink, Spark Streaming, Java, Scala, Python, Spark, Kafka, Snowflake, AWS Big Data Services, Redshift, CI/CD, Agile, machine learning, distributed microservices, Microsoft Azure, Google Cloud, SQL, DynamoDB, OpenSearch, UNIX/Linux, shell scripting
1d
Save
Mark Applied
Hide
Data Engineer
Irving, Texas, United States
$141k-$144k/yr HybridFull Time
CVS Health
CVS HealthNYSE: CVS: Provides retail pharmacy, health insurance, and pharmacy benefit management services.
2+ YOEMaster's degree in a related field and 2 years of experience required, including cloud platforms, SAS or SQL, visualization, Spark, ETL, production deployment, backend services, code reviews, and software collaboration.
Azure, Amazon Web Services (AWS), Google Cloud Platform (GCP), SAS, SQL, Power BI, Tableau, Spark, PySpark, Scala, Python, Java, Hadoop, HDFS
This job has expired