PwC
Posted 1w ago

IN_Senior Associate – AWS Data Engineer – Data and Analytics – Advisory – Bangalore

PwC
Bengaluru or Bangalore
OnsiteFull Time
Responsibilities
  • developing pipelines
  • building frameworks
  • managing scalability
Requirements
  • Requires 5+ years in data engineering or integration
  • Expert SQL
  • Python
  • PySpark
  • ETL pipelines
  • AWS Glue
  • Step Functions
  • Lambda, DMS
  • Production readiness, and high-volume data processing
Technical tools mentioned
Amazon Web Services (AWS)SQLPythonPySparkAWS GlueAWS Step FunctionsAWS LambdaAWS DMSAuroraSAP ODPPostgreSQLKafkaApache AirflowApache HadoopAzure Data FactoryDatabricks Unified Data Analytics Platform

Job description

Line of Service

Advisory

Industry/Sector

Not Applicable

Specialism

Data, Analytics & AI

Management Level

Senior Associate

Job Description & Summary

At PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design and develop robust data solutions for clients. They play a crucial role in transforming raw data into actionable insights, enabling informed decision-making and driving business growth.

In data engineering at PwC, you will focus on designing and building data infrastructure and systems to enable efficient data processing and analysis. You will be responsible for developing and implementing data pipelines, data integration, and data transformation solutions.

Why PWC

At PwC, you will be part of a vibrant community of solvers that leads with trust and creates distinctive outcomes for our clients and communities. This purpose-led and values-driven work, powered by technology in an environment that drives innovation, will enable you to make a tangible impact in the real world. We reward your contributions, support your wellbeing, and offer inclusive benefits, flexibility programmes and mentorship that will help you thrive in work and life. Together, we grow, learn, care, collaborate, and create a future of infinite experiences for each other. Learn more about us.

At PwC, we believe in providing equal employment opportunities, without any discrimination on the grounds of gender, ethnic background, age, disability, marital status, sexual orientation, pregnancy, gender identity or expression, religion or other beliefs, perceived differences and status protected by law. We strive to create an environment where each one of our people can bring their true selves and contribute to their personal growth and the firm’s growth. To enable this, we have zero tolerance for any discrimination and harassment based on the above considerations.

Job Description & Summary: 

We are looking for an experienced AWS Data Engineer with strong expertise in SQL, Python, and PySpark to design, build, and optimize scalable data ingestion and ETL/ELT pipelines. The role requires hands-on experience with AWS services such as Glue, Step Functions, Lambda, and DMS, along with a strong understanding of data engineering best practices, production readiness, and high-volume data processing. 

 

Responsibilities: 

  • Develop ETL/ELT pipelines using AWS services such as Glue, Step Functions, Lambda, and DMS. 

  • Implement direct-to-Aurora ingestion strategies for snapshots and delta loads. 

  • Ensure referential integrity, correct processing order, idempotency, and recovery mechanisms. 

  • Build monitoring, validation, and reconciliation frameworks for production data pipelines. 

  • Manage scalability and throughput during high-volume EOD ingestion workloads. 

Mandatory skill sets: 

  • Expert-level SQL development and performance optimization. 

  • Strong proficiency in SQL, Python, and PySpark. 

  • Hands-on experience with ETL pipelines and orchestration frameworks. 

  • Solid understanding of ACID-compliant data ingestion principles. 

  • Experience with schema evolution and CDC (Change Data Capture) patterns. 

  • Experience owning production readiness and end-to-end solution delivery. 

  • Exposure to cloud-native architectures and AWS tools such as Glue, Step Functions, Lambda, and DMS. 

Preferred skill sets: 

  • Experience with SAP ODP or enterprise data replication technologies. 

  • Familiarity with distributed PostgreSQL systems. 

  • Exposure to event-driven architectures such as Kafka. 

Years of experience required: 

Experience: 5–8 years 

  • Minimum 3 years of relevant work experience, typically reflecting 5+ years in data engineering, data integration, or related roles. 

  • Proven track record of building resilient and idempotent ingestion frameworks. 

  • Experience supporting high-volume, time-sensitive processing, especially EOD workloads. 

  • Strong operational mindset, including monitoring, alerting, and maintaining runbooks. 

Education Qualification:

BE, B.Tech, ME, M.Tech, MBA, MCA or equivalent qualification preferred

Education (if blank, degree and/or field of study not specified)

Degrees/Field of Study required: Bachelor of Engineering, Bachelor of Technology, MBA (Master of Business Administration)

Degrees/Field of Study preferred:

Certifications (if blank, certifications not specified)

Required Skills

Amazon Web Services (AWS), Data Engineering

Optional Skills

Accepting Feedback, Accepting Feedback, Active Listening, Agile Scalability, Amazon Web Services (AWS), Analytical Thinking, Apache Airflow, Apache Hadoop, Azure Data Factory, Communication, Creativity, Data Anonymization, Data Architecture Development, Database Administration, Database Management System (DBMS), Database Optimization, Database Security Best Practices, Databricks Unified Data Analytics Platform, Data Engineering, Data Engineering Platforms, Data Infrastructure, Data Integration, Data Lake, Data Modeling, Data Pipeline {+ 27 more}

Desired Languages (If blank, desired languages not specified)

Travel Requirements

Available for Work Visa Sponsorship?

Government Clearance Required?

Job Posting End Date

August 24, 2026

About PwC

Global provider of professional audit, tax, and consulting services.

Year founded
1998
Employees
364000
Organization type
Private
Subsidiaries
Headquarters
GB

Similar jobs

AWS Data Engineer roles near Bengaluru, Karnataka
2w
Save
Mark Applied
Hide
Sr AWS Data Engineer
Bangalore, Karnataka, India
HybridFull Time
Solventum
SolventumNYSE: SOLV: Provides medical technology and health information software solutions.
8+ YOEBachelor's in CS or related,8+ years building AWS cloud data platforms,experience with EMR,S3,Glue,RDS/Aurora,Redshift,Postgres,DynamoDB,Python,PySpark,Scala,SQL,Terraform,and AWS Lambda.
EMR, S3, Glue, RDS, Aurora, Redshift, PostgreSQL, DynamoDB, Python, PySpark, Scala, SQL, Terraform, AWS Lambda
3w
Save
Mark Applied
Hide
T&T- Engineering - Manager - AWS Data engineer - Multiple Location
Bengaluru or Pune or Hyderabad or Bhubaneswar or Chennai or Coimbatore
HybridFull Time
Deloitte
Deloitte: Professional services firm providing audit, consulting, and advisory services.
8+ YOE8+ years AWS data engineering experience with Glue, S3, Lambda, IAM, Step Functions; strong Python (PySpark) and SQL skills; experience with Redshift/Snowflake, EMR, data lakes, and IaC; client management and delivery experience.
AWS Glue, S3, Lambda, IAM, Step Functions, Kinesis, Kafka, SQS, Redshift, Snowflake, Glue Data Catalog, Athena, EMR, PySpark, SQL, Spark, Terraform, CloudFormation
1y
Save
Mark Applied
Hide
AWS Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
DataZymes
DataZymes: Data analytics platforms for the pharmaceutical industry.
3+ YOE3-8 years of experience with AWS services, data warehousing, ETL, and a degree in Computer Science or related field.
AWS S3, AWS Glue, PySpark, AWS EMR, Python, SQL, AWS RDS, AWS Redshift, AWS DynamoDB
4mo
Save
Mark Applied
Hide
Lead AWS Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
London Stock Exchange Group
London Stock Exchange GroupLondon Stock Exchange: LSEG: Provides financial market infrastructure and global data analytics services.
Design scalable data pipelines using Python/Spark; manage AWS data services; collaborate with stakeholders; ensure code quality and observability.
Python, Apache Spark, AWS (Glue, EMR, Lambda, S3), Apache Iceberg, CI/CD tools, Great Expectations, Deequ