PwC
Posted 1y ago

IN_Senior Associate_Cloud Data Engineer_Data and Analytics_Advisory_Pan India

PwC
Bengaluru, India
OnsiteFull Time
Responsibilities
  • Designing data pipelines
  • Implementing data ingestion
  • Optimizing Spark job performance
Requirements
  • 4-7 years of experience in data engineering with a strong focus on cloud environments
  • Proficiency in PySpark or Spark, and proven experience with data ingestion
  • Transformation, and data warehousing
Technical tools mentioned
AWSAzureGCPPySparkSparkSQLPython

Job description

Line of Service

Advisory

Industry/Sector

Not Applicable

Specialism

Data, Analytics & AI

Management Level

Senior Associate

Job Description & Summary

At PwC, our people in data and analytics engineering focus on leveraging advanced technologies and techniques to design and develop robust data solutions for clients. They play a crucial role in transforming raw data into actionable insights, enabling informed decision-making and driving business growth.

In data engineering at PwC, you will focus on designing and building data infrastructure and systems to enable efficient data processing and analysis. You will be responsible for developing and implementing data pipelines, data integration, and data transformation solutions.


*Why PWC

At PwC, you will be part of a vibrant community of solvers that leads with trust and creates distinctive outcomes for our clients and communities. This purpose-led and values-driven work, powered by technology in an environment that drives innovation, will enable you to make a tangible impact in the real world. We reward your contributions, support your wellbeing, and offer inclusive benefits, flexibility programmes and mentorship that will help you thrive in work and life. Together, we grow, learn, care, collaborate, and create a future of infinite experiences for each other. Learn more about us.

At PwC, we believe in providing equal employment opportunities, without any discrimination on the grounds of gender, ethnic background, age, disability, marital status, sexual orientation, pregnancy, gender identity or expression, religion or other beliefs, perceived differences and status protected by law. We strive to create an environment where each one of our people can bring their true selves and contribute to their personal growth and the firm’s growth. To enable this, we have zero tolerance for any discrimination and harassment based on the above considerations. "

Responsibilities: 

We are seeking skilled and dynamic Cloud Data Engineers specializing in AWS, Azure, Databricks, and GCP. The ideal candidate will have a strong background in data engineering, with a focus on data ingestion, transformation, and warehousing. They should also possess excellent knowledge of PySpark or Spark, and a proven ability to optimize performance in Spark job executions. 

 Key Responsibilities: 

 - Design, build, and maintain scalable data pipelines for a variety of cloud platforms including AWS, Azure, Databricks, and GCP. 
- Implement data ingestion and transformation processes to facilitate efficient data warehousing. 
- Utilize cloud services to enhance data processing capabilities: 
 - AWS: Glue, Athena, Lambda, Redshift, Step Functions, DynamoDB, SNS. 

 - Azure: Data Factory, Synapse Analytics, Functions, Cosmos DB, Event Grid, Logic Apps, Service Bus. 

 - GCP: Dataflow, BigQuery, DataProc, Cloud Functions, Bigtable, Pub/Sub, Data Fusion. 

- Optimize Spark job performance to ensure high efficiency and reliability. 
- Stay proactive in learning and implementing new technologies to improve data processing frameworks. 
- Collaborate with cross-functional teams to deliver robust data solutions. 
- Work on Spark Streaming for real-time data processing as necessary. 

 Qualifications: 

- 4-7 years of experience in data engineering with a strong focus on cloud environments. 
- Proficiency in PySpark or Spark is mandatory. 
- Proven experience with data ingestion, transformation, and data warehousing. 
- In-depth knowledge and hands-on experience with cloud services(AWS/Azure/GCP): 
- Demonstrated ability in performance optimization of Spark jobs. 
- Strong problem-solving skills and the ability to work independently as well as in a team. 
- Cloud Certification (AWS, Azure, or GCP) is a plus. 
- Familiarity with Spark Streaming is a bonus.  

Mandatory skill sets: 

Python, Pyspark, SQL with (AWS or Azure or GCP)

Preferred skill sets: 

Python, Pyspark, SQL with (AWS or Azure or GCP) 

Years of experience required: 

4-7 years 

Education qualification: 

  • BE/BTECH, ME/MTECH, MBA, MCA 

Education (if blank, degree and/or field of study not specified)

Degrees/Field of Study required: Master of Engineering, Master of Business Administration, Bachelor of Engineering, Bachelor of Technology

Degrees/Field of Study preferred:

Certifications (if blank, certifications not specified)

Required Skills

PySpark, Python (Programming Language), Structured Query Language (SQL)

Optional Skills

Accepting Feedback, Accepting Feedback, Active Listening, Agile Scalability, Amazon Web Services (AWS), Analytical Thinking, Apache Hadoop, Azure Data Factory, Communication, Creativity, Data Anonymization, Database Administration, Database Management System (DBMS), Database Optimization, Database Security Best Practices, Data Engineering, Data Engineering Platforms, Data Infrastructure, Data Integration, Data Lake, Data Modeling, Data Pipeline, Data Quality, Data Transformation, Data Validation {+ 19 more}

Desired Languages (If blank, desired languages not specified)

Travel Requirements

Available for Work Visa Sponsorship?

Government Clearance Required?

Job Posting End Date

About PwC

Global provider of professional audit, tax, and consulting services.

Year founded
1998
Employees
364000
Organization type
Private
Subsidiaries
Headquarters
GB

Similar jobs

Cloud Data Engineer roles near Bengaluru, null
2w
Save
Mark Applied
Hide
Cloud Data Engineer - AI/ML
Bengaluru, Karnataka, India
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Design and build large-scale data and AI platforms, collaborate across data, application, and ML teams to enable generative AI adoption and enterprise solutions.
4w
Save
Mark Applied
Hide
Cloud Data Engineer
Bangalore or Hyderabad
OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
Bachelor's degree or equivalent experience; experience with Python/Java/Scala, Spark or Hadoop, and Google Cloud Platform; experience with data warehouses, ETL/ELT, NoSQL/MongoDB, Terraform/Ansible/Jenkins; strong communication and critical thinking.
Python, Java, Scala, Spark, Hadoop, Google Cloud Platform (GCP), NoSQL, MongoDB, SparkML, PostgreSQL, Oracle, Alloy DB, IaC, CICD, Terraform, Ansible, Jenkins
1mo
Save
Mark Applied
Hide
Data Engineering Senior Analyst
Bengaluru, Karnataka, India
HybridFull Time
The Cigna Group
The Cigna GroupNYSE: CI: Provides health insurance and pharmacy benefit management services.
5+ YOE5+ years in software/data engineering with cloud-native data solutions; expertise in AWS services, Terraform, Spark/Glue/Databricks, SQL, Python, CI/CD; experience mentoring teams and working with Gen-AI environments.
AWS Lambda, Amazon S3, IAM, KMS, API Gateway, Terraform, Spark, AWS Glue, Databricks, SQL, Python, Jenkins, GitHub Actions, GitLab CI, Azure AI, AWS Bedrock, GCP, Cursor, Jira
6mo
Save
Mark Applied
Hide
Cloud Data Engineer - BLR
Bangalore, Karnataka, India
OnsiteFull Time
Photon: A technology providing cloud and data services.
6+ YOE6-9 years of experience; required skills SQL, Python, Snowflake, AWS; nice-to-have CI/CD, communication, and development tooling experience.
SQL, Python, Snowflake, AWS, CI/CD, DevOps tools
3w
Save
Mark Applied
Hide
DBT Cloud Data Engineer
Bangalore or Mumbai or Pune
RemoteFull Time
Capgemini
CapgeminiEuronext Paris: CAP: Provides global IT consulting and digital transformation services.
8+ YOE8+ years IT experience with minimum 6 years in data warehouse/ELT on cloud, advanced dbt expertise, cloud data platform experience (Snowflake/BigQuery/Redshift), Python and PySpark proficiency, CI/CD and IaC knowledge, strong communication.
dbt, Snowflake, BigQuery, AWS Redshift, ADLS, S3, AWS Lambda, Azure Functions, Step Functions, Cloud Run, Azure Databricks, Azure Data Factory, Azure Synapse Analytics, AWS Glue, AWS EMR, Dataflow, Dataproc, Airflow, dbt Cloud, Python, PySpark, CI/CD
1mo
Save
Mark Applied
Hide
Salesforce Data Cloud Developer
Bengaluru, Karnataka, India
OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
5+ YOEDeep knowledge of Salesforce Data Cloud/CDP, SQL, data transformation, API integration (REST/SOAP), and data modeling; 5+ years experience; Bachelor of Computer Science.
Salesforce Data Cloud, Salesforce Marketing Cloud, Salesforce Sales Cloud, SQL, REST, SOAP
1mo
Save
Mark Applied
Hide
Azure Cloud & Data ETL Developer
Bengaluru or Hyderabad
OnsiteFull Time
CGI
CGINYSE: GIB: Provides information technology and business consulting services.
6+ YOE6+ years experience in data engineering/ETL with Azure Data services, strong SQL and Python/PySpark skills, experience with hybrid on-premise-to-Azure integrations and CI/CD.
Azure Data Factory, Azure Databricks, Azure Synapse Analytics, SQL, Azure Data Lake Storage (ADLS), Python, PySpark, Azure DevOps, GitHub Actions
3w
Save
Mark Applied
Hide
AI Data Engineering and Sr Cloud Specialist
Bengaluru, Karnataka, India
OnsiteFull Time
Bosch
Bosch: Global manufacturer of automotive and industrial engineering technology.
6+ YOEDesign and maintain scalable data pipelines using Python, Apache Spark, Hadoop, Kafka, Airflow and Azure; strong SQL/NoSQL and SRE skills; PhD/MTech/BE in Computer Science and ~6 years experience.
Python, Apache Spark, Hadoop, Kafka, Airflow, Azure, SQL, NoSQL, TensorFlow, PyTorch, Tableau, Matplotlib, Docker, Kubernetes