Mastercard
Posted 1mo ago

Data Engineer ||

Mastercard
Pune, Maharashtra, India
OnsiteFull Time
Responsibilities
  • developing pipelines
  • building ETL
  • ensuring quality
Requirements
  • 3+ years data engineering experience with PySpark and Python
  • Cloud platform familiarity (AWS/Azure/GCP)
  • ETL/ELT pipeline development
  • SQL and data modeling
  • Git and CI/CD
  • Bachelor's degree or equivalent experience
Technical tools mentioned
PySparkPythonS3GlueData FactoryDatabricksSQLAirflowdbtStep FunctionsPurviewAtlanLake FormationDockerTerraformGitCI/CD

Job description

Our Purpose

Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential.

Title and Summary

Data Engineer ||

Who is Mastercard?
Mastercard is a global technology company in the payments industry. Our mission is to connect and power an inclusive, digital economy that benefits everyone, everywhere by making transactions safe, simple, smart, and accessible. Using secure data and networks, partnerships and passion, our innovations and solutions help individuals, financial institutions, governments, and businesses realize their greatest potential.
Our decency quotient, or DQ, drives our culture and everything we do inside and outside of our company. With connections across more than 210 countries and territories, we are building a sustainable world that unlocks priceless possibilities for all.

Overview
The Mastercard Services Technology team is looking for a Data Engineer to help drive our mission of unlocking the potential of data assets by improving how we manage, process, store, and access large-scale data across both cloud and on-premise environments. This role will contribute to building scalable and reliable data solutions while supporting engineering standards and best practices in the Big Data ecosystem.
We are looking for a hands-on and motivated engineer with experience in PySpark, cloud platforms, and modern data engineering practices, who is eager to learn, grow, and collaborate with others. The individual will work closely with senior engineers and cross-functional teams to develop and maintain scalable data pipelines and cloud-native data solutions.
This role is ideal for engineers who enjoy solving data challenges, building efficient data pipelines, learning new technologies, and contributing to a collaborative engineering culture. Familiarity with AI-assisted development tools that improve engineering productivity and code quality will be an added advantage.

Role
• Develop and maintain scalable, cloud-based data pipelines and platforms using PySpark, Python, and modern data engineering practices.
• Build reliable ETL/ELT workflows to ingest, transform, and process data from multiple systems for analytics and business use cases.
• Contribute to the design and implementation of modular and maintainable data engineering solutions under the guidance of senior engineers.
• Write clean, efficient, and testable code while following established engineering standards and best practices.
• Participate in code reviews, debugging, performance optimization, and troubleshooting activities to improve platform reliability.
• Collaborate with cross-functional teams including product managers, data analysts, data scientists, and engineering teams to deliver data solutions.
• Support data quality, governance, and operational excellence initiatives including monitoring, lineage, and access management.
• Work with cloud-based data services and orchestration tools to support scalable and efficient data processing workflows.
• Participate in Agile ceremonies, sprint planning, feature estimation, and team discussions.
• Continuously learn new technologies, frameworks, and tools to improve technical and professional skills.
• Contribute to documentation, operational support, and knowledge-sharing within the team.
• Follow established security, compliance, and development processes while delivering high-quality solutions.

All About You
• 3+ years of hands-on experience in data engineering with working knowledge of PySpark and Python.
• Experience developing and maintaining data pipelines, ETL workflows, and batch/stream processing solutions.
• Familiarity with cloud platforms such as AWS, Azure, or GCP and exposure to services like S3, Glue, Data Factory, Databricks, or equivalent.
• Good understanding of SQL, data modeling concepts, and database fundamentals.
• Basic understanding of modern data architecture concepts such as data lakes, lakehouse, and medallion architecture.
• Familiarity with version control systems (e.g., Git), CI/CD concepts, and testing practices.
• Strong problem-solving skills with the ability to work collaboratively in a team environment.
• Good communication skills and willingness to learn from peers and senior engineers.
• Bachelor’s degree in Computer Science, Engineering, or related field—or equivalent practical experience.
• Comfortable working in Agile/Scrum development environments.
• Self-motivated, curious, and eager to learn modern data engineering technologies and practices.

Good to Have
• Exposure to orchestration and workflow tools such as Airflow, dbt, or Step Functions.
• Familiarity with data governance and cataloging concepts/tools such as Purview, Atlan, or Lake Formation.
• Basic exposure to containerization or infrastructure automation tools such as Docker or Terraform.
• Understanding of data quality, monitoring, and observability practices.
• Relevant cloud or data engineering certifications will be an added advantage.
• Exposure to machine learning data pipelines or MLOps concepts is a plus.

Corporate Security Responsibility


All activities involving access to Mastercard assets, information, and networks comes with an inherent risk to the organization and, therefore, it is expected that every person working for, or on behalf of, Mastercard is responsible for information security and must:

  • Abide by Mastercard’s security policies and practices;

  • Ensure the confidentiality and integrity of the information being accessed;

  • Report any suspected information security violation or breach, and

  • Complete all periodic mandatory security trainings in accordance with Mastercard’s guidelines.




About Mastercard

Global payment processing and financial technology service provider.

Similar jobs

Data Engineer roles near Pune, Maharashtra
2h
Save
Mark Applied
Hide
Data Engineer
Bangalore or Pune
OnsiteFull Time
Siemens Healthineers
Siemens HealthineersXetra: SHL: Manufacturer of medical diagnostic and imaging equipment.
10+ YOEBachelor's degree in computer science, data science, or related field; 10 years overall experience including 6 years in data engineering, data integration, ETL/ELT, pipelines, cloud platforms, SnapLogic, Kafka, Python, and SQL.
SnapLogic, AWS Glue, Amazon Web Services (AWS), Kafka, Confluent Cloud, Python, SQL, Microsoft Azure, Google Cloud Platform (GCP), CI/CD
6h
Save
Mark Applied
Hide
Data Engineer 1
Pune, Maharashtra, India
HybridFull Time
Cummins
CumminsNYSE: CMI: Manufacturer of engines, generators, and power systems.
Bachelor's degree or equivalent experience; experience with data integration, ETL/ELT, pipelines, data modeling, SQL, data quality, governance, cloud platforms, and Agile cross-functional work.
SPARK, Scala, Java, Map-Reduce, Hive, Hbase, Kafka, SQL, Hadoop, Cassandra, MongoDB, Accumulo, DynamoDB, DevOps, Scrum, Kanban, IoT, ETL, ELT, GenAI, AI/ML
11h
Save
Mark Applied
Hide
Data Engineer 1
Pune, Maharashtra, India
OnsiteFull Time
Cummins
CumminsNYSE: CMI: Manufacturer of diesel engines and power generation systems.
Bachelor's degree or equivalent experience; experience with ETL/ELT, data pipelines, transformations, data modeling, SQL, data quality, metadata, governance, cloud platforms, and Agile collaboration.
SQL, SPARK, Scala, Java, Map-Reduce, Hive, Hbase, Kafka, ETL, ELT, DevOps, Scrum, Kanban, Microsoft Excel
14h
Save
Mark Applied
Hide
Azure Senior Data Engineer
Pune, Maharashtra, India
OnsiteFull Time
HCLTech
HCLTechNational Stock Exchange of India: HCLTECH: Global technology providing digital, engineering, and cloud services.
Requires strong Azure, Databricks, ADF, Synapse, ETL, SQL, and relational database expertise; bachelor's degree in a related field; Python, Spark, client engagement, and Agile experience preferred.
Microsoft Azure, T-SQL, SSIS, SSAS, SSRS, Azure Data Factory (ADF), Azure Databricks, Azure Synapse Analytics, Azure DevOps, Azure Storage, Data Lake, ETL, SQL, Python, Spark APIs, Agile, Scrum
14h
Save
Mark Applied
Hide
Data Engineer
Pune or Bengaluru
OnsiteFull Time
Ascendion
Ascendion: AI-native digital engineering and software development services provider.
Requires Azure Data Factory, Azure SQL, ADLS Gen2, Python, JSON, Git, data reconciliation, source-to-target validation, and data quality testing; Azure Key Vault and DevOps are preferred.
Azure Data Factory, ADF, Azure SQL, ADLS Gen2, Python, JSON, Git, Pentaho KJB, Pentaho KTR, Azure Key Vault, Self-Hosted Integration Runtime, ARM, Azure DevOps
16h
Save
Mark Applied
Hide
Lead Associate - Data / AI
Pune, Maharashtra, India
HybridFull Time
Davies
Davies: Insurance claims management and legal professional services provider.
3+ YOE3+ MgmtBachelor's degree in computer science or related field; 3–5 years managing BI or data teams; experience building data pipelines and architectures with ADF; strong communication and process improvement skills.
Azure Data Factory, Microsoft Fabric, Microsoft SQL Server, Microsoft Power BI
23h
Save
Mark Applied
Hide
Data Engineer
Pune, Maharashtra, India
OnsiteFull Time
Barclays
BarclaysLondon Stock Exchange: BARC: Global bank providing retail, corporate, and investment financial services.
Experience with PySpark, AWS data technologies, Snowflake, DBT, advanced SQL/PL SQL, ETL pipelines, data warehousing, and cloud-based data platforms; strong stakeholder and analytical skills.
PySpark, AWS, AWS Glue, Amazon S3, AWS Lambda, AWS Lake Formation, Amazon Athena, Snowflake, Parquet, Apache Iceberg, JSON, CSV, DBT, SQL, PL/SQL, Immuta, Alation, Apache Airflow, Snowflake Tasks, Ab Initio, NoSQL
23h
Save
Mark Applied
Hide
D&T Lead - Data Engineer
Pune, Maharashtra, India
OnsiteFull Time
Aramex
AramexDubai Financial Market: ARMX: International express delivery and global logistics services provider.
5+ YOEBachelor's degree in computer science, engineering, or related field; 5–8 years of data engineering or BI experience; strong SQL and Python; ETL/ELT, databases, cloud platforms, data pipelines, and troubleshooting experience.
SQL, Python, Azure, AWS, GCP, Git, JIRA, Confluence