Iris Software
Posted 1mo ago

Lead Data Engineer

Iris Software
Noida, Uttar Pradesh, India
OnsiteFull Time
Responsibilities
  • developing pipelines
  • overseeing pipelines
  • ensuring quality
Requirements
  • Bachelor's in CS/IT,8+ years in data migration/ETL,proficiency in Python,advanced SQL,PySpark,Spark,SSIS and AWS data services
  • Experience with data modelling,CI/CD and strong communication
Technical tools mentioned
PythonAdvanced SQLSparkPySparkAWS LambdaAWS GlueAmazon EMRDynamoDBAmazon S3Amazon RedshiftAmazon EC2Amazon RDSAWS DMSAWS Step FunctionsAmazon CloudWatchAWS Systems ManagerAmazon EventBridgeBoto3AirflowGitHubGitLabBitbucketKafkaKinesisPower BIPostgreSQLDBTTerraformTypeScriptSSISDatabricks

Job description

Why Join Iris?
Are you ready to do the best work of your career at one of India’s Top 25 Best Workplaces in IT industry? Do you want to grow in an award-winning culture that truly values your talent and ambitions?
Join Iris Software — one of the fastest-growing IT services companies — where you own and shape your success story.

 


About Us  
At Iris Software, our vision is to be our client’s most trusted technology partner, and the first choice for the industry’s top professionals to realize their full potential.

With over 4,300 associates across India, U.S.A, and Canada, we help our enterprise clients thrive with technology-enabled transformation across financial services, healthcare, transportation & logistics, and professional services.

Our work covers complex, mission-critical applications with the latest technologies, such as high-value complex Application & Product Engineering, Data & Analytics, Cloud, DevOps, Data & MLOps, Quality Engineering, and Business Automation.


Working with Us
At Iris, every role is more than a job — it’s a launchpad for growth.

Our Employee Value Proposition, “Build Your Future. Own Your Journey.” reflects our belief that people thrive when they have ownership of their career and the right opportunities to shape it.

We foster a culture where your potential is valued, your voice matters, and your work creates real impact. With cutting-edge projects, personalized career development, continuous learning and mentorship, we support you to grow and become your best — both personally and professionally.

Curious what it’s like to work at Iris? Head to this video for an inside look at the people, the passion, and the possibilities. Watch it here.

Job Description

Must Have -

• Python & Advanced SQL
• Spark / distributed processing / PySpark
• Cloud platforms (AWS) – Modules like Lambda, Glue, EMR, Dynamo DB, S3, Redshift, EC2, RDS, DMS, Step Functions, Cloud Watch, Systems Manager, Event Bridge, DMS, Boto3 API’s
• Data pipeline automation

• Data governance & quality. Preferred if worked on AWS Data Zone.

• Any orchestrator like Airflow, Step Functions

• Knowledge of any git repository like GitHub, GitLab, Bitbucket.

 

Nice to have -
• Kafka/Kinesis streaming
• Basic knowledge of AI
• Power BI

• Postgre SQL

• DBT

• Experience with CI/CD implementation using Terraform or TypeScript.

 

Job Summary: 

We are seeking a Lead Data Engineer who is able to develop and maintain ETL Pipelines preferably on AWS cloud. 
Should be able to lead a project that includes working with client and come up with suggestions as and when required.
Oversee data pipelines, data lakes, data warehouses, and ETL/ELT processes.
Establish data quality standards, monitor system health and drive continuous improvements.
Should be proficient in Python, advanced SQL and PySpark. 
Sound understanding of Spark architecture. 
Basic understanding of AI is essential.
Knowledge of DBT, Postgre database and Power BI would be an added advantage.

Key Responsibilities: 
• Hands on in Python and advanced SQL. 
• Data modelling with good understanding of dimensions and facts.
• Good understanding of Spark eco-system and PySpark programming. 
• Able to convert existing SSIS workflows to ETL pipelines running on AWS. That includes knowledge of Lambda, Glue or EMR, Systems Manager, Event Bridge, EC2, DMS
• Knowledge of Dynamo DB, Redshift, RDS (Preferably Postgre) 
• Sound knowledge of AWS Boto3 API’s 
• Awareness of reporting tool like Power BI. 
• Monitor and troubleshoot migration processes, resolving issues in a timely manner. 
• Work closely with business stakeholders, developers, and administrators to understand data requirements. 
• Good understanding of a Data Lakehouse 
• Understanding of CI/CD process with exposure to TypeScript and any Git repository like GitHub, GitLab or Bitbucket
• Document migration procedures, processes, and best practices. 

Required Skills and Qualifications: 
• Bachelor’s degree in Computer Science, Information Technology, or related field. 
8+ years of experience in data migration or similar roles. 
• Strong proficiency in SSIS. 
• Proven experience in ETL tools like SSIS and data migration pipelines build on AWS. 
• Familiarity with data modeling and schema design principles. 
• Ability to analyze and resolve complex data-related issues. 
• Strong analytical and problem-solving abilities. 
• Excellent communication and collaboration skills. 
• Ability to work independently and in a team-oriented environment.

Preferred: 
• Familiarity with Agile project management methodologies. 

 

Mandatory Competencies

Data Science and Machine Learning - Data Science and Machine Learning - Python
Data Science and Machine Learning - Data Science and Machine Learning - Apache Spark
Data & AI - Data Engineering - Data Quality & Validation
Big Data - Big Data - Pyspark
Database - Database Programming - SQL
Data Science and Machine Learning - Data Science and Machine Learning - Databricks
Beh - Communication and collaboration

Perks and Benefits for Irisians
Iris provides world-class benefits for a personalized employee experience. These benefits are designed to support financial, health and well-being needs of Irisians for a holistic professional and personal growth. Click here to view the benefits.

About Iris Software

Provides software engineering and IT consulting services to enterprises.

Year founded
1991
Employees
4300
Organization type
Private
Latest investment
Raised $948.00k Grant (2008)
Headquarters
US

Similar jobs

Data Engineer roles near Noida, Uttar Pradesh
6h
Save
Mark Applied
Hide
Senior Snowflake Data Engineer
Noida, Uttar Pradesh, India
HybridFull Time
Ibex
IbexNASDAQ: IBEX: Provides technology-enabled customer experience and business process outsourcing services.
7+ YOERequires 7+ years in data engineering, Snowflake, advanced SQL, Python, Snowpark, ELT/ETL, dbt, cloud platforms, Git, CI/CD, testing, and data modeling.
Snowflake, SQL, Python, Snowpark, Dynamic Tables, Streams, Tasks, dbt, AWS, Azure, GCP, Git, CI/CD, Fivetran, Kafka, Airflow, Terraform, Snowflake Iceberg Tables, Snowflake Cortex, Cortex Analyst, Cortex Search, Cortex LLM
10h
Save
Mark Applied
Hide
Associate Platform Services - Data Engineer
Pune or Gurgaon
HybridFull Time
ZS
ZS: Global management consulting and technology firm for healthcare and life sciences.
1+ YOEBachelor's degree required; 1–2 years of development experience, ETL, SQL, Python, data modeling, data warehousing, data products, cloud platforms, and strong analytical and communication skills.
ZAIDYN, SQL, Python, AWS, Azure, Hadoop, Spark, PySpark, Informatica, Talend, SSIS
16h
Save
Mark Applied
Hide
Data Engineer-Senior II
Bengaluru or Gurugram or Mumbai
OnsiteFull Time
FedEx
FedExNYSE: FDX: Global provider of courier, logistics, and transportation services.
4+ YOEBachelor's degree in a relevant quantitative or technical field required; master's preferred. Requires 4–7 years' experience, Python/PySpark/SAS, cloud, data engineering, SQL, ETL, data modeling, and English fluency.
Python, PySpark, SAS, Hadoop, Hive, Spark, Azure, Azure Data Factory (ADF), Azure Data Lake Storage (ADLS), Azure DevOps, Databricks, Delta Lake, Docker, CI/CD, Kubernetes, Terraform, Octopus, SQL, Ab Initio, Informatica, DataStage, Power BI
20h
Save
Mark Applied
Hide
Data Engineer-Senior II
Bengaluru or Mumbai or Gurugram
OnsiteFull Time
FedEx
FedExNYSE: FDX: Global provider of express delivery and logistics services.
4+ YOEBachelor's degree in a relevant discipline; 4–7 years' experience; Python, PySpark, SAS, SQL, data pipelines, cloud platforms, ETL, data modeling, and distributed data technologies.
Python, PySpark, SAS, Hadoop, Hive, Spark, Microsoft Azure, Azure Data Factory (ADF), ADLS Storage, Azure DevOps, Databricks, Delta Lake, Workflows, Docker, CI/CD, Kubernetes, Terraform, Octopus, SQL, Ab Initio, Informatica, DataStage, Power BI
1d
Save
Mark Applied
Hide
Technical Lead - Data Engineer (Data&AI)
Gurgaon, Haryana, India
OnsiteFull Time
Srijan
Srijan: Develops digital platforms and engineering solutions for global brands.
5+ YOE5+ years in data engineering with large-scale data warehouses and lakes, CDC, batch and streaming processing, multiple cloud platforms, advanced SQL, Python, APIs, Airflow, Airbyte, CI/CD, Docker, Kubernetes, and MLOps.
Databricks, Apache Spark, Delta Lake, Snowflake, Amazon Web Services (AWS), Microsoft Azure, Apache Airflow, Airbyte, Docker, Kubernetes, Splunk, Datadog, Dynatrace, Git, GitLab, Python, FastAPI, Flask, Apache Kafka, Spark Streaming, SQL, REST APIs, LLMs
1d
Save
Mark Applied
Hide
Data Engineer-Data Platforms-Google
Gurgaon, Haryana, India
HybridFull Time
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Bachelor's degree required; master's preferred. Experience with Google Cloud data platforms, batch and real-time pipelines, data migration, data layer design, and tools including Airflow, dbt, Spark, Python, and Scala.
Google DataProc, Google DataFlow, Google PubSub, Google BigQuery, Google BigTable, Google Cloud Spanner, Google CloudSQL, Google AlloyDB, Google Cloud Storage, Apache Airflow, dbt, Spark, Python, Scala, Hadoop, Apache Beam, Google Cloud Scheduler, Cloud Composer
1d
Save
Mark Applied
Hide
Lead Data Engineer - Azure + Fabric
Gurugram, Haryana, India
HybridFull Time
EXL
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
8+ YOERequires 8+ years of data engineering experience, Azure or Microsoft Fabric expertise, Spark/PySpark, Python, SQL, Delta Lake, ETL/ELT, data quality, and a bachelor's degree in a relevant field.
Microsoft Azure, Microsoft Fabric, Microsoft OneLake, Microsoft Fabric Lakehouse, Microsoft Fabric Warehouse, Microsoft Fabric Data Pipelines, Notebooks, Spark, PySpark, Python, SQL, Delta Lake, Power BI, Microsoft Purview, Microsoft Azure DevOps, Git, GCP, CSV, Excel, Microsoft SharePoint, Synapse Pipelines, CI/CD, RBAC
1d
Save
Mark Applied
Hide
Snr Data Engineer
Noida, Uttar Pradesh, India
OnsiteFull Time
Alight
AlightNYSE: ALIT: Provides cloud-based HR, payroll, and benefits administration services.
4+ YOERequires 4–8 years of ETL experience, AWS and Hadoop expertise, Scala or Python, PySpark, Spark, SQL, orchestration tools, data warehousing, CI/CD, GitHub, and strong analytical and communication skills.
Apache Spark, Cloudera, AWS Step Functions, AWS Glue, AWS Lambda, Amazon S3, Amazon Redshift, Hadoop, HDFS, Apache Hive, Apache Kafka, Amazon EMR, Kinesis, PySpark, Spark SQL, Parquet, ORC, Apache Airflow, Control-M, Scala, HiveQL, Impala, SQL, Python, Shell, CI/CD, GitHub, Docker, Amazon ECS, Amazon EKS