Iris Software
Posted 19h ago

Data Engineer Python - Senior Engineer

Iris Software
Noida, Uttar Pradesh, India
OnsiteFull Time
Responsibilities
  • designing pipelines
  • architecting workflows
  • mentoring engineers
Requirements
  • Requires Python
  • PySpark
  • Apache Kafka
  • Databricks
  • Delta Lake
  • Snowflake or SQL
  • Data pipeline architecture
  • Streaming
  • Workflow orchestration
  • Data quality, and stakeholder collaboration
Technical tools mentioned
PythonPySparkApache KafkaDatabricks WorkflowsDelta LakeDatabricksSnowflakeAmazon KinesisApache AirflowSQL

Job description

Why Join Iris?
Are you ready to do the best work of your career at one of India’s Top 25 Best Workplaces in IT industry? Do you want to grow in an award-winning culture that truly values your talent and ambitions?
Join Iris Software — one of the fastest-growing IT services companies — where you own and shape your success story.

 


About Us  
At Iris Software, our vision is to be our client’s most trusted technology partner, and the first choice for the industry’s top professionals to realize their full potential.

With over 4,300 associates across India, U.S.A, and Canada, we help our enterprise clients thrive with technology-enabled transformation across financial services, healthcare, transportation & logistics, and professional services.

Our work covers complex, mission-critical applications with the latest technologies, such as high-value complex Application & Product Engineering, Data & Analytics, Cloud, DevOps, Data & MLOps, Quality Engineering, and Business Automation.


Working with Us
At Iris, every role is more than a job — it’s a launchpad for growth.

Our Employee Value Proposition, “Build Your Future. Own Your Journey.” reflects our belief that people thrive when they have ownership of their career and the right opportunities to shape it.

We foster a culture where your potential is valued, your voice matters, and your work creates real impact. With cutting-edge projects, personalized career development, continuous learning and mentorship, we support you to grow and become your best — both personally and professionally.

Curious what it’s like to work at Iris? Head to this video for an inside look at the people, the passion, and the possibilities. Watch it here.

Job Description

Mandatory Skills:

PySpark, Apache Kafka, Databricks Workflows, Delta Lake on Databricks

Key Responsibilities:

  • Design scalable data engineering solutions using PySpark and modern distributed data processing frameworks.
  • Define data ingestion, transformation, and processing architectures aligned with business and analytical objectives.
  • Design and optimize Snowflake or Delta Lake on Databricks solutions to support enterprise-scale data platforms.
  • Lead implementation of high-performance batch and streaming data pipelines.
  • Design and optimize event-driven data architectures using Apache Kafka or Amazon Kinesis.
  • Define data streaming standards, integration frameworks, and scalable processing patterns.
  • Architect workflow orchestration solutions using Apache Airflow or Databricks Workflows.
  • Establish monitoring, scheduling, and operational controls for reliable pipeline execution.
  • Drive data quality, validation, reconciliation, and governance practices across data engineering solutions.
  • Design data engineering solutions following modern Lakehouse architecture principles, data observability practices, and platform engineering standards to improve scalability, reliability, and operational visibility.
  • Drive development of business-focused data products by improving data quality, discoverability, usability, documentation, and trusted data consumption across analytical platforms.
  • Promote responsible use of AI-assisted engineering capabilities to improve development productivity, testing, documentation, and engineering quality.
  • Review data pipeline designs and implementations to ensure adherence to engineering, scalability, and performance standards.
  • Troubleshoot complex data processing, workflow, and streaming platform issues through detailed root cause analysis.
  • Mentor team members on PySpark, Snowflake, Delta Lake, Kafka, Kinesis, Airflow, and data engineering best practices.
  • Collaborate with various teams and stakeholders to support end-to-end data platform delivery.

Behavioral Competencies:

  • Demonstrates strong ownership while driving data engineering excellence.
  • Collaborate effectively with various teams and business stakeholders to ensure smooth delivery.
  • Promotes quality-focused engineering through proactive validation, optimization, and continuous improvement.
  • Apply strong analytical thinking to evaluate complex data engineering and platform challenges.
  • Demonstrate adaptability while managing evolving technologies, data ecosystems, and business requirements.
  • Communicates effectively regarding delivery status, risks, dependencies, and improvement opportunities.
  • Maintains high attention to detail across data architecture, pipeline design, testing, and implementation activities.
  • Encourages continuous improvement in data engineering practices and platform operations.
  • Supports knowledge sharing and mentoring to strengthen team capabilities.
  • Balances scalability, performance, reliability, and business priorities while driving delivery excellence.
  • Promotes innovation by adopting modern data engineering practices, platform engineering principles, and AI-assisted development approaches to improve engineering productivity and solution quality.

Mandatory Competencies

Data Science and Machine Learning - Data Science and Machine Learning - Python
Data Science and Machine Learning - Data Science and Machine Learning - Apache Spark
Database - Database Programming - SQL
Data & AI - Data Engineering - Data Quality & Validation
Big Data - Big Data - Pyspark
Data & AI - Data Engineering - Apache Kafka
Data Science and Machine Learning - Data Science and Machine Learning - Databricks
Beh - Communication and collaboration

Perks and Benefits for Irisians
Iris provides world-class benefits for a personalized employee experience. These benefits are designed to support financial, health and well-being needs of Irisians for a holistic professional and personal growth. Click here to view the benefits.

About Iris Software

Provides software engineering and IT consulting services to enterprises.

Year founded
1991
Employees
4300
Organization type
Private
Latest investment
Raised $948.00k Grant (2008)
Headquarters
US

Similar jobs

Data Engineer roles near Noida, Uttar Pradesh
4h
Save
Mark Applied
Hide
Senior Snowflake Data Engineer
Noida, Uttar Pradesh, India
HybridFull Time
Ibex
IbexNASDAQ: IBEX: Provides technology-enabled customer experience and business process outsourcing services.
7+ YOERequires 7+ years in data engineering, Snowflake, advanced SQL, Python, Snowpark, ELT/ETL, dbt, cloud platforms, Git, CI/CD, testing, and data modeling.
Snowflake, SQL, Python, Snowpark, Dynamic Tables, Streams, Tasks, dbt, AWS, Azure, GCP, Git, CI/CD, Fivetran, Kafka, Airflow, Terraform, Snowflake Iceberg Tables, Snowflake Cortex, Cortex Analyst, Cortex Search, Cortex LLM
9h
Save
Mark Applied
Hide
Associate Platform Services - Data Engineer
Pune or Gurgaon
HybridFull Time
ZS
ZS: Global management consulting and technology firm for healthcare and life sciences.
1+ YOEBachelor's degree required; 1–2 years of development experience, ETL, SQL, Python, data modeling, data warehousing, data products, cloud platforms, and strong analytical and communication skills.
ZAIDYN, SQL, Python, AWS, Azure, Hadoop, Spark, PySpark, Informatica, Talend, SSIS
15h
Save
Mark Applied
Hide
Data Engineer-Senior II
Bengaluru or Gurugram or Mumbai
OnsiteFull Time
FedEx
FedExNYSE: FDX: Global provider of courier, logistics, and transportation services.
4+ YOEBachelor's degree in a relevant quantitative or technical field required; master's preferred. Requires 4–7 years' experience, Python/PySpark/SAS, cloud, data engineering, SQL, ETL, data modeling, and English fluency.
Python, PySpark, SAS, Hadoop, Hive, Spark, Azure, Azure Data Factory (ADF), Azure Data Lake Storage (ADLS), Azure DevOps, Databricks, Delta Lake, Docker, CI/CD, Kubernetes, Terraform, Octopus, SQL, Ab Initio, Informatica, DataStage, Power BI
19h
Save
Mark Applied
Hide
Data Engineer-Senior II
Bengaluru or Mumbai or Gurugram
OnsiteFull Time
FedEx
FedExNYSE: FDX: Global provider of express delivery and logistics services.
4+ YOEBachelor's degree in a relevant discipline; 4–7 years' experience; Python, PySpark, SAS, SQL, data pipelines, cloud platforms, ETL, data modeling, and distributed data technologies.
Python, PySpark, SAS, Hadoop, Hive, Spark, Microsoft Azure, Azure Data Factory (ADF), ADLS Storage, Azure DevOps, Databricks, Delta Lake, Workflows, Docker, CI/CD, Kubernetes, Terraform, Octopus, SQL, Ab Initio, Informatica, DataStage, Power BI
1d
Save
Mark Applied
Hide
Senior Data Engineer
Gurgaon, Haryana, India
OnsiteFull Time
Srijan
Srijan: Develops digital platforms and engineering solutions for global brands.
4+ YOERequires 4+ years with Big Data technologies, Databricks, ETL/ELT, batch ingestion, Python, Git, CI/CD, SQL, data modeling, warehousing, integration, documentation, and cross-functional communication.
Databricks, Apache Spark, SQL, Python, Nutter, Unittest, Pytest, Git, CI/CD, Microsoft Azure, Power BI, AWS
1d
Save
Mark Applied
Hide
Data Engineer-Data Platforms-Google
Gurgaon, Haryana, India
HybridFull Time
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Bachelor's degree required; master's preferred. Experience with Google Cloud data platforms, batch and real-time pipelines, data migration, data layer design, and tools including Airflow, dbt, Spark, Python, and Scala.
Google DataProc, Google DataFlow, Google PubSub, Google BigQuery, Google BigTable, Google Cloud Spanner, Google CloudSQL, Google AlloyDB, Google Cloud Storage, Apache Airflow, dbt, Spark, Python, Scala, Hadoop, Apache Beam, Google Cloud Scheduler, Cloud Composer
1d
Save
Mark Applied
Hide
Lead Data Engineer - Azure + Fabric
Gurugram, Haryana, India
HybridFull Time
EXL
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
8+ YOERequires 8+ years of data engineering experience, Azure or Microsoft Fabric expertise, Spark/PySpark, Python, SQL, Delta Lake, ETL/ELT, data quality, and a bachelor's degree in a relevant field.
Microsoft Azure, Microsoft Fabric, Microsoft OneLake, Microsoft Fabric Lakehouse, Microsoft Fabric Warehouse, Microsoft Fabric Data Pipelines, Notebooks, Spark, PySpark, Python, SQL, Delta Lake, Power BI, Microsoft Purview, Microsoft Azure DevOps, Git, GCP, CSV, Excel, Microsoft SharePoint, Synapse Pipelines, CI/CD, RBAC
1d
Save
Mark Applied
Hide
Snr Data Engineer
Noida, Uttar Pradesh, India
OnsiteFull Time
Alight
AlightNYSE: ALIT: Provides cloud-based HR, payroll, and benefits administration services.
4+ YOERequires 4–8 years of ETL experience, AWS and Hadoop expertise, Scala or Python, PySpark, Spark, SQL, orchestration tools, data warehousing, CI/CD, GitHub, and strong analytical and communication skills.
Apache Spark, Cloudera, AWS Step Functions, AWS Glue, AWS Lambda, Amazon S3, Amazon Redshift, Hadoop, HDFS, Apache Hive, Apache Kafka, Amazon EMR, Kinesis, PySpark, Spark SQL, Parquet, ORC, Apache Airflow, Control-M, Scala, HiveQL, Impala, SQL, Python, Shell, CI/CD, GitHub, Docker, Amazon ECS, Amazon EKS