EXL
Posted 9mo ago

Associate - Data Engineer-Data Engineering-Big Data Engineering

EXL
India
HybridFull Time
Responsibilities
  • Designing pipelines
  • Optimizing queries
  • Maintaining models
Requirements
  • 5-7 years data engineering with 2-3 years Databricks
  • Spark
  • Delta Lake
  • Cloud platforms
  • SQL
  • ETL
  • CI/CD
  • DevOps
  • MLOps a plus
  • Data governance
Technical tools mentioned
DatabricksApache SparkPySparkScalaSQLDelta LakeLakehouseCloud platformsAirflowADFDBX WorkflowsCI/CDGitMLflowMLOpsData governance

Job description

Required Skills & Experience

  • 5–7 years of experience in data engineering, with at least 2–3 years of Databricks hands-on experience.
  • Strong expertise in Apache Spark (PySpark/Scala/SQL) and distributed data processing.
  • Solid experience with Delta Lake, Lakehouse architecture, and data modeling.
  • Hands-on experience with at least one cloud platform: Azure Data Lake, AWS S3, or GCP BigQuery/Storage.
  • Strong proficiency in SQL for data manipulation and performance tuning.
  • Experience with ETL frameworks, workflow orchestration tools (Airflow, ADF, DBX Workflows).
  • Good understanding of CI/CD, Git-based workflows, and DevOps practices.
  • Exposure to MLOps and MLflow is a strong plus.
  • Knowledge of data governance, cataloging, and security frameworks.


 

Responsibilities

Required Skills & Experience

  • 5–7 years of experience in data engineering, with at least 2–3 years of Databricks hands-on experience.
  • Strong expertise in Apache Spark (PySpark/Scala/SQL) and distributed data processing.
  • Solid experience with Delta Lake, Lakehouse architecture, and data modeling.
  • Hands-on experience with at least one cloud platform: Azure Data Lake, AWS S3, or GCP BigQuery/Storage.
  • Strong proficiency in SQL for data manipulation and performance tuning.
  • Experience with ETL frameworks, workflow orchestration tools (Airflow, ADF, DBX Workflows).
  • Good understanding of CI/CD, Git-based workflows, and DevOps practices.
  • Exposure to MLOps and MLflow is a strong plus.

Qualifications

Bachelor's/Master's in Engineering 0-2 years

About EXL

Provides data analytics and digital operations solutions to businesses.

Similar jobs

Data Engineer roles
5h
Save
Mark Applied
Hide
Job Posting Title Sr. Data Engineer – Enterprise Knowledge Engineering
Hyderabad, Telangana, India
OnsiteFull Time
Amgen
AmgenNASDAQ: AMGN: Develops and manufactures biotechnology medicines for serious diseases.
8+ YOEBachelor’s or master’s degree in computer science, IT, or related field and 8–13 years’ experience. Requires modern data platforms, cloud, unstructured content, search, vector databases, AI, governance, and regulated-industry expertise.
Databricks, Apache Spark, Delta Lake, Unity Catalog, AWS, Microsoft Word, Collibra, Model Context Protocol (MCP)
13h
Save
Mark Applied
Hide
Senior Data Engineer
Hyderabad or Gurgaon
OnsiteFull Time
S&P Global
S&P GlobalNYSE: SPGI: Provides financial data, analytics, and credit ratings worldwide.
5+ YOEMinimum 5 years in data or database engineering; experience with NoSQL, cloud platforms, data pipelines, Python, SQL, RESTful APIs, orchestration, version control, CI/CD, and distributed systems.
MarkLogic, MongoDB, Couchbase, Cassandra, DynamoDB, Elasticsearch, OpenSearch, AWS, Azure, GCP, GitHub, GitLab, Bitbucket, Azure Repos, Python, SQL, FastAPI, Flask, Django, Scrum, Kanban
15h
Save
Mark Applied
Hide
Databricks, Pyspark, DBT, Autosys-Data Engineer, Airflow
Chennai, Tamil Nadu, India
OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesNational Stock Exchange of India: TCS: Global provider of IT services, consulting, and business solutions.
12+ YOERequires a Bachelor of Engineering and 12–16 years of experience. Desired skills include Databricks, PySpark, dbt, and Autosys.
Databricks, PySpark, dbt, Autosys, Airflow
16h
Save
Mark Applied
Hide
GCP Senior Data Engineer (141058)
Hyderabad or Chennai
HybridFull Time
HCLTech
HCLTechNational Stock Exchange of India: HCLTECH: Global provider of information technology services and software consulting.
5+ YOEBachelor’s or master’s degree in computer science, IT, or related field; 5+ years of data engineering experience; strong GCP, Python, SQL, data pipeline, warehousing, modeling, and stakeholder collaboration skills.
Google Cloud Platform (GCP), Dataflow, BigQuery, Pub/Sub, Cloud Storage, Cloud SQL, Firestore, Python, SQL, Hadoop, Spark, Machine Learning, Artificial Intelligence (AI), DevOps, CI/CD
16h
Save
Mark Applied
Hide
Lead/Senior Domino Data Lab with Python
Gurgaon or Pune or Coimbatore or Chennai or Hyderabad or Jaipur or Bangalore
OnsiteFull Time
EPAM Systems
EPAM SystemsNYSE: EPAM: Provides global digital platform engineering and software development services.
5+ YOERequires 5 years of software or data engineering experience, Domino Data Lab and Python expertise, REST API development, MLOps, Git, Docker, CI/CD, cloud platforms, databases, and data integration.
Domino Data Lab, Python, Pandas, NumPy, Scikit-learn, FastAPI, Flask, Django, Git, Docker, AWS, Azure, GCP, Node.js, Kubernetes, Jupyter, CI/CD
17h
Save
Mark Applied
Hide
Azure Databricks + Pyspark
Hyderabad or Chennai
OnsiteFull Time
Cognizant
CognizantNasdaq: CTSH: Provides IT consulting and digital business process services.
6+ YOERequires 6+ years with Azure Databricks, PySpark, Spark/Scala, Python, Kafka, Data Factory, SQOOP, Unix, shell scripting, RDBMS, data architecture, testing, and technical documentation.
Azure Databricks, PySpark, Apache Spark, Scala, SQOOP, Microsoft Azure Data Factory, Python, Apache Kafka, Oracle, Microsoft SQL Server, Netezza, IBM Db2, Unix, Shell scripting, Agile
17h
Save
Mark Applied
Hide
Associate Data Engineer
Chennai, Tamil Nadu, India
OnsiteFull Time
AstraZeneca
AstraZenecaLondon Stock Exchange: AZN: Researches, develops, and manufactures prescription medicines for major diseases.
5+ YOERequires Python, Pandas, PySpark, Postman, SQL, AWS Redshift, S3, EMR, DBeaver, and secure file transfer tools; strong troubleshooting, documentation, and collaboration skills. Five to eight years preferred.
Python, PyCharm, Pandas, PySpark, Postman, SQL, DBeaver, AWS, Amazon Redshift, Amazon S3, Amazon EMR, WinSCP, Git
17h
Save
Mark Applied
Hide
Senior Data Engineer – Python, SQL & AI-Augmented Engineering
Hyderabad, Telangana, India
OnsiteFull Time
L'Oréal
L'OréalEuronext Paris: OR: Manufactures and sells personal care, skincare, and cosmetic products globally.
5+ YOERequires 5+ years of data engineering, expert Python and SQL, 3+ years of Airflow, cloud processing experience, AI coding assistant use, and a relevant bachelor's or master's degree.
Python, SQL, GitHub Copilot, Cursor, ChatGPT, Claude Code, PySpark, BigQuery, Snowflake, Apache Airflow, Google Cloud Dataflow, Apache Beam, Dataform, Salesforce, Tealium, Braze, Google Ads, Spark, Dataproc, ELT, ETL, Data Lakehouse, Reverse ETL, Git, dbt, CI/CD