Capco
Posted 2mo ago

Data Engineer - Python and Azure Data Bricks

Capco
Bengaluru or Chennai or Gurugram or Hyderabad or Pune
OnsiteFull Time
Responsibilities
  • analyzing data
  • designing forecasts
  • deploying models
Requirements
  • Expertise in Python and SQL
  • Statistical modeling
  • Time-series forecasting
  • Machine learning
  • Databricks/Azure
  • Spark/PySpark
  • Geospatial analytics, and model explainability
Technical tools mentioned
PythonSQLPandasNumPyScikit-learnStatsmodelsProphetScipyGeoPandasShapelyDatabricksAzure Data & Analytics ecosystemSparkPySparkAzure MLMLOpsSHAPLIME

Job description

Job Title: Data Engineer - Python and Azure Data Bricks

About Us

“Capco, a Wipro company, is a global technology and management consulting firm. Awarded with Consultancy of the year in the  British Bank Award and has been ranked Top 100 Best Companies for Women in India 2022 by Avtar & Seramount. With our presence across 32 cities across globe, we support 100+ clients across banking, financial and Energy sectors. We are recognized for our deep transformation execution and delivery. 

WHY JOIN CAPCO?

You will work on engaging projects with the largest international and local banks, insurance companies, payment service providers and other key players in the industry. The projects that will transform the financial services industry.

MAKE AN IMPACT

Innovative thinking, delivery excellence and thought leadership to help our clients transform their business. Together with our clients and industry partners, we deliver disruptive work that is changing energy and financial services.

#BEYOURSELFATWORK

Capco has a tolerant, open culture that values diversity, inclusivity, and creativity.

CAREER ADVANCEMENT

With no forced hierarchy at Capco, everyone has the opportunity to grow as we grow, taking their career into their own hands.

DIVERSITY & INCLUSION

We believe that diversity of people and perspective gives us a competitive advantage.

 

Job Description

 

ole Overview

We are looking for a highly skilled Data Engineer with strong expertise in Python and Azure Databricks to design, build, and optimize scalable data pipelines and modern data platforms. The ideal candidate will have hands-on experience in cloud-based data engineering, big data processing, and data transformation using Azure services and Databricks.

Key Responsibilities

  • Design, develop, and maintain scalable data pipelines using Python and Azure Databricks.
  • Build and optimize ETL/ELT workflows for ingesting, transforming, and processing large datasets.
  • Develop data solutions using Apache Spark within the Databricks environment.
  • Integrate data from multiple structured and unstructured sources.
  • Ensure data quality, reliability, security, and governance across data platforms.
  • Collaborate with Data Scientists, Analysts, Architects, and Business stakeholders to deliver data solutions.
  • Implement performance tuning and optimization for large-scale data processing workloads.
  • Develop and maintain data models, metadata management, and documentation.
  • Monitor and troubleshoot production data pipelines and resolve data-related issues.
  • Support CI/CD implementation and deployment automation for data engineering solutions.

Required Skills & Qualifications

  • Bachelor's or Master's degree in Computer Science, Engineering, Information Technology, or a related field.
  • 4+ years of experience in Data Engineering.
  • Strong hands-on programming experience in Python.
  • Expertise in:
    • Azure Databricks
    • Apache Spark (PySpark)
    • SQL
    • Delta Lake
  • Experience with Azure data services:
    • Azure Data Factory (ADF)
    • Azure Data Lake Storage (ADLS)
    • Azure Synapse Analytics
    • Azure Key Vault
  • Strong understanding of data warehousing concepts and dimensional modeling.
  • Experience building and maintaining batch and real-time data pipelines.
  • Knowledge of Git, CI/CD pipelines, and DevOps practices.
  • Experience working in Agile development environments.

If you are keen to join us, you will be part of an organization that values your contributions, recognizes your potential, and provides ample opportunities for growth. For more information, visit www.capco.com. Follow us on Twitter, Facebook, LinkedIn, and YouTube.

 

About Capco

Global consultancy providing business and technology services.

Similar jobs

Data Scientist roles near Bengaluru, Karnataka
1h
Save
Mark Applied
Hide
Senior Data Scientist
Bangalore, Karnataka, India
OnsiteFull Time
HCLTech
HCLTechNational Stock Exchange of India: HCLTECH: Global technology providing digital, engineering, and cloud services.
Experience leading advanced data science initiatives, developing predictive and machine learning models, applying statistical and deep learning techniques, and collaborating across teams.
4h
Save
Mark Applied
Hide
Senior Data Scientist
Bangalore, Karnataka, India
OnsiteFull Time
Tredence
Tredence: Providing AI-driven data science and business analytics solutions.
5+ YOERequires 5+ years in data science or analytics, Python and SQL, machine learning, quantitative problem-solving, and forecasting or pricing optimization experience. A relevant master's, doctorate, or engineering degree is preferred.
SQL, Python, pandas, scikit-learn, statsmodels, Spark, Hadoop, Tableau, Power BI, matplotlib, seaborn, PuLP, OR-Tools, Gurobi
10h
Save
Mark Applied
Hide
Lead Data Scientist
Bangalore, Karnataka, India
OnsiteFull Time
London Stock Exchange Group
London Stock Exchange GroupLondon Stock Exchange: LSEG: Provides financial market infrastructure and global data analytics services.
8+ YOERequires 8–12 years in data science or analytics, advanced AI/ML expertise, Python and ML frameworks, cloud and MLOps experience, stakeholder leadership, and a relevant master's or engineering degree.
Python, TensorFlow, PyTorch, Scikit-learn, R, SQL, AWS, Azure, MLOps, CI/CD, NLP, LLMs, Retrieval-Augmented Generation (RAG)
10h
Save
Mark Applied
Hide
Data Scientist, Knowledge Management
Bengaluru or Bangalore
HybridFull Time
eBay
eBayNASDAQ: EBAY: Global online marketplace for buying and selling diverse products.
4+ YOE4+ years in analytics or data science; SQL and Python proficiency; experience with eCommerce search, recommendations, knowledge graphs, LLMs, prompt engineering, RAG, and ML model evaluation.
SQL, Python, LLMs, retrieval-augmented generation (RAG)
20h
Save
Mark Applied
Hide
Senior Data Scientist
Bengaluru, Karnataka, India
OnsiteFull Time
Kraft Heinz
Kraft HeinzNASDAQ: KHC: Manufacturer and global marketer of food and beverage products.
5+ YOEMaster's degree in a related quantitative field and 5+ years in predictive, time-series, and statistical modeling. Requires Python, R, scikit-learn, SQL, relational databases, and cloud services experience.
Python, R, scikit-learn, SQL, relational databases, AWS, Azure, Google Cloud
22h
Save
Mark Applied
Hide
Data Scientist-Artificial Intelligence
Bangalore, Karnataka, India
HybridFull Time
IBM
IBMNew York Stock Exchange: IBM: Global technology providing enterprise software, cloud, and consulting.
Bachelor's degree required; advanced analytics, Python, AI frameworks, NLP/ML, cloud platforms, databases, and AI model deployment experience. Master's degree preferred.
TensorFlow, PyTorch, Keras, Hugging Face, Github Copilot, Amazon Code Whisperer, Python, Kubernetes, AWS, Azure, GCP, SQL, Postgres, DB2, MongoDB, Backbone.js, AngularJS, React.js, Ember.js, Bootstrap, JQuery, SciKit Learn, Pandas, Matplotlib, Linux, Windows, iOS, Android
1d
Save
Mark Applied
Hide
Data Scientist
Bangalore, Karnataka, India
HybridFull Time
EXL
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
3+ YOEBachelor's or master's degree in a quantitative field, 3+ years of data science or NLP experience, Python and machine learning expertise, advanced RAG knowledge, and familiarity with cloud platforms.
Python, Pandas, NumPy, Scikit-learn, Jupyter, SQL Server, Spark, NLTK, Microsoft GitHub, Jira, GPT, LLAMA, Mistral, FLAN T5, AWS, Microsoft Azure, Databricks, SQL
1d
Save
Mark Applied
Hide
Staff Data Scientist, NIRA
Bengaluru or Bangalore or Asia or United States or Europe or Asia or North America
HybridFull Time
Tilt
Tilt: Offers mobile credit products using alternative financial data.
7+ YOERequires 7+ years of predictive modeling expertise, preferably in credit risk or actuarial settings, with Python, SQL, ML libraries, end-to-end project ownership, stakeholder management, and leadership communication skills.
Python, SQL, scikit-learn, LightGBM, XGBoost, Claude, Codex, TensorFlow, PyTorch, CNN, RNN