Infotel
Posted 2mo ago

Data Engineer - Python

Infotel
Bengaluru, Karnataka, India
OnsiteFull Time
Responsibilities
  • developing ETL
  • validating specifications
  • implementing pipelines
Requirements
  • 3+ years in corporate/institutional banking IT
  • Strong Python, ETL and data‑warehouse experience
  • Oracle/PL/SQL
  • Familiarity with PySpark
  • Data mapping
  • AML data domains
  • EMR/DevOps practices, and stakeholder coordination
Technical tools mentioned
PythonPySparkScalaOracle DatabasePL/SQLpandaspyarrowsqlalchemycx_OracleoracledbpolarsduckdbdbtKafkaRabbitMQApache AirflowPrefectMicrosoft Azure Data FactoryApache SparkAWS EMRMicrosoft Azure SynapseOracle Cloud Infrastructure (OCI)DockerKubernetesGitJenkinsGitHub ActionsTerraformApache AtlasCollibrapytestunittestPrometheusGrafanaOCI MonitoringSQL*LoaderOracle Data PumpExternal Tablescx_OracleGitJiraMicrosoft Azure Boards

Job description

 

The candidate should also have good knowledge of ETL design and be able to develop optimized code. Working experience on PySpark and Scala programming is an added advantage.

 

Responsibilities

·         Conduct detailed validation of functional specifications (eventually contribute to functional specifications if needed).

·         Initiate, build, and contribute to technical specifications.

·         Perform technical and/or data analysis to elaborate technical Specifications documents for the different IT stakeholders (IT 2S data provider, CIB datahub, AML dev teams)

·         Assist the technical stakeholders to validate the solutions and validate the technical tests planned.

·         Coordinate with the different stakeholders to implement the changes/evolutions in the delay, cost and quality expected.

·         Control the completion of the technical tests and associated deliverables.

·         Support and contribute to the releases organization.

·         Ensure data ingestion controls and technical tests automation developments are done according to expectations

  • Implement DevOps practices to ensure efficient and reliable deployment of data pipelines and ETL processes,

Direct Responsibilities

  • Understand business requirement from business analysts, users and should have analytical mind to understand existing process and purpose better solutions
  • Work on TSD designs, development, testing, deployment, support
  • Suggest and implement innovative approach.
  • Should be adaptable to new technology or methodology

 

Contributing Responsibilities

  • Contribute towards knowledge sharing initiatives with other team members
  • Contribute documentation of solutions and configurations of the models
  •  

Technical & Behavioral Competencies

Mandatory

    • 3+ years of experience in Corporate and Institutional Banking IT, with a full understanding of the Corporate Banking and/or Securities Services activity.
    • Good understanding of AML monitoring tools and data needed for AML detection models.
    • Good understanding of Data Analysis and Data Mapping processes.
    • Extensive experience in working with functional and technical teams, defining requirements (mainly technical specification), establishing technical strategies, and leading the full life cycle delivery of projects.
    • Experience in Data-Warehouse architectural design providing efficient solutions in Compliance AML data domains.
    • Good Experience in Python developments, Oralce PL/SQL development
    • Excellent communication skills with the ability to explain complex technical issues in a simple concise manner.
    • Strong coordination and organizational skills.
    • Multi-tasking capabilities

 

All these qualifications are a plus:

    • Knowledge of Corporate Banking and Securities Services transactional data sources, flowing through the Compliance and Regulatory frameworks is a plus.  
    • Knowledge of Swift message and/or MX message formats and relevance to AML monitoring 
    • Experienced in implementing various data lineage mechanisms to meet regulatory requirements.

Success in the role is heavily dependent on the ability to show leadership, proactivity, and work cooperatively with both functional and technical teams, onshore and offshore

Specific Qualifications:

Python, Oracle,

Skills Referential (Required knowledge, skills and abilities)

Technical Skills:

  • Programming & Scripting (Primary)
    • Advanced Python (3.x) – OOP, typing, async, performance profiling
    • Familiarity with Python data‑engineer libraries: pandas, pyarrow, sqlalchemy, cx_Oracle, oracledb,polars, duckdb
    • Shell scripting (bash, PowerShell) for automation and orchestration
  • Oracle Database Expertise (Primary)
    • Oracle Database (11g/12c/19c/21c) administration basics
    • SQL proficiency: complex queries, analytic functions, hierarchical queries, PL/SQL development
    • Data modeling (ER, dimensional) and schema design for OLTP & OLAP
    • Performance tuning: indexing, partitioning, optimizer hints, AWR/ASH analysis
    • Oracle Data Pump, SQL*Loader, External Tables
  • Data Integration & ETL (Primary)
    • Design and implementation of ETL/ELT pipelines in Python (e.g., polars,pandas, pySpark, dbt)
    • Knowledge of messaging/streaming (Kafka, RabbitMQ) for real‑time ingestion
    • Data orchestration platforms: Apache Airflow, Prefect, or Azure Data Factory
  • Big Data & Distributed Processing (Secondary)
    • Working knowledge of Apache Spark (PySpark) and its integration with Oracle
    • Experience with cloud‑based big‑data services (AWS EMR, Azure Synapse, GCP Dataproc)
  • Cloud & DevOps (Secondary)
    • Oracle Cloud Infrastructure (OCI) services: Autonomous DB, Object Storage, Functions
    • Containerization (Docker) and orchestration (Kubernetes) for scalable pipelines
    • CI/CD pipelines (Git, Jenkins, GitHub Actions) for automated testing and deployment
    • Infrastructure‑as‑Code tools (Terraform, OCI Resource Manager)
  • Data Quality & Governance (Primary)
    • Implementing data validation, profiling, and cleansing in Python
    • Familiarity with data lineage, metadata management, and catalog tools (Apache Atlas, Collibra)
    • Understanding of GDPR, CCPA, and other data‑privacy regulations
  • Testing & Monitoring (Primary)
    • Unit & integration testing frameworks (pytest, unittest) for data pipelines
    • Monitoring & alerting (Prometheus, Grafana, OCI Monitoring) of ETL jobs and database health
  • Version Control & Collaboration (Secondary)
    • Proficient with Git (branching, pull‑requests, code reviews)
    • Agile methodologies (Scrum/Kanban) and ticketing systems (Jira, Azure Boards)

Behavioral Skills:

    • Ability to collaborate / Teamwork
    • Communication skills - oral & written
    • Creativity & Innovation / Problem solving
    • Ability to share / pass on knowledge

Education Level: Bachelor Degree or equivalent

Location: Bangalore

 

About Infotel

Provides IT consulting and enterprise software for database management.

Similar jobs

Data Engineer roles near Bengaluru, Karnataka
10h
Save
Mark Applied
Hide
Data Engineer
Bangalore, Karnataka, India
OnsiteFull Time
Zensar
ZensarNational Stock Exchange of India: ZENSARTECH: Global technology firm providing digital transformation and infrastructure services.
10+ YOERequires 10–15 years of experience with data pipelines, Google BigQuery, data lakes, large data models and queries, SQL, CDP integrations, streaming pipelines, Looker, and data reconciliation.
Google BigQuery, Oracle CDP, Adobe CDP, Web SDK, SQL, Adobe CJA, Oracle CX Analytics, Looker
1d
Save
Mark Applied
Hide
Senior Data Engineer
Bengaluru, Karnataka, India
HybridFull Time
Atlassian
AtlassianNASDAQ: TEAM: Develops software for team collaboration and project management.
6+ YOE6+ years in data engineering or related fields; STEM bachelor's or master's degree or equivalent experience; Python, advanced SQL, relational databases, ETL, data modeling, quality, visualization, optimization, and analytics experience.
Python, SQL, ETL, ELT
1d
Save
Mark Applied
Hide
Data Specialist - R01568669
Bangalore, Karnataka, India
OnsiteFull Time
Brillio
Brillio: Digital technology services and AI-led transformation partner.
5+ YOEBachelor’s or Master’s in computer science, engineering, or related field; 5+ years’ data engineering experience; expertise in Spark, Python, Scala, SQL, cloud platforms, ETL, APIs, storage, and version control.
SQL, Google BigQuery, Google Cloud Dataproc, Python, Data Catalog, Composer, Google Cloud Dataflow, Cloud Trace, Cloud Logging, Google Cloud Storage, Data Fusion, PL/SQL, T-SQL, Apache Spark, PySpark, AWS EMR, Apache Hadoop, Apache Hive, PostgreSQL, Scala, UNIX, Databricks, API, Amazon S3, Elasticsearch, Apache Airflow, Autosys, Git, SVN, HTML
1d
Save
Mark Applied
Hide
Sr.Data Engineering
Bengaluru, Karnataka, India
OnsiteFull Time
Bosch
Bosch: Global manufacturer of automotive and industrial engineering technology.
6+ YOE6–8 years of data engineering experience with Python, PySpark, SQL, ETL/ELT, Azure Databricks, Azure Data Factory, OpenShift, GitHub Actions, Grafana, and data pipeline development.
Python, PySpark, Azure Databricks, Azure Data Factory, Grafana, OpenShift, HELM, GitHub Actions, SQL Server, Azure SQL Server, Azure Key Vault, Kafka, Azure Event Hubs, Kubernetes
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
3+ YOERequires 3+ years of experience, including 4 years with Databricks, 15 years of full-time education, and expertise in data engineering, generative and agentic AI, prompt engineering, and AI evaluation.
Databricks
1d
Save
Mark Applied
Hide
Data Engineer-Senior II
Bengaluru or Gurugram or Mumbai
OnsiteFull Time
FedEx
FedExNYSE: FDX: Global provider of courier, logistics, and transportation services.
4+ YOEBachelor's degree in computer science, engineering, mathematics, statistics, or similar; 4–7 years' experience; Python, PySpark, SAS, SQL, data modeling, cloud, ETL, and distributed data technologies.
Python, PySpark, SAS, Hadoop, Hive, Spark, Azure, Azure Data Factory, Azure Data Lake Storage, Azure DevOps, Databricks, Delta Lake, Docker, Kubernetes, Terraform, Octopus, SQL, Ab Initio, Informatica, DataStage, Power BI
1d
Save
Mark Applied
Hide
AWS Data Engineer
Gurugram or Bengaluru
HybridFull Time
EXL
EXLNASDAQ: EXLS: Provides data analytics and digital operations solutions to businesses.
5+ YOEBachelor's degree and 5+ years as a data engineer; advanced SQL, Snowflake, ETL, CI/CD, Python, data modeling, and AWS services including S3 and Glue.
AWS, SQL, Snowflake, ETL, CI/CD, Amazon S3, Amazon Athena, AWS Glue, Amazon EMR, Apache Spark, Python
2d
Save
Mark Applied
Hide
Data Engineer - Axion
Sunnyvale or Washington or San Diego or Fort Walton Beach or Ann Arbor or London or Stuttgart or Munich or Stockholm or Bangalore or Seoul or Tokyo
$200k-$275k/yr OnsiteFull Time
Applied Intuition
Applied Intuition: Developing software and simulation infrastructure for autonomous vehicles.
5+ YOERequires 5+ years of relevant experience, modern ML infrastructure, large-scale GPU jobs, data software, microservices or databases, U.S. citizenship, and eligibility for security clearance.
React, TypeScript, Python, Golang, Docker, Kubernetes, Opensearch, Postgres, LLMs, VLMs, GPU