📋 External Recruiting Agencies

TalPro is a recruitment and staffing agency that sources candidates for other companies, rather than acting as a direct employer for the roles advertised.

This company was flagged and excluded from default search results. Proceed with caution.

TalPro
Posted 9mo ago

Data Software Engineer – Spark, Python, Databricks (L2 / L4)

TalPro
Bengaluru, Karnataka, India
₹21-₹42/yrHybridFull Time
Responsibilities
  • design systems
  • develop apps
  • build pipelines
Requirements
  • Data Software Engineer with Spark, Python, and Databricks
  • L2 (4–5 yrs) or L4 (8–12 yrs)
  • Hybrid in Bangalore
  • FTE
Technical tools mentioned
Apache SparkPythonHadoopHiveImpalaKafkaRabbitMQNoSQLHBaseCassandraMongoDBAWS DatabricksAzure Databricks

Job description


Data Software Engineer – Spark, Python, Databricks (L2 / L4)

Experience:

  • L2: 4–5 Years

  • L4: 8–12 Years

Mode: FTE (Full-Time Employment)

Job Location: Bangalore

Work Mode: Hybrid

Notice Period:

  • L2: Immediate to 15 days

  • L4: Immediate to 15 days  

Drive Type: F2F

Drive Location: Bangalore

CTC Band:

  • L2: Up to 21 LPA

  • L4: Up to 42 LPA


Role Overview

We are hiring Data Software Engineers with strong expertise in Apache Spark, Python, and AWS/Azure Databricks. The ideal candidates will have deep Big Data engineering experience, strong distributed systems knowledge, and the ability to work on complex end-to-end data platforms at scale.


Key Responsibilities

Big Data Engineering

  • Design and build distributed data processing systems using Spark and Hadoop.

  • Develop and optimize Spark applications, ensuring performance and scalability.

  • Create and manage ETL/ELT pipelines for large-scale data ingestion and transformation.

Streaming & Event Processing

  • Build and manage real-time streaming systems using Spark Streaming or Storm.

  • Work with Kafka / RabbitMQ for event-driven ingestion and messaging patterns.

Cloud & Databricks Engineering

  • Develop & optimize workloads on AWS Databricks or Azure Databricks.

  • Perform cluster management, job scheduling, performance tuning, and automation.

Data Integration & Storage

  • Integrate data from diverse sources: RDBMS (Oracle, SQL Server), ERP, file systems.

  • Work with query engines like Hive and Impala.

  • Experience with NoSQL stores: HBase, Cassandra, MongoDB.

Programming & Scripting

  • Strong hands-on coding in Python for data transformations and automations.

  • Strong SQL skills for data validation, tuning, and complex queries

Team Leadership (L4)

  • Provide technical leadership and mentoring to junior engineers.

  • Drive solution design for Big Data platforms end-to-end

Ways of Working

  • Work in Agile teams, participate in sprint ceremonies and planning.

  • Collaborate with engineering, data science, and product teams.


Required Skills & Expertise (Both L2 & L4)

  • Apache Spark – Expert level (core, SQL, streaming)

  • Python – Strong hands-on

  • Distributed computing fundamentals

  • Hadoop ecosystem: Hadoop v2, MapReduce, HDFS, Sqoop

  • Streaming systems: Spark Streaming / Storm

  • Messaging: Kafka or RabbitMQ

  • SQL – Advanced (joins, stored procedures, query optimization)

  • NoSQL: HBase, Cassandra, MongoDB

  • ETL frameworks & data pipeline design

  • Hive / Impala querying

  • Performance tuning of Spark jobs

  • AWS or Azure Databricks

  • Experience working in Agile


Experience & Level Mapping


L2 – Mid-Level (4–5 Yrs)

  • Skills: Spark, Python, AWS

  • Notice Period: Immediate – 20 Days

  • CTC Band: Up to 21 LPA

L4 – Senior-Level (8–12 Yrs)

  • Skills: Spark, Python, Azure Databricks

  • Notice Period: 15 Days (Nov joiners) OR Jan joiners

  • CTC Band: Up to 42 LPA




Similar jobs

Data Software Engineer roles near Bengaluru, Karnataka
4w
Save
Mark Applied
Hide
Lead Software Engineer - Data
Bangalore, Karnataka, India
OnsiteFull Time
EPAM Systems
EPAM SystemsNYSE: EPAM: Provider of digital platform engineering and software development services.
8+ YOE8+ years in Big Data with expertise in Apache Spark, Hadoop ecosystem, Python, cloud native services (Azure/AWS) and team leadership; strong SQL, ETL, messaging and NoSQL experience.
Apache Spark, Python, Databricks, Microsoft Azure, Amazon Web Services, Hadoop, MapReduce, HDFS, Sqoop, Kafka, RabbitMQ, Hive, Impala, SQL, HBase, Cassandra, MongoDB, Apache Storm, Spark-Streaming
1mo
Save
Mark Applied
Hide
Senior Software Engineer, Data
Bengaluru, Karnataka, India
HybridFull Time
Logward
Logward: German supply-chain software serving shippers and logistics providers with AI-powered workflow automation.
Proven backend and data engineering experience with Python/TypeScript/Go, databases, ETL, orchestration, and building scalable, production-ready data platforms.
Python, TypeScript, Go, PostgreSQL, MySQL, MongoDB, Redis, JSON, XML, CSV, EDI, RESTful API, Apache Spark, Apache Flink, Apache Airflow, Kubernetes, Apache Iceberg, Hudi, Delta Lake
22h
Save
Mark Applied
Hide
Grid Technologies
Bangalore, Karnataka, India
OnsiteFull Time
Siemens Energy
Siemens EnergyFrankfurt Stock Exchange: ENR: Global energy technology driving the energy transition.
5+ YOEBachelor's degree preferred in a related field and 5–7 years of experience in data analytics, business intelligence, reporting, and visualization. Requires Alteryx, Snowflake, SQL, Power BI or Tableau, data governance, and stakeholder skills.
Alteryx, Snowflake, Tableau, Microsoft Power BI, SQL, Python, R, ETL, ELT, DevOps, Agile
1w
Save
Mark Applied
Hide
Software Engineer, Data Science
Bengaluru, Karnataka, India
HybridFull Time
LinkedIn: The world's largest professional network.
2+ YOEBachelor’s degree in a quantitative discipline, 2+ years working with large datasets, SQL and relational database experience, and programming experience in R, Python, Java, Scala, or PHP.
SQL, R, Python, Java, Scala, PHP, Spark, Hive, ETL, Hadoop, Presto, Pig, Tableau, D3, JavaScript, Unix, Git
3w
Save
Mark Applied
Hide
Data Scientist/Software Developer
Bangalore, Karnataka, India
HybridFull Time
Xylem
XylemNYSE: XYL: Global water technology providing sustainable water solutions.
2+ YOEBachelor's degree in a relevant technical field and 2–4 years of experience in software development, data engineering, or data science. Requires Python, Kedro, SQL, troubleshooting, analytics, and cross-team collaboration.
Python, Kedro, SQL, Databricks, Power BI
1mo
Save
Mark Applied
Hide
Principal Data Systems Software Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
Oracle Corporation
Oracle CorporationNYSE: ORCL: Cloud infrastructure and enterprise software solutions provider.
3+ YOEAdvanced software development experience (3+ years), strong algorithms, data structures, distributed systems; hands-on Java/J2EE, REST API, microservices, Oracle Database; scripting and SQL experience; cloud platform experience preferred.
Java, J2EE, REST API, microservices, Oracle Database, OCI, AWS, Azure, GCP, SQL, C/C++, JavaScript
2mo
Save
Mark Applied
Hide
Software Engineer – Data Platform
Bengaluru, Karnataka, India
HybridFull Time
ContiTech
ContiTech: Industrial materials manufacturer producing rubber and thermoplastic systems for mining, energy, construction, mobility, and manufacturing customers.
5+ YOEBachelor's in CS or related; 5+ years software development building production-grade data ingestion pipelines, APIs, CDC, strong in Scala/Java, Python, advanced SQL, Git, CI/CD, and hyperscaler (Azure) experience.
Scala, Java, Python, SQL, Git, CI/CD, Debezium Server, rclone, Theobald Extract Universal, Unity Catalog, Databricks, Azure, Salesforce, Microsoft SharePoint, Kerberos, APIs, HTTP
9mo
Save
Mark Applied
Hide
Lead AWS Software Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
FTSE Russell
FTSE Russell: Global financial index and benchmark provider serving asset owners, asset managers, ETF providers, investment banks, and retail investors.
Design and build scalable data pipelines in Python and Spark; manage AWS data stack; collaborate with product teams; ensure data quality and observability.
Python, Apache Spark, AWS S3, AWS Glue, AWS EMR, AWS Lambda, Apache Iceberg, Git, CI/CD, Databricks, Great Expectations, Deequ