Assembly
Posted 3w ago

Data Engineer

Assembly
Bengaluru, Karnataka, India
OnsiteFull Time
Responsibilities
  • designing pipelines
  • optimizing pipelines
  • troubleshooting issues
Requirements
  • 5+ years data engineering experience with PySpark and Databricks
  • Strong ETL and performance optimization skills, and effective cross-team communication
Technical tools mentioned
PySparkDatabricksGCPAWSCI/CD

Job description

At Assembly, we help brands find the change to fuel business growth. We are an award-winning global brand performance agency, home to 1,600 talented people across 25 offices globally. We create unique data, technology and media solutions that enable faster and smarter problem solving and an inspired, collaborative workplace culture.

 

At Assembly we embody three core values: Show Up - actively contribute to a space of personal and collective growth; Make Change - embrace obstacles as opportunities, taking intentional steps to drive positive change; and Win Well - approach success with integrity, responsibility, and a commitment to collaboration, understanding that the journey is as important as the destination. Together, we create an environment that fosters continuous learning, adaptability, and a shared passion for making a meaningful impact.

 

We stay ahead of what’s next, providing fresh insights to spark new ideas. We’re a trusted partner to our clients, working behind the scenes to bring imagination, depth, and clarity to their biggest challenges—in entertainment, technology, lifestyle, sports, and gaming. Together, we create with confidence

 

We are looking for a skilled Data Engineer with 5–7 years of experience to join our team. The ideal candidate will have hands-on expertise in ETL pipelines, data transformations, and performance optimization, with strong problem-solving skills and the ability to communicate effectively across technical and non-technical teams.

· Design, build, and maintain scalable ETL pipelines for large-scale data processing.

· Work extensively with PySpark (intermediate to advanced) and Databricks to implement transformations and data workflows.

· Optimize data pipelines for performance, scalability, and cost-efficiency.

· Collaborate with cross-functional teams to understand requirements and deliver high-quality data solutions.

· Troubleshoot, debug, and resolve data processing and pipeline issues.

· Implement best practices for data quality, governance, and performance tuning.

· 5–7 years of experience in data engineering.

· Strong proficiency in PySpark (intermediate to advanced).

· Hands-on experience with Databricks.

· Strong understanding of ETL processes and data transformations.

· Experience with performance optimization in data pipelines.

· Excellent communication and problem-solving skills

 

 

Must-Have Skills

· 5–7 years of experience in data engineering.

· Strong proficiency in PySpark (intermediate to advanced).

· Hands-on experience with Databricks.

· Strong understanding of ETL processes and data transformations.

· Experience with performance optimization in data pipelines.

· Excellent communication and problem-solving skills.

Good-to-Have Skills

· Experience with cloud platforms (GCP, AWS).

· Knowledge of modern data architectures, data lakes, and warehousing solutions.

· Exposure to CI/CD for data engineering projects.

What We Offer

· Opportunity to work with large-scale data and cutting-edge cloud technologies.

· Collaborative and innovative work environment.

· Professional growth and upskilling opportunities.

  • Annual Leave in number of 15 allotted to all employees beginning of every calendar year.
  • Sick Leave in number of 12 is allotted effective DOJ and beginning of ever calendar year.
  • Other Leaves-Maternity Leave & Paternity Leaves, Birthday Leave Entitlement
  • Dedicated L&D Budget for all Teams to upskill & get certified
  • All employees are entitled for Group Personal Accident Cover & Life Cover Insurance.
  • Insurance coverage for the entire family (Employee + up to 7 dependents - Self, Spouse, up to 4 children, and Parents)
  • Monthly Cross Team Lunch
  • Rewards and Recognition program-Employee of the month, Star Performer, Tenure Celebration & many more

Assembly is an advocate for equal opportunity in the workplace. We are committed to ensuring equal opportunities regardless of race, colour, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability and gender identity. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. If you have a disability or special need that requires accommodation, please let us know.

At Assembly, we have a responsibility to bring impact into our every day. This means we must always look for ways in which to be conscious citizens in our roles to support society and environmental sustainability. We encourage employees to; be a conscious citizen by actively participating in our organisation's sustainability efforts, help us promote environmentally friendly practices within the workplace, collaborate with community organisations and stakeholders to support initiatives aligned with our company's values, participate in volunteer activities that benefit the community. Employees are also encouraged to make suggestions and evaluate our business practices to identify areas for improvement in social and environmental performance. Employees at Assembly demonstrate commitment to sustainability and inclusivity in their actions and behaviors.

About Assembly

A global omnichannel media agency merging data and technology.

Year founded
2022
Employees
1700
Organization type
Public
Headquarters
US

Similar jobs

Data Engineer roles near Bengaluru, Karnataka
1h
Save
Mark Applied
Hide
Snowflake Senior Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
HCLTech
HCLTechNational Stock Exchange of India: HCLTECH: Global technology providing digital, engineering, and cloud services.
Proficiency in Snowflake, SQL, and Python; experience with data modeling and ETL; strong analytical, problem-solving, communication, and collaboration skills. SnowPro Core preferred.
Snowflake, SQL, Python
10h
Save
Mark Applied
Hide
Data Engineer Sr. Consultant
Bengaluru, Karnataka, India
OnsiteFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
Strong Varicent configuration and development, rules and calculations, integrations, SQL, data engineering, APIs, ETL/ELT, transformation, validation, troubleshooting, and communication skills.
Varicent, SQL, APIs, ETL, ELT
10h
Save
Mark Applied
Hide
Data Engineering Snowflake Lead Engineer _ Vice President _Data Engineering
Bengaluru, Karnataka, India
OnsiteFull Time
Morgan Stanley
Morgan StanleyNYSE: MS: Global financial services firm providing investment and wealth management.
8+ YOERequires 8+ years in data engineering or a related field, with Python, Snowflake Cortex, Databricks, SQL/NoSQL, data modeling, cloud platforms, and data pipeline experience.
Snowflake, Python, SQL/PLSQL, Snowflake Cortex, Apache Spark, Hadoop, SQL, NoSQL, Databricks, AWS, Azure, Kafka, Git, Jupyter, Tableau, Power BI, Collibra
10h
Save
Mark Applied
Hide
Senior Consultant | Databricks | Mumbai | Engineering
Mumbai or Bangalore
OnsiteFull Time
Deloitte
Deloitte: Professional services firm providing audit, consulting, and advisory services.
5+ YOERequires 5+ years of relevant data engineering experience, Databricks, PySpark, SQL, cloud platform, data architecture, ETL/ELT, DevOps, troubleshooting, and communication skills.
Databricks, PySpark, SQL, Azure, AWS, GCP, Delta Lake, Delta Live Tables (DLT), Unity Catalog, Databricks Workflows, Serverless Compute, GitHub, CI/CD, DevOps, Genie, Mosaic AI, Databricks Data Intelligence Platform, Python
10h
Save
Mark Applied
Hide
Lead Engineer - Data
Bangalore or Chennai
OnsiteFull Time
Kyndryl
KyndrylNYSE: KD: Manages and modernizes mission-critical IT infrastructure systems.
10+ YOE10+ years in data engineering and warehousing with Oracle, SQL/PLSQL, ODI, SAP BODS, Python, Dataiku, Unix/Linux, and TWS; experience in banking risk and regulatory reporting, production support, and data operations.
Oracle, SQL/PLSQL, Python, Business Object, Oracle Data Integrator (ODI), TWS, Shell Scripting, Data IKU, Dataiku, SAP BODS, Unix/Linux, AWS, Azure, Power BI, MicroStrategy, Business Objects, Jira, ServiceNow, Git, DevOps, Microsoft, Google, Amazon
20h
Save
Mark Applied
Hide
Data Engineer
Bengaluru or India
OnsiteFull Time
Accenture
AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
3+ YOEMinimum 3 years of experience, 15 years of full-time education, and Databricks expertise. Requires Python, SQL, Spark/PySpark, cloud data engineering, ETL/ELT, orchestration, streaming, and performance tuning skills.
Databricks Unified Data Analytics Platform, PySpark, Apache Spark, AWS, Microsoft Azure, Google Cloud Platform (GCP), Python, SQL, Amazon S3, Azure Blob Storage, Google BigQuery, Amazon Redshift, Apache Airflow, Parquet, Avro, JSON, Apache Kafka, dbt (Data Build Tool), Informatica, Talend, Matillion
20h
Save
Mark Applied
Hide
Data Engineer, YouTube Business Organization
Bengaluru, Karnataka, India
OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
3+ YOEBachelor's degree or equivalent practical experience, 3+ years coding and building data pipelines, dimensional models, data infrastructure, and exploratory queries; SQL and data platform experience required.
Flume, Dataflow, Apache Spark, SQL, Extract, Transform, Load (ETL)
22h
Save
Mark Applied
Hide
Software Engineer - Data Engineer - Bangalore, India
Bangalore, Karnataka, India
OnsiteFull Time, Contract
Societe Generale
Societe GeneraleEuronext Paris: GLE: Global provider of retail, investment, and private banking services.
5+ YOERequires at least 5 years of data engineering and 3 years of Azure experience, with expertise in ETLs, SQL, databases, data modeling, Azure services, scripting, and data architectures.
Microsoft Office, Microsoft Azure, Microsoft Azure DevOps, CI/CD, Azure Synapse, Databricks, PowerShell, Python, Bash, Microsoft Fabric, Microsoft Power BI, Microsoft Power Automate, SQL, SQL Server, ETLs