HackerRank
Posted 2mo ago

Lead Data Engineer

HackerRank
Bengaluru, Karnataka, India
HybridFull Time
Responsibilities
  • evolving platform
  • building datasets
  • owning features
Requirements
  • 6+ years data engineering experience (2+ years lead)
  • Expertise with OLAP (StarRocks/ClickHouse/Druid)
  • Data lakes (Apache Hudi/Iceberg/Delta Lake)
  • Trino/Presto
  • Apache Spark
  • Apache Ranger, AWS, and strong cross-functional communication
Technical tools mentioned
StarRocksClickHouseDruidApache HudiApache IcebergDelta LakeTrinoPrestoApache SparkApache RangerRedshiftAWS

Job description

HackerRank helps companies like NVIDIA, Amazon, and Microsoft hire and upskill the next generation of developers based on skills, not pedigree. Our platform is trusted by over 2,500 of the world’s most innovative companies to build strong engineering teams ready for what’s next.

Software has entered an era where humans and AI build side by side. As this shift accelerates, the definition of strong technical talent is changing. We give companies better ways to identify and invest in next-generation skills.

People at HackerRank care deeply about the impact of their work and sweat the small details so our customers can be wildly successful with products they genuinely love to use. We move with urgency and believe great outcomes come from high standards.

About the role

HackerRank's data platform is at an inflection point. We've completed a multi-year modernisation - migrating from Redshift to StarRocks + Apache Hudi - and cut export latencies from 25 seconds to under 5 seconds. The infrastructure groundwork is done. Now we're building the AI-native data layer that will power revenue-generating features like natural language querying for HackerRank for Work customers.

As Lead Data Engineer, you'll be a senior individual contributor at the heart of the data organisation - owning complex platform decisions, collaborating cross-functionally with AI, product, and go-to-market teams, and shipping data-driven features that directly drive revenue. This is a greenfield opportunity to shape the next phase of data at HackerRank.

What you will do

  • Own and evolve the data platform - StarRocks (OLAP), Apache Hudi (Data Lake), Trino, Spark, and Apache Ranger - ensuring performance, reliability, and security at scale.
  • Build the next-gen AI-optimised data layer: clean, structured datasets that power natural language querying and AI add-on features for HackerRank for Work customers.
  • Own in-product data features - exports, insights dashboards, interview analytics, and the self-serve Custom Reports interface.
  • Enable self-service pipelines for internal teams (AI platform, analytics, go-to-market), reducing ad-hoc data requests and scaling data access across the org.
  • Enforce robust data security - access controls, Apache Ranger policies, and confidence-scoring guardrails for AI-generated outputs.
  • Lead technical design reviews and define engineering standards for the data team.
  • Partner with PMs and business stakeholders to proactively identify and scope AI-enabled data use cases.

Who you are

  • 6+ years of data engineering experience, with at least 2 years in a senior or lead capacity.
  • Deep hands-on expertise with OLAP databases - StarRocks, ClickHouse, Druid, or similar.
  • Strong experience with data lake technologies - Apache Hudi, Iceberg, or Delta Lake.
  • Proficient with distributed query engines (Trino / Presto) and batch/streaming compute with Apache Spark.
  • Solid understanding of data security, RBAC, and access control tools like Apache Ranger.
  • Comfortable working in a hybrid AWS + open-source self-managed environment.
  • Strong communicator who can translate technical decisions for non-technical stakeholders and drive cross-functional projects independently.

Even better if you have

  • Hands-on experience with AI/LLM-adjacent data work - confidence scoring, agentic pipelines, RAG architectures, or vector stores.
  • Prior exposure to agentic workflows and understanding how to operationalise emerging AI concepts at production scale.
  • Experience scaling data infrastructure at a SaaS or B2B product company.
  • Familiarity with natural language querying interfaces or building data products for end-customer consumption.

You will thrive in this role if

  • You're energised by working on a platform that's both technically mature and still has enormous greenfield ahead of it.
  • You don't wait for a PM to hand you a roadmap - you proactively connect data capabilities to business outcomes.
  • You care as much about how other teams use data as you do about the pipelines that produce it.
  • You're genuinely curious about AI and want to be close to where data and intelligence intersect.
  • You thrive in lean, cross-functional environments where your decisions have visible, company-wide impact.

Want to learn more about HackerRank? Check out HackerRank.com to explore our products, solutions and resources, and dive into our story and mission here.

HackerRank is a proud equal employment opportunity and affirmative action employer. We provide equal opportunity to everyone for employment based on individual performance and qualification. We never discriminate based on race, religion, national origin, gender identity or expression, sexual orientation, age, marital, veteran, or disability status. All your information will be kept confidential according to EEO guidelines. 

Linkedin | X | Blog | Instagram | Life@HackerRank

Notice to prospective HackerRank job applicants:

  • Our Recruiters use @hackerrank.com email addresses.
  • We never ask for payment or credit check information to apply, interview, or work here.

About HackerRank

Platform for evaluating and hiring developers through skill assessments.

Year founded
2009
Employees
360
Organization type
Private
Latest investment
Raised $60.00M Series D (2022) — led by Susquehanna Growth Equity
Subsidiaries
Headquarters
US

Similar jobs

Data Engineer roles near Bengaluru, Karnataka
14h
Save
Mark Applied
Hide
Staff Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
Sandisk
SandiskNasdaq: SNDK: Designs and manufactures flash memory and data storage products.
6+ YOEBachelor's degree in Computer Science, Engineering, or related field and 6+ years in data engineering. Requires PySpark, Spark SQL, Azure, Databricks, SQL, data modeling, Git, and CI/CD expertise.
PySpark, Spark SQL, Databricks, Azure, Azure Data Factory (ADF), HVR, Fivetran, Git, Airflow, Delta Lake, Unity Catalog, Kafka, Event Hub, S/4 HANA, BDC, SQL, CI/CD
14h
Save
Mark Applied
Hide
Engineering-L2-Bengaluru-Analyst-Software Engineering
Bengaluru, Karnataka, India
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
1+ YOERequires 1–3 years of experience, relevant bachelor's or master's degree or equivalent, Python or Java, SQL, production data pipeline experience, distributed processing, data modeling, quality controls, and software engineering practices.
Python, Java, SQL, Apache Spark, JSON, Avro, Parquet, CI/CD
19h
Save
Mark Applied
Hide
Principal Data Engineer
Bengaluru, Karnataka, India
HybridFull Time
Atlassian
AtlassianNASDAQ: TEAM: Develops software for team collaboration and project management.
12+ YOE2+ MgmtRequires 12+ years in data engineering, 2+ years leading data engineering teams, large-scale systems experience, and expertise with streaming pipelines, Spark, Airflow, AWS data services, and data infrastructure.
SQL, Spark, Redshift, Airflow, AWS, Athena, EMR, Flink, Hive, Kafka, Databricks, Kappa, Lambda, Master Data Management (MDM)
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
NTT DATA
NTT DATATokyo Stock Exchange: 9613: Global provider of business and technology consulting and IT services.
Advanced Apache Spark/PySpark, AWS S3, Glue, EMR, and Lambda skills; expert PostgreSQL; strong Python and SQL; experience building scalable batch and streaming data pipelines.
Apache Spark, PySpark, Amazon S3, AWS Glue, Amazon EMR, AWS Lambda, PostgreSQL, Python, SQL
1d
Save
Mark Applied
Hide
Data Engineer- Azure Databricks Senior Associate
Bangalore, Karnataka, India
OnsiteFull Time
PwC
PwC: Global network providing professional audit, tax, and advisory services.
4+ YOERequires 4+ years of experience, a listed bachelor's-level qualification, Azure data engineering, Databricks, Hadoop, Spark, SQL, data extraction, testing, Unix scripting, and DevOps expertise.
Azure ADLS, Databricks, Data Flows, HDInsight, Azure Analysis Services, Hadoop, Spark, SQL, Unix Shell Scripting, Git, CI/CD Frameworks, Jenkins, GitLab, Code Pipeline, Code Build, Code Commit, Storm, SparkStreaming, Mahout, SparkML, H2O, Azure Functions, Python
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
Advanced Apache Spark/PySpark, AWS S3, Glue, EMR and Lambda, PostgreSQL, Python, and SQL expertise; experience designing scalable batch and streaming pipelines, ETL, optimization, and data governance.
Apache Spark, PySpark, Amazon S3, AWS Glue, Amazon EMR, AWS Lambda, PostgreSQL, Python, SQL
1d
Save
Mark Applied
Hide
Senior Associate | Data Engineering | Bengaluru | Engineering as a Service/ Operate
Bengaluru, Karnataka, India
OnsiteFull Time
Deloitte
Deloitte: Professional services firm providing audit, consulting, and advisory services.
3+ YOEBachelor's degree in a relevant field and 3+ years in data or software engineering. Requires ETL/ELT, large-scale datasets, SQL, Python, cloud platforms, data architecture, analytical, and collaboration skills.
SQL, Python, Azure, AWS, Google Cloud, Databricks, Spark, PySpark, Snowflake, BigQuery, Airflow, Azure Data Factory, Cloud Composer
3d
Save
Mark Applied
Hide
Lead Data Engineer - R01568425
Bangalore, Karnataka, India
OnsiteFull Time
Brillio
Brillio: Digital technology services and AI-led transformation partner.
4+ YOERequires 4–6 years of ETL and Informatica MDM experience, advanced SQL, Python, data warehousing, data modeling, PL/SQL, T-SQL, stored procedures, and a bachelor's degree in a relevant field.
SQL, Python, Informatica MDM, PL/SQL, T-SQL, Stored Procedures, Azure Data Factory, AWS Glue, Spark, Hadoop