This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

GlobalLogic
Posted 2mo ago

Data Engineer with Bigdata, Spark, Hive and Airflow | Basic knowledge working with Kubernetes IRC274090

GlobalLogic
Noida, Uttar Pradesh, India
HybridFull Time
Responsibilities
  • building pipelines
  • optimizing pipelines
  • processing data
Requirements
  • 8+ years on Big Data platforms
  • Expertise with Spark, Hive, Airflow, Kafka
  • Scala/Python
  • Complex SQL/Hive queries
  • Streaming and batch processing
  • Cloud (AWS/GCP/Azure) and data warehouses (Redshift/Snowflake/BigQuery)
Technical tools mentioned
AirflowApache SparkHivePythonScalaKafkaOozieORCAVROParquetAmazon S3AWS RedshiftSnowflakeBigQueryAWSGCPAzureEMREKSECSElasticsearchHadoop MapReduceApache FlinkKubernetes

Job description

Description

Roku Data Engineering Project – To work on batch and streaming data generated from Roku TVs and other streaming devices



Requirements

 

1. Data engineer with 8+ years of hands – on experience (Principle Engineer I ) working on Big Data Platforms
2. Experience building and optimizing Big data data pipelines and data sets ranging from Data ingestion to Processing to Data Visualization.
3. Good Experience in writing and optimizing Spark Jobs, Spark SQL etc. Should have worked on both batch and steaming data processing
4. Good experience in any one programming language -Scala/Python , Python preferred.
5. Experience in writing and optimizing complex Hive and SQL queries to process huge data. good with UDFs, tables, joins,Views etc
6. Experience in using Kafka or any other message brokers
7. Configuring, monitoring and scheduling of jobs using Oozie and/or Airflow
8. Processing streaming data directly from Kafka using Spark jobs, expereince in Spark- streaming is must
9. Should be able to handling different file formats (ORC, AVRO and Parquet) and unstructured data
10. Should have experience with any one No SQL databases like Amazon S3 etc
11. Should have worked on any of the Data warehouse tools like AWS Redshift or Snowflake or BigQuery etc
12. Work expereince on any one cloud AWS or GCP or Azure

Good to have skills:

1. Experience in AWS cloud services like EMR, S3, Redshift, EKS/ECS etc
2. Experience in GCP cloud services like Dataproc, Google storage etc
3. Experience in working with huge Big data clusters with millions of records
4. Experience in working with ELK stack, specially Elasticsearch
5. Experience in Iceberg, Hadoop MapReduce, Apache Flink, Kubernetes etc




Job responsibilities

1. Data engineer with 8+ years of hands on experience working on Big Data Platforms
2. Experience building and optimizing Big data data pipelines and data sets ranging from Data ingestion to Processing to Data Visualization.
3. Good Experience in writing and optimizing Spark Jobs, Spark SQL etc. Should have worked on both batch and steaming data processing
4. Good experience in any one programming language -Scala/Python , Python preferred.
5. Experience in writing and optimizing complex Hive and SQL queries to process huge data. good with UDFs, tables, joins,Views etc
6. Experience in using Kafka or any other message brokers
7. Configuring, monitoring and scheduling of jobs using Oozie and/or Airflow
8. Processing streaming data directly from Kafka using Spark jobs, expereince in Spark- streaming is must
9. Should be able to handling different file formats (ORC, AVRO and Parquet) and unstructured data
10. Should have experience with any one No SQL databases like Amazon S3 etc
11. Should have worked on any of the Data warehouse tools like AWS Redshift or Snowflake or BigQuery etc
12. Work expereince on any one cloud AWS or GCP or Azure

Good to have skills:

1. Experience in AWS cloud services like EMR, S3, Redshift, EKS/ECS etc
2. Experience in GCP cloud services like Dataproc, Google storage etc
3. Experience in working with huge Big data clusters with millions of records
4. Experience in working with ELK stack, specially Elasticsearch
5. Experience in Hadoop MapReduce, Apache Flink, Kubernetes etc



What we offer

Exciting Projects: We focus on industries like High-Tech, communication, media, healthcare, retail and telecom. Our customer list is full of fantastic global brands and leaders who love what we build for them.

Collaborative Environment: You Can expand your skills by collaborating with a diverse team of highly talented people in an open, laidback environment — or even abroad in one of our global centers or client facilities!

Work-Life Balance: GlobalLogic prioritizes work-life balance, which is why we offer flexible work schedules, opportunities to work from home, and paid time off and holidays.

Professional Development: Our dedicated Learning & Development team regularly organizes Communication skills training(GL Vantage, Toast Master),Stress Management program, professional certifications, and technical and soft skill trainings.

Excellent Benefits: We provide our employees with competitive salaries, family medical insurance, Group Term Life Insurance, Group Personal Accident Insurance , NPS(National Pension Scheme ), Periodic health awareness program, extended maternity leave, annual performance bonuses, and referral bonuses.

Fun Perks: We want you to love where you work, which is why we host sports events, cultural activities, offer food on subsidies rates, Corporate parties. Our vibrant offices also include dedicated GL Zones, rooftop decks and GL Club where you can drink coffee or tea with your colleagues over a game of table and offer discounts for popular stores and restaurants!


About GlobalLogic

GlobalLogic, a Hitachi Group Company, is a trusted digital engineering partner to the world’s largest and most forward-thinking companies. Since 2000, we’ve been at the forefront of the digital revolution – helping create some of the most innovative and widely used digital products and experiences. Today we continue to collaborate with clients in transforming businesses and redefining industries through intelligent products, platforms, and services.

About GlobalLogic

Digital product engineering and software development services provider.

Similar jobs

Data Engineer roles near Noida, Uttar Pradesh
1d
Save
Mark Applied
Hide
Lead - Data Engineer
Hyderabad or Navi Mumbai or Gurgaon or Hyderabad
HybridFull Time
Jacobs
JacobsNYSE: J: Global provider of professional engineering and technical services.
10+ YOERequires 10–12 years in data engineering, warehousing, or big data; expertise in cloud platforms, Microsoft Fabric, Databricks, ETL/ELT, data architecture, databases, governance, security, and enterprise-scale solutions.
Microsoft Azure, AWS, Google Cloud Platform (GCP), Microsoft Fabric, Lakehouses, Data Pipelines, Notebooks, Dataflows, Semantic Models, OneLake, Databricks, PySpark, Delta Lake, Unity Catalog, Databricks Workflows, Structured Streaming, SQL, NoSQL, Apache Spark, REST APIs, Web Services, Event Streaming, Power BI, Tableau, CI/CD, DevOps, Infrastructure as Code (IaC)
1d
Save
Mark Applied
Hide
Sr. Data Engineer
Gurugram, Haryana, India
OnsiteFull Time
Ascendion
Ascendion: AI-native digital engineering and software development services provider.
4+ YOERequires 4+ years of data engineering experience, strong SQL and Python, ETL/ELT and pipeline development, cloud platforms, data warehouses, Spark, Airflow, Git, CI/CD, and Agile experience.
SQL, Python, AWS, Microsoft Azure, GCP, Snowflake, BigQuery, Amazon Redshift, Azure Synapse, Apache Spark, PySpark, Apache Airflow, Git, CI/CD
1d
Save
Mark Applied
Hide
Lead Data Engineer (50356)
Gurugram, Haryana, India
OnsiteFull Time
Incedo
Incedo: Providing digital transformation and AI technology services.
7+ YOERequires 7–9 years of relevant experience, a B.Tech, B.E., M.Tech, or MCA, and expertise in AWS data platforms, pipelines, ETL, security controls, Apache Spark, Python, Java, and SQL.
Amazon Web Services (AWS), AWS Glue, AWS Redshift, AWS Lambda, Apache Spark, Python, Java, SQL
1d
Save
Mark Applied
Hide
Lead Data Engineer - Data Engineering 4C
Noida, Uttar Pradesh, India
HybridFull Time
Genpact
GenpactNYSE: G: Provides business process management and digital transformation services.
Bachelor's degree in a related field, strong Azure data engineering experience, Python and SQL proficiency, ETL/data pipeline expertise, Snowflake experience, and knowledge of cloud architecture, CI/CD, and DevOps.
Microsoft Azure, Azure Data Factory, Databricks, Azure Data Lake, Azure SQL Database, Snowflake, SQL, Python, Power BI, CI/CD, DevOps
1d
Save
Mark Applied
Hide
Senior Data Engineer
Greater Noida, Uttar Pradesh, India
RemoteFull Time
TaskUs
TaskUs: Provides outsourced customer experience and digital services for technology companies.
5+ YOE5+ years in data engineering with expert Python, SQL, and PySpark skills; experience with lakehouse platforms, dbt, orchestration, infrastructure as code, DuckDB, Kubernetes, and streaming technologies.
Python, SQL, PySpark, dbt, Databricks, Delta Lake, Snowflake, Medallion Architecture, Apache Airflow, Prefect, Kafka, Kinesis, Spark Streaming, Terraform, CloudFormation, DuckDB, S3, Azure Blob, Parquet, Iceberg, Kubernetes (K8s), CI/CD
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru or Gurugram or Chennai or Hyderabad or Kolkata
OnsiteFull Time
Accenture
AccentureNYSE: ACN: Global provider of management consulting and technology services.
3+ YOERequires 3+ years of experience, 15 years of full-time education, Databricks proficiency, and experience with cloud data engineering, PySpark or Spark, Python, SQL, ETL, data pipelines, and Big Data.
Databricks Unified Data Analytics Platform, PySpark, Spark, AWS, Microsoft Azure, GCP, S3, Blob Storage, BigQuery, Redshift, Apache Airflow, Parquet, Avro, JSON, Python, SQL, Kafka, dbt, Informatica, Talend, Matillion
1d
Save
Mark Applied
Hide
Senior Snowflake Data Engineer
Noida, Uttar Pradesh, India
HybridFull Time
Ibex
IbexNASDAQ: IBEX: Provides technology-enabled customer experience and business process outsourcing services.
7+ YOERequires 7+ years in data engineering, Snowflake, advanced SQL, Python, Snowpark, ELT/ETL, dbt, cloud platforms, Git, CI/CD, testing, and data modeling.
Snowflake, SQL, Python, Snowpark, Dynamic Tables, Streams, Tasks, dbt, AWS, Azure, GCP, Git, CI/CD, Fivetran, Kafka, Airflow, Terraform, Snowflake Iceberg Tables, Snowflake Cortex, Cortex Analyst, Cortex Search, Cortex LLM
2d
Save
Mark Applied
Hide
Associate Platform Services - Data Engineer
Pune or Gurgaon
HybridFull Time
ZS
ZS: Global management consulting and technology firm for healthcare and life sciences.
1+ YOEBachelor's degree required; 1–2 years of development experience, ETL, SQL, Python, data modeling, data warehousing, data products, cloud platforms, and strong analytical and communication skills.
ZAIDYN, SQL, Python, AWS, Azure, Hadoop, Spark, PySpark, Informatica, Talend, SSIS
This job has expired