Amazon
Posted 1mo ago

Data Engineer, IN Data Engineering & Analytics

Amazon
Bengaluru, Karnataka, India
OnsiteFull Time
Responsibilities
  • designing infrastructure
  • building pipelines
  • optimizing pipelines
Requirements
  • 3+ years data engineering experience
  • Strong SQL and ETL skills
  • Programming in Python/Java/Scala/NodeJS
  • Experience with data modeling
  • Warehousing, and large datasets
  • Strong communication and business acumen
Technical tools mentioned
Redshift SQLSpark SQLPythonJavaScalaNodeJSRedshiftS3AWS GlueEMRKinesisFireHoseLambdaIAMHadoopHiveSpark

Job description

Description

Amazon IN Platform Development team is looking to hire a rock star Data/BI Engineer to build for pan Amazon India businesses.

Amazon India is at the core of hustle @ Amazon WW today and the team is charted with democratizing data access for the entire marketplace & add productivity. That translates to owning the processing of every Amazon India transaction, for which the team is organized to have dedicated business owners & processes for each focus area. The Data Engineer will play a key role in contributing to the success of each focus area, by partnering with respective business owners and leveraging data to identify areas of improvement & optimization. He / She will build deliverables like business process automation, payment behavior analysis, campaign analysis, fingertip metrics, failure prediction etc. that provide edge to business decision making AND can scale with growth. The role sits in the sweet spot between technology and business worlds AND provides opportunity for growth, high business impact and working with seasoned business leaders.

An ideal candidate will be someone with sound technical background in data domain – storage / processing / analytics, has solid business acumen and a strong automation / solution oriented thought process. Will be a self-starter who can start with a business problem and work backwards to conceive & devise best possible solution. Is a great communicator and at ease on partnering with business owners and other internal / external teams. Can explore newer technology options, if need be, and has a high sense of ownership over every deliverable by the team. Is constantly obsessed with customer delight & business impact / end result and ‘gets it done’ in business time.

Key job responsibilities
- Design, implement and support an data infrastructure for analytics needs of large organization
- Adopt and drive AI-native practices while building data solutions. Build AI-powered products for business requirements
- Interface with other technology teams to extract, transform, and load data from a wide variety of data sources using Redshift SQL, Spark SQL and other AWS big data technologies
- Optimize data pipelines, build and manage Python-based automations/tools for data pipelines
- Be enthusiastic about building deep domain knowledge about Amazon’s business
- Must possess strong verbal and written communication skills, be self-driven, and deliver high quality results in a fast-paced environment
- Enjoy working closely with your peers, stakeholders and customers.
- Help continually improve ongoing reporting and analysis processes, automating or simplifying self-service support for customers
- Explore and learn the latest AWS technologies to provide new capabilities and increase efficiency

About the team
India Data Engineering and Analytics (IDEA) team is central data engineering team for Amazon India. Our vision is to simplify and accelerate data driven decision making for Amazon India by providing cost effective, easy & timely access to high quality data. We achieve this by building UDAI (Unified Data & Analytics Infrastructure for Amazon India) which serves as a central data platform and provides data engineering infrastructure, ready to use datasets and self-service reporting capabilities. Our core responsibilities towards India marketplace include a) providing systems(infrastructure) & products that allow ingestion, storage, processing and querying of data b) building AI-powered products to transform the data space c) building ready-to-use datasets for easy and faster access to the data d) automating standard business analysis / reporting/ dashboarding e) empowering business with self-service AI and non-AI tools for deep dives & insights seeking.

Basic Qualifications

- 3+ years of data engineering experience
- Experience with SQL
- Experience with data modeling, warehousing and building ETL pipelines
- Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS
- Experience as a data engineer or related specialty (e.g., software engineer, business intelligence engineer, data scientist) with a track record of manipulating, processing, and extracting value from large datasets

Preferred Qualifications

- Experience with AWS technologies like Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions
- Experience with big data technologies such as: Hadoop, Hive, Spark, EMR
- Experience working on and delivering end to end projects independently
- Experience with non-relational databases / data stores (object storage, document or key-value stores, graph databases, column-family databases)

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

About Amazon

Global online retail and cloud computing technology provider.

Similar jobs

Data Engineer roles near Bengaluru, Karnataka
12h
Save
Mark Applied
Hide
Staff Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
Sandisk
SandiskNasdaq: SNDK: Designs and manufactures flash memory and data storage products.
6+ YOEBachelor's degree in Computer Science, Engineering, or related field and 6+ years in data engineering. Requires PySpark, Spark SQL, Azure, Databricks, SQL, data modeling, Git, and CI/CD expertise.
PySpark, Spark SQL, Databricks, Azure, Azure Data Factory (ADF), HVR, Fivetran, Git, Airflow, Delta Lake, Unity Catalog, Kafka, Event Hub, S/4 HANA, BDC, SQL, CI/CD
13h
Save
Mark Applied
Hide
Engineering-L2-Bengaluru-Analyst-Software Engineering
Bengaluru, Karnataka, India
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
1+ YOERequires 1–3 years of experience, relevant bachelor's or master's degree or equivalent, Python or Java, SQL, production data pipeline experience, distributed processing, data modeling, quality controls, and software engineering practices.
Python, Java, SQL, Apache Spark, JSON, Avro, Parquet, CI/CD
17h
Save
Mark Applied
Hide
Principal Data Engineer
Bengaluru, Karnataka, India
HybridFull Time
Atlassian
AtlassianNASDAQ: TEAM: Develops software for team collaboration and project management.
12+ YOE2+ MgmtRequires 12+ years in data engineering, 2+ years leading data engineering teams, large-scale systems experience, and expertise with streaming pipelines, Spark, Airflow, AWS data services, and data infrastructure.
SQL, Spark, Redshift, Airflow, AWS, Athena, EMR, Flink, Hive, Kafka, Databricks, Kappa, Lambda, Master Data Management (MDM)
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
NTT DATA
NTT DATATokyo Stock Exchange: 9613: Global provider of business and technology consulting and IT services.
Advanced Apache Spark/PySpark, AWS S3, Glue, EMR, and Lambda skills; expert PostgreSQL; strong Python and SQL; experience building scalable batch and streaming data pipelines.
Apache Spark, PySpark, Amazon S3, AWS Glue, Amazon EMR, AWS Lambda, PostgreSQL, Python, SQL
1d
Save
Mark Applied
Hide
Data Engineer- Azure Databricks Senior Associate
Bangalore, Karnataka, India
OnsiteFull Time
PwC
PwC: Global network providing professional audit, tax, and advisory services.
4+ YOERequires 4+ years of experience, a listed bachelor's-level qualification, Azure data engineering, Databricks, Hadoop, Spark, SQL, data extraction, testing, Unix scripting, and DevOps expertise.
Azure ADLS, Databricks, Data Flows, HDInsight, Azure Analysis Services, Hadoop, Spark, SQL, Unix Shell Scripting, Git, CI/CD Frameworks, Jenkins, GitLab, Code Pipeline, Code Build, Code Commit, Storm, SparkStreaming, Mahout, SparkML, H2O, Azure Functions, Python
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
Advanced Apache Spark/PySpark, AWS S3, Glue, EMR and Lambda, PostgreSQL, Python, and SQL expertise; experience designing scalable batch and streaming pipelines, ETL, optimization, and data governance.
Apache Spark, PySpark, Amazon S3, AWS Glue, Amazon EMR, AWS Lambda, PostgreSQL, Python, SQL
1d
Save
Mark Applied
Hide
Senior Associate | Data Engineering | Bengaluru | Engineering as a Service/ Operate
Bengaluru, Karnataka, India
OnsiteFull Time
Deloitte
Deloitte: Professional services firm providing audit, consulting, and advisory services.
3+ YOEBachelor's degree in a relevant field and 3+ years in data or software engineering. Requires ETL/ELT, large-scale datasets, SQL, Python, cloud platforms, data architecture, analytical, and collaboration skills.
SQL, Python, Azure, AWS, Google Cloud, Databricks, Spark, PySpark, Snowflake, BigQuery, Airflow, Azure Data Factory, Cloud Composer
3d
Save
Mark Applied
Hide
Lead Data Engineer - R01568425
Bangalore, Karnataka, India
OnsiteFull Time
Brillio
Brillio: Digital technology services and AI-led transformation partner.
4+ YOERequires 4–6 years of ETL and Informatica MDM experience, advanced SQL, Python, data warehousing, data modeling, PL/SQL, T-SQL, stored procedures, and a bachelor's degree in a relevant field.
SQL, Python, Informatica MDM, PL/SQL, T-SQL, Stored Procedures, Azure Data Factory, AWS Glue, Spark, Hadoop