Amazon
Posted 4d ago

Data Engineer, Conversational Shopping Data Engineering

Amazon
Bengaluru, Karnataka, India
OnsiteFull Time
Responsibilities
  • designing data models
  • building pipelines
  • analyzing data
Requirements
  • Requires 1+ year of data engineering experience
  • Data modeling
  • Data warehousing
  • ETL pipelines
  • Query and scripting languages, and experience with databases
  • Analytics, or big data technologies
Technical tools mentioned
RedshiftSpectrumEMRETLSQLNoSQLAmazon S3AWS LambdaAWS GlueHadoopSparkHiveImpalaPythonKornShellPL/SQLDDLMDXHiveQLSparkSQLScalaInformaticaODISSISBODIDatastage

Job description

Description

As a Data Engineer, you will be working on building and maintaining complex data pipelines, assemble large and complex datasets to generate business insights and to enable data driven decision making and support the rapidly growing and dynamic business demand for data.
Data Engineers help us design, manage, and continuously enhance our Analytics needs. You will develop and deliver analytics applications including metrics generation, metrics correlation, modeling, and many other use cases to help improve process effectiveness, customer experience, and automation. As a DE, you will deal with technical aspects of data warehousing (Redshift, Spectrum, EMR, ETL), infrastructure and build data pipelines, tools, and reports that enable program managers, analysts, BIE’s, solution architects, and executives to design and deliver bench marking services for Amazon’s business units. We want to you to feel welcomed, included and valued right from the start. We know that your experiences will help us build a better world

Key job responsibilities
• Design data schema and operate internal data warehouses and SQL/NOSQL database systems.
• Design data models, implement, automate, optimization and monitor data pipelines
• Own the design, development and maintenance of ongoing metrics, reports, analyses, dashboards, etc. to drive key business decisions
• Analyze and solve problems at their root, stepping back to understand the broader context
• Manage Redshift/Spectrum/EMR infrastructure, and drive architectural plans and implementation for future data storage, reporting, and analytic solutions
• Work on different AWS technologies such as S3, Redshift, Lambda, Glue, etc.. and Explore and learn the latest AWS technologies to provide new capabilities and increase efficiency
• Work on data lake platform and different components in the data lake such as Hadoop, Amazon S3 etc.
• Work on SQL technologies on Hadoop such as Spark. Hive, Impala etc.
• Recognize and adopt best practices in reporting and analysis: data integrity, test design, analysis, validation, and documentation.
• Must possess verbal and written communication skills, be self-driven, and deliver high quality results in a fast-paced environment.
• Conduct rapid prototyping and proof of concepts
• Conceptualize and develop automation tools for bench marking data collection and analytics
• Interface with other technology teams to extract, transform, and load data from a wide variety of data sources using SQL and AWS big data technologies

A day in the life
0

Basic Qualifications

- 1+ years of data engineering experience
- Experience with data modeling, warehousing and building ETL pipelines
- Experience with one or more query language (e.g., SQL, PL/SQL, DDL, MDX, HiveQL, SparkSQL, Scala)
- Experience with one or more scripting language (e.g., Python, KornShell)

Preferred Qualifications

- Experience with big data technologies such as: Hadoop, Hive, Spark, EMR
- Experience with any ETL tool like, Informatica, ODI, SSIS, BODI, Datastage, etc.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

About Amazon

Global online retail and cloud computing technology provider.

Similar jobs

Data Engineer roles near Bengaluru, Karnataka
14h
Save
Mark Applied
Hide
Staff Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
Sandisk
SandiskNasdaq: SNDK: Designs and manufactures flash memory and data storage products.
6+ YOEBachelor's degree in Computer Science, Engineering, or related field and 6+ years in data engineering. Requires PySpark, Spark SQL, Azure, Databricks, SQL, data modeling, Git, and CI/CD expertise.
PySpark, Spark SQL, Databricks, Azure, Azure Data Factory (ADF), HVR, Fivetran, Git, Airflow, Delta Lake, Unity Catalog, Kafka, Event Hub, S/4 HANA, BDC, SQL, CI/CD
15h
Save
Mark Applied
Hide
Engineering-L2-Bengaluru-Analyst-Software Engineering
Bengaluru, Karnataka, India
OnsiteFull Time
Goldman Sachs
Goldman SachsNYSE: GS: Global investment banking, securities, and investment management firm.
1+ YOERequires 1–3 years of experience, relevant bachelor's or master's degree or equivalent, Python or Java, SQL, production data pipeline experience, distributed processing, data modeling, quality controls, and software engineering practices.
Python, Java, SQL, Apache Spark, JSON, Avro, Parquet, CI/CD
19h
Save
Mark Applied
Hide
Principal Data Engineer
Bengaluru, Karnataka, India
HybridFull Time
Atlassian
AtlassianNASDAQ: TEAM: Develops software for team collaboration and project management.
12+ YOE2+ MgmtRequires 12+ years in data engineering, 2+ years leading data engineering teams, large-scale systems experience, and expertise with streaming pipelines, Spark, Airflow, AWS data services, and data infrastructure.
SQL, Spark, Redshift, Airflow, AWS, Athena, EMR, Flink, Hive, Kafka, Databricks, Kappa, Lambda, Master Data Management (MDM)
1d
Save
Mark Applied
Hide
Senior Associate | Data Engineering | Bengaluru | Engineering as a Service/ Operate
Bengaluru, Karnataka, India
OnsiteFull Time
Deloitte
Deloitte: Professional services firm providing audit, consulting, and advisory services.
3+ YOEBachelor's degree in a relevant field and 3+ years in data or software engineering. Requires ETL/ELT, large-scale datasets, SQL, Python, cloud platforms, data architecture, analytical, and collaboration skills.
SQL, Python, Azure, AWS, Google Cloud, Databricks, Spark, PySpark, Snowflake, BigQuery, Airflow, Azure Data Factory, Cloud Composer
1d
Save
Mark Applied
Hide
Data Engineer- Azure Databricks Senior Associate
Bangalore, Karnataka, India
OnsiteFull Time
PwC
PwC: Global network providing professional audit, tax, and advisory services.
4+ YOERequires 4+ years of experience, a listed bachelor's-level qualification, Azure data engineering, Databricks, Hadoop, Spark, SQL, data extraction, testing, Unix scripting, and DevOps expertise.
Azure ADLS, Databricks, Data Flows, HDInsight, Azure Analysis Services, Hadoop, Spark, SQL, Unix Shell Scripting, Git, CI/CD Frameworks, Jenkins, GitLab, Code Pipeline, Code Build, Code Commit, Storm, SparkStreaming, Mahout, SparkML, H2O, Azure Functions, Python
1d
Save
Mark Applied
Hide
Data Engineer
Bengaluru, Karnataka, India
OnsiteFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
Advanced Apache Spark/PySpark, AWS S3, Glue, EMR and Lambda, PostgreSQL, Python, and SQL expertise; experience designing scalable batch and streaming pipelines, ETL, optimization, and data governance.
Apache Spark, PySpark, Amazon S3, AWS Glue, Amazon EMR, AWS Lambda, PostgreSQL, Python, SQL
1d
Save
Mark Applied
Hide
Data Engineer Senior Consultant
Bengaluru or India
OnsiteFull Time
NTT DATA
NTT DATATokyo Stock Exchange: 9613: Global provider of business and technology consulting and IT services.
Hands-on data platform development using SQL and Python, experience navigating large codebases, ownership of homegrown platforms, and developing data warehouse systems from the ground up.
SQL, Python, Data Warehouse, Database
3d
Save
Mark Applied
Hide
Data Specialist - R01568669
Bangalore, Karnataka, India
OnsiteFull Time
Brillio
Brillio: Digital technology services and AI-led transformation partner.
5+ YOEBachelor’s or Master’s in computer science, engineering, or related field; 5+ years’ data engineering experience; expertise in Spark, Python, Scala, SQL, cloud platforms, ETL, APIs, storage, and version control.
SQL, Google BigQuery, Google Cloud Dataproc, Python, Data Catalog, Composer, Google Cloud Dataflow, Cloud Trace, Cloud Logging, Google Cloud Storage, Data Fusion, PL/SQL, T-SQL, Apache Spark, PySpark, AWS EMR, Apache Hadoop, Apache Hive, PostgreSQL, Scala, UNIX, Databricks, API, Amazon S3, Elasticsearch, Apache Airflow, Autosys, Git, SVN, HTML