Amazon
Posted 2mo ago

Senior Data Engineer, AWS Analytics Engineering

Amazon
Seattle, Washington, United States
$155k-$209k/yrOnsiteFull Time
Responsibilities
  • operating pipelines
  • designing architecture
  • mentoring engineers
Requirements
  • 7+ years data engineering experience with SQL
  • Data modeling, ETL
  • MPP databases
  • Distributed data systems
  • Programming in Python/Java/Scala/NodeJS
  • Experience with AWS data services and mentoring peers
Technical tools mentioned
SQLPythonJavaScalaNodeJSAmazon RedshiftAWS EMRAWS GlueAWS LambdaHadoopHiveSparkLLMs

Job description

Description

The AWS Analytics Engineering (AAE) organization is the analytics backbone of AWS — we build and operate the data platform that powers business decisions across more than 150 AWS services. Every insight surfaced to AWS product leadership, from service adoption trends to revenue drivers, flows through systems our team designs, builds, and maintains.

We operate at massive scale — processing petabytes of data daily through thousands of jobs consisting of transformations, reporting queries, ingestions, and infrastructure management scripts. Our engineers work directly with source systems to procure data, convert it into structured formats, build large-scale processing pipelines, design analytical data models, and maintain infrastructure with the highest security and compliance standards.

We are seeking a Senior Data Engineer to join our team. This individual will own 3-5 data domains end-to-end, operate hundreds of pipelines, and drive architectural improvements that impact how AWS leadership makes decisions. You will partner with service teams across AWS to design data contracts, build ingestion flows, and deliver analytical models that serve the entire organization.

The ideal candidate is a technical leader who thrives in ambiguity, takes a long-term architectural view, and consistently delivers exemplary solutions. You are an expert with SQL, ETL, and data processing, with experience leveraging cloud-based data services such as AWS EMR, Glue, Redshift, and Lambda. The candidate should have hands-on experience with AI/ML technologies, including LLMs, and a strong understanding of designing and building Agentic Frameworks — including autonomous agents, multi-agent orchestration, and tool integration. You are comfortable with ambiguity in a fast-paced environment, able to think big while paying careful attention to detail, and passionate about building data platforms using AI to accelerate the next generation of analytics at AWS scale

Key job responsibilities
Identify limitations and opportunities in data processing tools, drive improvements and innovation, define data processing guidelines, and ensure best practices in all pipelines designed and reviewed. For example: redesigning ingestion frameworks to handle new AWS service telemetry data, or building reusable transformation patterns adopted across multiple teams.

Define and own data architecture at the team level — ensuring architecture effectively matches business problems and data challenges with security, scalability, and cost effectiveness. Show good judgment making technical trade-offs between short-term technology needs and long-term business needs.

Produce exemplary code — solutions that are easily usable by customers, inventive, secure, easily maintainable, appropriately scalable, and extensible. Build solutions that are easy for others to contribute to. Work to simplify, optimize, and remove bottlenecks.

Define and own infrastructure architecture at the team level. Anticipate data management and access patterns, evolve the technology stack to remove bottlenecks, and deliver systems that are secure, scalable, and long lasting. Define team-level guidelines and best practices for infrastructure management and automation.

Solve complex ambiguous problems — for example, designing cross-domain data models that unify billing, usage, and service telemetry data, or combining multiple datasets to solve problems that couldn't be solved before. Spot areas that might lead to customer confusion, data misinterpretation, or gaps in data contracts.

Effectively split project work into parallel tasks that can be performed by themselves and others and reassembled successfully. Drive to completion projects with dependencies on peers or other teams.

Influence related teams' data architecture and software design. Provide technical assessments for promotions. Actively mentor and develop others. Build consensus when confronted with discordant views.

Drive data engineering best practices — Data Discovery, Naming Conventions, Operational Excellence, Data Security. Ensure team's data is auditable, available, and accessible.

Proactively fix data architecture deficiencies and propose larger projects which may require the work of other teams. Drive improvements through code review, design discussions, team planning, and operational reviews.

Participate in on-call rotation and own operational health of data systems — establish monitoring, alarming, runbooks, and SLA tracking. Drive continuous improvement in reliability and incident response.

Basic Qualifications

- 7+ years of data engineering experience
- Experience with data modeling, warehousing and building ETL pipelines
- Experience with SQL
- Experience in at least one modern scripting or programming language, such as Python, Java, Scala, or NodeJS
- Experience mentoring team members on best practices
- Experience with MPP databases such as Amazon Redshift
- Experience building/operating highly available, distributed systems of data extraction, ingestion, and processing of large data sets

Preferred Qualifications

- Experience with big data technologies such as: Hadoop, Hive, Spark, EMR
- Experience operating large data warehouses
- Experience providing technical leadership and mentoring other engineers for best practices on data engineering
- Bachelor's degree in computer science, engineering, analytics, mathematics, statistics, IT or equivalent
- Knowledge of distributed systems as it pertains to data storage and computing

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status, disability, or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process, including support for the interview or onboarding process, please visit https://amazon.jobs/content/en/how-we-hire/accommodations for more information. If the country/region you’re applying in isn’t listed, please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience, qualifications, and location. Amazon also offers comprehensive benefits including health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance and option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage), 401(k) matching, paid time off, and parental leave. Learn more about our benefits at https://amazon.jobs/en/benefits.



USA, WA, Seattle - 154,600.00 - 209,100.00 USD annually

About Amazon

Multinational technology focused on e-commerce and cloud computing.

Similar jobs

Data Engineer roles near Seattle, Washington
2d
Save
Mark Applied
Hide
Data Engineer, Growth
San Francisco or Seattle
$190k-$240k/yr HybridFull Time
Superhuman
Superhuman: Private AI productivity platform for people and teams, combining writing, collaborative docs, email, and proactive AI agents.
3+ YOERequires 3+ years building production data pipelines, proficiency in SQL, Python, Spark, and modern data platforms, plus experience with machine learning workflows, data modeling, orchestration, CI/CD, and data quality.
Spark, Databricks, Google, Meta, LinkedIn, SQL, Python, Delta Lake, dbt, Snowflake, Databricks Workflows, Airflow, Git, Claude Code, Codex, Google Ads
3d
Save
Mark Applied
Hide
Staff Data Engineer, Analytics
San Francisco or New York City or Los Angeles or Seattle or United Kingdom or Ireland or Poland or Germany or Australia
$207k-$290k/yr HybridFull Time
Whatnot
Whatnot: Live shopping marketplace connecting buyers and sellers.
7+ YOERequires 7+ years building data warehouses or distributed data systems; expertise in data modeling, modern data tooling, cloud warehouses, Python or SQL, CI/CD, and infrastructure-as-code.
Kafka, Debezium, dbt, Spark, Flink, Dagster, Airflow, Monte Carlo, Great Expectations, Snowflake, BigQuery, Redshift, Python, SQL, CI/CD, infrastructure-as-code
3d
Save
Mark Applied
Hide
Senior Data Engineer
Glendale or Seattle or New York City or Bristol or Burbank or New York City or Seattle or Bristol
$142k-$199k/yr OnsiteFull Time
The Walt Disney Company
The Walt Disney CompanyNYSE: DIS: A leading global entertainment and media conglomerate.
5+ YOERequires 5+ years of data engineering experience, RDBMS and SQL expertise, programming with Python or PySpark, ETL/orchestration experience, data warehousing knowledge, and a bachelor's degree or equivalent experience.
SQL Server, MySQL, Oracle, Python, PySpark, Airflow, Nifi, Amazon S3, Databricks, SQL, RDBMS, ETL
4d
Save
Mark Applied
Hide
Staff Data Engineer
Denver or Portland or Seattle or Springfield or Toronto or Bangalore
$134k-$175k/yr HybridFull Time
DAT
DAT: Private freight marketplace and logistics software serving shippers, brokers, carriers, and transportation analysts.
7+ YOE7+ years of engineering experience building large-scale distributed systems; bachelor's degree in a technical discipline; expertise in data engineering, cloud services, architecture, and software engineering practices.
dbt, Snowflake, Kafka, Airflow, Terraform, Tableau, MySQL, Oracle, Python, SQL, AWS, AWS S3, EMR, Glue, Apache Airflow, Agile, Data Mesh, CQRS, Saga
4d
Save
Mark Applied
Hide
Sr. Data Engineer, Enterprise - Slack
San Francisco or Seattle or McLean
$173k-$260k/yr OnsiteFull Time
Slack
Slack: Slack is a privately owned, Salesforce-controlled workplace communication platform serving businesses with messaging, automation, integrations, and AI tools.
5+ YOERequires 5+ years in data architecture, modeling, master data, or metadata management; strong Python and SQL; relational modeling, ETL optimization, Airflow, governance, SDLC, Agile, and U.S. citizenship or permanent residency.
Python, SQL, Airflow, Spark SQL, AWS S3, AWS EMR, Apache Pinot, Hadoop, Flink, Scala, AWS, GCP, Azure
4d
Save
Mark Applied
Hide
Sr. Data Engineer, Enterprise - Slack
San Francisco or Seattle or McLean
$173k-$260k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: The #1 AI CRM driving customer success together.
5+ YOERequires 5+ years in data architecture, modeling, master or metadata management; Python and SQL; relational and dimensional modeling; ETL optimization; Airflow; governance; and U.S. citizenship or lawful permanent residency.
Airflow, Python, SQL, Spark SQL, AWS, Amazon S3, Amazon EMR, Apache Pinot, Hadoop, Flink, Scala, GCP, Azure
6d
Save
Mark Applied
Hide
Azure Fabric Data Engineer
Bellevue, Washington, United States
$80k-$115k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesBSE: 532540: Global leader in IT services, consulting, and business solutions.
8+ YOERequires 8+ years of experience, advanced SQL, Python/PySpark, Azure data technologies, Microsoft Fabric, data modeling, CI/CD, Azure DevOps, Git, customer-facing delivery, and strong communication skills.
PySpark, Data Warehouse, Microsoft Fabric, Azure Data Factory, Azure DevOps, Git, SQL, OneLake, Lakehouse, Delta Tables, Hightouch, Braze, Microsoft Power BI, PL/SQL
6d
Save
Mark Applied
Hide
Senior Data Engineer
Chicago or Austin or Dallas or New York City or San Francisco or Seattle
$147k-$181k/yr RemoteFull Time
JLL Technologies
JLL TechnologiesNYSE: JLL: 's commercial real estate technology division provides software, consulting, and managed services to property owners and occupiers.
5+ YOERequires 5+ years of data engineering experience, Python, SQL, PySpark, cloud data pipelines, data contracts or quality governance, Delta Lake or sharing technologies, API services, and mentoring experience.
Python, SQL, Azure, Delta Lake, Delta Sharing, PySpark, Spark UI, Databricks, Unity Catalog, Azure Data Factory, Event Hubs, CI/CD