Bill & Melinda Gates Foundation
Posted 1w ago

Senior Data Engineer, IT Enterprise Data Solutions

Bill & Melinda Gates Foundation
Seattle or Washington
$157k-$236k/yrOnsiteFull Time
Responsibilities
  • building pipelines
  • modeling data
  • monitoring platforms
Requirements
  • Bachelor’s degree or equivalent experience
  • 7+ years of enterprise data engineering experience
  • Expertise in Azure data services
  • Databricks
  • Snowflake
  • Dbt Cloud, SQL, and Python
  • Experience with production platforms
  • Security
  • Governance, CI/CD, and AI data use cases
Technical tools mentioned
AzureDatabricksSnowflakedbt CloudAzure Data FactoryFivetranSQLPythonSparkPySparkScalaDatabricks notebooksSnowparkPower BICollibraJiraScrumDevSecOpsCI/CDMicrosoft ExcelJSONXMLCSV

Job description

The Foundation

We are the largest nonprofit fighting poverty, disease, and inequity around the world. Founded on a simple premise: people everywhere, regardless of identity or circumstances, should have the chance to live healthy, productive lives. We believe our employees should reflect the rich diversity of the global populations we aim to serve. We provide an exceptional benefits package to employees and their families which include comprehensive medical, dental, and vision coverage with no premiums, generous paid time off, paid family leave, foundation-paid retirement contribution, regional holidays, and opportunities to engage in several employee communities. As a workplace, we’re committed to creating an environment for you to thrive both personally and professionally.

The Team

As part of the IT Enterprise Data Solutions (EDS) department, the Data Engineering team’s mission is to lead on data technology and utilization, enabling informed decision-making and strategic insights across the foundation. We are committed to providing a robust, secure, and scalable data platform that empowers data-driven initiatives and cultivate collaboration. Working with our business operations and foundation strategy program partners, the team supports the management of the EDW, Enterprise Data Platform, building and integration of data systems, shared data exchange, data platform operations and data architecture support.

Your Role

As a Senior Data Engineer, you will design, build, operate, and continuously improve secure, scalable, and reliable data solutions on the foundation’s Modern Data Platform. You will be a senior hands-on engineer and domain expert who translates business and technical requirements into production-ready pipelines, data products, models, and platform capabilities.

You will work with data and AI engineers, business systems analysts, product owners, data analysts, BI engineers, security specialists, and service partners. Your work will improve the availability, quality, lineage, and usability of data for reporting, search, AI-assisted knowledge discovery, retrieval-augmented generation, and other data-driven decision-making across the foundation and its affiliates.

*This is a Seattle based role.

What You’ll Do

  • Data Engineering & Solution Delivery

Design, develop, test, deploy, and support configuration-driven ELT/ETL pipelines that acquire data from enterprise applications, databases, APIs, file repositories, partner data sources, and shared data exchanges.

Build reusable ingestion and transformation patterns for structured, semi-structured and unstructured data, including relational data, JSON, XML, CSV, Excel, and document metadata.

Develop curated data products and dimensional models using star and snowflake techniques to support enterprise reporting, semantic models, analytics, and downstream operational use cases.

Implement transformations using SQL, Python, Spark/PySpark, Databricks notebooks, dbt Cloud, Snowpark, and related cloud-native engineering tools.

Contribute hands-on to modernization initiatives, including migration from legacy data warehouse and ETL patterns to lakehouse, cloud warehouse, and ELT-based architectures.

  • Platform Reliability, Security & Operations

Monitor and optimize pipelines, storage, compute, and databases for performance, scalability, availability, and cost efficiency.

Implement data quality controls, reconciliation, exception handling, observability, logging, alerting, and service-level measures for critical data flows.

Lead or support incident triage, root-cause analysis, restoration, and follow-through on preventive actions for production data services.

Apply secure engineering practices including managed identity, role-based access, encryption, secrets management, least-privilege access, and sensitive-data handling requirements.

Implement and maintain CI/CD, automated testing, version control, release management, and environment promotion practices for data workloads.

  • AI, Search & Advanced Analytics Enablement

Engineer governed, high-quality data and document pipelines that support enterprise search, AI-assisted knowledge discovery, retrieval-augmented generation, analytics, and data science use cases.

Prepare and curate source data, metadata, permissions, lineage, and quality signals needed for reliable AI solutions while preserving source-system security and access policies.

Partner with Knowledge Management, AI, Business Intelligence, and Information Security teams to define fit-for-purpose data products for model grounding, evaluation, and responsible use.

Help establish repeatable patterns for document ingestion, chunking-ready content preparation, semantic metadata, and controlled data exchange across internal teams and affiliates.

Evaluate emerging data and AI platform capabilities through proofs of concept and recommend adoption based on value, security, sustainability, and architectural fit.

  • Engineering Excellence & Collaboration

Conduct design and code reviews, contribute to engineering standards, and promote reusable patterns for data modeling, orchestration, testing, observability, and documentation.

Collaborate with architects and product teams to clarify requirements, assess trade-offs, estimate delivery, and convert solution designs into actionable implementation plans.

Mentor data engineers and contingent staff through pairing, technical guidance, review, and knowledge sharing; may coordinate work across small delivery efforts.

Create and maintain technical documentation, runbooks, data mappings, lineage, and operational procedures; contribute metadata and descriptions for synchronization with Collibra.

Work effectively in Agile delivery teams using Scrum, Jira, and DevSecOps practices while communicating clearly with technical and non-technical partners.

Your Experience

  • Bachelor’s degree in computer science, data engineering, information systems, or a related field, or equivalent combination of education and experience.

  • Typically 7+ years of relevant data engineering experience in enterprise environments, including designing, delivering, and operating production data platforms and pipelines.

  • Strong hands-on expertise with Azure data services, Databricks, Snowflake, and dbt Cloud; experience with Azure Data Factory and Fivetran is strongly preferred.

  • Advanced proficiency in SQL and Python, plus practical experience with Spark/PySpark or Scala and cloud scripting or automation.

  • Experience implementing lakehouse, cloud data warehouse, ELT/ETL, and medallion or layered data architecture patterns.

  • Demonstrated experience with dimensional modeling, data normalization, slowly changing dimensions, point-in-time history, and enterprise data quality practices.

  • Experience building CI/CD pipelines for data workloads, automated testing, infrastructure-as-code or configuration-as-code, and modern version control workflows.

  • Strong understanding of data security, privacy, access controls, encryption, lineage, retention, and governance in cloud data environments.

  • Experience supporting analytics and semantic-layer platforms such as Power BI, and working with metadata/governance platforms such as Collibra, is preferred.

  • Demonstrated ability to enable AI through data, including an understanding of the architecture, quality, metadata, permissions, lineage, scalability, and evaluation requirements for search, RAG, and other AI use cases.

  • Proven ability to diagnose production issues, lead technical problem solving, and implement short-, medium-, and long-term corrective actions.

  • Experience working with vendors or distributed engineering teams and reviewing deliverables for alignment with enterprise standards and maintainability.

  • Excellent written and verbal communication skills, sound judgment, and the ability to explain complex technical concepts to varied audiences.

  • A demonstrated commitment to diversity, equity, inclusion, and respectful collaboration.

*Must have unrestricted work authorization in the country where this position is located. 

The Foundation does not provide immigration-related sponsorship for this role. This includes direct company sponsorship and any work authorization requiring a written submission or other immigration support from the company (eg: H-1B, O-1, L-1,  E, OPT, STEM-OPT, CPT, TN, J-1, etc.). 

#LI-SC2

The salary range for this role is $157,400 to $236,000 USD. We recognize high-wage market differences in Seattle and Washington D.C., where our offices are located. The range for this role in these locations is $173,100 to $259,700 USD. As a mission-driven organization, we strive to balance competitive pay with our mission. New hires salaries are typically between the range minimum and the salary range midpoint. Actual placement in the range will depend on a candidate’s job-related skills, experience, and expertise, as evaluated during the interview process.

Hiring Requirements

As part of our standard hiring process for new employees, employment will be contingent upon successful completion of a background check.

Candidate Accommodations

We’re committed to providing an inclusive and accessible hiring experience for all candidates. If you have a disability or medical condition and need an accommodation at any stage of the application or interview process—such as an ASL interpreter, alternative interview format, or physical accessibility support—we’re happy to help. Please contact [email protected] with the position number and a brief description of your accommodation needs. Requests will be handled confidentially.

Inclusion Statement

We are dedicated to the belief that all lives have equal value. We strive for a global and cultural workplace that supports ever greater diversity, equity, and inclusion — of voices, ideas, and approaches — and we support this diversity through all our employment practices.

All applicants and employees who are drawn to serve our mission will enjoy equality of opportunity and fair treatment without regard to race, color, age, religion, pregnancy, sex, sexual orientation, disability, gender identity, gender expression, national origin, genetic information, veteran status, marital status, and prior protected activity.

About Bill & Melinda Gates Foundation

Nonprofit foundation funding health, poverty-reduction, education, and economic-opportunity programs worldwide.

Similar jobs

Data Engineer roles near Seattle, Washington
7h
Save
Mark Applied
Hide
Data Engineer, Growth
San Francisco or Seattle
$190k-$240k/yr HybridFull Time
Superhuman
Superhuman: Private AI productivity platform for people and teams, combining writing, collaborative docs, email, and proactive AI agents.
3+ YOERequires 3+ years building production data pipelines, proficiency in SQL, Python, Spark, and modern data platforms, plus experience with machine learning workflows, data modeling, orchestration, CI/CD, and data quality.
Spark, Databricks, Google, Meta, LinkedIn, SQL, Python, Delta Lake, dbt, Snowflake, Databricks Workflows, Airflow, Git, Claude Code, Codex, Google Ads
23h
Save
Mark Applied
Hide
Senior Data Engineer, Selling Partner Agentic Interfaces Data Products
Bengaluru or Seattle or Karnataka
OnsiteFull Time
Amazon
AmazonNASDAQ: AMZN: Multinational technology focused on e-commerce and cloud computing.
7+ YOERequires 7+ years of data engineering, 5+ years of programming, AWS services experience, data modeling, warehousing, ETL pipelines, and technical leadership or architecture experience.
Amazon Kinesis, Amazon EMR, Apache Spark, AWS Glue, Amazon Redshift, Amazon Athena, Amazon S3, AWS Lake Formation, AWS Lambda, AWS Step Functions, Amazon SageMaker, Amazon EC2, Amazon QuickSight, Amazon Quick Suite, Apache Kafka, Java, Scala, Python, Bash, Perl, SQL
1d
Save
Mark Applied
Hide
Staff Data Engineer, Analytics
San Francisco or New York City or Los Angeles or Seattle or United Kingdom or Ireland or Poland or Germany or Australia
$207k-$290k/yr HybridFull Time
Whatnot
Whatnot: Live shopping marketplace connecting buyers and sellers.
7+ YOERequires 7+ years building data warehouses or distributed data systems; expertise in data modeling, modern data tooling, cloud warehouses, Python or SQL, CI/CD, and infrastructure-as-code.
Kafka, Debezium, dbt, Spark, Flink, Dagster, Airflow, Monte Carlo, Great Expectations, Snowflake, BigQuery, Redshift, Python, SQL, CI/CD, infrastructure-as-code
1d
Save
Mark Applied
Hide
Senior Data Engineer
Glendale or Seattle or New York City or Bristol or Burbank or New York City or Seattle or Bristol
$142k-$199k/yr OnsiteFull Time
The Walt Disney Company
The Walt Disney CompanyNYSE: DIS: A leading global entertainment and media conglomerate.
5+ YOERequires 5+ years of data engineering experience, RDBMS and SQL expertise, programming with Python or PySpark, ETL/orchestration experience, data warehousing knowledge, and a bachelor's degree or equivalent experience.
SQL Server, MySQL, Oracle, Python, PySpark, Airflow, Nifi, Amazon S3, Databricks, SQL, RDBMS, ETL
2d
Save
Mark Applied
Hide
Staff Data Engineer
Denver or Portland or Seattle or Springfield or Toronto or Bangalore
$134k-$175k/yr HybridFull Time
DAT
DAT: Private freight marketplace and logistics software serving shippers, brokers, carriers, and transportation analysts.
7+ YOE7+ years of engineering experience building large-scale distributed systems; bachelor's degree in a technical discipline; expertise in data engineering, cloud services, architecture, and software engineering practices.
dbt, Snowflake, Kafka, Airflow, Terraform, Tableau, MySQL, Oracle, Python, SQL, AWS, AWS S3, EMR, Glue, Apache Airflow, Agile, Data Mesh, CQRS, Saga
2d
Save
Mark Applied
Hide
Sr. Data Engineer, Enterprise - Slack
San Francisco or Seattle or McLean
$173k-$260k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: The #1 AI CRM driving customer success together.
5+ YOERequires 5+ years in data architecture, modeling, master or metadata management; Python and SQL; relational and dimensional modeling; ETL optimization; Airflow; governance; and U.S. citizenship or lawful permanent residency.
Airflow, Python, SQL, Spark SQL, AWS, Amazon S3, Amazon EMR, Apache Pinot, Hadoop, Flink, Scala, GCP, Azure
2d
Save
Mark Applied
Hide
Sr. Data Engineer, Enterprise - Slack
San Francisco or Seattle or McLean
$173k-$260k/yr OnsiteFull Time
Slack
Slack: Slack is a privately owned, Salesforce-controlled workplace communication platform serving businesses with messaging, automation, integrations, and AI tools.
5+ YOERequires 5+ years in data architecture, modeling, master data, or metadata management; strong Python and SQL; relational modeling, ETL optimization, Airflow, governance, SDLC, Agile, and U.S. citizenship or permanent residency.
Python, SQL, Airflow, Spark SQL, AWS S3, AWS EMR, Apache Pinot, Hadoop, Flink, Scala, AWS, GCP, Azure
3d
Save
Mark Applied
Hide
Azure Fabric Data Engineer
Bellevue, Washington, United States
$80k-$115k/yr OnsiteFull Time
Tata Consultancy Services
Tata Consultancy ServicesBSE: 532540: Global leader in IT services, consulting, and business solutions.
8+ YOERequires 8+ years of experience, advanced SQL, Python/PySpark, Azure data technologies, Microsoft Fabric, data modeling, CI/CD, Azure DevOps, Git, customer-facing delivery, and strong communication skills.
PySpark, Data Warehouse, Microsoft Fabric, Azure Data Factory, Azure DevOps, Git, SQL, OneLake, Lakehouse, Delta Tables, Hightouch, Braze, Microsoft Power BI, PL/SQL