Google
Posted 2w ago

Staff Software Engineer, AI/ML, Google Distributed Cloud, Storage

Google
Kirkland, Washington, United States
$207k-$300k/yrOnsiteFull Time
Responsibilities
  • designing storage
  • optimizing throughput
  • integrating hardware
Requirements
  • Bachelor's or equivalent
  • 8 years software development
  • 5 years ML design/infrastructure
  • 5 years distributed systems and systems programming (C++, Go, or Rust)
  • Leadership and advanced degree preferred
Technical tools mentioned
C++GoRust

Job description

Minimum qualifications:

  • Bachelor’s degree or equivalent practical experience.
  • 8 years of experience in software development.
  • 5 years of experience with ML design and ML infrastructure (e.g., model deployment, model evaluation, data processing, debugging, fine tuning).
  • 5 years of experience in systems architecture, including building or maintaining distributed systems or large-scale storage architectures.
  • 5 years of experience in systems programming (e.g., C++, Go, or Rust).

Preferred qualifications:

  • Master’s degree or PhD in Engineering, Computer Science, or a related technical field.
  • 8 years of experience with data structures and algorithms.
  • 3 years of experience in a technical leadership role leading project teams and setting technical direction.
  • 3 years of experience working in a complex, matrixed organization involving cross-functional, or cross-business projects.

About the job

Google's software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend well beyond web search. We're looking for engineers who bring fresh ideas from all areas, including information retrieval, distributed computing, large-scale system design, networking and data storage, security, artificial intelligence, natural language processing, UI design and mobile; the list goes on and is growing every day. As a software engineer, you will work on a specific project critical to Google’s needs with opportunities to switch teams and projects as you and our fast-paced business grow and evolve. We need our engineers to be versatile, display leadership qualities and be enthusiastic to take on new problems across the full-stack as we continue to push technology forward.

The AI Storage team within Google Distributed Cloud (GDC) builds the foundational data layer that powers next-generation machine learning and generative AI workloads at the edge and in air-gapped environments. AI models require massive throughput and ultra-low latency; our team tackles the complex challenge of delivering high-performance, massively scalable storage systems directly to customer data centers. We focus on optimizing the data pipeline to keep GPUs and accelerators fully saturated, ensuring that enterprise and public-sector customers can run advanced AI on their most sensitive data without compromising on sovereignty, security, or speed.

Google Cloud accelerates every organization’s ability to digitally transform its business and industry. We deliver enterprise-grade solutions that leverage Google’s cutting-edge technology, and tools that help developers build more sustainably. Customers in more than 200 countries and territories turn to Google Cloud as their trusted partner to enable growth and solve their most critical business problems.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $207000 - $300000 (USD) + 20% bonus target + equity + benefits

Learn more about benefits at Google.

Responsibilities

  • Design and develop scalable, distributed File and Object storage solutions that serve as the critical foundational backbone for complex AI/ML workloads within the GDC environment.
  • Engineer advanced solutions optimized for massive throughput, specifically enabling high-frequency model checkpointing and the ultra-low latency data access required for AI training and inference.
  • Drive technical execution with external partners (e.g., VAST Data) to seamlessly integrate industry-leading, high-performance storage hardware with Google’s distributed software ecosystem.
  • Take full life-cycle responsibility for core storage services, ensuring uncompromising security, data durability, and high availability across disconnected edge and on-premises data center environments.
  • Partner closely with AI infrastructure, compute, and networking teams to architect system-wide improvements, eliminate Input/Output bottlenecks, and deliver a unified "cloud-anywhere" experience.
Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.

About Google

Provides online search, advertising, cloud computing, and consumer electronics.

Year founded
1998
Employees
190000
Organization type
Public
Headquarters
US

Similar jobs

Software Engineer roles near Kirkland, Washington
2h
Save
Mark Applied
Hide
Principal Software Engineer - AI Experiences, Copilot
Mountain View or Redmond
HybridFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Develops software, services, devices, and cloud computing solutions.
4+ YOEBachelor's degree in computer science or related technical field and 4+ years of coding experience, or equivalent. Backend APIs, scalable cloud services, distributed systems, and AI experience preferred.
C, C++, C#, Java, JavaScript, Python, GraphQL, Protobuf, Thrift, WebSocket, SSE, WebRTC, Azure, AWS, GCP, RDBMS
5h
Save
Mark Applied
Hide
Senior Software Engineer - Release
College Park or Boulder or Bothell
$146k-$209k/yr HybridFull Time
IonQ
IonQNYSE: IONQ: Develops and sells trapped-ion quantum computers and cloud services.
8+ YOERequires 8+ years in release or build engineering, hands-on release automation and artifact management, CI/CD pipeline experience, standards adoption across teams, and a bachelor's degree or equivalent practical experience.
Artifactory, GitHub Packages, CI/CD, FPGA
5h
Save
Mark Applied
Hide
Senior Software Engineer, Network Tooling
San Francisco or Sunnyvale or Bellevue or Seattle
$170k-$205k/yr OnsiteFull Time
Crusoe
Crusoe: Provides energy-efficient cloud infrastructure powered by stranded and renewable energy.
5+ YOE5+ years building production backend systems; proficiency in Go, Rust, C++, Java, or Python; distributed systems expertise; scalable API or service design; CI/CD and infrastructure as code experience; mentoring ability.
Go, Rust, C++, Java, Python, Terraform, Ansible, BGP, SDN, gNMI, gRPC, NETCONF, Kubernetes, CI/CD, infrastructure as code (IaC)
6h
Save
Mark Applied
Hide
Senior Lead Software Engineer - Infrastructure Platforms
Seattle, Washington, United States
$176k-$260k/yr OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years of applied software engineering experience, advanced programming skills, system design and cloud-native experience, and expertise using enterprise-authorized AI-assisted development tools securely.
AWS EKS, GCP GKE, Oracle, Microsoft SQL Server, Cockroach, Software Development Life Cycle, AI-assisted development
7h
Save
Mark Applied
Hide
Staff Software Engineer - Protect
San Francisco or Seattle or New York City or Washington, D.C. or London or Amsterdam
$208k-$274k/yr HybridFull Time
Plaid
Plaid: Provides financial data connectivity and payment infrastructure via APIs.
8+ YOETypically 8+ years building and operating backend or distributed systems at scale, with hands-on coding, strong product judgment, and experience translating ambiguous customer problems into measurable deliverables.
Jupyter, Apache Spark, Tecton
7h
Save
Mark Applied
Hide
Senior Software Engineer, Platform & Infrastructure - Riot Technology
Los Angeles or Mercer Island or Redwood City
$162k-$227k/yr OnsiteFull Time
Riot Games
Riot Games: Developing and publishing competitive multiplayer video games.
3+ YOEBachelor’s degree or equivalent experience; 3+ years in software engineering, infrastructure, platform engineering, or SRE; production distributed systems, Kubernetes, cloud, infrastructure-as-code, CI/CD, GPU infrastructure, and Python.
Kubernetes, AWS, GCP, Python, CI/CD, infrastructure-as-code, MLOps, HPC, Unreal
7h
Save
Mark Applied
Hide
Software Engineer – Commerce
Seattle, Washington, United States
$110k-$145k/yr HybridFull Time
Truveta
Truveta: Provides a clinical data platform for medical research insights.
3+ YOERequires 3+ years building production software, backend or cloud services, Kubernetes-hosted services, and modern cloud-native platforms. Requires a computer science degree and experience with C#, Python, or Java.
C#, Python, Java, Kubernetes, AWS, GCP, Azure, APIs
8h
Save
Mark Applied
Hide
Staff Software Engineer, Fullstack-MarTech
New York City or San Francisco or United States or Los Angeles or Seattle or Austin or Washington, D.C. or Philadelphia or San Diego or Chicago or Atlanta or Salt Lake City or Australia or Canada or South America
$170k-$250k/yr HybridFull Time
Flex
Flex: A financial platform allowing renters to split monthly rent payments.
8+ YOERequires 8+ years software development, 6+ years Java, 3+ years TypeScript and React, plus AWS, scalable APIs, event-driven systems, web performance, and cross-functional communication experience.
Java, Spring Boot, TypeScript, React, AWS, Spring, Gradle, JUnit, JVM, EKS, Aurora RDS, ElastiCache, DynamoDB, REST APIs, Braze, Singular, AppsFlyer, Branch, Google Tag Manager, Amplitude, LaunchDarkly, Meta, Google, TikTok, Apple, Snowflake, GitHub Actions, git, DataDog, CDK, Terraform, React Native