Google
Posted 1w ago

Senior Software Engineer, Site Reliability Engineering

Google
Seattle or Kirkland
$174k-$252k/yrOnsiteFull Time
Responsibilities
  • improving services
  • designing systems
  • monitoring availability
Requirements
  • Bachelor's degree or equivalent experience
  • 5 years of software development
  • 3 years designing and troubleshooting distributed systems, and 2 years leading projects and providing technical leadership

Job description

In accordance with Washington state law, we are highlighting our comprehensive benefits package, which is available to all eligible US based employees. Benefits for this role include:
  • Health, dental, vision, life, disability insurance
  • Retirement Benefits: 401(k) with company match
  • Paid Time Off: 20 days of vacation per year, accruing at a rate of 6.15 hours per pay period for the first five years of employment
  • Sick Time: 40 hours/year (increased to 69 hours/year for Seattle) including 5 discretionary sick days per instance
  • Maternity Leave (Short-Term Disability + Baby Bonding): 28-30 weeks
  • Baby Bonding Leave: 18 weeks
  • Holidays: 13 paid days per year

Note: By applying to this position you will have an opportunity to share your preferred working location from the following: Seattle, WA, USA; Kirkland, WA, USA.

Minimum qualifications:

  • Bachelor’s degree in Computer Science, Engineering, a related field, or equivalent practical experience.
  • 5 years of experience with software development in one or more programming languages.
  • 3 years of experience in designing, analyzing, and troubleshooting large-scale distributed systems.
  • 2 years of experience leading projects and providing technical leadership.

Preferred qualifications:

  • Master's degree in Computer Science or Engineering.

About the job

Site Reliability Engineering (SRE) is what you get when you treat operations as if it’s a software problem. Our mission is to progress, protect, and provide for the software and systems behind all of Google’s public services - Search, Ads, Gmail, Android, YouTube, and AppEngine, to name just a few - with an ever-watchful eye on their availability, latency, performance, and capacity.

This is an unusual job, unlike others in the industry. Like traditional operations groups, we keep important, revenue-critical systems up and running despite hurricanes, bandwidth outages, and configuration problems. Unlike traditional operations groups, we also have full access to and authority to fix, extend, and scale the code to keep it working and harden it against all the vagaries of the Internet. We hire people from both systems and software backgrounds. Strong candidates will have experience with both.

Just as what we do is unique, where we do it is unique too. At Google, we have the good fortune to have developed many interesting systems ranging from planet-spanning databases to near real-time scalable data warehousing to fault-tolerant datastream joining. In SRE, we flip between the fine-grained detail of disk driver I/O scheduling to the big picture of continental-level service capacity, across a range of systems and a user population measured in billions. We own those products in production. We drive reliability and performance across massive scale by mastering the full depth of the stack. We literally do learn something new every day - usually surprising things - that have the potential to transform the lives of billions of our users around the world.

Behind everything our users see online is the architecture built by the Technical Infrastructure team to keep it running. From developing and maintaining our data centers to building the next generation of Google platforms, we make Google's product portfolio possible. We're proud to be our engineers' engineers and love voiding warranties by taking things apart so we can rebuild them. We keep our networks up and running, ensuring our users have the best and fastest experience possible.

Individual pay is determined by factors including job-related skills, experience, and relevant education or training.

US: $174000 - $252000 (USD) + 15% bonus target + equity + benefits

Learn more about benefits at Google.

Responsibilities

  • Engage in and improve the whole lifecycle of services—from inception and design, through to deployment, operation and refinement.
  • Support services before they go live through activities such as system design consulting, developing software platforms and frameworks, capacity planning and launch reviews.
  • Maintain services once they are live by measuring and monitoring availability, latency and overall system health.
  • Scale systems sustainably through mechanisms like automation, and evolve systems by pushing for changes that improve reliability and velocity.
  • Practice sustainable incident response and blameless postmortems.
Google is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. See also Google's EEO Policy and EEO is the Law. If you have a disability or special need that requires accommodation, please let us know by completing our Accommodations for Applicants form.

About Google

Global technology specializing in internet-related services and products.

Year founded
1998
Employees
181000
Organization type
Private
Latest investment
Raised $84.75B Funding Round (2026) — led by Berkshire Hathaway
Headquarters
US

Similar jobs

Software Engineer roles near Seattle, Washington
1d
Save
Mark Applied
Hide
Software Engineer
Menlo Park or Seattle
$219k-$301k/yr OnsiteFull Time
Meta
MetaNASDAQ: META: Builds technologies that help people connect, find communities, and grow businesses.
10+ YOEBachelor's degree or equivalent experience, 10+ years in networking or infrastructure software, 4+ years designing production dataplane/control-plane systems, Kubernetes networking, C/C++, scripting, and test automation.
Kubernetes, C, C++, Python, Shell Scripting, DPDK, eBPF/XDP, AF_XDP, io_uring, RDMA/RoCEv2, Linux, TC/iptables/nftables
1d
Save
Mark Applied
Hide
Principal Software Engineer
Redmond, Washington, United States
$143k-$275k/yr HybridFull Time
Microsoft
MicrosoftNASDAQ: MSFT: Multinational technology providing software, cloud, and AI solutions.
6+ YOEBachelor's degree and 6+ years of coding experience required; preferred experience includes distributed computing, AI infrastructure, GPU programming, performance optimization, and technical leadership.
C, C++, C#, Java, JavaScript, Python, Kubernetes, Docker, Volcano, SLURM, CUDA, NCCL, PyTorch, InfiniBand, NVLink
1d
Save
Mark Applied
Hide
Senior Software Engineer II
Livingston or New York City or Sunnyvale or Bellevue
$182k-$242k/yr OnsiteFull Time
CoreWeave
CoreWeaveNasdaq: CRWV: Specialized cloud provider for large-scale AI and machine learning.
3+ YOE3+ years in systems, platform, infrastructure, or production engineering; strong Kubernetes and Linux experience; proficiency in Go, C/C++, Rust, or Bash; experience with secure runtimes, virtualization, GPU workloads, and performance tuning.
Kubernetes, Kata Containers, gVisor, KubeVirt, QEMU, Go, C, C++, Rust, Bash, Linux
1d
Save
Mark Applied
Hide
Principal Software Engineer, ML Platform
Seattle or Denver or Portland or Springfield or Toronto or Bangalore
$257k-$321k/yr HybridFull Time
DAT
DAT: Private freight marketplace and logistics software serving shippers, brokers, carriers, and transportation analysts.
Extensive ML, data infrastructure, or platform engineering experience; hands-on real-time ML platform expertise; strong Python skills; Kafka, Kubernetes, cloud, DevOps, MLOps, and team leadership experience.
Python, Go, TypeScript, Node, gRPC, Kubernetes, Kafka, Snowflake, Avro, JSON, Iceberg, Terraform, CI/CD, SQL, dbt, Helm, Postgres, RisingWave
1d
Save
Mark Applied
Hide
Staff Software Engineer, Build Services - Teamfight Tactics
Los Angeles or Mercer Island
$193k-$269k/yr OnsiteFull Time
Riot Games
Riot Games: Developer and publisher of video games and esports.
6+ YOERequires 6+ years in software development, 4+ years with C++/Java, 3+ years with Unreal Engine and Horde, 3+ years with Perforce, build and CI/CD experience, and a computer science bachelor's degree or equivalent knowledge.
C++, Java, Unreal Engine, Horde, Perforce, C#, CI/CD, Slack
1d
Save
Mark Applied
Hide
Staff Software Engineer, Infrastructure
New York City or San Francisco or Seattle or Boston or Washington or Chicago
$230k-$270k/yr HybridFull Time
Maven Clinic
Maven Clinic: Private virtual clinic providing women’s and family healthcare to employers and consumers.
8+ YOEBachelor's or master's degree in computer science or equivalent experience; 8+ years in backend development and platform architecture; expertise in distributed systems, cloud platforms, containers, orchestration, and multiple programming languages.
AWS, Google Cloud, Azure, Docker, Kubernetes, Java, Python, Go
1d
Save
Mark Applied
Hide
Senior Software Engineer, Developer Experience (DevX)
Menlo Park or New York City or Bellevue or Washington or Toronto
$196k-$230k/yr HybridFull Time
Robinhood
RobinhoodNASDAQ: HOOD: Provides brokerage, crypto, advisory, and banking services.
Experience with developer infrastructure, CI/CD pipelines, or large-scale build systems; Go or Python proficiency; distributed build systems; Kubernetes; and strong problem-solving skills.
Bazel, GitHub, Go, Python, Kubernetes
1d
Save
Mark Applied
Hide
Staff Software Engineer (Platform Core Team)
Seattle, Washington, United States
$175k-$220k/yr HybridFull Time
Félix
Félix: U.S. fintech remittance platform serving Latino immigrants with WhatsApp-based money transfers to Latin America.
8+ YOE8+ years building complex distributed systems; backend expertise in Python, Go, or Java; REST, gRPC, SQL/NoSQL, GCP, Kubernetes, or Terraform; architectural leadership and CI/CD experience; advanced English.
React, HTTP, HTML/DOM, JavaScript, CSS, AJAX, Python, Flask, FastAPI, REST APIs, Redis, SQL, gRPC, Google Cloud, Firestore, Cloud Storage, Pub/Sub, Google Artifact Registry, Cloud Run, BigQuery, Cloud Composer, Kubernetes, Terraform, NewRelic, PagerDuty, Go, Java, NoSQL, Google Cloud Platform (GCP), CI/CD