Rootly
Posted 8mo ago

Senior Platform Engineer

Rootly
Canada
HybridFull Time
Responsibilities
  • enhance observability
  • own pipelines
  • build automation
Requirements
  • 5+ years in SRE/Platform/DevOps or Infrastructure Engineering
  • 5+ years writing production software
  • Strong cloud, distributed systems, observability
  • Production-grade tooling
  • Experience with CI/CD and scaling
Technical tools mentioned
Ruby on RailsPostgreSQLRedisSidekiqAWS LambdaKafkaDynamoDBTurboStimulusViewComponentsAWSTerraform

Job description

About Rootly


At Rootly, we are on a mission to be the go-to way companies respond when things go wrong, helping every organization be more reliable. We do this by building an industry-leading incident management platform that allows companies around the world consistently and quickly resolve incidents. We are not simply transforming an industry, we are carving an entirely new +$B segment ourselves and need incredible talent to achieve this ambitious goal together.

Customers love Rootly. Some of the fastest growing companies around the world such as NVIDIA, Figma, Canva, Tripadvisor, Squarespace and more rely on Rootly to power their critical incident management process. They obsess over our delightful enterprise-ready platform and unique partnership model. See why our customers have reviewed us 5 stars on G2.

Investors love Rootly. We are backed by some of the most respected funds in the world from Y Combinator to operators like the CTO of Dropbox and GitHub. We'd be happy to disclose our entire funding and profitability picture live during the interview. As a culture we relentlessly put transparency first. We conduct monthly financial reviews as a team so everyone has a pulse on the health of the business and publish what we are building in our weekly changelog.

About the Role


This is a rare opportunity to join Rootly as an early engineer and fundamentally shape our trajectory. You’ll work on infrastructure that underpins incident response and on-call for some of the most forward-thinking teams in the world. This is not a traditional ops role - we’re looking for software engineers who love infrastructure, developer experience, think in systems, and are obsessed with building tools that make the entire engineering org more reliable, performant, and scalable.

We move with speed, taste, and impact. At Rootly, engineers are expected to take initiative and deliver high-leverage work—whether it’s building automation to remove toil, improving our observability stack, or making our services more resilient. You’ll work closely with product engineers and customers alike to ensure we’re always one step ahead. If you thrive in environments where ownership is real, excellence is expected, and reliability is non-negotiable, this is the place for you.

What You’ll Do


  • Embed with product teams to enhance observability, reliability, and performance of their services.
  • Own our CI/CD pipelines, observability tooling, monitoring systems, and incident response processes.
  • Build tools and automation to eliminate manual toil, improve engineering velocity and developer experience, and improve system reliability.
  • Collaborate deeply across engineering to understand systems at the code level and surface cross-cutting reliability, performance, and scaling concerns early.
  • Architect and scale our infrastructure, ensuring best-in-class performance, availability, and operational excellence.
  • Drive capacity planning efforts to ensure our infrastructure is resilient and scalable as we grow.
  • Define and manage SLOs and error budgets in partnership with Engineering teams who own production services.
  • Act as a strong voice for reliability, performance, and scalability across the engineering organization.


What You'll Need


You don’t need a fancy degree or a resume full of logos. What matters is your ability to execute, influence, and inspire. If the following sounds like you, we want to talk:

Minimum Qualifications


  • 5+ years of experience in an SRE, Platform, DevOps, or Infrastructure Engineering role.
  • 5+ years of experience writing software in a production environment.
  • Strong technical knowledge of cloud infrastructure, distributed systems, and reliability practices.
  • Strong understanding of observability, performance tuning, and scaling strategies.
  • Deep familiarity with incident response, monitoring, and CI/CD systems.
  • Hands-on experience supporting web or RPC services at meaningful scale.
  • You write code to solve infrastructure problems—not shell scripts alone, but production-grade software.


Preferred Qualifications


  • You have a big-picture systems mindset and a proactive approach to reliability.
  • You’ve embedded with product teams and influenced design and architecture decisions.
  • You’re comfortable taking ownership of complex problems—and seeing them through.
  • Experience with Ruby and Go is a plus.


Our Tech Stack


  • Backend: Ruby on Rails, PostgreSQL, Redis, Sidekiq, AWS Lambdas, Kafka, DynamoDB
  • Frontend: Turbo, Stimulus, ViewComponents
  • Infrastructure: AWS, Terraform (IaC)


Why Rootly?


We’re not just another startup. We’re building something category-defining and want teammates who crave ownership, love solving hard problems, and thrive in a high-bar, high-impact environment.

Here’s what you can expect when you join Rootly:

  • Competitive compensation and early equity in a fast-growing, venture-backed company.
  • Comprehensive medical, dental, and vision coverage.
  • 3 weeks of vacation, plus unlimited sick and mental health days, and a company-wide end-of-year shutdown to recharge.
  • $500 stipend for home office setup.
  • Unlimited token usage and access to AI tools
  • A fast-moving, high-impact environment where your leadership and ideas directly shape the future of the company.

If this sounds like the kind of challenge and opportunity you’re looking for, apply now and let’s build something great together.

Rootly is an equal opportunity employer. We aim to create an environment where every team member at Rootly feels like they belong so they can have a greater impact on our business and customers. We do not discriminate on the basis of race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

About Rootly

A mission-driven software focused on reliable incident response.

Employees
75
Organization type
Private
Latest investment
Raised $12.00M Series A (2023) — led by Renegade Partners, XYZ Venture Capital, Gradient Ventures
Headquarters
US

Similar jobs

Senior Platform Engineer roles
2mo
Save
Mark Applied
Hide
Senior Platform Engineer
Calgary, Alberta, Canada
$100k-$140k/yr RemoteFull Time
Wagepoint
Wagepoint: Cloud-based payroll software for small businesses.
6+ YOE6+ years cloud infrastructure, 2+ years Azure, Kubernetes, Terraform, CI/CD, OpenTelemetry, SOC/fintech security.
Azure, AWS, Kubernetes, Docker, Terraform, GitHub, Azure DevOps, OpenTelemetry
4mo
Save
Mark Applied
Hide
Senior Platform Engineer
Toronto, Ontario, Canada
$115k-$140k/yr HybridFull Time
Guidepoint
Guidepoint: Connects business leaders with global subject-matter experts.
5+ YOE5+ years software development and 7+ years DevOps/systems engineering; strong Python/Go; Kubernetes (AKS); Terraform/Terragrunt; ArgoCD; Datadog; AI agent tooling; on-call and incident response experience.
Python, Go, Kubernetes, AKS, Terraform, Terragrunt, ArgoCD, Datadog, Backstage
4mo
Save
Mark Applied
Hide
Senior Platform Engineer, Machine Learning
Toronto, Ontario, Canada
$143k-$200k/yr HybridFull Time
Movable Ink
Movable Ink: Provides AI-powered content personalization for digital marketing campaigns.
5+ YOE5+ years in software engineering focusing on distributed systems; strong Python; experience with Spark/Flink/Beam; cloud (AWS/Azure/GCP); Kubernetes, Kafka, gRPC; Docker, CI/CD; excellent problem-solving and leadership.
Python, Apache Spark, Apache Flink, Apache Beam, AWS, Azure, GCP, Kubernetes, Kafka, gRPC, Docker, GitHub Actions, CI/CD
4mo
Save
Mark Applied
Hide
Senior Platform Engineer
Toronto or United States
RemoteFull Time
Quantiphi
Quantiphi: Provides applied AI research and data engineering services.
8+ YOEDesign, build, and scale GenAI and multi-GPU infrastructure; Slurm, OpenShift/Kubernetes, NVIDIA GPU stack; IaC with Terraform/Helm; cloud/on-prem GPU clusters.
Slurm, OpenShift, Kubernetes, CUDA, cuDNN, NCCL, Triton, TensorRT, Terraform, Helm, Ansible, GCP, Azure, AWS, OCI
5mo
Save
Mark Applied
Hide
Senior Platform Engineer
Canada
$140k-$160k/yr RemoteFull Time
Lillio
Lillio: Provides a management platform for early childhood education centers.
5+ YOESenior Platform Engineer with 5+ years in software engineering/devops; experience AWS, Postgres, Rails, GraphQL, React; observability tooling; Canada applicants only.
AWS, PostgreSQL, Heroku, Redis, Terraform, Ruby on Rails, GraphQL, React, Datadog, New Relic, Papertrail, Sentry, Prometheus
2y
Save
Mark Applied
Hide
Senior Platform Implementation Engineer
London, Ontario, Canada
RemoteFull Time
RedIron
RedIron: Integrates and modernizes software systems for retail brands.
Bachelor's Degree or Diploma in computer science or related field; strong communication and project management skills; proficiency in Linux, SQL, Jira, Git/BitBucket, AWS, OCI.
Linux, SQL, Jira, Git, BitBucket, AWS, OCI, Jenkins, Bamboo, Maven
2mo
Save
Mark Applied
Hide
Senior Data Platform Developer
Calgary, Alberta, Canada
HybridFull Time
Helcim
Helcim: Provides transparent payment processing and POS solutions for small businesses.
6+ YOE6+ years in data engineering/analytics, strong data pipelines, production data products, cross-functional collaboration, mentorship, and architectural decision making.
BigQuery, dbt, Airflow, Airbyte, SQL, Python
2mo
Save
Mark Applied
Hide
Senior Data Platform Engineer
United States or Canada
$152k-$198k/yr RemoteFull Time
1Password
1Password: Software for secure password management and digital identity protection.
7+ YOEDesigns, builds, and operates scalable, streaming data systems with governance and quality controls; 7+ years in data platforms; cloud, big data, and IaC experience; strong cross-functional collaboration.
Databricks, Snowflake, Redshift, AWS, GCP, Terraform, Ansible, EC2, ECS, EKS, Kafka, Kinesis, Protobuf, Apache Iceberg, Dagster, dbt, Prometheus, Grafana, Datadog, Python, Go, Scala, Java