This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

🏢 Offshore Consulting Shops

Xpert Development is a software development services provider that builds custom technology solutions for client businesses; the job is for engineering roles supporting these client projects.

This company was flagged and excluded from default search results. Proceed with caution.

Xpert Development
Posted 3mo ago

Senior DevOps & Site Reliability Engineer

Xpert Development
United States
$165k-$190k/yrRemoteFull Time
Responsibilities
  • own releases
  • manage incidents
  • improve observability
Requirements
  • 4+ years in SRE/DevOps/Production Engineering
  • Strong Kubernetes, Terraform, AWS
  • Hybrid code-literate mindset
  • On-call experience
  • English proficiency for runbooks
  • Ability to patch backend code
Technical tools mentioned
KubernetesTerraformAWSObservability tooling

Job description

About the role

As our first SRE & first engineer in the US, you will own the platform’s stability and releases, especially during PST hours.

You are the perfect "bridge" profile: part system administrator, part software engineer. You don't just manage infrastructure; you understand the code running on it. You will operate with high autonomy, making critical decisions during incidents and ensuring that our production environment is state-of-the-art, secure, and resilient.

You’ll report to our Lead DevOps Engineer, Pierre, and your main mission will be:

  • Own US coverage for releases and incidents as the first responder during PST hours.

  • Bridge infra and code by working hand-in-hand with our DevOps team on Kubernetes, Terraform, and AWS, while being able to read and patch Elixir code to unblock yourself without waiting for a backend engineer.

  • Drive incident response end-to-end, managing triage, mitigation, and blameless post-mortems with real follow-through.

  • Improve the platform’s operability by defining SLOs, tuning alerts to reduce toil, and pushing observability (metrics, logs, tracing) where it’s lacking.

  • Transfer operational knowledge from France to the US by authoring runbooks and documenting procedures so local teams are empowered to act when something breaks.

  • Support compliance and security in our regulated medical-device environment, maintaining HIPAA-aligned controls and an audit-ready infrastructure.

About the profile

Sonio is a mission-driven company, so interest in our mission is critical. Other requirements are:

  • 4+ years of experience in SRE, DevOps, or Production Engineering, including significant on-call experience on a 24/7 product

  • You possess a hybrid "code-literate" mindset, acting as an infrastructure expert who can also navigate a backend codebase to triage and patch issues independently.

  • You bring strong technical foundations in Kubernetes, Terraform, and AWS, along with the ability to architect and tune your own observability signals.

  • You are highly autonomous and comfortable making technical decisions with limited supervision, which is essential given the timezone difference with France.

  • You maintain operational rigor and stay calm under pressure, with the written English skills necessary to produce high-quality runbooks and handle async handoffs.

Location: where you can cover for PST timezone (not necessarily only in the US)

Salary: $165,000 -190,000 + 10% bonus

Benefits:

⚕️Health Insurance (Medical plan, vision, dental) - up to 30,000$ per year + FSA & HSA

👵 401(k) - up 4% of your salary matched

⛑️ Life Insurance - covering 2 times your salary, up to $200k

🐣 An attractive Parental Policy for primary and secondary caregivers

🏝️ 20 PTO + 1 week offered between Christmas and New Year

🖥️ Offices in Boston (HQ) & New York (incl. free breakfast, drinks & gym)

⏰ Flexible hours & remote policies

🚎 Commuter Benefits

✈️ One offsite per year in France & regular team building with US team

🚀 Ongoing trainings and continuous opportunities for professional growth and development, specifically unlimited access to coaching

We move fast and aspire to be transparent over the process - our objective is that the process from the first chat to an offer is no longer than a month.

Similar jobs

Senior DevOps & Site Reliability Engineer roles
2mo
Save
Mark Applied
Hide
Senior DevOps Engineer/Site Reliability Engineer
United States
$165k-$215k/yr RemoteFull Time
Stellar Cyber
Stellar Cyber: Unified platform for automated cyber threat detection and response.
5+ YOE5+ years in DevOps/SRE/Platform Engineering; Kubernetes, Docker; cloud production experience; IaC with Terraform/Helm; CI/CD tooling; Linux, networking, distributed systems; observability stack; Python/Go/Bash; on-call; data platforms; AI-driven tooling; strong collaboration; East Coast residence.
Kubernetes, Docker, Terraform, Helm, Prometheus, Grafana, Loki, Alertmanager, Elastic Stack, Kafka, Spark, Elasticsearch, Redis, MongoDB, ArgoCD, GitHub Actions, Python, Go, Bash
This job has expired