This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

Knotch
Posted 3mo ago

DevOps Engineer (GCP)

Knotch
Canada or United States
$100k-$150k/yrRemoteFull Time
Responsibilities
  • designing infrastructure
  • developing pipelines
  • managing infrastructure
Requirements
  • 5+ years in DevOps/Cloud/SRE with ≥3 years in GCP
  • Hands-on Terraform, CI/CD
  • Kubernetes/Docker
  • Monitoring with Prometheus/Grafana
  • Strong documentation and collaboration skills
Technical tools mentioned
TerraformGitHub ActionsArgoCDKubernetesDockerHelmPrometheusGrafanaGoogle Cloud Platform (GCP)AWSSnowflake

Job description

About the Role

As we evolve into an AI-native platform powered by agentic systems and large-scale data pipelines, the reliability, scalability, and observability of our infrastructure becomes mission-critical. We’re not just deploying services — we’re operating complex, production-grade AI systems that enterprise clients depend on every day.


As a DevOps Engineer, you’ll take part in building and scaling the foundation that powers everything we ship — from our core platform to our AI agents. You’ll work across infrastructure, CI/CD, observability, and security to ensure our systems are fast, resilient, and cost-efficient.


This role goes beyond "keeping the lights on". You’ll help define how Knotch operates as an AI-first company, shaping infrastructure strategy, enabling developer velocity, and ensuring our systems scale alongside rapid product innovation. If you want to own infrastructure in a high-impact environment, work closely with engineering teams across the stack, and directly influence how production systems are built and operated, this is *that* role. If you've made it this far, please kindly input this code with your application: DEVOPS-ENG-2026.


Responsibilities

  • Design, build, and maintain scalable, secure, and highly available infrastructure across pre-production and production environments.
  • Develop and manage CI/CD pipelines to enable fast, reliable, and repeatable deployments across multiple environments.
  • Own infrastructure as code (IaC) practices using tools like Terraform to ensure consistency and reproducibility.
  • Manage environment lifecycle (development, staging, production), including promotion workflows and configuration management.
  • Partner closely with Engineering, Data, and AI teams to support system performance, reliability, and scalability.
  • Implement and maintain monitoring, logging, and alerting systems to ensure high visibility into system health and performance.
  • Optimize infrastructure for cost, performance, and reliability, especially for compute- and data-intensive AI workloads.
  • Support Kubernetes-based deployments and container orchestration for distributed systems.
  • Contribute to security best practices across infrastructure, including IAM, networking, and application-level protections.
  • Create dashboards and reporting systems to provide visibility into system performance, uptime, and operational metrics.
  • Document architecture, operational processes, and infrastructure decisions to support knowledge sharing and onboarding.
  • Act as a DevOps/SRE partner across teams, helping troubleshoot issues and improve system reliability.


Qualifications
You have a minimum 5+ years of experience in DevOps, Cloud, SRE, or Infrastructure Engineering roles within SaaS, PaaS, or cloud-native environments, with at least 3 years of experience working in GCP cloud environments.

Must Haves

  • Prior experience in growth-stage and/or startup environement scaling from $10M to $20M+ ARR with a lean team.
  • Recent experience with Google Cloud Provider (GCP) is required, including IAM, networking, and data services.
  • Hands-on experience with Infrastructure as Code tools such as Terraform.
  • Experience building and maintaining CI/CD pipelines (GitHub Actions, ArgoCD, or similar).
  • Solid experience with Kubernetes, Docker, and containerized environments.
  • Familiarity with deployment tools such as Helm.
  • Experience with monitoring and observability tools like Prometheus and Grafana.
  • Strong understanding of system reliability, scalability, and performance optimization.
  • Ability to work across multiple systems and priorities in a dynamic environment.
  • Strong documentation and communication skills, with attention to clarity and detail.

Nice-to-Haves (not mandatory)

  • Supplementary experience supporting AI/ML or data-intensive workloads in production environments.
  • Familiarity with workflow orchestration or data pipeline tools.
  • Experience with cost optimization strategies for cloud infrastructure.
  • Exposure to security frameworks and compliance best practices.
  • Experience working with distributed or globally deployed systems.

How to be Successful

  • Infrastructure ownership mindset: You have experience building services in GCP from scratch. You take responsibility for system reliability, performance, and scalability — not just deployments.
  • Strong DevOps fundamentals: You understand CI/CD, IaC, observability, and containerization deeply and apply best practices consistently.
  • Systems thinking: You think holistically about how services interact, scale, and fail — and design accordingly.
  • Collaboration-first approach: You work closely with engineers across disciplines to enable velocity and reliability.
  • Pragmatic decision-making: You balance speed, cost, and reliability without over-engineering.
  • Operational excellence: You prioritize monitoring, alerting, and incident response as core parts of system design
  • Adaptability: You thrive in fast-moving environments and can context switch effectively across priorities.
  • Continuous improvement mindset: You proactively identify gaps and improve systems, processes, and tooling over time.

Why Join Knotch

We offer a unique opportunity to build and scale infrastructure that powers a truly AI-native platform. Our stack includes modern tools like Kubernetes, Terraform, AWS, Snowflake, and cutting-edge AI systems, all supported by a team actively building at the intersection of data, AI, and enterprise SaaS.


You’ll have real ownership and impact, influencing how systems are designed, deployed, and operated across the company. The work is highly consequential: the infrastructure you build will directly support production AI systems used by enterprise clients. Most importantly, you’ll be part of a team where infrastructure is not an afterthought — it’s a core part of how we innovate and scale.

The expected salary range for this opportunity is $100,000–$150,000 CAD plus our other perks and benefits, depending on skills and experience.


Our Benefits and Perks

Knotch is a fully remote company. Candidates may work from anywhere within Canada or the U.S., with a mandatory EST working time zone. Some of our other great benefits include:

  • Comprehensive medical, dental, and vision insurance eligibility
  • 401(k) plan
  • Unlimited PTO
  • 10+ company-paid holidays
  • A daily company-wide break, and more!


Equal Opportunity Employer

Knotch is a US-based equal opportunity employer. We strive to provide equal opportunities in all of our processes, including our hiring and employee experience. We pride ourselves on our three values: transparency, relentlessness, and inclusiveness.

 

We commit to daily work towards leading with empathy, reducing bias through periodic training, and engaging with and uplifting communities of marginalized groups. We condemn all forms of racism and discrimination on the basis of race, religion, ethnicity, nationality, gender identity, sexual orientation, age, marital status, pregnancy or parenthood status, veteran status, disability status, or any other identifier. We encourage all employees, clients, investors, candidates, vendors, and friends of Knotch to deliver honest feedback directly or anonymously so that we may always seek to improve as an organization dedicated to diversity, equity, inclusion, and belonging.


About Knotch

We're a growth-stage technology company helping brands optimize content performance and apply AI to modern marketing. Our culture is fast-paced, entrepreneurial, and highly adaptable—we move quickly, test ideas, and evolve with our customers.


Through the Knotch platform, brands can measure the impact of their content, identify what drives results, and continuously improve performance across channels. By combining performance data, strategic insights, and AI-driven capabilities, we help marketing teams make smarter decisions and get more impact from every piece of content they produce.

About Knotch

Helping enterprise brands measure and optimize content performance.

Year founded
2013
Employees
60
Organization type
Private
Latest investment
Raised $20.00M Series B (2019) — led by New Enterprise Associates
Headquarters
US

Similar jobs

DevOps Engineer roles
15h
Save
Mark Applied
Hide
Lead DevOps Engineer
United States
$165k-$205k/yr RemoteFull Time
Planet Depos
Planet Depos: Global provider of court reporting and litigation support services.
5+ YOE2+ MgmtRequires 5+ years in DevOps, SRE, or platform roles, 2+ years leading engineers, hands-on AWS, Terraform or equivalent IaC, CI/CD, containers, relational databases, and SOC 2 or similar compliance work.
Amazon Web Services (AWS), Terraform, CI/CD, Amazon ECS, AWS Fargate, Amazon ECR, Amazon CloudWatch, SQL Server, Postgres
15h
Save
Mark Applied
Hide
DevOps Engineer (USA)
Stamford, Connecticut, United States
HybridFull Time
Trexquant
Trexquant: Quantitative hedge fund developing systematic trading strategies.
2+ YOERequires 2+ years in DevOps, infrastructure, or SRE; Kubernetes, CI/CD, Linux administration, shell scripting, configuration tooling, observability platforms, and cross-functional collaboration experience.
Kubernetes, GitLab CI, Linux, Prometheus, Grafana, ELK
16h
Save
Mark Applied
Hide
Staff DevOps Engineer
State College, Pennsylvania, United States
OnsiteFull Time
Minitab
Minitab: Sells statistical analysis and process improvement software for businesses.
7+ YOERequires 7+ years in software engineering, DevOps, or systems administration; a bachelor's degree or equivalent experience; CI/CD, automation, containers, cloud, scripting, project leadership, and English proficiency.
Jenkins, GitLab CI/CD, GitHub Actions, Ansible, Puppet, Docker, Kubernetes, AWS, Azure, GCP, Python, Bash
16h
Save
Mark Applied
Hide
Lead DevOps/AIOps Engineer
Columbia, Maryland, United States
$130k-$155k/yr RemoteFull Time
Blend360
Blend360: Provides data science, AI, and marketing consulting services.
7+ YOERequires 7+ years in DevOps, cloud, platform engineering, or MLOps; strong GCP, BigQuery, CI/CD, Kubernetes, Terraform, security, observability, and Python or Bash experience.
Google Cloud Platform (GCP), Terraform, BigQuery, Cloud Storage, Dataflow, Pub/Sub, Dataproc, Cloud Composer, Vertex AI, Git, Kubernetes, Google Kubernetes Engine (GKE), Python, Bash, MLflow, Kubeflow, Docker, Amazon Web Services (AWS), Microsoft Azure
16h
Save
Mark Applied
Hide
DevOps Engineer
Tirana or United Kingdom or Denmark or United States
OnsiteFull Time
Dexi
Dexi: Digital commerce intelligence platform for market monitoring and analytics.
2+ YOE2+ years of Linux administration, scripting or programming, IP networking, internet protocols, AWS or Azure administration, and containerization experience.
AWS, Azure, IaaS, IDS/IPS, CI/CD, Linux, DNS, SSH, HTTP/HTTPS, FTP, DHCP
22h
Save
Mark Applied
Hide
DevOps Engineer (Data & AI Platform)
Mexico City or Los Angeles or United States or Dominican Republic or Mexico or Ukraine
$1124k-$1406k/yr RemoteFull Time
SimplePractice
SimplePractice: Practice management software for health and wellness professionals.
3+ YOERequires 3+ years in DevOps, SRE, or infrastructure engineering; AWS, Terraform, Docker, Kubernetes, CI/CD, data platforms, Python or Bash, observability, and MLOps/LLMOps experience.
Terraform, Docker, Kubernetes, AWS, CI/CD, Git, Airflow, Kafka, Spark, Python, Bash, MLflow, SageMaker, Kubeflow, Outerbounds, Metaflow
23h
Save
Mark Applied
Hide
DevOps Engineer - Level 03 - FULLY CLEARED with POLYGRAPH REQUIRED
Columbia, Maryland, United States
OnsiteFull Time
Constellation Technologies
Constellation Technologies: Provides technical services and cyber solutions for government agencies.
N/A
1d
Save
Mark Applied
Hide
DevOps Engineer, Senior
Chantilly, Virginia, United States
$78k-$176k/yr OnsiteFull Time
Booz Allen Hamilton
Booz Allen HamiltonNYSE: BAH: Consulting and technology services for government and commercial clients
5+ YOERequires 5+ years deploying AWS cloud resources, Kubernetes, infrastructure as code, container security, TS/SCI clearance with polygraph, and a bachelor's degree. Must obtain an approved security certification within 4 months.
AWS, Kubernetes, ArgoCD, Flux, Java, Python, Prometheus, Grafana, Elasticsearch, Kibana, FluentD, Azure, Jenkins, Git, TeamCity, JSON, REST, XML, Bash, PowerShell, Groovy, Ruby, YAML, Linux, UNIX, Microsoft Azure DevOps
This job has expired