Worth AI
Posted 2w ago

Senior DevOps Engineer, Infrastructure & Reliability

Worth AI
Atlanta or United States
RemoteFull Time
Responsibilities
  • writing terraform
  • tuning kubernetes
  • automating infrastructure
Requirements
  • 8+ years in DevOps/SRE or infrastructure engineering with Kubernetes, AWS
  • Terraform
  • CI/CD ownership
  • Incident response experience, and strong distributed systems knowledge
Technical tools mentioned
AWSEKSRDSMSKS3LambdaIAMVPCKubernetesArgoCDTerraformGitHub ActionsDataDogPostgreSQLKafkaRedisBashPythonTypeScriptJavaScript

Job description

Worth AI, a leader in the computer software industry, is looking for a Senior DevOps Engineer to join our Infrastructure team with a singular mission: to make our systems faster, more reliable, and more resilient while making life dramatically easier for engineers shipping software. 

This is a hands-on build role. You will spend most of your time writing Terraform, tuning Kubernetes workloads, automating things that are currently manual, and shipping infrastructure changes to production. You'll join a small platform team with an established roadmap and existing patterns, and a strong voice in how the work gets built.

  • Implement scalable Infrastructure-as-Code patterns using tools like Terraform to standardize cloud provisioning and reduce configuration drift.
  • Own and evolve our Kubernetes platform (EKS or self-managed), ensuring workloads are secure, scalable, and resilient by default.
  • Optimize CI/CD pipelines to improve deployment frequency, reduce lead time, and increase confidence in releases.
  • Design and enforce secure networking, IAM, and secrets management strategies across environments.
  • Improve observability by refining metrics, logs, and tracing using tools like DataDog, ensuring actionable insight into system health.
  • Optimize cloud cost efficiency through rightsizing, autoscaling strategies, and architectural improvements.
  • Implement disaster recovery planning, backup strategies, and multi-region resilience initiatives.
  • Refactor brittle or manually managed infrastructure into automated, testable, and reproducible systems.
  • Introduce new infrastructure tooling or architectural shifts and drive adoption through documentation, workshops, and hands-on support.
  • Partner with engineering teams to eliminate friction in CI/CD, deployments, and cloud environments.
  • Communicate technical trade-offs clearly across engineering and product stakeholders, balancing speed with safety.

Technology Stack

  • Cloud & Infrastructure: AWS (EKS, RDS, MSK, S3, Lambda, IAM, VPC)
    Containerization & Orchestration: Kubernetes, ArgoCD
    Infrastructure-as-Code: Terraform
    CI/CD: GitHub Actions
    Monitoring & Observability: DataDog
    Data & Messaging: PostgreSQL, Kafka, Redis
    Languages (as needed): Bash, Python, TypeScript, JavaScript

Requirements

  • 8+ years in DevOps, SRE, or infrastructure engineering.
  • Proven experience designing and operating production Kubernetes environments at scale.
  • Deep hands-on expertise with AWS infrastructure and cloud networking.
  • Strong experience building and maintaining Terraform modules across large cloud environments.
  • Demonstrated ownership of CI/CD systems and measurable improvement of DORA metrics.
  • Experience leading incident response processes and driving meaningful postmortem outcomes.
  • Strong understanding of distributed systems, event-driven architectures (Kafka), and database performance (PostgreSQL).
  • Proven ability to modernize legacy infrastructure and eliminate manual operational toil.
  • Track record of taking a scoped infrastructure project from an ambiguous starting point to production without needing daily direction.
  • Demonstrated ability to build trust across teams while raising the reliability bar.

Success Metrics

  • System Reliability: Maintain or exceed defined SLO/SLA targets with reduced incident frequency and duration.
  • Infrastructure Stability: Reduce production incidents caused by misconfiguration, manual processes, or infrastructure drift.
  • Operational Efficiency: Increase the percentage of infrastructure managed through code and automation.
  • Cost Optimization: Improve cloud cost efficiency without sacrificing reliability or performance.

Bonus Points (Nice to Have)

  • Experience coding applications
  • Experience operating high-throughput Kafka clusters (MSK or self-managed).
  • Strong background in database performance tuning (PostgreSQL, Redis).
  • Experience implementing autoscaling strategies for high-traffic systems.
  • Familiarity with service mesh technologies.
  • Experience building internal developer platforms (IDP).
  • Background in security best practices (zero-trust networking, policy-as-code).
  • Experience with multi-region or globally distributed systems.
  • Experience introducing platform-wide reliability frameworks (SLOs, error budgets, chaos testing).

All Remote Hires will be required to travel to Orlando, Florida at least twice per year for Town Halls and team collaboration, in addition to orientation in Orlando.

Benefits

  • Health Care Plan (Medical, Dental & Vision)
  • Retirement Plan (401k)
  • Life Insurance
  • Flexible Paid Time Off
  • 9 paid Holidays
  • Family Leave
  • Remote
  • Hybrid work (for Orlando Associates)
  • Free Food & Snacks (Orlando)
  • Wellness Resources

About Worth AI

Automating business underwriting and onboarding for financial institutions.

Year founded
2023
Employees
62
Industries
Organization type
Private
Latest investment
Raised $25.00M Seed (2025) — led by TTV Capital
Headquarters
US

Similar jobs

DevOps Engineer roles near Atlanta, Georgia
21h
Save
Mark Applied
Hide
Senior DevOps Developer (Cloud Automation)
Suwanee, Georgia, United States
OnsiteFull Time
Mujin
Mujin: Develops intelligent robot controllers and software for industrial automation.
5+ YOEBachelor's or master's degree or equivalent experience, plus 5+ years in cloud platforms, Kubernetes, Docker, CI/CD, Terraform, ArgoCD, Ansible, Linux, and networking.
AWS, Google Cloud Platform (GCP), Microsoft Azure, Kubernetes, Go, Docker, ArgoCD, Terraform, Ansible, Linux, GitOps, CI/CD, HSA, HDHP, FSA
1d
Save
Mark Applied
Hide
Senior DevOps Engineer
Philadelphia or Boston or New York City or Baltimore or Washington or Charlotte or Raleigh-Durham or Atlanta or Chicago or Connecticut or Delaware or Florida or Georgia or Illinois or Indiana or Massachusetts or Maryland or Michigan or North Carolina or New Jersey or New York or Ohio or Pennsylvania or Tennessee or Virginia
$100k-$145k/yr HybridFull Time
HealthVerity
HealthVerity: Privacy-protected healthcare data exchange and patient identity platform.
Requires deep distributed systems and cloud architecture knowledge, strong AWS experience, scalable IaC, security expertise, serverless and container orchestration experience, production debugging, and cloud cost optimization.
Infrastructure-as-Code (IaC), AWS, Docker Compose, Kubernetes
2d
Save
Mark Applied
Hide
DevOps Engineer
Atlanta, Georgia, United States
HybridFull Time
CapTech
CapTech: Provides technology and management consulting services to large enterprises.
Requires source control, programming, scripting, CI, infrastructure automation, Agile, and distributed-systems troubleshooting experience. Strong communication and collaboration skills required.
Subversion, Git, Java, .NET, ANT, Artifactory, Groovy, Maven, MSBuild, Nexus, NuGet, Jenkins, TFS, TeamCity, Bamboo, Terraform, CloudFormation, Ansible, Chef, Puppet, Kubernetes, Amazon ECS, Amazon Web Services, Azure, OpenShift
2d
Save
Mark Applied
Hide
DevOps Developer
Atlanta, Georgia, United States
HybridFull Time
Euna Solutions
Euna Solutions: Provides purpose-built cloud software for public sector organizations.
3+ YOE3+ years in DevOps, cloud operations, or infrastructure engineering; production ownership; Azure, GitHub Actions, CI/CD, IaC, communication, and AI-generated code validation experience.
Microsoft Azure, Microsoft Azure App Service, Microsoft Azure SQL, Microsoft Azure Storage, Microsoft Azure Virtual Networks, Microsoft Azure Monitor, GitHub, GitHub Actions, ARM Templates, FTP, SFTP, PowerShell, Python, Azure DevOps, Wiz, Claude, MCP
3d
Save
Mark Applied
Hide
Sr. DevOps Engineer
Atlanta, Georgia, United States
$134k-$224k/yr RemoteFull Time
Omnissa
Omnissa: Provides AI-driven digital workspace and endpoint management software solutions.
6+ YOEUS citizenship and 6+ years of DevOps or cloud engineering experience. Requires Python, PowerShell, Bash, Terraform or Ansible, cloud platforms, containers, cloud security, and federal security framework knowledge.
Terraform, Ansible, Python, PowerShell, Bash, AWS, CI/CD, FedRAMP, NIST
2w
Save
Mark Applied
Hide
DevOps Engineer
Oakland or Alpharetta
$93k-$202k/yr HybridFull Time
Delta Dental of California
Delta Dental of California: Provides dental insurance plans and oral health benefit coverage.
4+ YOERequires 4+ years of experience, a bachelor's degree or equivalent experience, cloud infrastructure expertise, GitHub Actions, Terraform, scripting, Docker, CI/CD, and strong troubleshooting skills.
Microsoft Azure, Amazon Web Services, Google Cloud Platform, Oracle Cloud, GitHub Actions, YAML, Terraform, Shell, Python, Docker, WebLogic, ReactJS, Spring Boot, Java, Node.js, Apache HTTP Server, GitHub Enterprise Cloud, AKS, App Service, ACE, APIM, AFD, ADF, Power Apps, App Gateway, Jenkins, CI/CD
2w
Save
Mark Applied
Hide
DevOps Engineer with Java
Atlanta, Georgia, United States
OnsiteContract
Delan Associates
Delan Associates: Provides engineering and professional services to government agencies.
Requires Java, Spring, Hibernate, AWS, CI/CD, Terraform, Docker, Kubernetes, scripting, monitoring, production troubleshooting, Agile, DevOps, and mandatory banking experience.
Java, Spring, Hibernate, Amazon Web Services (AWS), Amazon EC2, Amazon S3, Amazon RDS, AWS Lambda, AWS CloudFormation, Amazon CloudWatch, Jenkins, GitLab CI/CD, AWS CodePipeline, Terraform, Docker, Kubernetes, Python, Bash, PowerShell, ELK Stack
2w
Save
Mark Applied
Hide
Azure DevOps Engineer
Alpharetta, Georgia, United States
HybridFull Time
Cognizant
CognizantNASDAQ: CTSH: Provides IT consulting and technology services to global enterprises.
10+ YOE10+ years in DevOps/Release Engineering with Azure DevOps, CI/CD, Terraform/Bicep/ARM, Kubernetes/AKS, ArgoCD/GitOps, scripting (PowerShell/Bash/Python), and strong troubleshooting skills.
Azure DevOps, Terraform, Bicep, PowerShell, ARM templates, ArgoCD, Kubernetes, AKS, Helm, Docker, ACR, Git, Python, Bash, YAML