18 cloud reliability engineer jobs at 18 companies in Roswell, GA

1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer, Google Cloud
Atlanta or Milpitas
$240k-$250k/yr HybridFull Time
Saviynt
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
Go (Golang), Python, Kubernetes, GCP, AWS, Azure, Kafka, RMQ, NATS, Google Pub/Sub, GitLab CI, ArgoCD, Prometheus, Grafana, ELK stack, Datadog, Envoy, Istio, MySQL, PostgresSQL
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Alpharetta, Georgia, United States
OnsiteFull Time
Morgan Stanley
Morgan StanleyNYSE: MS: Provides global investment banking, wealth management, and advisory services.
5+ YOE5+ years production experience; strong scripting (Python, Perl, Shell, Ruby, Java, C#); DB2/Sybase/Oracle, Autosys, CI/CD, containers/VMs, Splunk/IP Soft/Sockeye, Jenkins/Train; cloud (Azure/AWS); BS in CS/Engineering required.
Python, Perl, Shell, Ruby, Java, C#, DB2, Sybase, Oracle, Autosys, Splunk, IP Soft, Sockeye, Jenkins, Train, Azure, AWS, MQ, UNIX, Linux, Windows
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Alpharetta or United States
$129k-$161k/yr RemoteFull Time
Priority Technology Holdings
Priority Technology HoldingsNASDAQ: PRTH: Provides integrated payment processing and banking-as-a-service solutions.
5+ YOE5+ years software/systems engineering (including 3+ years SRE), expertise in distributed systems, cloud (AWS preferred), CI/CD, observability, incident management, Java/Node.js/JavaScript, and database experience.
AWS, CI/CD, Java, Node.js, JavaScript
5d
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta or Alpharetta
OnsiteFull Time
Incident IQ
Incident IQ: Workflow management software for K-12 school district operations.
5+ YOE5+ years SRE/DevOps experience, strong systems fundamentals, SLI/SLO and observability experience, incident management, automation and cloud skills, proficient with modern SRE tooling and AI-accelerated execution.
Grafana, PromQL, Grafana Alloy, Prometheus, Datadog, OpenTelemetry, SigNoz, Uptrace, Tempo, Grafana Faro, k6, PagerDuty, Locust, JMeter, Python, Go, Bash, Terraform, Ansible, Kubernetes, Amazon Web Services (AWS), Google Cloud Platform (GCP), Azure, GitOps, eBPF, Grafana Beyla, OpenTelemetry eBPF Instrumentation, .NET, Real User Monitoring (RUM)
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer I
Boston or Seattle or Atlanta
$134k-$215k/yr HybridFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
7+ YOEBachelor's in CS/Engineering, 7+ years software engineering experience, expertise in distributed systems, Kubernetes, cloud (Azure/AWS/GCP), observability, Kafka, Terraform/Pulumi, and experience with agentic AI/LLM tooling preferred.
Kubernetes, Terraform, Pulumi, Kafka, Grafana, Datadog, New Relic, MySQL, Cassandra, PostgreSQL, Azure, AWS, GCP
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Atlanta, Georgia, United States
OnsiteFull Time
Florence Healthcare
Florence Healthcare: Digital platform for managing clinical trial documents and workflows.
4+ YOESRE with 4+ years in SRE/DevOps, cloud-native architectures, Linux, observability, IaC (Terraform), AI-assisted tooling; AWS experience; strong collaboration.
AWS, Terraform, CI/CD, Linux, Observability, AI-assisted tooling
1mo
Save
Mark Applied
Hide
Sr. Site Reliability Engineer
Berkeley Heights or Alpharetta or Sunnyvale
$128k-$216k/yr OnsiteFull Time
Fiserv
FiservNew York Stock Exchange: FI: Provides financial technology and payment processing services to institutions.
5+ YOE5+ years production experience with AWS, Kubernetes, and Linux; strong Terraform, CI/CD (GitHub Actions), Docker, GitHub, RDBMS/Document storage, and scripting (Python/Bash/Node/Ruby); experience designing scalable cloud systems.
Amazon Web Services, Kubernetes, GitHub Actions, Terraform, New Relic, Dynatrace, Datadog, Docker, GitHub, Python, Bash, Node, Ruby on Rails
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta, Georgia, United States
$152k-$162k/yr OnsiteFull Time
Unum Group
Unum GroupNYSE: UNM: Provides insurance and employee benefits to businesses and individuals.
5+ YOEBachelor's in Computer Science/Engineering plus 5+ years experience; expertise with observability, cloud-native architectures, incident response, automation scripting, CI/CD, infrastructure-as-code, and collaboration in DevOps environments.
Dynatrace, AWS CloudWatch, Datadog, Grafana, Amplitude, AWS, Python, Bash, PowerShell, GitHub, GitLab, Bitbucket, GitHub Actions, Jenkins, Azure DevOps, Terraform, AWS CloudFormation, Ansible
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer – Unified Observability
Atlanta, Georgia, United States
OnsiteFull Time
NCR Voyix
NCR VoyixNYSE: VYX: Provides checkout software and kiosks for retailers and restaurants.
10+ YOE10+ years SRE/Platform/Cloud experience; expertise with Azure, GCP, Kubernetes (AKS,GKE); observability tools (Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic); Terraform; Python/Go/PowerShell; bachelor\u0002s degree or equivalent.
Azure, Google Cloud Platform, Kubernetes, AKS, GKE, Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic, ServiceNow, Terraform, Python, Go, PowerShell
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (AWS)
Jacksonville or Atlanta or Milwaukee
HybridFull Time
FIS
FISNYSE: FIS: Provides technology solutions for merchants, banks, and capital markets
7+ YOE7+ years SRE/Cloud Engineering experience with hands-on AWS, Linux, CI/CD (Jenkins/Harness), Docker/Kubernetes (EKS), Terraform, scripting (Python/Bash/Shell), monitoring tools, on-call experience, and required AWS certifications.
AWS, EC2, EKS, RDS, S3, KMS, Secrets Manager, IAM, Route53, Security Groups, Linux, Git, Docker, Kubernetes, OpenShift, Helm, Jenkins, Harness, Terraform, Python, Bash, Shell, Dynatrace, CloudWatch, Splunk, Prometheus, Grafana, CheckMarx, SonarQube, Maven, Node, Artifactory, FlyWay, KeyFactor, HashiCorp Vault, CyberArk, SNOW, Jira, Confluence, Oracle DB, PostgreSQL DB, Postgres DB, Redis, SFTP, Tivoli
1mo
Save
Mark Applied
Hide
Software Reliability Engineer
Atlanta, Georgia, United States
$84k-$151k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
2+ YOEBachelor's degree (or equivalent) plus experience, DevOps/SRE experience with CI/CD, cloud-native platforms, containerization, automation, monitoring and incident troubleshooting. Familiarity with languages (C, C#, Java, Perl, Python, Go), CI/CD and DevOps tools is preferred.
C, C#, Java, Perl, Python, Go, Shell, Jenkins, Cloudbees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Dynatrace, Grafana, Prometheus, Terraform, CI/CD, APM, DevOps
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior / Staff)
Atlanta, Georgia, United States
$130k-$165k/yr OnsiteFull Time
Satine Technologies
Satine Technologies: Provides cybersecurity, software engineering, and DevSecOps services.
5+ YOE5+ years SRE/DevOps experience; strong Kubernetes and Terraform; cloud fundamentals; Linux; US citizenship or LPR.
Kubernetes, Terraform, IaC, AWS, Azure, GCP, Linux, Prometheus, Grafana, ELK, GitOps, ArgoCD, Flux
2mo
Save
Mark Applied
Hide
Site Reliability Engineering Lead
Florida or Chicago or Boca Raton or Alpharetta
$118k-$220k/yr RemoteFull Time
LexisNexis Risk Solutions
LexisNexis Risk SolutionsNYSE: RELX: Provides data and analytics for risk management and compliance.
Lead SRE teams; implement infrastructure as code and DevOps practices; manage production reliability; cloud (AWS/Azure); Kubernetes and Docker; security tooling; incident management; FinOps cost optimization; collaboration with cross-functional teams.
Amazon Web Services, Microsoft Azure, Kubernetes, Docker, GitHub Advanced Security, Qualys, Wiz, Trufflehog
1mo
Save
Mark Applied
Hide
Lead Infrastructure Engineer
Charlotte or Atlanta
OnsiteFull Time
Truist
TruistNYSE: TFC: Offers personal banking, business lending, and investment management services.
10+ YOEBachelor's in CS/Engineering/IS, 10+ years infrastructure engineering experience with advanced knowledge of cloud, network, database, storage, platform, automation, and enterprise-scale reliability.
Python, FastAPI, Java, Kubernetes, OpenShift, Open Policy Agent (OPA), Rego, Apache Kafka, Confluent, GitLab CI/CD, Backstage, HashiCorp Vault, OpenTelemetry, Prometheus, Splunk, Jaeger, Tempo, AWS, Azure, Terraform, Ansible, Cosign, CycloneDX, GitLab Duo, GitHub Copilot
2mo
Save
Mark Applied
Hide
Site Reliability Engineering Lead
Chicago or Boca Raton or Alpharetta or Florida or Illinois
$118k-$220k/yr RemoteFull Time
RELX
RELXLondon Stock Exchange: REL: Provides information-based analytics and decision tools for professional customers.
Lead SRE teams; implement Infrastructure as Code; manage cloud environments (AWS/Azure); containerization (Docker, Kubernetes); CI/CD; security tooling; production support; FinOps cost optimization.
AWS, Azure, EC2, ECS, AKS, Docker, Kubernetes, Qualys, Wiz, Trufflehog, GitHub Advanced Security
1mo
Save
Mark Applied
Hide
Staff Platform Engineer (Fully Remote)
Atlanta or United States
$175k-$200k/yr RemoteFull Time
PadSplit
PadSplit: Operates a marketplace for affordable-living and room rentals.
Extensive experience with distributed systems, cloud infrastructure (AWS), backend architecture, Django, and Postgres; strong systems design, reliability, and mentoring skills required.
Django, AWS, Postgres
1w
Save
Mark Applied
Hide
Principal Engineer, Full Stack
Jacksonville or Atlanta
OnsiteFull Time
Intercontinental Exchange
Intercontinental ExchangeNYSE: ICE: Operates global financial exchanges, clearing houses, and mortgage technology.
10+ YOE10+ years software or site reliability engineering experience; strong Java and React (TypeScript) skills; Kubernetes, ArgoCD, CI/CD, observability, and cloud-native experience; Bachelor's degree or equivalent.
Spring, React, TypeScript, Kubernetes, ArgoCD, Prometheus, Grafana, Jaeger, OpenTelemetry (OTEL), Istio, Kiali, Crossplane, Java, Postgres, PL/SQL, Microsoft Azure DevOps, Microsoft TFS, Jira, Git, webpack, Claude Code, Codex, OpenCode
2mo
Save
Mark Applied
Hide
Slack Proactive Monitoring Engineer
Indianapolis or Chicago or Atlanta or Seattle
$75k-$114k/yr HybridFull Time
Salesforce
SalesforceNYSE: CRM: Sells cloud-based customer relationship management and business software solutions.
2+ YOE2+ years in technical support, site reliability, or related operations; strong observability skills; cloud SaaS knowledge; SQL/log analysis; excellent communication; AI-assisted workflows experience.
Grafana, Splunk, Datadog, PagerDuty, Slack API, Slack Workflows, Bolt framework, Python, JavaScript, Bash

Explore Jobs

Expand Your Job Search