35 cloud reliability engineer jobs at 29 companies in Georgia

1d
Save
Mark Applied
Hide
Sr. Cloud Operations Reliability Engineer (SRE)
Georgia, United States
RemoteFull Time
NextGen Healthcare
NextGen Healthcare: Provides cloud-based electronic health record and practice management software.
10+ YOE10+ years in cloud operations/SRE with GCP/AWS, SLO/SLI, incident response, IaC, Kubernetes, observability, and mentoring experience.
Google Cloud Platform (GCP), AWS, Terraform, Deployment Manager, CloudFormation, Kubernetes, Grafana, Prometheus, Cloud Monitoring, Datadog, New Relic, Python, Bash, Go, CI/CD
3mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE) - AI Platform & Cloud
Alpharetta, Georgia, United States
OnsiteFull Time
Morgan Stanley
Morgan StanleyNYSE: MS: Global financial services firm providing investment and wealth management.
5+ YOESenior SRE with 5+ years production experience; programming in Python/Go/Java; Kubernetes, Docker, cloud (AWS/Azure/Google), IaC (Terraform/Helm/CloudFormation/Ansible); monitoring (Prometheus/Grafana/ELK/Datadog); networking and GPU/AI compute experience.
Kubernetes, AWS, Azure, Google, API, REST, Python, Go, Java, Docker, Terraform, Helm, CloudFormation, Ansible, Prometheus, Grafana, ELK, EFK, Datadog, Open Telemetry, Loki, Cortex, Kafka, Spark, Flink, SQL, Redis, Snowflake, Slurm, ModelOps, ML Ops, LLM Op
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Atlanta, Georgia, United States
$110k-$209k/yr RemoteFull Time
AbbVie
AbbVieNYSE: ABBV: Develops and sells innovative pharmaceutical and biopharmaceutical medicines.
7+ YOE7+ years in site reliability engineering / information security, cloud platforms (AWS/GCP/Azure), CI/CD, Kubernetes, GitOps, Linux/Windows administration; strong security practices.
AWS, GCP, Azure, Kubernetes, Docker, GitOps, Terraform, Crossplane, ArgoCD, Helm, Prometheus, Grafana, OpenTelemetry, Jenkins, GitHub Actions, Azure DevOps, Python, Go, NoSQL, Relational Databases
2w
Save
Mark Applied
Hide
Principal Site Reliability Engineer, Google Cloud
Atlanta or Milpitas
$240k-$250k/yr HybridFull Time
Saviynt
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
Go (Golang), Python, Kubernetes, GCP, AWS, Azure, Kafka, RMQ, NATS, Google Pub/Sub, GitLab CI, ArgoCD, Prometheus, Grafana, ELK stack, Datadog, Envoy, Istio, MySQL, PostgresSQL
3mo
Save
Mark Applied
Hide
Site Reliability Engineer
St. Louis or Alpharetta
HybridFull Time
Equifax
EquifaxNew York Stock Exchange: EFX: Provides consumer credit reporting and data analytics services globally.
5+ YOEBS in CS or related technical field; 5-7 years in software engineering, systems, database or networking; 2+ years in public cloud; Python/Bash/Java/Go/JavaScript; Terraform and CI/CD; on-call; strong problem solving.
Terraform, AWS, GCP, Jenkins, Docker, Kubernetes, Ansible, Chef, Python, Bash, Java, Go, JavaScript, Node.js
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Alpharetta, Georgia, United States
OnsiteFull Time
Morgan Stanley
Morgan StanleyNYSE: MS: Provides global investment banking, wealth management, and advisory services.
5+ YOE5+ years production experience; strong scripting (Python, Perl, Shell, Ruby, Java, C#); DB2/Sybase/Oracle, Autosys, CI/CD, containers/VMs, Splunk/IP Soft/Sockeye, Jenkins/Train; cloud (Azure/AWS); BS in CS/Engineering required.
Python, Perl, Shell, Ruby, Java, C#, DB2, Sybase, Oracle, Autosys, Splunk, IP Soft, Sockeye, Jenkins, Train, Azure, AWS, MQ, UNIX, Linux, Windows
2mo
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta, Georgia, United States
$117k-$209k/yr OnsiteFull Time
Autodesk
AutodeskNASDAQ: ADSK: Developing software for architecture, engineering, and entertainment industries.
5+ YOE5+ years DevOps/SRE with cloud apps; Linux admin; AWS; scripting; IaC; CI/CD; monitoring; on-call; U.S. citizenship or permanent residency.
Docker, Kubernetes, Terraform, CloudFormation, Jenkins, Git, AWS, CI/CD, Splunk, Dynatrace, New Relic, Grafana
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Alpharetta or United States
$129k-$161k/yr RemoteFull Time
Priority Technology Holdings
Priority Technology HoldingsNASDAQ: PRTH: Provides integrated payment processing and banking-as-a-service solutions.
5+ YOE5+ years software/systems engineering (including 3+ years SRE), expertise in distributed systems, cloud (AWS preferred), CI/CD, observability, incident management, Java/Node.js/JavaScript, and database experience.
AWS, CI/CD, Java, Node.js, JavaScript
3w
Save
Mark Applied
Hide
Customer Reliability Engineer, Airflow
United States or San Francisco or Boston or Atlanta or Austin or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
4+ YOEData engineering background, 4 years Python, 1 year Airflow administration/DAG creation, Kubernetes/Docker experience, cloud provider (AWS/GCP/Azure) experience, troubleshooting, strong communication, and mentoring experience.
Apache Airflow, Python, Kubernetes, Docker, AWS, GCP, Azure, SQL, PostgreSQL, Databricks, Snowflake, Redshift, dbt, Zoom
1mo
Save
Mark Applied
Hide
GOV Site Reliability Engineer
United States or Kansas or Washington or California or Texas or Illinois or North Carolina or Colorado or Massachusetts or Pennsylvania or Virginia or Oregon or Nevada or Hawaii or New York or Georgia or Ohio or Arizona or Seattle or San Francisco or New York City
$110k-$183k/yr RemoteFull Time
Veeam
Veeam: Data resilience and security for hybrid cloud environments
3+ YOE3+ years in software engineering with 1+ year in SRE/Platform/DevOps, cloud experience (Azure or comparable), observability (Prometheus, Grafana, OpenTelemetry, ELK), IaC (Terraform/Terragrunt/Pulumi), Kubernetes, CI/CD tooling, programming in TypeScript/JS, Go, Java, or C#, and experience in compliance-oriented environments.
VDC, Prometheus, Grafana, OpenTelemetry, ELK stack, Terraform, Terragrunt, Pulumi, Kubernetes, GitHub Actions, Azure DevOps, GitLab CI, ArgoCD, TypeScript, JS, Go, Java, C#, Azure Government, AWS GovCloud
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Atlanta, Georgia, United States
OnsiteFull Time
Florence Healthcare
Florence Healthcare: Digital platform for managing clinical trial documents and workflows.
4+ YOESRE with 4+ years in SRE/DevOps, cloud-native architectures, Linux, observability, IaC (Terraform), AI-assisted tooling; AWS experience; strong collaboration.
AWS, Terraform, CI/CD, Linux, Observability, AI-assisted tooling
2w
Save
Mark Applied
Hide
Sr. Site Reliability Engineer
Berkeley Heights or Alpharetta or Sunnyvale
$128k-$216k/yr OnsiteFull Time
Fiserv
FiservNew York Stock Exchange: FI: Provides financial technology and payment processing services to institutions.
5+ YOE5+ years production experience with AWS, Kubernetes, and Linux; strong Terraform, CI/CD (GitHub Actions), Docker, GitHub, RDBMS/Document storage, and scripting (Python/Bash/Node/Ruby); experience designing scalable cloud systems.
Amazon Web Services, Kubernetes, GitHub Actions, Terraform, New Relic, Dynatrace, Datadog, Docker, GitHub, Python, Bash, Node, Ruby on Rails
19h
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta, Georgia, United States
$152k-$162k/yr OnsiteFull Time
Unum Group
Unum GroupNYSE: UNM: Provides insurance and employee benefits to businesses and individuals.
5+ YOEBachelor's in Computer Science/Engineering plus 5+ years experience; expertise with observability, cloud-native architectures, incident response, automation scripting, CI/CD, infrastructure-as-code, and collaboration in DevOps environments.
Dynatrace, AWS CloudWatch, Datadog, Grafana, Amplitude, AWS, Python, Bash, PowerShell, GitHub, GitLab, Bitbucket, GitHub Actions, Jenkins, Azure DevOps, Terraform, AWS CloudFormation, Ansible
5d
Save
Mark Applied
Hide
Site Reliability Engineer, Linux (Remote)
Austin or North Dakota or Montana or Maine or New Mexico or New Hampshire or Kentucky or Alabama or Ohio or Nebraska or Louisiana or South Carolina or Illinois or Texas or Nevada or Hawaii or Georgia or Missouri or Iowa or Mississippi or Tennessee or Colorado or North Carolina or Minnesota or Kansas or Wyoming or Wisconsin or West Virginia or Rhode Island or Delaware or Vermont or Arkansas or Utah or Oregon or South Dakota or Florida or Pennsylvania or Michigan or Indiana or Idaho or Arizona or Oklahoma
$127k-$182k/yr RemoteFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
5+ YOESTEM degree or equivalent experience,5+ years Linux production experience,Python or Ruby,Ansible,system debugging,cloud/on-prem experience;Kubernetes and FedRAMP knowledge preferred;must be a U.S. Person.
Python, Ruby, Ansible, Kubernetes, Docker, Go, C
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer II
Louisville or Atlanta or Lehi
OnsiteFull Time
Waystar
WaystarNASDAQ: WAY: Provides cloud-based healthcare payment and revenue cycle management software.
7+ YOE7+ years in SRE/DevOps or infrastructure engineering; cloud platforms (AWS, GCP, Azure); Kubernetes; IaC (Terraform, CloudFormation); observability tools (Prometheus, Grafana, Splunk); CI/CD; data platforms/ETL; Python and PowerShell; AI tooling experience.
AWS, GCP, Azure, Kubernetes, Infrastructure as Code, Terraform, CloudFormation, Prometheus, Grafana, Splunk, CI/CD, Python, PowerShell
1d
Save
Mark Applied
Hide
Senior Site Reliability Engineer – Unified Observability
Atlanta, Georgia, United States
OnsiteFull Time
NCR Voyix
NCR VoyixNYSE: VYX: Provides checkout software and kiosks for retailers and restaurants.
10+ YOE10+ years SRE/Platform/Cloud experience; expertise with Azure, GCP, Kubernetes (AKS,GKE); observability tools (Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic); Terraform; Python/Go/PowerShell; bachelor\u0002s degree or equivalent.
Azure, Google Cloud Platform, Kubernetes, AKS, GKE, Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic, ServiceNow, Terraform, Python, Go, PowerShell
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (AWS)
Jacksonville or Atlanta or Milwaukee
HybridFull Time
FIS
FISNYSE: FIS: Provides technology solutions for merchants, banks, and capital markets
7+ YOE7+ years SRE/Cloud Engineering experience with hands-on AWS, Linux, CI/CD (Jenkins/Harness), Docker/Kubernetes (EKS), Terraform, scripting (Python/Bash/Shell), monitoring tools, on-call experience, and required AWS certifications.
AWS, EC2, EKS, RDS, S3, KMS, Secrets Manager, IAM, Route53, Security Groups, Linux, Git, Docker, Kubernetes, OpenShift, Helm, Jenkins, Harness, Terraform, Python, Bash, Shell, Dynatrace, CloudWatch, Splunk, Prometheus, Grafana, CheckMarx, SonarQube, Maven, Node, Artifactory, FlyWay, KeyFactor, HashiCorp Vault, CyberArk, SNOW, Jira, Confluence, Oracle DB, PostgreSQL DB, Postgres DB, Redis, SFTP, Tivoli
4w
Save
Mark Applied
Hide
Software Reliability Engineer
Atlanta, Georgia, United States
$84k-$151k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
2+ YOEBachelor's degree (or equivalent) plus experience, DevOps/SRE experience with CI/CD, cloud-native platforms, containerization, automation, monitoring and incident troubleshooting. Familiarity with languages (C, C#, Java, Perl, Python, Go), CI/CD and DevOps tools is preferred.
C, C#, Java, Perl, Python, Go, Shell, Jenkins, Cloudbees, Ansible, Chef, Puppet, Docker, Kubernetes, AppDynamics, Splunk, Dynatrace, Grafana, Prometheus, Terraform, CI/CD, APM, DevOps
1w
Save
Mark Applied
Hide
Platform Delivery & Reliability Engineer (Remote) - 29337
Colorado Springs or Columbia or San Antonio or Boise or Greenville or Augusta
$120k-$195k/yr RemoteFull Time
HII
HIINYSE: HII: Builds naval ships and provides global defense technology solutions.
7+ YOEMust obtain U.S. security clearance; 7+ years' relevant experience (varies by degree); deep Kubernetes, IaC, cloud, CI/CD, scripting, SRE practices, troubleshooting and delivery experience across distributed systems.
Kubernetes, Terraform, AWS, Azure, GCP, GitLab CI, Go, Python, Bash, Spark, Trino, Presto, Kafka, NiFi, Iceberg, Delta, Hudi, YouTrack, Nexus
2mo
Save
Mark Applied
Hide
Site Reliability Engineering Lead
Florida or Chicago or Boca Raton or Alpharetta
$118k-$220k/yr RemoteFull Time
LexisNexis Risk Solutions
LexisNexis Risk SolutionsNYSE: RELX: Provides data and analytics for risk management and compliance.
Lead SRE teams; implement infrastructure as code and DevOps practices; manage production reliability; cloud (AWS/Azure); Kubernetes and Docker; security tooling; incident management; FinOps cost optimization; collaboration with cross-functional teams.
Amazon Web Services, Microsoft Azure, Kubernetes, Docker, GitHub Advanced Security, Qualys, Wiz, Trufflehog

Explore Jobs

Expand Your Job Search