108 site reliability jobs at 85 companies in Georgia

6d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year; automation, monitoring, CI/CD, cloud computing, scripting, incident management, and system reliability skills required.
CI/CD, Kubernetes, AWS
6d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year. Requires automation, monitoring, scripting, cloud computing, CI/CD, incident management, and system reliability skills.
CI/CD, Kubernetes, AWS
2mo
Save
Mark Applied
Hide
VP of Site Reliability
Atlanta, Georgia, United States
RemoteFull Time
Titan
Titan: AI platform providing secure models and agents for banking.
10+ YOE10+ years in engineering with 5+ years building SRE/platform operations for enterprise or regulated markets; hands-on leader; incident response and on-call expertise.
1w
Save
Mark Applied
Hide
Site Reliability Engineer
San Mateo or Arizona or California or Colorado or Florida or Georgia or Illinois or Nevada or North Carolina or Oregon or Texas or Utah or Washington
$140k-$150k/yr RemoteFull Time
VyncaCare
VyncaCare: Offers palliative care services and advance care planning technology.
3+ YOE3+ years SRE/DevOps experience, strong AWS and Terraform skills, Kubernetes and Helm experience, observability and incident response knowledge, bachelor's or equivalent, on-call participation, East Coast hours.
AWS, Terraform, Kubernetes, Helm, Prometheus, Grafana, Datadog, CloudWatch, SigNoz, OpenTelemetry, ArgoCD, Flux, PostgreSQL, MySQL, Redshift, ClickHouse, AWS Secrets Manager, HashiCorp Vault, Snowflake, Python, Go, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Alpharetta, Georgia, United States
OnsiteFull Time
Morgan Stanley
Morgan StanleyNYSE: MS: Global financial services firm providing investment and wealth management.
5+ YOEMinimum 5 years production experience; strong scripting (Python, Perl, Shell, Ruby, Java, C#); DB2/Sybase/Oracle, Autosys, Jenkins/Train, Splunk/IP Soft/Sockeye; cloud (Azure/AWS); BS in CS/Engineering required.
Python, Perl, Shell, Ruby, Java, C#, DB2, Sybase, Oracle, Autosys, Jenkins, Train, Splunk, IP Soft, Sockeye, Azure, AWS, UNIX, Linux, Windows
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Alpharetta, Georgia, United States
OnsiteFull Time
Morgan Stanley
Morgan StanleyNYSE: MS: Provides global investment banking, wealth management, and advisory services.
5+ YOE5+ years production experience; strong scripting (Python, Perl, Shell, Ruby, Java, C#); DB2/Sybase/Oracle, Autosys, CI/CD, containers/VMs, Splunk/IP Soft/Sockeye, Jenkins/Train; cloud (Azure/AWS); BS in CS/Engineering required.
Python, Perl, Shell, Ruby, Java, C#, DB2, Sybase, Oracle, Autosys, Splunk, IP Soft, Sockeye, Jenkins, Train, Azure, AWS, MQ, UNIX, Linux, Windows
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta, Georgia, United States
$152k-$162k/yr OnsiteFull Time
Unum Group
Unum GroupNYSE: UNM: Provides insurance and employee benefits to businesses and individuals.
5+ YOEBachelor's in Computer Science/Engineering plus 5+ years experience; expertise with observability, cloud-native architectures, incident response, automation scripting, CI/CD, infrastructure-as-code, and collaboration in DevOps environments.
Dynatrace, AWS CloudWatch, Datadog, Grafana, Amplitude, AWS, Python, Bash, PowerShell, GitHub, GitLab, Bitbucket, GitHub Actions, Jenkins, Azure DevOps, Terraform, AWS CloudFormation, Ansible
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta or Alpharetta
OnsiteFull Time
Incident IQ
Incident IQ: Workflow management software for K-12 school district operations.
5+ YOE5+ years SRE/DevOps experience, strong systems fundamentals, SLI/SLO and observability experience, incident management, automation and cloud skills, proficient with modern SRE tooling and AI-accelerated execution.
Grafana, PromQL, Grafana Alloy, Prometheus, Datadog, OpenTelemetry, SigNoz, Uptrace, Tempo, Grafana Faro, k6, PagerDuty, Locust, JMeter, Python, Go, Bash, Terraform, Ansible, Kubernetes, Amazon Web Services (AWS), Google Cloud Platform (GCP), Azure, GitOps, eBPF, Grafana Beyla, OpenTelemetry eBPF Instrumentation, .NET, Real User Monitoring (RUM)
2w
Save
Mark Applied
Hide
Site Reliability Engineer
United States or Atlanta or Chicago
$100k-$120k/yr HybridFull Time
Origami Risk
Origami Risk: Cloud-based risk management and insurance software solutions.
5+ YOE5+ years SRE experience, strong incident management, observability tooling (New Relic, Data Dog, SumoLogic), cloud (AWS/Azure), coding (JavaScript,.NET,C#,+SQL), CI/CD and IaC familiarity, strong communication and problem-solving.
New Relic, Data Dog, SumoLogic, JavaScript, .NET, C#, SQL, SQL Server, AWS, Azure, Windows, CI/CD, Infrastructure as Code (IaC)
1d
Save
Mark Applied
Hide
Site Reliability Engineering Lead
Atlanta or Raleigh or Charlotte
OnsiteFull Time
Truist
TruistNYSE: TFC: Offers personal banking, business lending, and investment management services.
7+ YOEBachelor's degree and 7+ years of software development experience, with expertise in distributed systems, Kubernetes, automation scripting, incident management, cloud-native architecture, and reliability engineering.
Kubernetes, Python, Go, PowerShell, Ansible, Splunk, Dynatrace, Linux/Unix, AI, AIOps, CI/CD
1w
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Atlanta or United States
$120k-$175k/yr RemoteFull Time
PrizePicks
PrizePicks: Operates a daily fantasy sports and player prediction platform.
5+ YOE5+ years reliability engineering experience; cloud (AWS/Azure/GCP), IaC (Terraform/Crossplane), Kubernetes, Python/Ruby/Go, monitoring tools, incident response, SLO governance, strong cross-functional and debugging skills.
AWS, Azure, GCP, Terraform, Crossplane, Python, Ruby, Go, Kubernetes, Grafana, New Relic, Datadog, Windows, Mac
2w
Save
Mark Applied
Hide
Site Reliability Engineer II
Georgia or California or Pittsburgh or Pennsylvania
$72k-$119k/yr RemoteFull Time
LexisNexis Risk Solutions
LexisNexis Risk SolutionsNYSE: RELX: Provides data and analytics for risk management and compliance.
1+ YOE1+ years experience in DevOps/SRE/cloud engineering, Azure preferred; scripting (PowerShell,Bash,Python); IaC (Terraform,ARM); Git; networking basics; Bachelor's preferred.
Azure, AWS, GCP, PowerShell, Bash, Python, Terraform, ARM templates, Git, GitHub
2w
Save
Mark Applied
Hide
Site Reliability Engineer II
Georgia or California or Pittsburgh or Pennsylvania
$72k-$119k/yr RemoteFull Time
RELX
RELXLondon Stock Exchange: REL: Provides information-based analytics and decision tools for professional customers.
1+ YOE1–3 years in DevOps/SRE/cloud roles; familiarity with Azure (preferred), networking basics, scripting (PowerShell/Bash/Python), IaC (Terraform/ARM), Git; bachelor's preferred.
Azure, AWS, GCP, PowerShell, Bash, Python, Terraform, ARM templates, Git, GitHub
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Alpharetta or United States
$129k-$161k/yr RemoteFull Time
Priority Technology Holdings
Priority Technology HoldingsNASDAQ: PRTH: Provides integrated payment processing and banking-as-a-service solutions.
5+ YOE5+ years software/systems engineering (including 3+ years SRE), expertise in distributed systems, cloud (AWS preferred), CI/CD, observability, incident management, Java/Node.js/JavaScript, and database experience.
AWS, CI/CD, Java, Node.js, JavaScript
1mo
Save
Mark Applied
Hide
Sr Site Reliability Engineer
Atlanta, Georgia, United States
$178k-$205k/yr HybridFull Time
Workday
Workday: A provider of cloud-based enterprise software and services focused on managing people and finances.
5+ YOEBachelor's degree plus 5 years experience. 5+ years with Ansible, Terraform, Packer, Python, Shell Scripting, Kafka, CI/CD tools (GIT, Maven/Gradle, Jenkins), Docker, Kubernetes, and cloud provisioning; strong monitoring and automation skills.
Ansible, Terraform, Packer, Python, Shell Scripting, Kafka, GIT, Maven, Gradle, Jenkins, Docker, Kubernetes
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Atlanta or Uniontown
HybridFull Time
QGenda
QGenda: Provides healthcare workforce management and physician scheduling software.
7+ YOE7+ years SRE/DevOps experience, BS in CS or equivalent, advanced scripting skills, strong AWS (Lambda, EC2, ECS/EKS, S3, SNS, SQS, RDS, Redshift, ElastiCache) and Terraform experience, Docker, Kubernetes, CI/CD, Git, observability tools, and on-call availability.
AWS Lambda, Amazon EC2, AWS ECS, AWS EKS, Kubernetes, AWS S3, AWS SNS, AWS SQS, AWS RDS, Amazon Redshift, AWS ElastiCache, Docker, Terraform, Git, AWS CodeBuild, Jenkins, TeamCity, Datadog, Amazon CloudWatch, PagerDuty, Claude, GitHub Copilot
1w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Austin or Atlanta
$100k-$115k/yr OnsiteFull Time
Atlanticus
AtlanticusNASDAQ: ATLC: Provides credit cards and lending solutions for underserved consumers.
5+ YOERequires 5+ years supporting production applications, Java, AWS, Kubernetes, Docker, Datadog or Splunk, CI/CD, Python or Bash, Linux, cloud troubleshooting, and incident management experience.
AWS, Amazon EKS, Amazon EC2, ALB/NLB, Amazon RDS, IAM, Amazon Route 53, Amazon CloudWatch, Amazon S3, VPC, Datadog, Splunk, Docker, Kubernetes, Jenkins, GitHub Actions, Argo CD, MySQL, Oracle, Python, Bash, Linux, Helm, Terraform, Prometheus, Grafana, OpenTelemetry, Karpenter, Cluster Autoscaler, Java, JVM
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (AWS)
Jacksonville or Atlanta or Milwaukee
HybridFull Time
FIS
FISNYSE: FIS: Provides technology solutions for merchants, banks, and capital markets
7+ YOE7+ years SRE/Cloud Engineering experience with hands-on AWS, Linux, CI/CD (Jenkins/Harness), Docker/Kubernetes (EKS), Terraform, scripting (Python/Bash/Shell), monitoring tools, on-call experience, and required AWS certifications.
AWS, EC2, EKS, RDS, S3, KMS, Secrets Manager, IAM, Route53, Security Groups, Linux, Git, Docker, Kubernetes, OpenShift, Helm, Jenkins, Harness, Terraform, Python, Bash, Shell, Dynatrace, CloudWatch, Splunk, Prometheus, Grafana, CheckMarx, SonarQube, Maven, Node, Artifactory, FlyWay, KeyFactor, HashiCorp Vault, CyberArk, SNOW, Jira, Confluence, Oracle DB, PostgreSQL DB, Postgres DB, Redis, SFTP, Tivoli
1w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Atlanta or Austin
$100k-$115k/yr OnsiteFull Time
Atlanticus
AtlanticusNASDAQ: ATLC: Provide credit products and financial services to underserved consumers.
5+ YOERequires 5+ years supporting production applications, Java, AWS, Kubernetes, Docker, Datadog, Splunk, CI/CD, databases, Python or Bash, Linux, networking, and incident management experience.
Amazon Web Services (AWS), Amazon EKS, Amazon EC2, ALB, NLB, Amazon RDS, IAM, Amazon Route 53, Amazon CloudWatch, Amazon S3, Amazon VPC, Java, Datadog, Splunk, Docker, Kubernetes, Jenkins, GitHub Actions, Argo CD, MySQL, Oracle, Python, Bash, Linux, Helm, Terraform, Prometheus, Grafana, OpenTelemetry, Karpenter, Cluster Autoscaler, AI-assisted development tools, agentic AI systems
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer I
Boston or Seattle or Atlanta
$134k-$215k/yr HybridFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
7+ YOEBachelor's in CS/Engineering, 7+ years software engineering experience, expertise in distributed systems, Kubernetes, cloud (Azure/AWS/GCP), observability, Kafka, Terraform/Pulumi, and experience with agentic AI/LLM tooling preferred.
Kubernetes, Terraform, Pulumi, Kafka, Grafana, Datadog, New Relic, MySQL, Cassandra, PostgreSQL, Azure, AWS, GCP