86 reliability engineer jobs at 67 companies in Griffin, GA

1mo
Save
Mark Applied
Hide
Reliability Engineer
Atlanta or Costa Mesa
$126k-$167k/yr OnsiteFull Time
Anduril Industries
Anduril Industries: Defense technology building autonomous military hardware and software.
5+ YOE5+ years reliability/design/test engineering experience; BS in mechanical/aerospace/systems engineering; experience with safety-critical defense/aerospace systems, MIL standards, FMEA/FTA, Weibull analysis; eligible for U.S. Secret clearance.
MIL-HDBK-217Plus, MIL-HDBK-217, MIL-HDBK-472, MIL-STD-810, MIL-STD-461, MIL-STD-516C, MIL-STD-1629, MIL-STD-1916, FMEA, FMECA, Fault Tree Analysis (FTA), Weibull analysis, HALT, HASS, MRL
1mo
Save
Mark Applied
Hide
Reliability Engineer
Newnan, Georgia, United States
OnsiteFull Time
Mauser Packaging Solutions
Mauser Packaging Solutions: Manufactures and reconditions sustainable industrial rigid packaging containers.
10+ YOEBachelor's in Mechanical Engineering (or equivalent experience), 10 years related experience (industrial maintenance preferred), strong facility/equipment knowledge, project and financial acumen, and willingness to travel extensively.
1mo
Save
Mark Applied
Hide
Network Reliability Engineer
Austin or Atlanta or Denver or Seattle or Washington
HybridFull Time
Cloudflare
CloudflareNYSE: NET: Provides security and performance services for internet properties.
3+ YOE3+ years network/site reliability engineering experience; BA/BS in CS or equivalent; experience with Saltstack, Ansible, Chef, NX-OS/JUNOS/EOS/Cumulus/Sonic; Linux administration; iproute2, Traffic Control, Devlink; software development in Go and Python; AI/LLM tooling experience.
Saltstack, Ansible, Chef, NX-OS, JUNOS, EOS, Cumulus, Sonic, LLM, iproute2, Traffic Control, Devlink, Go, Python, AirFlow, Temporal, FRR, Bird, GoBGP, C, C++, rust, Linux, Linux kernel, Prometheus, Grafana, Thanos, Clickhouse, Kubernetes, Docker, Consul
2w
Save
Mark Applied
Hide
Reliability Engineer, Mechanical, NA (Design)
Shackelford County or Texas or Atlanta or Abilene or Dallas or Phoenix or Ashburn or Wisconsin
HybridFull Time
Vantage Data Centers
Vantage Data Centers: Provides hyperscale data center campuses for cloud and AI providers.
2+ YOEMechanical reliability engineer for data center cooling systems; 2–3 years critical facility experience preferred; bachelor’s degree preferred; experience with commissioning, maintenance program design, RCA, and technical support.
1d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless network services and mobile communication devices.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year. Requires automation, monitoring, scripting, cloud computing, CI/CD, incident management, and system reliability skills.
CI/CD, Kubernetes, AWS
1d
Save
Mark Applied
Hide
Engineer, Site Reliability
Atlanta or Frisco
$85k-$153k/yr OnsiteFull Time
T-Mobile
T-MobileNASDAQ: TMUS: Provides wireless voice, data, and mobile internet services.
3+ YOEBachelor's degree and 3 years of related experience, or advanced degree and 1 year; automation, monitoring, CI/CD, cloud computing, scripting, incident management, and system reliability skills required.
CI/CD, Kubernetes, AWS
3d
Save
Mark Applied
Hide
Senior Site Reliability Engineer (SRE)
Atlanta or United States
$120k-$175k/yr RemoteFull Time
PrizePicks
PrizePicks: Operates a daily fantasy sports and player prediction platform.
5+ YOE5+ years reliability engineering experience; cloud (AWS/Azure/GCP), IaC (Terraform/Crossplane), Kubernetes, Python/Ruby/Go, monitoring tools, incident response, SLO governance, strong cross-functional and debugging skills.
AWS, Azure, GCP, Terraform, Crossplane, Python, Ruby, Go, Kubernetes, Grafana, New Relic, Datadog, Windows, Mac
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta or Alpharetta
OnsiteFull Time
Incident IQ
Incident IQ: Workflow management software for K-12 school district operations.
5+ YOE5+ years SRE/DevOps experience, strong systems fundamentals, SLI/SLO and observability experience, incident management, automation and cloud skills, proficient with modern SRE tooling and AI-accelerated execution.
Grafana, PromQL, Grafana Alloy, Prometheus, Datadog, OpenTelemetry, SigNoz, Uptrace, Tempo, Grafana Faro, k6, PagerDuty, Locust, JMeter, Python, Go, Bash, Terraform, Ansible, Kubernetes, Amazon Web Services (AWS), Google Cloud Platform (GCP), Azure, GitOps, eBPF, Grafana Beyla, OpenTelemetry eBPF Instrumentation, .NET, Real User Monitoring (RUM)
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer I
Boston or Seattle or Atlanta
$134k-$215k/yr HybridFull Time
Axon
AxonNASDAQ: AXON: Develops weapons and software for law enforcement and safety.
7+ YOEBachelor's in CS/Engineering, 7+ years software engineering experience, expertise in distributed systems, Kubernetes, cloud (Azure/AWS/GCP), observability, Kafka, Terraform/Pulumi, and experience with agentic AI/LLM tooling preferred.
Kubernetes, Terraform, Pulumi, Kafka, Grafana, Datadog, New Relic, MySQL, Cassandra, PostgreSQL, Azure, AWS, GCP
2mo
Save
Mark Applied
Hide
Senior Reliability Engineer
South San Francisco or Atlanta
$150k-$180k/yr HybridFull Time
AeroVect
AeroVect: Develops autonomous driving software for airport logistics vehicles.
5+ YOE5–8 years reliability engineering in hardware-focused field; strong FMEA/FTA/RBD; environmental and accelerated life testing; cross-domain systems; data analysis; excellent communication.
FMEA, FTA, RBD, Weibull Analysis, environmental testing, DFR
2mo
Save
Mark Applied
Hide
System Reliability Engineer, III - VI
Tucker, Georgia, United States
OnsiteFull Time
Georgia Transmission
Georgia Transmission: Plans, builds and maintains Georgia's high-voltage power grid.
4+ YOEBS Electrical Engineering; multiple E-level experience in power utility; PE/EIT applicable but not required; proficient in Excel, PowerBI, data tools; strong communication and planning skills.
Microsoft Excel, Power BI, Microsoft Access, Microsoft Word, PowerPoint
2mo
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Atlanta, Georgia, United States
OnsiteFull Time
Florence Healthcare
Florence Healthcare: Digital platform for managing clinical trial documents and workflows.
4+ YOESRE with 4+ years in SRE/DevOps, cloud-native architectures, Linux, observability, IaC (Terraform), AI-assisted tooling; AWS experience; strong collaboration.
AWS, Terraform, CI/CD, Linux, Observability, AI-assisted tooling
1mo
Save
Mark Applied
Hide
Sr Site Reliability Engineer
Atlanta, Georgia, United States
$178k-$205k/yr HybridFull Time
Workday
Workday: A provider of cloud-based enterprise software and services focused on managing people and finances.
5+ YOEBachelor's degree plus 5 years experience. 5+ years with Ansible, Terraform, Packer, Python, Shell Scripting, Kafka, CI/CD tools (GIT, Maven/Gradle, Jenkins), Docker, Kubernetes, and cloud provisioning; strong monitoring and automation skills.
Ansible, Terraform, Packer, Python, Shell Scripting, Kafka, GIT, Maven, Gradle, Jenkins, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Atlanta or Uniontown
HybridFull Time
QGenda
QGenda: Provides healthcare workforce management and physician scheduling software.
7+ YOE7+ years SRE/DevOps experience, BS in CS or equivalent, advanced scripting skills, strong AWS (Lambda, EC2, ECS/EKS, S3, SNS, SQS, RDS, Redshift, ElastiCache) and Terraform experience, Docker, Kubernetes, CI/CD, Git, observability tools, and on-call availability.
AWS Lambda, Amazon EC2, AWS ECS, AWS EKS, Kubernetes, AWS S3, AWS SNS, AWS SQS, AWS RDS, Amazon Redshift, AWS ElastiCache, Docker, Terraform, Git, AWS CodeBuild, Jenkins, TeamCity, Datadog, Amazon CloudWatch, PagerDuty, Claude, GitHub Copilot
1mo
Save
Mark Applied
Hide
Staff Site Reliability Engineer, Security
United States or Atlanta or Canada or Europe
RemoteFull Time
Stord
Stord: Cloud-based logistics platform for omnichannel fulfillment and warehousing
Deep GCP/GKE security, Dependabot/secret scanning, CI/CD supply-chain hardening, Terraform and IaC automation, proficiency with Python/Go/shell, and experience building security tooling and remediation workflows.
GCP, GKE, Istio, Cloud Armor, Cloudflare, GitHub, Dependabot, Terraform, Python, Go, shell, GitHub Actions, SLSA, sigstore, Log Explorer, BigQuery, OWASP ZAP, nmap
1w
Save
Mark Applied
Hide
Site Reliability Engineer III, DevEx
United States or Atlanta or Boston
$140k-$165k/yr RemoteFull Time
Flock Safety
Flock Safety: Sells AI-powered cameras and software for public safety surveillance.
Experience writing production Go or TypeScript, proficiency with Kubernetes, Helm, Terraform, GitHub Actions, and AWS; observability and CI/CD expertise; participates in on-call rotations.
Go, TypeScript, Kubernetes, Helm, Terraform, GitHub Actions, AWS
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Atlanta, Georgia, United States
$152k-$162k/yr OnsiteFull Time
Unum Group
Unum GroupNYSE: UNM: Provides insurance and employee benefits to businesses and individuals.
5+ YOEBachelor's in Computer Science/Engineering plus 5+ years experience; expertise with observability, cloud-native architectures, incident response, automation scripting, CI/CD, infrastructure-as-code, and collaboration in DevOps environments.
Dynatrace, AWS CloudWatch, Datadog, Grafana, Amplitude, AWS, Python, Bash, PowerShell, GitHub, GitLab, Bitbucket, GitHub Actions, Jenkins, Azure DevOps, Terraform, AWS CloudFormation, Ansible
4d
Save
Mark Applied
Hide
Site Reliability Engineer , Engineering Enablement (Remote)
Boston or Atlanta or Chicago or Washington or New York City or Herndon
$138k-$198k/yr RemoteFull Time
Cisco
CiscoNASDAQ: CSCO: Develops and sells networking hardware and cybersecurity software.
3+ YOEBachelor's+5 or Master's+3 experience; 3+ years writing production Python or Ruby; 3+ years managing Terraform/Ansible IaC; Unix/Linux experience; containerization and CI/CD experience; ability to operate at scale and participate in on-call rotations.
Python, Ruby, Terraform, Ansible, Unix/Linux, Docker, Kubernetes, DORA, SPACE
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer, Google Cloud
Atlanta or Milpitas
$240k-$250k/yr HybridFull Time
Saviynt
Saviynt: Provides AI-powered identity governance and cloud security platforms.
9+ YOE9+ years in platform/infra/SRE roles, deep Kubernetes and GCP expertise, strong Go and Python skills, experience with CI/CD, event-driven systems, observability, distributed systems, and building shared platform services.
Go (Golang), Python, Kubernetes, GCP, AWS, Azure, Kafka, RMQ, NATS, Google Pub/Sub, GitLab CI, ArgoCD, Prometheus, Grafana, ELK stack, Datadog, Envoy, Istio, MySQL, PostgresSQL
2w
Save
Mark Applied
Hide
Senior Site Reliability Engineer – Unified Observability
Atlanta, Georgia, United States
OnsiteFull Time
NCR Voyix
NCR VoyixNYSE: VYX: Provides checkout software and kiosks for retailers and restaurants.
10+ YOE10+ years SRE/Platform/Cloud experience; expertise with Azure, GCP, Kubernetes (AKS,GKE); observability tools (Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic); Terraform; Python/Go/PowerShell; bachelor\u0002s degree or equivalent.
Azure, Google Cloud Platform, Kubernetes, AKS, GKE, Grafana, Datadog, Prometheus, OpenTelemetry, Dynatrace, New Relic, ServiceNow, Terraform, Python, Go, PowerShell