23 site reliability engineer jobs at 16 companies in Smithfield, NC

3w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Raleigh, North Carolina, United States
HybridFull Time
SoftPro
SoftPro: Develops software for real estate closing and title insurance.
Experience with Microsoft Azure, Infrastructure as Code (Terraform, Ansible), automation (PowerShell, AZ CLI), Linux/Windows admin, containers (Docker,Kubernetes), observability, CI/CD, incident response, and collaboration skills.
Microsoft Azure, Terraform, Ansible, PowerShell, AZ CLI, Service Fabric, Kubernetes, Docker, Jira, TFS, Git, Azure Monitor, Application Insights, AWS CloudWatch, MS SQL Server, MySQL, MongoDb, Azure CosmosDb, NGINX, HAProxy, OpenID Connect (OIDC), OAuth 2.0, SAML
2mo
Save
Mark Applied
Hide
Site Reliability Engineer II
Raleigh or Nashville or South Carolina or Louisiana or Pennsylvania or Plain City or South Bend or Orlando or Detroit
OnsiteFull Time
Kastle Systems
Kastle Systems: Provides managed security and property technology for buildings and businesses
4+ YOE4+ years SRE/platform experience; Azure, Kubernetes, GitOps, Terraform/OpenTofu, observability and incident management; strong scripting (Python/Go/Bash) and communication skills.
ArgoCD, GitOps, Terraform, OpenTofu, Pulumi, AKS, Azure Container Registry, Azure Monitor, Cosmos DB, Key Vault, Azure Front Door, Kubernetes, Crossplane, Prometheus, Grafana, OpenTelemetry, ELK, OpenSearch, Python, Go, Bash, C#, SQL, Flux, LaunchDarkly, Flagsmith, Linux
3w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Raleigh, North Carolina, United States
$119k-$196k/yr HybridFull Time
Red Hat
Red Hat: Provides enterprise-grade open source software and cloud solutions.
5+ YOE5+ years operating production services on Kubernetes/OpenShift, 3+ years programming in Python/Go, 2+ years with cloud providers, Linux and networking knowledge, SRE principles and on-call experience.
Python, Go, Splunk, Splunk IM, Prometheus, Grafana, Catchpoint, DataDog, OpenShift, Kubernetes, OpenShift Pipelines, Tekton, GitLab, ArgoCD, GitOps, Operator SDK, Linux, AWS, Google, Azure
2mo
Save
Mark Applied
Hide
Site Reliability Engineer II
Falls Church or South Carolina or Raleigh or Nashville or Louisiana or Pennsylvania or Plain City or South Bend or Orlando or Detroit
OnsiteFull Time
Kastle Systems
Kastle Systems: Managed security services provider for commercial and residential properties.
4+ YOE4+ years SRE/Platform experience owning production systems. Hands-on with Azure/AKS, Kubernetes, Terraform/OpenTofu/Pulumi, GitOps/ArgoCD, observability (Prometheus/Grafana/OpenTelemetry/ELK), Python/Go/Bash, and feature-flag/CI/CD practices.
ArgoCD, Flux, Crossplane, LaunchDarkly, Flagsmith, Terraform, OpenTofu, Pulumi, Prometheus, Grafana, OpenTelemetry, ELK, OpenSearch, Python, Go, Bash, C#, SQL, AKS, Azure Container Registry, Azure Monitor, Cosmos DB, Key Vault, Azure Front Door, GitOps
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - HPC
Santa Clara or Austin or Durham
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
5+ YOE5+ years building/supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes); IaC and CI/CD proficiency; coding in Python/Go/Perl/Ruby; monitoring, capacity planning, and incident response skills.
Slurm, LSF, Kubernetes, Infrastructure as Code (IaC), CI/CD, AWS, GCP, OCI, Python, Go, Perl, Ruby
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer - HPC
Santa Clara or Durham or Austin
$152k-$288k/yr OnsiteFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
5+ YOEBS in CS or equivalent with 5+ years supporting critical services; experience with large-scale HPC clusters (Slurm, LSF, Kubernetes), IaC, CI/CD, multi‑cloud (AWS/GCP/OCI), and 2+ languages such as Python or Go.
Slurm, LSF, Kubernetes, AWS, GCP, OCI, Infrastructure as Code (IaC), CI/CD, Python, Go, Perl, Ruby, AIOps
2d
Save
Mark Applied
Hide
Site Reliability Engineer
Morrisville, North Carolina, United States
HybridFull Time
Varonis
VaronisNASDAQ: VRNS: Automates data security and protection across cloud environments.
Bachelor's degree or equivalent; experience building scalable, highly available production services; development experience in C#, Python, or Java; cloud and operational experience; strong analytical and communication skills.
C# .Net, Python, Java, Microsoft Azure, GCP, AWS, CI/CD
3w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Raleigh, North Carolina, United States
HybridFull Time
Fidelity National Financial
Fidelity National FinancialNYSE: FNF: Provides title insurance and real estate settlement services.
Experience with cloud operations, infrastructure as code, automation, containerization, observability, incident response, and collaboration with engineering teams.
Microsoft Azure, Terraform, Ansible, PowerShell, AZ CLI, CI/CD, Service Fabric, Kubernetes, Docker, Jira, DevOps, TFS, Git Repos, Azure Monitor, Application Insights, AWS CloudWatch, MS SQL Server, MySQL, MongoDb, Azure CosmosDb, NGINX, HAProxy, OpenID Connect (OIDC), OAuth 2.0, SAML
1mo
Save
Mark Applied
Hide
Senior Site Reliability Automation Engineer
Raleigh or Denver
HybridFull Time
Litera
Litera: AI-powered document and workflow software for legal professionals.
7+ YOE7+ years in SRE/DevOps/platform engineering; strong software development in Python, PowerShell, TypeScript, JavaScript, Java, or C#; experience with CI/CD, AWS and/or Microsoft Azure, production troubleshooting, root cause analysis, and automation.
Python, PowerShell, TypeScript, JavaScript, Java, C#, CI/CD pipelines, AWS, Microsoft Azure, Datadog, Dynatrace, Splunk, New Relic, Azure Monitor, Amazon CloudWatch
1mo
Save
Mark Applied
Hide
Principal Site Reliability Engineer
Durham, North Carolina, United States
OnsiteFull Time
Fidelity Investments
Fidelity Investments: Provides investment management, retirement planning, and brokerage services.
3+ YOEBachelor's in CS/IT/Engineering + 5 years SRE experience, or Master's + 3 years; experience with CI/CD, cloud (AWS/Azure), observability (Datadog, Splunk), performance testing, Python/Shell scripting.
Datadog, Splunk, Grafana, Java, JMeter, Cloud-test, Rush-hour, Python, Shell, Kubernetes, uDeploy, Jenkins Core, Ansible AWX, Terraform, Azure, AWS, AWS Route53, Azure Load Balancer, F5, AVI
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
Raleigh or Morrisville
OnsiteFull Time
IXL Learning
IXL Learning: Provides personalized digital learning platforms and educational resources.
6+ YOE6+ years SRE/software engineering experience; Bachelor's in computer science or related; experience with Java/C++/C, Python/Bash/Perl, AWS or GCP, Docker and Kubernetes, observability, troubleshooting, and on-call flexibility.
Java, C++, C, Python, Bash, Perl, AWS, GCP, Docker, Kubernetes
1mo
Save
Mark Applied
Hide
Site Reliability / DevOps Engineer
Raleigh, North Carolina, United States
$120k-$138k/yr OnsiteFull Time
eClerx
eClerxNational Stock Exchange of India: ECLERX: Provides business process management and data analytics services globally.
5+ YOE5+ years SRE/DevOps experience with public cloud (Azure preferred or AWS), observability (Datadog/Elasticsearch/Grafana), scripting (Python, Bash, PowerShell), containerization (Docker, Kubernetes), CI/CD and IaC (Terraform, Azure Bicep, CloudFormation), incident response and on-call participation.
Datadog, Elasticsearch, Grafana, Python, Bash, PowerShell, Docker, Kubernetes, Azure DevOps, GitLab CI/CD, GitHub Actions, Jenkins, Terraform, Azure Bicep, AWS CloudFormation, Gremlin, Chaos Mesh, Kafka, Azure Event Hubs, Linux, Git, Maven, Gradle, Artifactory, Cypress, Selenium, Cucumber, AG Grid, D3, GitHub Copilot, Claude, ChatGPT
4w
Save
Mark Applied
Hide
CaaS Private Site Reliability Lead Engineer - Vice President
Cary, North Carolina, United States
$125k-$185k/yr HybridFull Time
Deutsche Bank
Deutsche BankNew York Stock Exchange: DB: A global bank providing financial services to individuals and corporations.
Extensive experience with Kubernetes and Linux, observability, SLOs, automation, incident management, production platform operations, and mentoring engineers.
Kubernetes, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer (Senior or Staff)
Boston or Miami or New Jersey or New York City or Princeton or Raleigh or Washington or Toronto or North America
$127k-$249k/yr HybridFull Time
MongoDB
MongoDBNASDAQ: MDB: Cloud-based document database platform for software application development.
6+ YOE6+ years software development experience; proficiency in Python or Go; experience building and operating large-scale CI/CD pipelines; Kubernetes and cloud (AWS, Google Cloud Platform, Microsoft Azure) expertise; Linux and networking knowledge.
Argo Workflows, ArgoCD, Kubernetes, Python, Go, AWS, Google Cloud Platform (GCP), Microsoft Azure, Linux
1mo
Save
Mark Applied
Hide
Director, SRE - Enterprise Reliability
Cary or New York City
$150k-$210k/yr HybridFull Time
MetLife
MetLifeNYSE: MET: Global provider of insurance, annuities, and financial services.
8+ YOE8+ years in SRE/production support or engineering; experience leading SRE teams; expertise in observability, SLO/SLI design, incident leadership, and cross-team influence.
ServiceNow, Kubernetes, CI/CD, AIOps
2w
Save
Mark Applied
Hide
Software Engineer III, Site Reliability Engineering
Raleigh or Durham
$147k-$211k/yr OnsiteFull Time
Google
GoogleNASDAQ: GOOGL: Provides online search, advertising, cloud computing, and consumer electronics.
2+ YOEBachelor's degree or equivalent,2+ years software development experience; distributed systems experience and a Master's degree preferred; strong design, debugging, and communication skills.
1mo
Save
Mark Applied
Hide
Software Engineer - Cloud Platform / Reliability (Morrisville, NC, US)
Morrisville, North Carolina, United States
$113k-$168k/yr OnsiteFull Time
NetApp
NetAppNASDAQ: NTAP: Sells enterprise data storage and cloud management software.
2+ YOEExperience in software or site reliability engineering; strong Python/Shell scripting, Linux (RHEL/CentOS), Kubernetes, AWS, CI/CD, observability (Dynatrace, Grafana), and SQL/NoSQL experience.
Python, Shell, Kubernetes, Rancher, Dynatrace, Grafana, RHEL/CentOS, cron, Airflow, SQL, NoSQL, AWS, Linux, CI/CD, AI/ML, Generative AI
2mo
Save
Mark Applied
Hide
Director, Total Productive Maintenance (TPM) - (USA REMOTE)
Chaska or Chicago or Charleston or Little Rock or Owensboro or New York or Newark or Orlando or Nashville or Richmond or Baltimore or Indianapolis or Silver Spring or Columbia or Greenville or Miami or Louisville or Florence or Jersey City or Alexandria or Jackson or Washington or Boston or Durham or St. Louis or Charlotte or Bowling Green or Cambridge
$200k-$225k/yr RemoteFull Time
Danaher
DanaherNYSE: DHR: Develops scientific instruments and diagnostic tools for healthcare markets.
10+ YOE5+ MgmtBachelor’s in Engineering; 10+ years GMP maintenance/engineering/ops; 5+ years multi-site leadership; strong TPM, reliability, and CMMS/EAM knowledge; experience improving equipment reliability in regulated environments.
CMMS, EAM, RCM, RCA, FMEA, Industry 4.0, CMMS/EAM, Digital maintenance technologies