6 infrastructure reliability engineer jobs at 5 companies in Henderson, NC

1mo
Save
Mark Applied
Hide
Customer Reliability Engineer - Infrastructure
San Francisco or Boston or Washington D.C. or Raleigh or Pittsburgh or Philadelphia or New York City or Miami or Columbus or Austin or United States
$125k-$130k/yr RemoteFull Time
Astronomer
Astronomer: Managed data orchestration platform powered by Apache Airflow.
5+ YOE5+ years with large cloud infrastructures, 3+ years Kubernetes, production distributed systems on AWS/GCP/Azure, strong Linux, Python scripting, DevOps/CI/CD, observability/monitoring, and customer-facing troubleshooting.
Apache Airflow, AWS, Azure, CI/CD, GCP, Infrastructure as Code (IaC), Kubernetes, Linux, Python
3w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Raleigh, North Carolina, United States
HybridFull Time
SoftPro
SoftPro: Develops software for real estate closing and title insurance.
Experience with Microsoft Azure, Infrastructure as Code (Terraform, Ansible), automation (PowerShell, AZ CLI), Linux/Windows admin, containers (Docker,Kubernetes), observability, CI/CD, incident response, and collaboration skills.
Microsoft Azure, Terraform, Ansible, PowerShell, AZ CLI, Service Fabric, Kubernetes, Docker, Jira, TFS, Git, Azure Monitor, Application Insights, AWS CloudWatch, MS SQL Server, MySQL, MongoDb, Azure CosmosDb, NGINX, HAProxy, OpenID Connect (OIDC), OAuth 2.0, SAML
3w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
Raleigh, North Carolina, United States
HybridFull Time
Fidelity National Financial
Fidelity National FinancialNYSE: FNF: Provides title insurance and real estate settlement services.
Experience with cloud operations, infrastructure as code, automation, containerization, observability, incident response, and collaboration with engineering teams.
Microsoft Azure, Terraform, Ansible, PowerShell, AZ CLI, CI/CD, Service Fabric, Kubernetes, Docker, Jira, DevOps, TFS, Git Repos, Azure Monitor, Application Insights, AWS CloudWatch, MS SQL Server, MySQL, MongoDb, Azure CosmosDb, NGINX, HAProxy, OpenID Connect (OIDC), OAuth 2.0, SAML
3w
Save
Mark Applied
Hide
Site Reliability Engineering (SRE) & DevOps
Raleigh, North Carolina, United States
$91k-$135k/yr OnsiteFull Time
NTT DATA
NTT DATA: Global provider of IT and business consulting services.
5+ YOE5+ years experience; expertise in SRE/DevOps, cloud engineering on GCP, infrastructure automation, observability (SLI/SLO/Error Budgets), Terraform, Python, PowerShell, and Git-based workflows.
Google Cloud Platform (GCP), Terraform, GitHub, Git, Python, PowerShell, Kubernetes, Docker, Jenkins, Maven, Ant, Puppet, Gradle, Ansible, TeamCity, uDeploy, Dataproc, Java, C#
1mo
Save
Mark Applied
Hide
Senior System Architect, Infrastructure Reliability
Santa Clara or Westford or Austin or Durham or Redmond
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs GPU-accelerated computing and artificial intelligence hardware.
6+ YOE6+ years systems programming experience, BS/MS/PhD in CS or EE (or equivalent), expertise in CPU/GPU diagnostics, C++ and Python proficiency, experience with RCA, cluster managers (Slurm/LSF/Kubernetes).
C++, Python, Slurm, LSF, Kubernetes, NVIDIA DCGM, NVIDIA Management Library (NVML), CRIU, CUDA, /dev/mcelog, dmesg, journald
1mo
Save
Mark Applied
Hide
Senior System Architect, Infrastructure Reliability
Santa Clara or Austin or Westford or Durham or Redmond
$184k-$357k/yr HybridFull Time
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
6+ YOE6+ years systems programming experience; BS/MS/PhD in CS or EE (or equivalent); experience building RCA pipelines for HPC/cloud; deep CPU/GPU architecture knowledge; strong C++ and Python; familiarity with Slurm/LSF/Kubernetes.
C++, Python, Slurm, LSF, Kubernetes, CUDA, DCGM, NVML, CRIU, /dev/mcelog, dmesg, journald, Linux kernel