🔄 Staff Augmentation

Braves Technologies is a staff augmentation and offshoring firm that recruits and manages dedicated technical teams for external global clients.

This company was flagged and excluded from default search results. Proceed with caution.

Braves Technologies
Posted 1y ago

Site Reliability Engineer (DevOps) (Remote)

Braves Technologies
United States
RemoteFull Time
Responsibilities
  • Maintaining applications
  • Implementing improvements
  • Collaborating with teams
Requirements
  • Strong knowledge of SRE
  • DevOps
  • AWS Cloud, and monitoring tools
  • Experience with CI/CD tools and Agile methodologies
Technical tools mentioned
AWSNewrelicAppDynamicsDatadogSumologicSplunkGrafanaGitlabBamboo

Job description


Who we’re looking for?

A Site Reliability DevOps engineer working as part of the high-performing Operations team (SRE) growing their knowledge and skillset. Helps maintain existing business-critical applications and infrastructure while recommending technical and process improvements.


Our Company

Founded in 2003, Braves Technologies is helping global technology companies incubate their dedicated offshore software development teams in India. For the past 15+ years, Braves has been building Software Engineering, Game Development, and Customer Success teams for clients across the US and Australia.


For more information, you can visit https://www.bravestechnologies.com/


Our Culture

We are a team focused on high performance, high delivery, diverse thinking, and embodying a collaborative culture at all levels. We value and encourage learning throughout the organization. Every employee at Braves understands ownership and fulfils what is required. We align a perfect work-life balance.


Working days: Monday to Friday (Fixed weekend off)

Work Location: Remote


The role will interface and collaborate with multiple teams and stakeholders to deliver and maintain new and innovative world-class solutions. The Site Reliability engineer will:

  • Focus on the availability, security, scalability, and operational readiness of cloud-based solutions.
  • Implement and continuously improve development and delivery practices for our tools, platform, and patterns.
  • Implement technical initiatives and automation collaboratively and influence stakeholder conversations.
  • Enhance and further embed DevOps/SRE culture within the team and into wider delivery teams.
  • Enhance existing observability patterns by developing and enhancing dashboards, toolset etc.
  • Support delivery teams with new and innovative solutions that will delight customers.
  • Provide support for online applications to ensure they are available, scalable, resilient, and high performing.
  • Maintain the security and compliance of the production environment using PCI-DSS, ISM.
  • Develop and Enhance monitoring practices to improve the observability within technical environment.
  • Develop and Enhance dashboards to improve the observability for delivery teams
  • Manage components like Load Balancers, Proxies, DNS, API Gateways etchant automate the provisioning and deployments.
  • Assist delivery teams by ensuring application deployments are aligned with the operational practices.
  • Provide guidance on the application architecture, security, scalability, and operational readiness.
  • Contribute to application and infrastructure architectural designs and patterns.
  • Contribute to the continuous improvement of the SRE team processes and practices.
  • Assist delivery teams with understanding, managing, and optimizing cloud-based application costs using AWS Well-Architected Framework.
  • Contribute to the optimization of cloud-based costs.
  • Engage with internal stakeholders to meet their expectations.
  • Assist colleagues with building innovative working software and solutions for customers.
  • Evaluate and develop new capabilities enabling solutions.
  • Contribute to maintaining a safe, respectful, collaborative, and positive working environment.
  • Be a role model in culture and ways of working.


-Market and Environment

  • Excellent Troubleshooting abilities and Strong understanding of SRE and application support practices.
  • Strong knowledge of Monitoring and Logging tools like Newrelic, AppDynamics, Datadog, Sumologic, Splunk etc.
  • Strong Dashboarding skills in various tools like Grafana, Geckoboard etc.
  • Strong Experience in AWS Cloud and managing infrastructure components like Load Balancers, Proxies, DNS, API Gateways etc and automate the provisioning and deployments.
  • Strong understanding of SRE, DevOps, Continuous Delivery, Continuous Integration, “Infrastructure as code”, and related current practices and ideas.
  • Experience with various CI/CD tools like Gitlab, Bamboo etc.
  • Understanding of Agile methodologies and ITIL concepts.


What’s in it for you/Benefits of working for us:

Competitive Salary

Family Group Medical Health Insurance

Group Accidental Insurance

Leave encashments (Gross, not just base salary)

Regular Fun and Sports activities.

Birthday/Anniversary Celebrations

Other benefits like Gratuity, PF/VPF, maternity, etc.


Similar jobs

Site Reliability Engineer roles
3w
Save
Mark Applied
Hide
Site Reliability Engineer I
United States
$85k-$178k/yr RemoteFull Time
Indeed
Indeed: Online employment platform for job seekers and employers.
Bachelor's degree in CS or related field required; familiarity with AWS, Java, Kotlin, React.js; strong communication and continuous learning mindset.
AWS, Java, Kotlin, React.js
4w
Save
Mark Applied
Hide
Senior Site Reliability Engineer
United States
$160k-$208k/yr RemoteFull Time
Clover Health
Clover HealthNasdaq: CLOV: Provide Medicare Advantage plans and AI-powered clinical decision tools.
5+ YOE5+ years programming experience; proficiency in Python, Go, or shell scripting; Kubernetes, containerization, cloud platforms, Linux, networking, monitoring and SRE concepts.
Python, Go, Shell Scripting, Docker, containerd, Kubernetes, Helm, gRPC, Prometheus, GCP, Azure, AWS, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
United States
$150k-$200k/yr RemoteFull Time
Runpod
Runpod: Cloud platform for AI training and model deployment.
5+ YOE5+ years SRE or production engineering experience; strong Linux, networking, container, distributed systems, SLI/SLO, incident response, and scripting skills.
Prometheus, Grafana, Python, Go, Bash, Linux, Slack
8mo
Save
Mark Applied
Hide
FedRAMP Site Reliability Engineer (FedSRE) - CloudVision
Remote, United States
$101k-$161k/yr RemoteFull Time
Arista Networks
Arista NetworksNYSE: ANET: Provides cloud networking solutions and high-speed multilayer Ethernet switches.
5+ YOE5+ years software engineering; experience with FedRAMP SaaS or highly regulated systems; Python/Go; Bash scripting; distributed databases or large-scale SaaS; U.S. citizenship.
Golang, Python, Bash, Ansible, Pulumi, Kubernetes, GKE, Cloud Platforms
1y
Save
Mark Applied
Hide
Senior Site Reliability Engineer — AI Studio (Inference Platform)
Amsterdam or Berlin or London or Prague or United States
HybridFull Time
Nebius
NebiusNasdaq: NBIS: Builds cloud infrastructure and software for artificial intelligence development.
Deep fluency with Kubernetes, Prometheus, Grafana, Terraform, and infrastructure-as-code. Experience with GPU-heavy workloads and MLOps is preferred.
Kubernetes, Prometheus, Grafana, Terraform, Python, Bash