Kibo
Posted 4mo ago

DevOps Engineer

Kibo
Pune, Maharashtra, India
RemoteFull Time
Responsibilities
  • Manage clusters
  • Troubleshoot issues
  • Build infrastructure
Requirements
  • 8+ years DevOps/Developer Engineer
  • Ownership of production Kubernetes (EKS preferred)
  • Real-time troubleshooting
  • Terraform
  • On-call rotation
Technical tools mentioned
KubernetesEKSTerraformPrometheusGrafana

Job description


About This Role:


We are hiring a hands-on DevOps Engineer to manage and support production-grade cloud infrastructure for Kibo’s commerce platform. This role focuses on Kubernetes (EKS), Terraform, and real-time production troubleshooting in a 24/7 on-call environment.




ABOUT KIBO 




KIBO is a composable digital commerce platform for B2C, D2C, and B2B organizations who want to simplify the complexity in their businesses and deliver modern customer experiences.  KIBO is the only modular, modern commerce platform that supports experiences spanning B2B and B2C Commerce, Order Management, and Subscriptions. Companies like Ace Hardware, Zwilling, Jelly Belly, Nivel, and Honey Birdette trust Kibo to bring simplicity and sophistication to commerce operations and deliver experiences that drive value.   


KIBO's cutting-edge solution is MACH Alliance Certified and has been recognized by Forrester, Gartner, IDC, Internet Retailer, and TrustRadius. KIBO has been named a leader in The Forrester Wave™: Order Management Systems, Q1 2025 and in the IDC MarketScape report “Worldwide Enterprise Headless Digital Commerce Applications 2024 Vendor Assessment”.


By joining KIBO, you will be part of a team of Kibonauts all over the world in a remote-friendly environment. Whether your job is to build, sell, or support KIBO’s commerce solutions, we tackle challenges together with the approach of trust, growth mindset, and customer obsession. If you’re seeking a unique challenge with amazing growth potential, then come work with us!


 



WHAT YOU’LL DO 





  • Manage and operate production-grade Kubernetes clusters (EKS preferred), ensuring high availability and scalability

  • Troubleshoot real-time production issues across distributed systems and microservices

  • Diagnose and resolve issues such as:

    • Pod failures (CrashLoopBackOff, Pending, OOMKilled)

    • Node failures, autoscaling, and resource constraints

    • Networking, ingress, and service connectivity issues



  • Build, maintain, and debug infrastructure using Terraform (modules, remote state, locking, drift handling)

  • Implement and enhance monitoring & alerting systems using Prometheus, Grafana, and related tools

  • Perform root cause analysis (RCA) for incidents and drive permanent fixes to improve system reliability

  • Participate in a 24/7 on-call rotation, owning incidents and resolving them independently

  • Collaborate with engineering teams to improve system performance, resilience, and deployment processes

  • Automate deployments, infrastructure provisioning, and operational workflows to reduce manual effort

  • Ensure adherence to security best practices across infrastructure and deployments 


 



WHAT YOU’LL NEED 







  • 8 + Years of experience as a Developer Engineer, owning and operating production Kubernetes clusters (EKS preferred), including cluster health, scaling, and availability

  • Troubleshoot real-time production issues independently across microservices and distributed systems

  • Debug and resolve critical issues such as:

    • Pods stuck in CrashLoopBackOff, Pending, OOMKilled states

    • Node failures, node pressure, autoscaling issues

    • Service connectivity, ingress, and networking issues



  • Investigate and fix cluster-level issues including scheduling, resource constraints, and misconfigurations

  • Build and maintain infrastructure using Terraform, including:

    • Writing and modifying modules

    • Managing remote state and locking

    • Handling drift and failed deployments

    • Design and implement reusable Terraform modules for scalable infrastructure

    • Troubleshoot and resolve Terraform apply failures and infrastructure inconsistencies in production



  • Monitor system health using Prometheus, Grafana, and logging tools, and proactively identify issues

  • Perform root cause analysis (RCA) for production incidents and implement long-term fixes

  • Handle on-call incidents (24/7 rotation) and take full ownership until resolution

  • Work closely with development teams to improve system reliability, performance, and scalability

  • Automate operational tasks and improve deployment and infrastructure processes

  • Ensure security best practices across infrastructure, networking, and access controls

  • .

 




KIBO PERKS 






  • Flexible schedule and hybrid work setting 




  • Paid company holidays and global volunteer holiday 




  • Generous health, wellness, benefits, and time away programs 










  • Commitment to individual growth and development and opportunity for internal mobility 




  • Passionate, high-achieving teammates excited to help you succeed and learn 




  • Company-sponsored events and other activities  






At Kibo we celebrate and support all differences. Kibo is proud to be an equal opportunity workplace. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital, disability, and veteran status. 




 



About Kibo

Composable digital commerce platform for eCommerce and order management.

Year founded
2016
Employees
350
Organization type
Private
Latest investment
Raised $21.60M Private Equity (2016) — led by Vista Equity Partners
Headquarters
US

Similar jobs

DevOps Engineer roles in Maharashtra
1d
Save
Mark Applied
Hide
Senior/Lead DevOps Engineer - Azure
Gurgaon or Pune or Coimbatore or Chennai or Hyderabad or Bangalore
OnsiteFull Time
EPAM Systems
EPAM SystemsNYSE: EPAM: Provides global digital platform engineering and software development services.
8+ YOERequires 8+ years of DevOps experience, Azure DevOps, Terraform, AKS, Docker, Kubernetes, PowerShell, Python or Bash, Git/GitHub, monitoring tools, cloud security, and Agile/Scrum experience.
Microsoft Azure, Azure DevOps, Azure Kubernetes Service, Terraform, Kubernetes, Docker, PowerShell, Python, Bash, Git, GitHub, Azure Monitor, Grafana, Prometheus, Azure Key Vault, Azure Policy, Microsoft LinkedIn Learning
1d
Save
Mark Applied
Hide
DevOps
Pune, Maharashtra, India
OnsiteFull Time
Ascendion
Ascendion: AI-native digital engineering and software development services provider.
3+ YOERequires 3+ years of production Kubernetes, Terraform, and large-scale CI/CD experience, plus SRE, incident response, debugging, multi-cloud, and security best-practice knowledge.
Kubernetes, GKE, AKS, EKS, Terraform, Bitbucket, GitHub, Azure DevOps, Prometheus, Grafana, Databricks, Azure Databricks, Vertex AI, GCP, Azure, AWS, GitOps, IAM
1d
Save
Mark Applied
Hide
DevOps Engineer
Pune, Maharashtra, India
OnsiteFull Time
Global Payments
Global PaymentsNYSE: GPN: Provides payment technology and software solutions for global commerce.
2+ YOEBachelor's degree or equivalent technical experience, 2+ years in IT, scripting, CI/CD, infrastructure automation, cloud platforms, containers, monitoring, networking, security, and system architecture.
Terraform, Pulumi, CloudFormation, Kubernetes, ECS, Docker, Docker Warm, Python, Bash, PowerShell, AWS, Azure, GCP, Prometheus, Grafana, ELK stack
3d
Save
Mark Applied
Hide
Senior Associate- DevOps
Pune, Maharashtra, India
HybridFull Time
Davies
Davies: Insurance claims management and legal professional services provider.
5+ YOE5+ years DevOps/Cloud experience, degree or equivalent, strong CI/CD and IaC skills (Terraform/ARM/Bicep), Azure and monitoring experience, cloud security and regulated environment exposure.
Terraform, ARM, Bicep, Azure Monitor, App Insights, Prometheus, Grafana, Azure
4d
Save
Mark Applied
Hide
DevOps Engineer-Lifecycle Services
Chandigarh or Pune
HybridFull Time
Emerson
EmersonNYSE: EMR: Engineering industrial automation and software solutions for global industries.
3+ YOEBachelor's degree,3+ years DevOps or cloud operations experience,proficiency with Azure DevOps,CI/CD,pipeline troubleshooting,Kubernetes/AKS,Docker,Azure,PowerShell/Bash,and Linux/Windows administration.
Azure DevOps, Git, YAML, PowerShell, Bash, Kubernetes, AKS, Docker, Azure, Salesforce, Adobe Experience Manager, MS SQL, Oracle, Linux, Windows Server
4d
Save
Mark Applied
Hide
DevOps Engineer
Pune, Maharashtra, India
OnsiteFull Time
DataDirect Networks
DataDirect Networks: High-performance storage and data management for AI and HPC.
10+ YOE10+ years in infrastructure/platform/DevOps, strong Python or Go, hybrid cloud and bare-metal experience, Terraform/Helm/Kubernetes knowledge, cross-functional collaboration, and participation in on-call rotation.
Python, Go, Terraform, Helm, Kubernetes
4d
Save
Mark Applied
Hide
Lead DevOps Engineer - Tieto Caretech (m/f/d)
Pune, Maharashtra, India
HybridFull Time
Tietoevry
TietoevryNasdaq Helsinki: TIETO: A digital services and software providing IT solutions.
7+ YOE7+ years DevOps experience with Kubernetes and Terraform; expertise in CI/CD (Azure DevOps, GitHub, ArgoCD), observability and security, AI-assisted tooling, and enterprise deployment strategies for cloud-agnostic environments.
Kubernetes, Terraform, ArgoCD, Azure DevOps (ADO), GitHub, GitHub Actions, GitHub Copilot, Claude, .NET 10, C#, Angular, TypeScript, Prometheus, Grafana, OpenTelemetry
4d
Save
Mark Applied
Hide
Java Devops Engineer
Pune, Maharashtra, India
HybridFull Time
Citi
CitiNYSE: C: Global diversified financial services holding.
6+ YOE6+ years experience in DevOps/Applications Development with strong Java, Linux, CI/CD, containerization, Kubernetes, scripting, Oracle DB, and Git-based workflows.
Java, Maven, Gradle, Docker, Kubernetes, Enterprise Container System (ECS), Oracle, SQL, PL/SQL, GitHub, Git, GitHub Actions, Jenkins, Ansible, Harness, Shell scripting, AppDynamics, Kibana, YAML, JSON, XML, Python, Spring, Copilot, Claude, Devin AI