Shakudo
Posted 2mo ago

Senior DevOps Engineer — Customer Deployments & Infrastructure

Shakudo
Bengaluru or Toronto or San Francisco
OnsiteFull Time
Responsibilities
  • deploying software
  • troubleshooting infrastructure
  • designing deployments
Requirements
  • 5+ years in DevOps/Platform/SRE with strong Kubernetes and Helm expertise
  • Experience with AWS/GCP/Azure and Terraform
  • Customer deployments experience, and familiarity with Python/Go/Bash/TypeScript
  • Strong communication skills
Technical tools mentioned
KubernetesHelmAWSGCPAzureTerraformPythonGoBashTypeScriptArgo CDFluxGitOps

Job description

At Shakudo, we're building the world's first operating system for data and AI. We use the term "operating system" in the truest sense: just like iOS, Windows, or Linux, Shakudo's end-to-end OS provides ever-evolving, fully automated, best-in-class open-source components tailored to each business's unique needs.

We are seeking a Senior DevOps Engineer to join our Engineering team and take ownership of deploying, configuring, and operating Shakudo in customer environments. This is a hands-on infrastructure role for someone who can work across Kubernetes, Helm charts, cloud and on-premise environments, and act as a trusted technical advisor to customers — diagnosing problems, designing deployment architectures, and ensuring Shakudo runs reliably in production.

In this role, you will own the deployment lifecycle from architecture to operations: assessing customer infrastructure, deploying Shakudo into complex environments, resolving production issues, and turning recurring problems into product improvements. This is not a traditional internal DevOps role — it is a mix of DevOps engineering, Kubernetes platform engineering, and solution architecture where success is measured by deployment reliability, customer satisfaction, and operational excellence.



At Shakudo, we're building the world's first operating system for data and AI. We use the term "operating system" in the truest sense: just like iOS, Windows, or Linux, Shakudo's end-to-end OS provides ever-evolving, fully automated, best-in-class open-source components tailored to each business's unique needs.

We are seeking a Senior DevOps Engineer to join our Engineering team and take ownership of deploying, configuring, and operating Shakudo in customer environments. This is a hands-on infrastructure role for someone who can work across Kubernetes, Helm charts, cloud and on-premise environments, and act as a trusted technical advisor to customers — diagnosing problems, designing deployment architectures, and ensuring Shakudo runs reliably in production.

In this role, you will own the deployment lifecycle from architecture to operations: assessing customer infrastructure, deploying Shakudo into complex environments, resolving production issues, and turning recurring problems into product improvements. This is not a traditional internal DevOps role — it is a mix of DevOps engineering, Kubernetes platform engineering, and solution architecture where success is measured by deployment reliability, customer satisfaction, and operational excellence.



At Shakudo, we're building the world's first operating system for data and AI. We use the term "operating system" in the truest sense: just like iOS, Windows, or Linux, Shakudo's end-to-end OS provides ever-evolving, fully automated, best-in-class open-source components tailored to each business's unique needs.
We are seeking a Senior DevOps Engineer to join our Engineering team and take ownership of deploying, configuring, and operating Shakudo in customer environments. This is a hands-on infrastructure role for someone who can work across Kubernetes, Helm charts, cloud and on-premise environments, and act as a trusted technical advisor to customers — diagnosing problems, designing deployment architectures, and ensuring Shakudo runs reliably in production.
In this role, you will own the deployment lifecycle from architecture to operations: assessing customer infrastructure, deploying Shakudo into complex environments, resolving production issues, and turning recurring problems into product improvements. This is not a traditional internal DevOps role — it is a mix of DevOps engineering, Kubernetes platform engineering, and solution architecture where success is measured by deployment reliability, customer satisfaction, and operational excellence.


Responsibilities
  • Own the deployment and operation of Shakudo across customer Kubernetes environments
  • Design, develop, customize, and troubleshoot Helm charts for complex production deployments
  • Work deeply with Kubernetes primitives including deployments, stateful sets, services, ingress, storage classes, secrets, config maps, RBAC, network policies, CRDs, and operators
  • Debug Kubernetes issues across scheduling, networking, storage, permissions, DNS, ingress, certificates, and workload reliability
  • Build repeatable deployment patterns that work across different customer infrastructure environments
  • Assess customer infrastructure and recommend the right deployment architecture for Shakudo
  • Work with customer platform, DevOps, security, and infrastructure teams to deploy Shakudo into their environments
  • Support deployments across AWS, GCP, Azure, hybrid cloud, and on-premise Kubernetes clusters
  • Design for enterprise constraints such as private networking, IAM/RBAC, security controls, observability, compliance requirements, and restricted environments
  • Help customers make the right trade-offs across reliability, scalability, performance, cost, and operational complexity
  • Build and maintain infrastructure-as-code using tools such as Terraform and related cloud-native tooling
  • Operate cloud managed services that interface with Shakudo Kubernetes clusters, including databases, storage, networking, secrets, and identity services
  • Support GPU infrastructure and specialized compute environments for data and AI workloads
  • Improve deployment automation, release processes, upgrade workflows, monitoring, and operational runbooks
  • Identify recurring deployment issues and turn them into product improvements, automation, or reusable patterns
  • Monitor, debug, and resolve production issues in customer environments
  • Lead root-cause analysis for infrastructure, deployment, and platform reliability issues
  • Execute product upgrades, maintenance windows, rollouts, and customer-specific configuration changes
  • Improve observability, alerting, logging, and operational visibility across deployments
  • Ensure customer environments are stable, secure, scalable, and maintainable
  • Act as a trusted technical advisor to customers during deployment and production operations
  • Explain infrastructure decisions clearly to both technical and non-technical stakeholders
  • Collaborate with Solution Engineering, Product Engineering, and Customer Engineering teams to translate customer requirements into robust deployment architectures
  • Document deployment designs, customer-specific configurations, best practices, and troubleshooting guides
  • Represent the voice of the customer internally and influence product and platform improvements


  • Qualifications
  • 5+ years of experience in DevOps, Platform Engineering, Infrastructure Engineering, SRE, or a related role
  • Strong hands-on experience with Kubernetes in production environments
  • Strong hands-on experience developing, maintaining, and troubleshooting Helm charts
  • Experience deploying and operating software in customer or enterprise environments
  • Experience with cloud platforms such as AWS, GCP, or Azure
  • Experience with infrastructure-as-code tools such as Terraform
  • Strong understanding of Kubernetes networking, storage, ingress, RBAC, secrets management, observability, and cluster operations
  • Ability to troubleshoot complex infrastructure issues across application, Kubernetes, cloud, and network layers
  • Familiarity with Python, Go, Bash, or TypeScript for automation and tooling
  • Strong communication skills and comfort working directly with customer technical teams
  • Ability to operate independently, make sound technical decisions, and drive deployments to completion


  • A Plus
  • Experience with data platforms, AI infrastructure, MLOps, or GPU workloads
  • Experience with Kubernetes operators, CRDs, GitOps, Argo CD, Flux, or similar deployment tooling
  • Experience with enterprise security requirements, private networking, identity providers, SSO, and compliance-driven environments
  • Experience deploying software into air-gapped, restricted, or customer-managed infrastructure
  • Prior experience in a customer-facing infrastructure, solution engineering, or solution architecture role
  • Contributions to open-source Kubernetes, DevOps, or infrastructure projects


  • Why Shakudo Stands Out

    • Work with cutting-edge technologies in machine learning and high-performance computing
    • Contribute to a platform that transforms how organizations leverage data and AI
    • Join a dynamic team that values innovation, efficiency, and diversity

     

    This is a work from office role based out of Bangalore (HSR Layout). Shakudo has offices in Toronto, San Francisco, and Bangalore.

    About Shakudo

    Develops an operating system for enterprise AI applications.

    Year founded
    2021
    Employees
    40
    Organization type
    Private
    Latest investment
    Raised $7.00M Series A (2025) — led by Wittington Ventures
    Headquarters
    CA

    Similar jobs

    DevOps Engineer roles near Bengaluru, Karnataka
    19h
    Save
    Mark Applied
    Hide
    Senior Engineer, DevOps - R01569948
    Bangalore, Karnataka, India
    HybridContract
    Brillio
    Brillio: Provides digital transformation and big data analytics consulting services.
    8+ YOERequires 8+ years in DevOps or platform automation, 5+ years building enterprise CI/CD frameworks, 3+ years with Terraform and Azure automation, and strong scripting, Git, security, and release governance skills.
    Microsoft Azure, Azure Database, Azure Analytics Services, Azure Data Catalog, Azure Identity, Azure Developer Tools, Azure Migration, Azure Integration, Azure Platform Administration (PaaS), Azure Event Hub, Azure Data Platform (DaaS), Azure Administration (IaaS), Kusto, Azure DevOps, GitHub Actions, Jenkins, GitLab CI, Terraform, Git, Bash, PowerShell, Python, OpenShift, Kubernetes, SAST
    22h
    Save
    Mark Applied
    Hide
    Lead Engineer - DevOps
    Bengaluru, Karnataka, India
    OnsiteFull Time
    Standard Chartered
    Standard CharteredLondon Stock Exchange: STAN: International group providing corporate, retail, and investment banking services.
    Engineering degree in computer science or information technology required. Requires networking, Linux, microservices, DevOps, infrastructure as code, automation, and stakeholder management experience; master's or MBA preferred.
    Azure DevOps, Kubernetes, Helm, Kibana, ELK, GitHub, Docker, Confluence, Linux
    1d
    Save
    Mark Applied
    Hide
    DevOps Engineer
    Bengaluru, Karnataka, India
    OnsiteFull Time
    Accenture
    AccentureNYSE: ACN: Global professional services firm providing consulting and technology solutions.
    3+ YOERequires 3+ years of ServiceNow ITSM experience, 15 years of full-time education, and knowledge of CI/CD, cloud platforms, container orchestration, infrastructure as code, automation, and security practices.
    ServiceNow IT Service Management (ITSM)
    1d
    Save
    Mark Applied
    Hide
    DevOps Engineer III
    Bangalore, Karnataka, India
    HybridFull Time
    MRI Software
    MRI Software: Provides property management and investment software for real estate.
    5+ YOE5+ years in DevOps or SRE, Azure/PaaS expertise, Docker and Kubernetes, infrastructure as code, CI/CD, Linux, cloud security, monitoring, SRE methods, documentation, and a relevant bachelor's or master's degree.
    Microsoft Azure, Terraform, Ansible, Docker, Kubernetes, Jenkins, GitHub Actions, Azure Pipelines, Unix, Linux, Git, GitHub, Infrastructure as Code (IaC), PowerShell, Bash, Azure DevOps Services
    1d
    Save
    Mark Applied
    Hide
    DevOps Engineer
    Bengaluru, Karnataka, India
    OnsiteFull Time
    Accenture
    AccentureNYSE: ACN: Global provider of management consulting and technology services.
    3+ YOERequires 3+ years in ServiceNow ITSM, 15 years of full-time education, CI/CD, cloud, container orchestration, infrastructure automation, and security knowledge.
    ServiceNow IT Service Management (ITSM), CI/CD
    1d
    Save
    Mark Applied
    Hide
    Lead I - DevOps Engineering
    Bengaluru, Karnataka, India
    OnsiteFull Time
    UST
    UST: Global provider of digital transformation and IT services.
    5+ YOERequires 5+ years of DevOps or cloud engineering experience, AWS, Terraform, GitLab CI/CD, EKS, Kubernetes, Docker, networking, Argo Rollouts, Python, and a BTech, MTech, or similar qualification.
    Helm, Amazon CloudWatch, Argo CD, Prometheus, Grafana, AWS, Terraform, Amazon EKS, Python, GitLab CI/CD, Kubernetes, Docker, AWS Transit Gateway, Argo Rollouts, AWS Step Functions, VPC
    2d
    Save
    Mark Applied
    Hide
    DevOps Engineer
    Bangalore, Karnataka, India
    OnsiteFull Time
    2K
    2KNASDAQ: TTWO: Publishes and develops global video game franchises and entertainment.
    4+ YOERequires 4+ years in DevOps, SRE, or platform engineering; cloud infrastructure, Kubernetes, IaC, CI/CD, observability, incident response, troubleshooting, and communication experience.
    AWS, GCP, Kubernetes, EKS, Terraform, GitHub Actions, Jenkins, Datadog, Splunk, Grafana, Jira, ServiceNow, ArgoCD, Python, Bash, Go
    2d
    Save
    Mark Applied
    Hide
    Devops Engineer
    Bengaluru, Karnataka, India
    HybridFull Time
    Arrive
    Arrive: Global mobility platform providing smart parking and transportation solutions.
    Hands-on GitHub Actions, scripting in Bash, Python, or Go, Kubernetes, ArgoCD, Helm Charts, and Sealed Secrets experience; strong independent problem-solving, communication, documentation, and collaboration skills.
    GitHub, GitHub Actions, Bash, Python, Go, Kubernetes, ArgoCD, Helm Charts, Umbrella Charts, Sealed Secrets, Agentic AI