Oracle
Posted 2mo ago

Site Reliability Developer 3

Oracle
Japan
OnsiteFull Time
Responsibilities
  • automating operations
  • resolving incidents
  • improving reliability
Requirements
  • Bachelor's or equivalent experience
  • Native Japanese and business English
  • 2-5+ years SRE/Systems/Cloud/DevOps experience supporting Linux production
  • Knowledge of cloud
  • Networking
  • Distributed systems, and scripting (Python/Java/Go/Shell)
  • Willingness for 24x7 on-call/shift work
Technical tools mentioned
Oracle Cloud Infrastructure (OCI)LinuxPythonJavaGoShell

Job description

Oracle Cloud Infrastructure (OCI) is building the next generation of cloud services to support mission-critical workloads for customers across Japan. As a Site Reliability Developer (IC3), you will help operate and improve the reliability, scalability, and performance of the Japan Sovereign Cloud platform. Working closely with software engineering, cloud operations, and global OCI teams, you will leverage software engineering principles to automate operations, resolve complex production issues, and enhance service resiliency.

Responsibilities

This role includes an initial hands-on operational learning period to understand 24x7 shift workflows, alerts, incidents, escalation paths, runbooks, and customer-impacting reliability risks. This role requires participation in a 24x7 shift rotation and collaboration across both Japanese and international teams. You will partner with shift teams to capture recurring operational issues, improve alert actionability, maintain operational documentation, and contribute practical fixes through tooling, automation, and process improvements.
 

Qualifications

- Bachelor’s degree in computer science, Engineering, Information Technology, or equivalent practical experience
- Native-level Japanese language proficiency and business-level English communication skills
- 2+ years of experience in Site Reliability Engineering, Systems Engineering, Cloud Operations, DevOps, or Software Development and experience supporting Linux-based production environments
- Knowledge of cloud computing, networking, distributed systems, and automation technologies
- Experience with scripting or programming languages such as Python, Java, Go, Shell, or similar
- Willingness to participate in a 24x7 on-call and shift-based operational support model
- Ability to learn day-to-day sovereign cloud operations, follow shift procedures, and identify recurring operational pain points
- Ability to improve runbooks, alert response guidance, and operational handoff quality

Qualifications

Career Level - IC3

Company

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing [email protected] or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

About Oracle

Provides cloud infrastructure and enterprise software for global businesses.

Similar jobs

Site Reliability Developer roles
14h
Save
Mark Applied
Hide
Site Reliability Engineer Intern (Global SRE) - 2027 Summer
San Jose or Los Angeles or New York City or London or Dublin or Paris or Berlin or Dubai or Jakarta or Seoul or Tokyo
OnsiteInternship, Full Time
TikTok
TikTok: Global short-form video hosting and social media platform.
Currently pursuing a bachelor's degree in computer science or related field; Unix/Linux, IP networking, and Python, Go, C, C++, or Java programming experience required.
Unix/Linux, IP networking, Python, Go, C, C++, Java
6d
Save
Mark Applied
Hide
Principal Site Reliability Engineer
London or Singapore or Tokyo or Houston or Boston
HybridFull Time
Veson Nautical
Veson Nautical: Develops enterprise software for global maritime freight management.
5+ YOEBachelor's degree or equivalent experience; 5+ years of GCP experience, production Kubernetes/GKE, Terraform, cloud networking, and Python, Go, or TypeScript programming skills.
Google Cloud Platform, Bigtable, Cloud SQL, Dataflow, Datastore, Google Kubernetes Engine (GKE), Google Cloud Storage (GCS), Google Cloud Key Management Service (KMS), Pub/Sub, Amazon Web Services, Kubernetes, Amazon Elastic Kubernetes Service (EKS), Terraform, Terragrunt, Atlantis, GitLab Pipelines, ArgoCD, Octopus Deploy, ElasticSearch, Kubernetes Operator, PostgreSQL, SQL Server, BigQuery, Splunk, Grafana, Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry, Claude, Amazon Bedrock, Gemini, Vertex AI, Python, Go, TypeScript, GitLab CI
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Tokyo, Tokyo, Japan
HybridFull Time
IFS
IFS: Provides enterprise software for asset-intensive and service industries.
Requires cloud platform, Docker, Kubernetes, Linux/Unix, Windows Server, Azure networking, Oracle DB, MS SQL Server, ITIL, ServiceNow, and Jira experience, plus fluent English and Japanese.
Microsoft Azure, Google Cloud Platform (GCP), Amazon Web Services (AWS), Docker, Azure Kubernetes Service (AKS), Linux/Unix, Windows Server, Azure VPN, Azure ExpressRoute, Cloud Service Routers, Oracle Database, Microsoft SQL Server, ITIL, ServiceNow, Jira Service Management
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Hyderabad or New York City or Chicago or London or Singapore or Tokyo or Hong Kong or Europe or United States or Asia-Pacific
HybridFull Time
Pico
Pico: Provides managed infrastructure and data services to financial markets.
Bachelor's degree or relevant experience; financial markets technology experience; Linux, networking, computer architecture, programming or scripting, customer service, communication, and collaborative teamwork skills.
Linux, Python, C, C++, Java
3w
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer, Electronic Colo Trading
Tokyo, Tokyo, Japan
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years Linux production administration and SRE experience, proficiency with Bash/Python/Go, strong performance tuning and incident management, bachelor's degree in CS or related, SRE certification preferred.
Bash, Python, Go, RHEL, Debian, Ubuntu
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Booking Services Search Group - Search Department (SED)
Tokyo, Tokyo, Japan
OnsiteFull Time
Rakuten Group
Rakuten GroupTokyo Stock Exchange: 4755: Provides online retail, banking, and telecommunications services globally.
8+ YOE8+ years IT experience; strong Linux, distributed systems and networking skills; experience with configuration management, observability, containers, automation and Java; excellent troubleshooting and collaboration skills.
Chef, Ansible, Prometheus, Grafana, Loki, Kubernetes, Shell, Python, Jenkins, Spark, Solr, Cassandra, Kafka, Ceph, S3, MinIO, Java, Git, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara or St. Louis or Bangalore or London or Paris or Melbourne or Taipei or Tokyo
OnsiteFull Time
Netskope
NetskopeNASDAQ: NTSK: Cloud-native cybersecurity and data protection platform for enterprises.
3+ YOEBachelor's in CS/Engineering or equivalent; 3+ years building/managing complex systems (including 1-2 years SRE); experience with cloud services, microservices, availability/performance optimization, debugging, and strong communication.
Python, C, C++, Go, Rust, Docker, Kubernetes, AWS, GCP, KVM, OpenNebula, OpenStack, TCP/IP
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty