This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

Oracle
Posted 1mo ago

Site Reliability Developer 4

Oracle
Japan
OnsiteFull Time
Responsibilities
  • designing automation
  • executing improvements
  • mentoring engineers
Requirements
  • Native Japanese and business-level English,5+ years SRE/Software/Cloud/DevOps experience,proficiency in Java/Python/Go/C++,strong Linux,cloud automation,observability,incident response and on-call experience
Technical tools mentioned
JavaPythonGoC++Linux

Job description

As a Senior Site Reliability Developer (IC4), you will play a key role in ensuring the availability, scalability, and operational excellence of OCI's Japan Sovereign Cloud services. You will design and implement automation, drive service reliability improvements, lead complex incident investigations, and partner with development teams to improve operational readiness. You will own and prioritize an SRD operational improvement backlog based on shift feedback, incident reviews, alert quality reviews, and business reliability requirements. 

 


 

Responsibilities

The role combines software engineering expertise with large-scale cloud operations and requires participation in a 24x7 shift rotation supporting critical cloud infrastructure. You will translate operational and business requirements into reliability plans, then execute improvements through tooling, automation, runbook updates, process changes, and cross-team coordination. You will also serve as a technical mentor for less experienced engineers and contribute to continuous improvement initiatives across the organization. You will collaborate with JP Sovereign Cloud and EU Sovereign Cloud teams to share operational practices and align reliability improvements where appropriate.

Qualifications

- Native-level Japanese language proficiency and business-level English communication skills
- 5+ years of experience in Site Reliability Engineering, Software Engineering, Cloud Infrastructure, DevOps, or related technical disciplines
- Proficiency in one or more programming languages such as Java, Python, Go, C++, or similar
- Experience with cloud platforms, infrastructure automation, observability, monitoring, and incident response practices
- Strong understanding of Linux systems administration, networking, storage, and performance optimization
- Demonstrated ability to troubleshoot complex cross-functional production issues and drive root cause analysis
- Ability to participate in a 24x7 shift rotation and provide technical leadership during critical service events
- Demonstrated ability to intake, triage, and prioritize operational issues raised by shift teams and convert them into executable improvement plans
- Experience improving alert quality, reducing alert noise, increasing actionability, and ensuring operational documentation supports timely incident response
- Ability to balance business requirements, technical feasibility, and operational risk when planning reliability improvements

Qualifications

Career Level - IC4

Company

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing [email protected] or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.

About Oracle

Provides cloud infrastructure and enterprise software for global businesses.

Similar jobs

Site Reliability Developer roles
6d
Save
Mark Applied
Hide
Site Reliability Engineer, Security Engineering
Sydney or Los Angeles or Singapore or New York City or London or Dublin or Paris or Berlin or Dubai or Jakarta or Seoul or Tokyo
OnsiteFull Time
TikTok
TikTok: Short-form mobile video and social media platform.
3+ YOEBachelor's degree in computer science or related field, 3+ years relevant experience, programming in Go, Java, or Python, web framework experience, Linux and networking knowledge, Kubernetes and SRE tooling experience.
Go, Java, Python, Gin, Django, Spring, Linux, TCP/IP, HTTP, Kubernetes, Ansible, Argo CD, Prometheus, Grafana
1w
Save
Mark Applied
Hide
Site Reliability Engineer (SRE)
United States or Japan or Europe or Africa or Asia or North America or South America or Australia
RemoteFull Time
Social Discovery Group
Social Discovery Group: Private social-technology building online communication platforms and investing in social-discovery startups worldwide.
3+ YOE3+ years in SRE, DevOps, systems administration, or build/release engineering; strong Linux, Kubernetes, containers, CI/CD, IaC, monitoring, networking, and Git skills; fluent Russian required.
Linux, Kubernetes, Docker, Podman, GitLab CI, Ansible, Terraform, Prometheus, Grafana, Zabbix, VictoriaMetrics, Git, AWS, GCP, RabbitMQ, AMQP, Cloudflare, Akamai, WAF, CDN, DNS, HTTP, HTTPS
2w
Save
Mark Applied
Hide
Principal Site Reliability Engineer
London or Singapore or Tokyo or Houston or Boston
HybridFull Time
Veson Nautical
Veson Nautical: Private maritime software serving shipowners, charterers, traders, and operators with commercial freight management solutions.
5+ YOEBachelor's degree or equivalent experience; 5+ years of GCP experience, production Kubernetes/GKE, Terraform, cloud networking, and Python, Go, or TypeScript programming skills.
Google Cloud Platform, Bigtable, Cloud SQL, Dataflow, Datastore, Google Kubernetes Engine (GKE), Google Cloud Storage (GCS), Google Cloud Key Management Service (KMS), Pub/Sub, Amazon Web Services, Kubernetes, Amazon Elastic Kubernetes Service (EKS), Terraform, Terragrunt, Atlantis, GitLab Pipelines, ArgoCD, Octopus Deploy, ElasticSearch, Kubernetes Operator, PostgreSQL, SQL Server, BigQuery, Splunk, Grafana, Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry, Claude, Amazon Bedrock, Gemini, Vertex AI, Python, Go, TypeScript, GitLab CI
2w
Save
Mark Applied
Hide
Site Reliability Engineer
Tokyo, Tokyo, Japan
HybridFull Time
IFS
IFS: Global enterprise software provider specializing in ERP and EAM.
Requires cloud platform, Docker, Kubernetes, Linux/Unix, Windows Server, Azure networking, Oracle DB, MS SQL Server, ITIL, ServiceNow, and Jira experience, plus fluent English and Japanese.
Microsoft Azure, Google Cloud Platform (GCP), Amazon Web Services (AWS), Docker, Azure Kubernetes Service (AKS), Linux/Unix, Windows Server, Azure VPN, Azure ExpressRoute, Cloud Service Routers, Oracle Database, Microsoft SQL Server, ITIL, ServiceNow, Jira Service Management
3w
Save
Mark Applied
Hide
Site Reliability Engineer
Hyderabad or New York City or Chicago or London or Singapore or Tokyo or Hong Kong or Europe or United States or Asia-Pacific
HybridFull Time
Pico
Pico: Private financial-markets technology providing trading infrastructure, connectivity, market data, software, and analytics to institutions.
Bachelor's degree or relevant experience; financial markets technology experience; Linux, networking, computer architecture, programming or scripting, customer service, communication, and collaborative teamwork skills.
Linux, Python, C, C++, Java
1mo
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer, Electronic Colo Trading
Tokyo, Tokyo, Japan
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services and investment banking firm.
5+ YOE5+ years Linux production administration and SRE experience, proficiency with Bash/Python/Go, strong performance tuning and incident management, bachelor's degree in CS or related, SRE certification preferred.
Bash, Python, Go, RHEL, Debian, Ubuntu
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Booking Services Search Group - Search Department (SED)
Tokyo, Tokyo, Japan
OnsiteFull Time
Rakuten Group
Rakuten GroupTokyo Stock Exchange: 4755: Japanese technology conglomerate operating e-commerce, fintech, and mobile services.
8+ YOE8+ years IT experience; strong Linux, distributed systems and networking skills; experience with configuration management, observability, containers, automation and Java; excellent troubleshooting and collaboration skills.
Chef, Ansible, Prometheus, Grafana, Loki, Kubernetes, Shell, Python, Jenkins, Spark, Solr, Cassandra, Kafka, Ceph, S3, MinIO, Java, Git, Linux
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Customer engagement platform for cross-channel marketing and analytics.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
This job has expired