This job has expired

This job posting is no longer active and is not accepting applications. Explore similar roles below!

Sony Interactive Entertainment
Posted 2w ago

Platform Support Site Reliability Engineer (SRE)

Sony Interactive Entertainment
Koto-ku, Tokyo, Japan
¥5600k-¥9300k/yrOnsiteFull Time
Responsibilities
  • automating operations
  • supporting engineers
  • improving reliability
Requirements
  • Experience designing
  • Building, or operating public-cloud services
  • Troubleshooting systems across infrastructure and applications, and automating operations with AI or scripts. Business-level Japanese and English communication required
Technical tools mentioned
TerraformCloudFormationInfrastructure as Code (IaC)PythonGoJavaScriptAPM

Job description

 

Why PlayStation?

PlayStation isn’t just the Best Place to Play — it’s also the Best Place to Work. Today, we’re recognized as a global leader in entertainment producing The PlayStation family of products and services including PlayStation®5, PlayStation®4, PlayStation®VR, PlayStation®Plus, acclaimed PlayStation software titles from PlayStation Studios, and more.

PlayStation also strives to create an inclusive environment that empowers employees and embraces diversity. We welcome and encourage everyone who has a passion and curiosity for innovation, technology, and play to explore our open positions and join our growing global team.

The PlayStation brand falls under Sony Interactive Entertainment, a wholly-owned subsidiary of Sony Group Corporation.

 

 

■職務内容 / Job Responsibilities

Platform Supportチームは、PlayStationのオンラインサービスを支えるコンテナ、ネットワーク、認証、監視などの共通基盤を対象に、プラットフォーム開発チームが提供する共通基盤の運用を担いながら、AIを活用した標準化・自動化を推進し、運用モデルそのものを継続的に改善しています。
また、共通基盤全体に対する深い知識と運用経験を活かし、共通基盤に不慣れな開発チームを支援し、PlayStation全体でベストプラクティスを実践できるようEnablementも行います。共通基盤を利用する開発者がより安全かつ効率的に開発・運用できる環境を提供することも重要なミッションです。

本ポジションでは、Platform Support Engineerとして、AIを活用した運用自動化やワークフロー改善、開発者体験(Developer Experience)の向上、さらには共通基盤の利用促進・標準化に取り組んでいただきます。
アプリケーション開発チーム、プラットフォーム開発チーム、運用チームと密接に連携しながら、よりスケーラブルで信頼性が高く、運用しやすいプラットフォームを実現するための仕組みづくりをリードしていただきます。
AIも活用して運用を変革したい方、複雑な技術課題をシンプルな仕組みに落とし込みたい方、そして組織全体の生産性向上やDeveloper Experienceの改善に情熱を持つエンジニアにとって、大きな裁量とインパクトを持って活躍できるポジションです。

【主な業務】

  • 共通基盤の運用を担い、定常業務の標準化・自動化、および運用ワークフローの継続的な改善を推進 
  • 新たに発生した運用業務や高度な判断を伴う運用業務は自ら実施しつつ、AIを活用した自動化によって再現可能な運用プロセスへ進化させたうえで、TOS(Technical Operations Support)などの実行チームへ移管し、スケーラブルな運用モデルを構築
  • 共通基盤を利用する社内エンジニアリングチームへの技術支援、課題調査、ベストプラクティスの展開
  • アプリケーション開発チームやプラットフォーム開発チームと連携し、信頼性・運用性・DeveloperExperienceを向上させるソリューションの企画・実装
  • USをはじめとするグローバルチームと協力し、Follow the Sun体制の整備や運用プロセスの共通化を通じて、日本に閉じないオペレーションモデルを構築


■組織・職場紹介 / About Team and Organization

PlayStationのオンラインサービスは世界中で1億人を超えるユーザーに利用されており、ゲーム、コマース、ソーシャル機能をはじめとするPlayStationエコシステム全体を支えています。
Platform Supportチームは、PlayStationのオンラインサービスを支えるコンテナ、ネットワーク、認証、監視などの共通基盤を対象に、プラットフォーム開発チームが提供する共通基盤の運用を担いながら、AIを活用した標準化・自動化を推進し、運用モデルそのものを継続的に改善しています。
私たちは、日本だけでなく、USやインドを含むグローバルメンバーと日常的に協力しながら、Follow the Sunを意識した運用体制や共通プロセスの整備を進めています。拠点を越えて知見を共有し、世界共通のベストプラクティスを作り上げることも、このチームの重要なミッションです。
オープンなコミュニケーション、継続的な学習、実践的な課題解決を大切にしながら、社内エンジニアのDeveloper Experienceと、世界中のPlayStationプレイヤーの体験を継続的に改善していくことを目指しています。

 


■求められるスキル・経験 / Required Skills

以下のいずれか一つ以上(MUST):技術領域

  • パブリッククラウド環境におけるサービスの設計・構築・運用経験
  • システム、インフラ、ネットワーク、アプリケーションにまたがる障害調査・トラブルシューティングの経験
  • AIやスクリプトを活用し、運用業務や定型業務の自動化・効率化を推進した経験

必須(MUST):コミュニケーション・協業

  • ビジネスレベルの日本語コミュニケーション力
  • 英語でグローバルチームと協働できるコミュニケーション力、マインドセット(文書・口頭)
  • 社内のエンジニアリングユーザーおよび部門横断の関係者と円滑に連携できる対人スキル

歓迎(WANT)

  • Observabilityプラットフォーム、監視ソリューション、APMツールの利用経験
  • Terraform、CloudFormation等を用いたInfrastructure as Code(IaC)の設計・実装経験
  • Python、Go、JavaScriptまたは類似言語を用いたソフトウェア、自動化、運用ツールの開発経験
  • 大規模分散システムの運用経験
  • クラウドネットワーク、DNS、トラフィック管理、証明書管理、ID/アクセス管理、構成管理、または関連するプラットフォーム技術の経験
  • 規制やコンプライアンス要件が重視される環境でのサービス運用経験
  • Business Process Outsourcing(BPO)、オフショア、Shared Serviceなどを活用し、業務を標準化・整理した上で他チームへ移管した経験
  • 技術資格、技術コミュニティ、オープンソースプロジェクト、その他関連活動への貢献実績
  • コンピューターサイエンス、ソフトウェアエンジニアリング、または関連分野の学士号、もしくは同等の実務経験

【求める人物像】

  • 課題や改善機会を主体的に見つけ、自らオーナーシップを持って解決を推進できる方
  • システム改善、運用の複雑性低減、反復的な手作業の削減にやりがいを感じる方
  • 課題分析やソリューション評価において、論理的かつ体系的に考えられる方
  • 異なる意見を尊重しながら、自身の提案や懸念を明確に伝えられる方
  • 社内ユーザー、同僚、部門横断のチームと良好な関係を築ける方
  • 新しい技術を素早く学び、変化する優先順位にも柔軟に対応できる方
  • 運用に対する強い当事者意識、信頼性へのこだわり、継続的な改善意欲を持つ方

■このポジションの魅力 / Why Join This Team

世界規模のオンラインサービスを支える技術に携わりながら、世界中のプレイヤーにより良い体験を届けるために、社内エンジニアリングチームを支援できるポジションです。
国際的なチームと協働し、プラットフォームの信頼性とオペレーショナルエクセレンスに貢献しながら、PlayStationの未来を支えるツール、プロセス、自動化の整備・改善に携わっていただきます。

 

 

Equal Opportunity Statement:

Sony is an Equal Opportunity Employer. All persons will receive consideration for employment without regard to gender (including gender identity, gender expression and gender reassignment), race (including colour, nationality, ethnic or national origin), religion or belief, marital or civil partnership status, disability, age, sexual orientation, pregnancy or maternity, trade union membership or membership in any other legally protected category.

We strive to create an inclusive environment, empower employees and embrace diversity. We encourage everyone to respond.

PlayStation is a Fair Chance employer and qualified applicants with arrest and conviction records will be considered for employment.

 

 

About Sony Interactive Entertainment

Develops and publishes PlayStation gaming hardware, software, and services.

Similar jobs

Site Reliability Engineer roles near Koto-ku, Tokyo
5d
Save
Mark Applied
Hide
Principal Site Reliability Engineer
London or Singapore or Tokyo or Houston or Boston
HybridFull Time
Veson Nautical
Veson Nautical: Develops enterprise software for global maritime freight management.
5+ YOEBachelor's degree or equivalent experience; 5+ years of GCP experience, production Kubernetes/GKE, Terraform, cloud networking, and Python, Go, or TypeScript programming skills.
Google Cloud Platform, Bigtable, Cloud SQL, Dataflow, Datastore, Google Kubernetes Engine (GKE), Google Cloud Storage (GCS), Google Cloud Key Management Service (KMS), Pub/Sub, Amazon Web Services, Kubernetes, Amazon Elastic Kubernetes Service (EKS), Terraform, Terragrunt, Atlantis, GitLab Pipelines, ArgoCD, Octopus Deploy, ElasticSearch, Kubernetes Operator, PostgreSQL, SQL Server, BigQuery, Splunk, Grafana, Grafana Tempo, OpenTelemetry, Cloud Armor Enterprise, OpsGenie, Renovate, Sentry, Claude, Amazon Bedrock, Gemini, Vertex AI, Python, Go, TypeScript, GitLab CI
6d
Save
Mark Applied
Hide
Site Reliability Engineer
Tokyo, Tokyo, Japan
HybridFull Time
IFS
IFS: Provides enterprise software for asset-intensive and service industries.
Requires cloud platform, Docker, Kubernetes, Linux/Unix, Windows Server, Azure networking, Oracle DB, MS SQL Server, ITIL, ServiceNow, and Jira experience, plus fluent English and Japanese.
Microsoft Azure, Google Cloud Platform (GCP), Amazon Web Services (AWS), Docker, Azure Kubernetes Service (AKS), Linux/Unix, Windows Server, Azure VPN, Azure ExpressRoute, Cloud Service Routers, Oracle Database, Microsoft SQL Server, ITIL, ServiceNow, Jira Service Management
1w
Save
Mark Applied
Hide
Site Reliability Engineer
Hyderabad or New York City or Chicago or London or Singapore or Tokyo or Hong Kong or Europe or United States or Asia-Pacific
HybridFull Time
Pico
Pico: Provides managed infrastructure and data services to financial markets.
Bachelor's degree or relevant experience; financial markets technology experience; Linux, networking, computer architecture, programming or scripting, customer service, communication, and collaborative teamwork skills.
Linux, Python, C, C++, Java
3w
Save
Mark Applied
Hide
Senior Lead Site Reliability Engineer, Electronic Colo Trading
Tokyo, Tokyo, Japan
OnsiteFull Time
JPMorgan Chase
JPMorgan ChaseNYSE: JPM: Global financial services firm providing banking and investment solutions.
5+ YOE5+ years Linux production administration and SRE experience, proficiency with Bash/Python/Go, strong performance tuning and incident management, bachelor's degree in CS or related, SRE certification preferred.
Bash, Python, Go, RHEL, Debian, Ubuntu
1mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer, Booking Services Search Group - Search Department (SED)
Tokyo, Tokyo, Japan
OnsiteFull Time
Rakuten Group
Rakuten GroupTokyo Stock Exchange: 4755: Provides online retail, banking, and telecommunications services globally.
8+ YOE8+ years IT experience; strong Linux, distributed systems and networking skills; experience with configuration management, observability, containers, automation and Java; excellent troubleshooting and collaboration skills.
Chef, Ansible, Prometheus, Grafana, Loki, Kubernetes, Shell, Python, Jenkins, Spark, Solr, Cassandra, Kafka, Ceph, S3, MinIO, Java, Git, Linux
1mo
Save
Mark Applied
Hide
Site Reliability Engineer
Santa Clara or St. Louis or Bangalore or London or Paris or Melbourne or Taipei or Tokyo
OnsiteFull Time
Netskope
NetskopeNASDAQ: NTSK: Cloud-native cybersecurity and data protection platform for enterprises.
3+ YOEBachelor's in CS/Engineering or equivalent; 3+ years building/managing complex systems (including 1-2 years SRE); experience with cloud services, microservices, availability/performance optimization, debugging, and strong communication.
Python, C, C++, Go, Rust, Docker, Kubernetes, AWS, GCP, KVM, OpenNebula, OpenStack, TCP/IP
2mo
Save
Mark Applied
Hide
Senior Site Reliability Engineer
New York City or Austin or Berlin or Bucharest or Chicago or Dubai or Jakarta or London or Paris or San Francisco or São Paulo or Singapore or Seoul or Sydney or Tokyo
HybridFull Time
Braze
BrazeNASDAQ: BRZE: Platform for personalized customer engagement and cross-channel messaging.
3+ YOE3+ years as a Software/DevOps/Site Reliability Engineer, strong Linux/Unix shell skills, programming experience in Ruby and/or Go, experience with Docker, Kubernetes, Terraform/Chef, and data stores like MongoDB, Redis, Kafka, or Postgres.
Ruby on Rails, Ruby, Go, Linux, Unix Shell, Docker, Kubernetes, Terraform, Chef, MongoDB, Redis, Kafka, Postgres, PagerDuty
2mo
Save
Mark Applied
Hide
Site Reliability Engineer, Siri Evaluation Reliability
Minato, Tokyo, Japan
OnsiteFull Time
Apple
AppleNASDAQ: AAPL: Designs and sells consumer electronics, software, and online services.
Experience in site reliability for distributed systems, resource management, session orchestration, on-call response, and observability for ML evaluation infrastructure.
This job has expired