Padmi
Oracle logo
Oracle

Cloud Infrastructure (OCI) · AI Database

Site Reliability Developer 5

JapanPosted 1 month ago
Software engineeringUnspecified
Apply at Oracle

Opens the source posting on eeho.fa.us2.oraclecloud.com

Source description

About the role

View original

This position requires deep expertise in distributed systems, cloud infrastructure, and software engineering, combined with the ability to collaborate effectively with senior engineering leaders across global OCI organizations. You will align operational practices with JP Sovereign Cloud, EU Sovereign Cloud, and global OCI reliability teams, and drive standardization where it improves service resiliency. The role includes participation in a 24x7 operational support model while serving as a key escalation point for high-severity incidents and strategic reliability improvements. You will sponsor improvements raised from shift operations, ensure recurring issues are addressed through durable fixes, and mentor senior engineers on Plan + Execution ownership. Qualifications

  • Native-level Japanese language proficiency and business-level English communication skills
  • 8+ years of experience in Site Reliability Engineering, Cloud Infrastructure Engineering, Software Development, or large-scale distributed systems operations
  • Extensive experience designing, operating, and improving highly available cloud platforms and mission-critical services
  • Expert-level proficiency in software development and automation using languages such as Java, Go, Python, or similar
  • Deep understanding of distributed systems architecture, networking, storage, observability, and service resiliency principles
  • Proven track record leading major incident response efforts, reliability programs, and cross-organizational technical initiatives
  • Ability to influence architecture, operational standards, and engineering best practices across multiple teams
  • Willingness to participate in a 24x7 shift rotation and act as a senior escalation resource for critical production events
  • Proven ability to define reliability strategy and operational standards across multiple teams or services
  • Experience translating business and operational requirements into prioritized reliability roadmaps with measurable outcomes
  • Ability to drive cross-sovereign collaboration, including shared operational practices and tooling alignment across JP Sovereign Cloud and EU Sovereign Cloud

One address, no account. We’ll tell you when matching roles go live.

More at Oracle

Related open roles

View all roles
Site Reliability Developer 5 at Oracle · Padmi