Jobs Companies Rakuten Senior Site Reliability (DevOps) Engineer

Über diese Senior Site Reliability (DevOps) Engineer Stelle bei Rakuten

Rakuten · Vor Ort · Singapore

Job Description:

Rakuten Group, Inc. is a global leader in internet services and has a diverse ecosystem spanning across e-commerce, fintech, communications and more serving approximately 1.8 billion members worldwide. Founded in Tokyo in 1997, the Group operates in over 30 countries and regions with more than 30,000 employees.
 

Based in Singapore's Central Business District, Rakuten Asia Pte. Ltd. serves as the regional headquarters for Asia, driving value through areas such as advertising product development, product strategy, and data management to support Rakuten Group's global ecosystem. Learn more at: https://global.rakuten.com/corp/

The Marketing Cloud Platform Department (MCPD) drives Rakuten's marketing product strategy, executes product development, and ensures successful implementation. We empower Rakuten's internal marketing teams by creating engaging, respectful, and cost-efficient marketing platforms that prioritize our customers. Leveraging the Rakuten Ecosystem, we offer comprehensive marketing solutions, including campaign management, multichannel communication, and personalization. As a team of over 150 experts across Japan, India, and Singapore, we pride ourselves on being a technology-driven organization that shares knowledge within the Rakuten Tech community.

As an Senior Site Reliability (DevOps) Engineer in MCPD, you will drive operational excellence by implementing best practices in observability, incident management, and automation. This role bridges engineering and operations, requiring both strong technical expertise and people management skills to build and maintain highly available systems that serve millions of Rakuten's customers globally.

Main Responsibilities:

  • Define and drive SRE strategy, including SLO/SLI frameworks, error budgets, and reliability targets aligned with business objectives and customer expectations

  • Establish and improve incident management processes, including on-call rotations, escalation procedures, and blameless post-mortem practices to minimize MTTR and prevent recurring issues

  • Collaborate with development teams to embed reliability practices into the software development lifecycle, advocating for design reviews, chaos engineering, and production readiness reviews

  • Design and implement comprehensive observability solutions (monitoring, logging, tracing, alerting) to provide actionable insights into system health and performance

  • Drive automation initiatives to reduce toil, improve deployment reliability, and enable self-service capabilities for engineering teams

  • Partner with Architecture and Platform teams to ensure infrastructure decisions support scalability, fault tolerance, and cost optimization goals

  • Manage capacity planning and performance optimization for critical marketing platforms handling high-volume campaign executions and real-time personalization

  • Report on reliability metrics, incident trends, and operational health to leadership, translating technical insights into business impact assessments

Required Qualifications:

  • 8+ years of experience in software engineering, DevOps, or site reliability engineering, with at least 3 years in a people management role

  • Proven track record of building and leading high-performing SRE or platform engineering teams in a distributed, multi-timezone environment

  • Deep expertise in cloud platforms (GCP preferred, AWS/Azure acceptable) including compute, networking, storage, and managed services

  • Strong knowledge of containerization and orchestration technologies (Kubernetes, Docker) and Infrastructure as Code (Terraform, Ansible)

  • Hands-on experience with observability tools and practices (Prometheus, Grafana, Datadog, ELK Stack, or similar) and defining meaningful SLOs/SLIs

  • Experience with CI/CD pipelines, deployment strategies (blue-green, canary), and release engineering best practices

  • Strong programming/scripting skills in languages such as Python, Go, or Java for automation and tooling development

  • Excellent communication skills with the ability to collaborate effectively across engineering, product, and business stakeholders

  • Strong incident management experience with demonstrated ability to lead high-pressure situations calmly and effectively

Rakuten is an equal opportunities employer and welcomes applications regardless of sex, marital status, ethnic origin, sexual orientation, religious belief, or age.

Bereit, sich bei Rakuten zu bewerben?
Bei Rakuten bewerben

Über Rakuten

In Japanese, Rakuten stands for ‘optimism.’ It means we believe in the future. It’s an understanding that, with the right mind-set, we can make the future better by what we do today. So we challenge ourselves to evolve, innovate and experiment, to create a better, brighter future for everyone. Today, our 70+ businesses span e-commerce, digital content, communications and fintech, bringing the joy of discovery to almost 1.3 billion members across the world. If you have any trouble logging in, please contact us here Rakuten Group, Inc.: [email protected] *Please read the Application Requirements(EN) / 募集要項(JP) before applying. Our Diversity & Inclusion Policy and Applica

Alle Jobs bei Rakuten ansehen →

Ähnliche Jobs

Riot Games
Principal Software Engineer - DevOps / Site Reliability Engineer
Riot Games
⚡ Früh bewerben Singapore Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 3 Tg.
Riot Games
Principal Software Engineer - ML Platform Engineer
Riot Games
⚡ Früh bewerben Singapore Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 3 Tg.
FWD Group
Senior Manager, DevOps & Production Support Engineer
FWD Group
⚡ Früh bewerben Hong Kong - Taikoo Shing (Grou... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 3 Tg.
Mastercard
Senior Platform Engineer
Mastercard
⚡ Früh bewerben Singapore Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 4 Tg.
Lightning AI
AI Platform Support Engineer (APAC)
Lightning AI
⚡ Früh bewerben Philippines; Singapore Hybrid
● Neu 👁 Gesehen ✓ Beworben vor 4 Tg.
OKX
Senior/Staff Engineer - Exchange Middle Platform
OKX
⚡ Früh bewerben Singapore, Singapore Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 4 Tg.
Workato
Senior Software Engineer (Golang - Platform team)
Workato
⚡ Früh bewerben Singapore Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 6 Tg.
Schonfeld
Platform Support Engineer
Schonfeld
⚡ Früh bewerben Hong Kong, Hong Kong; Singapor... Vor Ort
● Neu 👁 Gesehen ✓ Beworben vor 6 Tg.
BlackRock
Analyst, DevOps Engineer – Service Management Group
BlackRock
⚡ Früh bewerben Singapore, Singapore Hybrid
● Neu 👁 Gesehen ✓ Beworben vor 1 Wo.

Registrieren für Vorschläge, die auf die von Ihnen geöffneten Jobs und gespeicherten Suchen zugeschnitten sind.

Mehr Jobs bei Rakuten

Alle Jobs bei Rakuten ansehen →

Jetzt bewerben
🤖

Moment — langsam

JobsRadar wurde für echte Menschen gebaut, die eine schwere Zeit bei der Jobsuche haben — nicht für automatisierte Anfragen. Sie klicken viel zu schnell und sind jetzt vorübergehend blockiert.

Kommen Sie später wieder. Wenn Sie wirklich auf Jobsuche sind, stehen wir hinter Ihnen — verhalten Sie sich einfach wie ein Mensch.

Catch your next role the second it’s posted.

Create a free account and we’ll watch the boards for you — the instant a job matches your search, it lands in your inbox or Telegram. No digging, no refreshing.

Create free account

Free forever · takes 30 seconds · already have one?

Verschaffe dir einen Vorsprung bei der Jobsuche.

Tritt unserem Telegram-Kanal bei für das, was dir hilft, die Stelle zu bekommen — Gehaltsbenchmarks, den wöchentlichen Marktpuls und neue Feature-Drops. Kein Spam, nur Signal.

Dem Kanal beitreten — kostenlos