Senior Site Reliability Engineer, Booking Services Search Group - Search Department (SED)

rakuten

Tokyo, JPonsitePosted Jul 10, 2026
Posting intelligenceActively listedReposted 2×, possible evergreen/ghost posting

Skills

kubernetesprometheuscassandrajenkinsansiblegrafanapythonsparkkafkareactjava

About the role

Job Description:

Business Overview

Rakuten is one of the biggest marketplaces of Japan and is the largest internet ecosystem with a wide range of services ranging from e-Commerce, Travel, Banking, Fintech, Food Delivery, Golf, Insurance, Instant Messaging, Mobile Network etc. Rakuten has over 140 services globally. Our mission is to empower people and society through the internet while aiming at becoming the Global Innovation Company.

Department Overview

Search Department (SED) is part of the AI Engineering Supervisory Department (AIESD) under AI & Data Division (AIDD) in Rakuten. Search Department focuses on Search, Discovery, and Navigation experience for users of Rakuten. We in Search Department help demand meet supply through our services that are engineered for scale, performance, and ease of use. We support 20 plus Rakuten Businesses across three different continents, 50 plus user functions and over a billion active users annually generating over two Billion dollars in revenue directly from us. The Booking Service Search group is responsible for delivering mission critical Rakuten Travel and Leisure businesses.

Position:

Why We Hire

We are responsible for the reliability of large-scale distributed Search platform that empower one of the biggest e-commerce of Japan and enables Rakuten ecosystem services worldwide. Rakuten Search Department is constantly striving to challenge what existing search technologies can do and going beyond. We are looking for a highly motivated site reliability engineer with commitment to customer satisfaction and a strong desire to create highest-quality Search platform for Rakuten Services.

Position Details Responsibilities (includes but are not limited to): - Take part in the design and deployment of production environment for new clients, from hardware to service level - Actively monitor production system, including on-call - Quickly react to issues, report, triage, troubleshoot and assist peers - Handle operations like Production releases, including nighttime operations - Drive continuous improvement of our tooling and automation stacks through innovation and collaboration - Collaborate with product managers, software engineers, and quality assurance teams to identify testing requirements, design and conduct relevant plans (including performance, security, fire drills)

Mandatory Qualifications: - More than 8 years of work experience working in IT - Excellent problem solving and troubleshooting skills - Excellent teamwork and communication skills - Strong Linux experience with understanding of system performance and reliability - Experience analyzing and tuning performance of distributed systems - Experience deploying and using configuration management tools (e.g., Chef, Ansible) - Experience with observability tools (e.g., Prometheus, Grafana, Loki) - Experience with container technologies and orchestration (Kubernetes) - Experience automating tasks using shell scripts and/or Python - Strong understanding of computer networking and common protocols - Experience managing services built with Java

Desired Qualifications:

- Experience automating processes for software testing and deployment (Jenkins)

- Experience working with Spark, Solr, Cassandra, or Kafka

- Experience with cloud storage solutions (Ceph, S3, MinIO)

- Experience working with a Git-based workflow

- Software development experience

Other Information:

Additional information on English Qualification

English (Overall - 3 - Advanced)

#engineer #infrastructureengineer #applicationsengineer #aianddatadiv

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for DevOps / SRE roles in Japan varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for Japan medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.