TE

Senior Site Reliability Engineer

Tencent

Singapore, SGonsitePosted Jul 23, 2026
Posting intelligenceActively listed

Skills

prometheusgrafanapythonazureredismysqlawsgo

About the role

Business Unit Technology Engineering Group (TEG) is responsible for supporting the company and its business groups on technology and operational platforms, as well as the construction and operation of R&D management and data centers, TEG provides users with a full range of customer services. As the operator of the largest networking, devices, and data center in Asia,TEG also leads the Tencent Technology Committee in strengthening infrastructure R&D through internal and distributed open source collaboration, constructing new platforms and supporting business innovation.

What the Role Entails

1. Responsible for the operation and maintenance of overseas model services at Hunyuan, ensuring stable, reliable, and efficient service operations;

2. Responsible for capacity management and planning, resource cost optimization, ensuring reasonable online service capacity and improving resource efficiency;

3. Responsible for continuous integration and delivery, efficient and automated operational optimization, enhancing service stability and research and development efficiency;

4. Participate in the design of online systems and various service architectures, providing professional solutions for stability and architecture improvement;

5. Analyze and deeply explore the shortcomings of existing systems, data-driven to find weak points, and promote system optimization implementation and improvement;

6. Pay attention to industry front-end technology trends, explore technologies and directions for automation and intelligence in the operation and maintenance of complex business systems.

Who We Look For

1. Bachelor's degree or above, with 2 years or more experience in internet operations and maintenance;

2. Familiar with Linux operating system, with solid system management and network knowledge;

3. Familiar with deploying, configuring, and tuning components such as Nginx, Redis, MySQL;

4. Proficient in monitoring systems such as Zabbix, Prometheus, Grafana, real-time grasping the running status of overseas systems;

5. Proficient in at least one programming language (such as Python, Go, Shell, etc.), with experience in developing automated operational tools to meet the needs of complex and variable overseas operations and maintenance;

6. Familiar with mainstream public cloud operations and maintenance management overseas (such as AWS, Azure, etc.), with experience in containerization and microservices architecture, able to cope with the characteristics and differences of local cloud services;

7. Strong sense of work responsibility, good communication skills, learning ability, and team spirit;

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for DevOps / SRE roles in Singapore varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for Singapore medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.