Back to search
Tencent Linkedin · Posted 2d ago

IEG - SRE (3rd Party Contract -1 Year Renewable)

Federal Territory of Kuala Lumpur

Linkedin
Continue to application Add your email once, then Caio opens the original posting.

Indexed description

About Job

About Tencent

Tencent is an Internet-based platform company founded in Shenzhen, China, in 1998. We use technology to enrich the lives of Internet users and assist the digital upgrade of enterprises. Our mission is "Value for Users, Tech for Good". We embrace a culture of teamwork & creativity and are driven by our values - Integrity, Proactivity, Collaboration and Creativity.

We are rapidly expanding our international operations and are looking for top talent to propel us forward. Combining the results-oriented nature of a start-up with the resources of a profitable and leading Internet company, Tencent offers a unique opportunity for aspiring individuals to thrive.


Business Unit Introduction - Interactive Entertainment Group (IEG)

Responsible for the R&D, operation, and development of the company's interactive entertainment business including games and eSports. Through online gaming, live broadcasts, and offline eSports, IEG assists the company in leading the global interactive entertainment market to create better interactive entertainment content experiences for users.


Responsibilities

  • Responsible for daily SRE operations, including CI/CD, capacity planning, system monitoring, alerting, incident response, and troubleshooting.
  • Ensure high system availability, stability, performance, and security while continuously identifying opportunities to optimize infrastructure and reduce computing costs.
  • Design and develop automation tools and solutions to improve operational efficiency, system reliability, and engineering productivity.
  • Drive standardization and automation of operational processes, reducing manual effort and improving overall service quality.
  • Continuously improve operational SOPs, technical documentation, and troubleshooting guides, while promoting knowledge sharing across teams.
  • Collaborate closely with development and infrastructure teams to identify potential reliability risks and implement proactive solutions.
  • Participate in on-call rotations and respond to critical incidents when required to maintain system availability and business continuity.

Requirements

  • Bachelor’s degree or above in Computer Science, Information Technology, or a related field.
  • At least 2 years of relevant experience in SRE, DevOps, Cloud Infrastructure, System Operations, or a related field.
  • Strong understanding of Linux, TCP/IP, Kubernetes, databases, SQL, Shell scripting, and Python.
  • Hands-on experience with at least one major cloud platform, such as AWS, Tencent Cloud, or Microsoft Azure.
  • Experience with CI/CD pipelines, system monitoring, alerting, troubleshooting, and infrastructure automation.
  • Familiarity with AI-assisted development and troubleshooting tools, such as Claude Code and Codex, with the ability to leverage AI tools to improve development and operational efficiency.
  • Good command of both Chinese and English, with the ability to communicate effectively with cross-functional and technical teams.
  • Strong sense of ownership, problem-solving ability, and initiative, with a proactive mindset toward improving system reliability and operational efficiency.
  • Willingness and ability to participate in 24/7 on-call support for critical incidents when required.
Free. 20 seconds. No password. See every match in this search.

Create a free Caio profile to unlock more results and save your role and location preferences.

Unlock free search
Want help applying to roles like this? Search Caio for free. If repetitive applications get heavy, Managed Job Search adds supervised execution for $99/month.
View Managed Job Search