Senior / Expert Engineer, Site Reliability Engineering
Singapore · Full Time
Be the first to apply
- Experience
- 3+ yrs
- Salary
- —
- Openings
- 1
- Posted
- 5 hours ago
- Work mode
- In office
- Education
- Bachelor's Degree in Computer Science or related discipline
- Resume
- Required to apply
Where you'll work
Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.
Job description
About Shared Game Services (SGS)
Shared Game Services offers fundamental platform functionalities supporting all Garena-published games. These include critical services such as authentication systems, payment processing, and mobile SDKs. Our team is composed of dedicated product managers and software engineers who strive to deliver scalable, robust solutions capable of handling high traffic volumes while providing reliable and localized user experiences.
Role Overview
The Senior / Expert Site Reliability Engineer will engage deeply with the development lifecycle to understand and optimize each application component. The role involves ensuring product scalability, enhancing system stability, and boosting overall performance. Responsibilities include managing middleware, product and big data applications, conducting server deployments and tuning, creating automation to streamline workflows, and overseeing capacity and resource allocation.
Key Responsibilities
- Analyze and comprehend development workflows and application mechanisms to drive scalability, stability, and peak performance.
- Set up, operate, and maintain middleware, product, and big data services.
- Perform scheduled and unscheduled deployments, address server performance issues, and resolve technical problems effectively.
- Develop automation tools to improve operational workflows.
- Manage capacity planning and resource distribution.
- Conduct comprehensive stress tests across the system to identify bottlenecks and eliminate redundancies.
- Document regular operational procedures thoroughly.
Requirements
- Bachelor's degree or higher in Computer Science, Engineering, Information Systems, or relevant disciplines.
- Minimum of three years full-time experience specifically as a Site Reliability Engineer.
- Proficient with Linux OS such as Ubuntu and CentOS, with practical, hands-on experience.
- Proficient with Kubernetes and its related ecosystem in a hands-on capacity.
- Solid understanding of networking protocols like TCP/IP and DNS, and operating systems.
- Programming experience in at least one language: Bash scripting, Python, or Go.
- Strong analytical capabilities paired with excellent problem-solving skills, especially under pressure.
- Quick learner who works well collaboratively within teams.
- Highly attentive to detail, careful, and prudent in approach.
Minimum education
Bachelor's Degree