Site Reliability Engineer
Kuala Lumpur, Malaysia
DG Hero is a recruitment platform that connects digital talents in IT, product management, design, data, marketing, and sales with leading companies.
We review all applications and only shortlisted candidates will be contacted for further review. Complete details about the hiring company will be shared with shortlisted candidates. Apply today!
Site Reliability Engineer Job Descriptions
As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure.
This role emphasizes strong system architecture and design principles, focusing on key SRE practices such as Service Level Objectives (SLOs), Service Level Indicators (SLis), and the reduction of operational toil. You will collaborate closely with diverse teams to drive reliability improvements and foster a culture of continuous learning and accountability.
Key Responsibilities:
Design and implement resilient system architectures that support high availability and scalability.
Develop automation tools and scripts to enhance operational efficiency and reduce manual effort.
Define, track, and analyze SLOs and SLis to ensure reliability and performance meet business needs.
Conduct thorough post-mortem analyses following incidents, driving continuous improvement through root cause identification and solution implementation.
Collaborate with development and operations teams to establish best practices in system reliability and incident management.
Troubleshoot and resolve issues related to database performance, network connectivity, and deployment failures, including diagnosing problems at the underlying platform level (e.g., Kubernetes, virtual machines).
Ensure that issues are resolved within the stipulated Service Level Agreements (SLAs), maintaining high standards of service delivery.
Identify and troubleshoot performance bottlenecks across systems, providing actionable recommendations for enhancements.
Maintain detailed documentation of processes and incident responses to support knowledge sharing and compliance.
Qualifications:
Open to Malaysian only
Mandarin speaker is a must
Proficiency in programming languages such as Python, Golang, Java, or similar, focusing on operational efficiency.
Demonstrated experience in system architecture and design, prioritizing reliability, and scalability.
Strong understanding of SRE principles, including SLOs, SLis, toil reduction, and incident post-mortems.
Experience with cloud environments (e.g., AWS, Azure, Google Cloud) and their operational management.
Strong expertise in Linux system administration.
Proven experience in troubleshooting application support issues with a focus on performance and connectivity.
Familiarity with networking concepts and effective troubleshooting techniques.
Excellent problem-solving abilities and a proactive approach to operational challenges.
Ability to work independently while effectively collaborating within a team environment.
Preferred Skills:
Familiarity with monitoring tools and performance optimization techniques.
Experience in scripting or automation for system administration tasks.
Knowledge of networking concepts and troubleshooting methodologies.
Hands-on knowledge of cloud platforms (e.g., AWS, Azure, Google Cloud) and their services.
Familiarity with DevOps practices and frameworks, including CI/CD, infrastructure as code, and containerization.
Site Reliability Engineer Skills
Python, Golang, Java, SLOs, SLis, AWS, Azure, Google Cloud, DevOps, CI/CD, Service Level Objectives, Service Level Indicators
Site Reliability Engineer Salary
Average salary for Site Reliability Engineer ranges from RM 6,000 to RM 17,000 per month






How It Works
Stop Searching
Stop the endless job searching & apply. Instead, get interview invitations from companies who are interested in you.
Job Matching
Based on your profile, we match you with relevant jobs. Your contact information are hidden until you accept their interview invitations.
Always Free
On DG Hero, everything is free for our talents. We charge the companies when they found talents they like.


Intelligent Job Matching
DG Hero's intelligent job matching enables everyone to land better jobs without having to waste time searching & applying for jobs.
​
We match you with relevant jobs while protecting your privacy. It's a better way for everyone to land jobs easily!


Your Privacy Is Our Priority
We only match you to relevant jobs & companies, preventing irrelevant companies to search and view your personal information.
​
We never reveal your contact information including email address, phone number & address to anyone, unless with your permission, usually when you agree to attend an interview with the company.