Site Reliability Engineer (Bilingual - Mandarin) Job in CDA

Site Reliability Engineer (bilingual Mandarin)

Kuala Lumpur, M14, MY, Malaysia

Apply Now

Job Description

Site Reliability Engineer (SRE)

Overview

As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE practices such as Service Level Objectives (SLOs), Service Level Indicators (SLIs), and the reduction of operational toil. You will collaborate closely with diverse teams to drive reliability improvements and foster a culture of continuous learning and accountability.

Qualifications

Proficiency in Mandarin to effectively communicate with Mandarin-speaking clients and stakeholders. Strong programming skills in Python, Golang, Java, or similar, with a focus on operational efficiency. Experience with Bash/Shell scripting or automation for system administration tasks. Demonstrated expertise in system architecture and design, prioritizing reliability and scalability. Solid understanding of SRE principles: SLOs, SLIs, toil reduction, and incident post-mortems. Hands-on experience with cloud environments (AWS, Azure, Google Cloud). Strong expertise in Linux system administration. Proven experience in application support troubleshooting, focusing on performance and connectivity. Familiarity with networking concepts and effective troubleshooting techniques. Excellent problem-solving skills with a proactive approach to operational challenges. Ability to work independently while collaborating effectively within a team. Willingness to work in rotational shifts, with schedules communicated in advance.

Key Responsibilities

Monitor and maintain system performance to ensure stability and reliability of applications and infrastructure. Design and implement resilient system architectures supporting high availability and scalability. Develop automation tools and scripts to reduce manual effort and enhance efficiency. Define, track, and analyze SLOs and SLIs to align reliability with business needs. Conduct post-mortem analyses for incidents, identifying root causes and driving improvements. Collaborate with development and operations teams to establish best practices in reliability and incident management. Troubleshoot and resolve issues in databases, networking, deployments, and platforms (e.g., Kubernetes, virtual machines). Ensure resolution of issues within defined SLAs, maintaining service delivery standards. Identify and resolve performance bottlenecks in applications and infrastructure. Maintain detailed documentation of processes, incidents, and resolutions. Enhance monitoring solutions to detect and mitigate issues proactively. Assist in deployment and configuration of applications and services. Participate in on-call rotations, responding to critical incidents as needed. Analyze system logs and metrics to identify trends and potential improvements.

Preferred Skills

Familiarity with monitoring tools and performance optimization techniques. Knowledge of DevOps practices and frameworks, including: CI/CD pipelines Infrastructure as Code (IaC) Containerization (e.g., Docker, Kubernetes)
Job Types: Full-time, Permanent

Benefits:

Opportunities for promotion Professional development
Work Location: In person

Beware of fraud agents! do not pay money to get a job

MNCJobz.com will not be responsible for any payment made to a third-party. All Terms of Use are applicable.

Related Jobs

Site Reliability Engineer Based in Johor Bahru

Arvion Services

Johor Bahru, Johor

Apply Now
Site Reliability Engineer Based in Johor Bahru

Arvion Services

Johor Bahru, Johor

Apply Now

T&T Senior Consultant Innovation & Cloud Development Centre (Site Reliability Engineer) MY

Deloitte

Kuala Lumpur

Apply Now
R

Site Reliability Engineer (DevOps Consultant) Mandarin

Rationalz S&S (India) LLP

Kuala Lumpur - Johor Bahru, Johor

Apply Now

Job Detail

Job Id

JD1162654
Industry

Not mentioned
Total Positions

1
Job Type:

Full Time
Salary:

Not mentioned
Employment Status

Permanent
Job Location

Kuala Lumpur, M14, MY, Malaysia
Education

Not mentioned

Jobs by Function

Popular Job Skills

Popular Industries

Popular Cities

Jobseekers

Employers