Talent.com
This job offer is not available in your country.
Site Reliability Engineer

Site Reliability Engineer

GlintsKuala Lumpur, Malaysia
25 days ago
Job description

Overview

Recruitment Consultant at Glints Singapore Responsibilities

Monitor and maintain system performance to ensure the stability and reliability of applications and infrastructure. Design and implement resilient system architectures that support high availability and scalability. Develop automation tools and scripts to enhance operational efficiency and reduce manual effort. Define, track, and analyze SLOs and SLIs to ensure reliability and performance meet business needs. Conduct thorough post-mortem analyses following incidents, driving continuous improvement through root cause identification and solution implementation. Collaborate with development and operations teams to establish best practices in system reliability and incident management. Troubleshoot and resolve issues related to database performance, network connectivity, and deployment failures, including diagnosing problems at the underlying platform level. Ensure that issues are resolved within the stipulated Service Level Agreements (SLAs), maintaining high standards of service delivery. Identify and troubleshoot performance bottlenecks in applications and infrastructure, providing actionable recommendations for enhancements. Maintain detailed documentation of processes and incident responses to support knowledge sharing and compliance. Improve monitoring solutions to proactively identify and mitigate issues before they impact services. Assist in the deployment and configuration of new applications and services, ensuring adherence to best practices. Participate in on-call rotations and respond to critical incidents as they arise. Analyze system logs and metrics to identify trends and potential areas for improvement. Qualifications & Details

Seniority level : Mid-Senior level Employment type : Contract Job function : Information Technology Industry : Information Technology - IT Services and IT Consulting

#J-18808-Ljbffr

Create a job alert for this search

Reliability Engineer • Kuala Lumpur, Malaysia