Talent.com
Site Reliability Engineer (SRE) | KL

Site Reliability Engineer (SRE) | KL

Hunters International Sdn BhdKuala Lumpur, Kuala Lumpur, Malaysia
20 hours ago
Job description

About the job Site Reliability Engineer (SRE) | KL

Overview :

As a Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE practices such as Service Level Objectives (SLOs), Service Level Indicators (SLIs), and the reduction of operational toil. You will collaborate closely with diverse teams to drive reliability improvements and foster a culture of

continuous learning and accountability.

Responsibilities

  • Design and implement resilient system architectures that support high availability and scalability.
  • Develop automation tools and scripts to enhance operational efficiency and reduce manual effort.
  • Define, track, and analyze SLOs and SLIs to ensure reliability and performance meet business needs.
  • Conduct thorough post-mortem analyses following incidents, driving continuousimprovement through root cause identification and solution implementation.
  • Collaborate with development and operations teams to establish best practices in system reliability and incident management.
  • Troubleshoot and resolve issues related to database performance, network connectivity, and deployment failures, including diagnosing problems at the underlying platform level (e.g., Kubernetes, virtual machines).
  • Ensure that issues are resolved within the stipulated Service Level Agreements (SLAs),maintaining high standards of service delivery.
  • Identify and troubleshoot performance bottlenecks across systems, providing actionable recommendations for enhancements.
  • Maintain detailed documentation of processes and incident responses to support knowledge sharing and compliance.

Requirements :

  • Proficiency in Mandarin is a must, in order to liaise with stakeholders from China.
  • Proficiency in programming languages such as Python, Golang, Java, or similar, focusing on operational efficiency.
  • Demonstrated experience in system architecture and design, prioritizing reliability, and scalability.
  • Strong understanding of SRE principles , including SLOs, SLIs, toil reduction, and incident post-mortems.
  • Experience with cloud environments (e.g., AWS, Azure, Google Cloud) and their operational management.
  • Strong expertise in Linux system administration.
  • Proven experience in troubleshooting application support issues with a focus on performance and connectivity.
  • Familiarity with networking concepts and effective troubleshooting techniques.
  • Excellent problem-solving abilities and a proactive approach to operational challenges.
  • Ability to work independently while effectively collaborating within a team environment.
  • Familiarity with monitoring tools and performance optimization techniques.
  • Experience in scripting or automation for system administration tasks.
  • Knowledge of networking concepts and troubleshooting methodologies.
  • Hands-on knowledge of cloud platforms (e.g., AWS, Azure, Google Cloud) and their services.
  • Familiarity with DevOps practices and frameworks, including CI / CD, infrastructure as code, and containerization.
  • Location :

    To be based in client's site (Klang Valley)

    Remuneration :

    Up to MYR 19,000 (Based on relevant experience)

    Consultant in Charge

    May Chong | |

    This is a contract position with the possibility to be absorbed as a permanent staff.

    #J-18808-Ljbffr

    Create a job alert for this search

    Reliability Engineer • Kuala Lumpur, Kuala Lumpur, Malaysia

    Related jobs
    • Promoted
    Site Reliability Engineer (SRE)

    Site Reliability Engineer (SRE)

    Refine GroupKuala Lumpur, Kuala Lumpur, Malaysia
    Design and maintain scalable failover systems, backup strategies, and redundancy mechanisms across cloud and on-prem environments. Create and update disaster recovery documentation, runbooks, and re...Show moreLast updated: 9 days ago
    • Promoted
    Site Reliability Engineer III

    Site Reliability Engineer III

    Guidewire SoftwareKuala Lumpur, Kuala Lumpur, Malaysia
    At Guidewire, we make software that offers Property and Casualty (P&C) Insurance companies the tools to take care of their customers when they need it the most, whether that’s a time of crisis, a n...Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    FINEXUS GroupKuala Lumpur, Kuala Lumpur, Malaysia
    Location : FINEXUS Group, Kuala Lumpur, Federal Territory of Kuala Lumpur, Malaysia.Ensure high availability and reliability of IT systems, applications, and PCI DSS‑certified data centers, supporti...Show moreLast updated: 9 days ago
    • Promoted
    Lead Site Reliability Engineer

    Lead Site Reliability Engineer

    SWIFTKuala Lumpur, Kuala Lumpur, Malaysia
    We’re the world’s leading provider of secure financial messaging services, headquartered in Belgium.We are the way the world moves value – across borders, through cities and overseas.No other organ...Show moreLast updated: 29 days ago
    Site Reliability Engineer

    Site Reliability Engineer

    Unison GroupKuala Lumpur, Federal Territory of Kuala Lumpur, MY
    Quick Apply
    As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and op...Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Russell TobinKuala Lumpur, Kuala Lumpur, Malaysia
    Job Opportunity : Site Reliability Engineer (SRE) in Cyberjaya.Note : Only Malaysian locals or PR holders can apply.We are looking for a Site Reliability Engineer (SRE) to join our forward-thinking C...Show moreLast updated: 26 days ago
    • Promoted
    Site Reliability Engineer (SRE) / Devops Engineer

    Site Reliability Engineer (SRE) / Devops Engineer

    Unison Consulting Pte LtdKuala Lumpur, Kuala Lumpur, Malaysia
    As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and op...Show moreLast updated: 30+ days ago
    • Promoted
    Specialist, Site Reliability Engineer (SRE)

    Specialist, Site Reliability Engineer (SRE)

    TNG DigitalKuala Lumpur, Kuala Lumpur, Malaysia
    Specialist, Site Reliability Engineer (SRE).We are hiring for a Specialist, Site Reliability Engineer (SRE) to join our team. Role focuses on network administration, cloud infrastructure management,...Show moreLast updated: 23 days ago
    • Promoted
    • New!
    Site Reliability Engineer

    Site Reliability Engineer

    Encora Inc.Kuala Lumpur, Kuala Lumpur, Malaysia
    Kuala Lumpur, Federal Territory of Kuala Lumpur, Malaysia.Encora is a global digital engineering company specializing in AI, Cloud, and Data solutions to help enterprises become agile and adaptable...Show moreLast updated: 20 hours ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Tata Consultancy ServicesKuala Lumpur, Kuala Lumpur, Malaysia
    Talent Acquisition | Human Resource Executive | Tata Consultancy Service.Join Tata Consultancy Services, Asia Pacific and be part of an organization committed to sustainable development for our fut...Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Unison Consulting Pte LtdKuala Lumpur, Kuala Lumpur, Malaysia
    As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and op...Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Smart Teq Solution Sdn BhdKuala Lumpur, Kuala Lumpur, Malaysia
    Ensure all our infrastructure are running at optimal condition.Provide deployment, patches and update on all services that running on public cloud and on premise. Identify and resolve support ticket...Show moreLast updated: 30+ days ago
    • Promoted
    Senior Site Reliability Engineer (SRE)

    Senior Site Reliability Engineer (SRE)

    Ryt BankKuala Lumpur, Kuala Lumpur, Malaysia
    Senior Talent Acquisition Specialist @ Ryt Bank.We are Ryt Bank, a joint venture between YTL and the SEA Group, proudly awarded as one of the five digital banking license winners by BNM in Malaysia...Show moreLast updated: 30+ days ago
    • Promoted
    Lead Site Reliability Engineer

    Lead Site Reliability Engineer

    SwiftKuala Lumpur, Kuala Lumpur, Malaysia
    We’re the world’s leading provider of secure financial messaging services, headquartered in Belgium.We are the way the world moves value – across borders, through cities and overseas.No other organ...Show moreLast updated: 15 days ago
    • Promoted
    Site Reliability Engineer - Kuala Lumpur, Malaysia

    Site Reliability Engineer - Kuala Lumpur, Malaysia

    Kneat SolutionsKuala Lumpur, Kuala Lumpur, Malaysia
    Site Reliability Engineer – Kuala Lumpur, Malaysia.Kneat enables regulated organizations to move from paper-based validation to intelligent, digitized, paperless solutions.And we do it through the ...Show moreLast updated: 7 days ago
    • Promoted
    Lead Site Reliability Engineer

    Lead Site Reliability Engineer

    Swift SoftwareKuala Lumpur, Kuala Lumpur, Malaysia
    Lead Site Reliability Engineer page is loaded## Lead Site Reliability Engineerlocations : Kuala Lumpur, Malaysiatime type : Full timeposted on : Posted Todayjob requisition id : We’re the worl...Show moreLast updated: 30+ days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    Razer Inc.Kuala Lumpur, Kuala Lumpur, Malaysia
    Bangsar South, Federal Territory of Kuala Lumpur, Malaysia.Joining Razer will place you on a global mission to revolutionize the way the world games. Razer is a place to do great work, offering you ...Show moreLast updated: 9 days ago
    • Promoted
    Site Reliability Engineer

    Site Reliability Engineer

    iSoftStoneKuala Lumpur, Kuala Lumpur, Malaysia
    SoftStone – Federal Territory of Kuala Lumpur, Malaysia.A leading global technology group, renowned for its extensive ecosystem of digital services and platforms. With a strong presence in cloud com...Show moreLast updated: 30+ days ago