Site Reliability Engineer

5 days ago


Kuala Lumpur, Kuala Lumpur, Malaysia PERSOL APAC Full time

About the company:

We have partnered with a renowned global leader in information and communications technology (ICT) infrastructure and smart devices. They are providing full-stack, all-scenario solution for products and services carriers, enterprises, governments, and individual consumers worldwide.

Our client is looking for enthusiastic
Site Reliability Engineer
to contribute your talents in a wide range of scenarios and applications.

Responsibilities:

  • To be responsible for reliability, availability, user experience, capacity planning, toil reduction, process enhancement and digitalization of the cloud-based internet services.
  • Handle SRE role for assigned cloud services owning the KPIs for reliability, issue to resolution, service deployment, business continuity management, security policy planning, capacity planning, toil reduction through automation.
  • Introduce service governance initiatives based on latest technologies to consistently increase reliability and user experience components of mobile services on cloud to provide world class user experience with high reliability.
  • Effectively utilize our world class AIOPS and autonomous service governance platform to ideate new ways to streamline process, accuracy of alerts, time series-based trend analysis, anomaly detection, risk identifications.
  • Support platform/service expansions, migrations to new architectures, upgrades and drill activities across different technology domains.
  • Incorporate mature chaos engineering for risk identification, IPDRR for security, comprehensive automation frameworks to reduce ops effort to reach lowest possible level and make time, space for engineering related focus for the team.

Requirements and qualifications:

  • Bachelor/Master of computer science engineering or related majors
  • Have knowledge of Linux, Network, Database, Containers, Container management systems, etc.
  • Have knowledge of at least one programming language or scripting such as Java, Python, Shell, Ansible, Terraform
  • Have knowledge in big data analytics.
  • Explored new technology trends, opensource technologies, methodologies in internet service domain.

Interested candidates, who wish to apply for the advertised position, please click on "Apply Now". We regret that only shortlisted candidates will be notified.



  • Kuala Lumpur, Kuala Lumpur, Malaysia Kneat Full time

    Site Reliability Engineer – Kuala Lumpur, MalaysiaKneat enables regulated organizations to move from paper-based validation to intelligent, digitized, paperless solutions. And we do it through the ongoing development of a powerful, purpose-built software platform. In 2014, after eight years of intensive software development, we launched Kneat Gx—the...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Kneat Full time

    Site Reliability Engineer – Kuala Lumpur, MalaysiaKneat enables regulated organizations to move from paper-based validation to intelligent, digitized, paperless solutions. And we do it through the ongoing development of a powerful, purpose-built software platform. In 2014, after eight years of intensive software development, we launched Kneat Gx—the...


  • Kuala Lumpur, Kuala Lumpur, Malaysia VCB Malaysia Berhad Full time 144,000 - 156,000 per year

    Overview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...


  • Kuala Lumpur, Kuala Lumpur, Malaysia PeopleScope Full time 60,000 - 120,000 per year

    Site Reliability EngineerJob Description:Ability to debug scripts and automate routine tasks in OS, network, database or application servers. Coding experience beyond simple scripts; Experience in Devops process, programming knowledge in at least one of the following languages: Java, Python, or Go; Scripting skills in at least of the following:...


  • Kuala Lumpur, Kuala Lumpur, Malaysia FPT Software Malaysia Sdn. Bhd. Full time

    Key Responsibilities:Disaster Recovery Planning (DRP):Design and maintain scalable failover systems, backup strategies, and redundancy mechanisms across cloud and on-prem environments.Develop and update DR documentation, runbooks, and recovery playbooks for infrastructure and application layers.Business Continuity Testing:Plan, coordinate, and execute...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Abhidi Solution Private Limited Full time 120,000 - 180,000 per year

    Job Title: Site Reliability Engineer (SRE)Job Type: Permanent positionWork Location Kuala LumpurResponsibilities:Strong hands-on experience with VMware solutionsStrong experience with patch management for OS & middlewareExperience in VMware server templating/blueprints (RedHat & Windows)Experience with Infrastructure-as-Code, orchestration, configuration...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Encora Full time

    About the RoleWe are seeking a Site Reliability Engineer (SRE) to support and manage our Infrastructure-as-a-Service (IaaS) platform built on the VMware Cloud Foundation (VCF) stack. You will be responsible for maintaining, automating, and optimizing the VCF environment to ensure high availability, scalability, and operational efficiency. This role is ideal...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Encora Full time

    We are looking for a Site Reliability Engineer (SRE) to support and manage our Infrastructure-as-a-Service (IaaS) platform built on the VMware Cloud Foundation (VCF) stack. The role involves maintaining, automating, and optimizing the VCF environment, ensuring high availability, scalability, and operational efficiencyDay-to-Day ResponsibilitiesManage,...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Hunters International Full time 19,000 per year

    Overview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Guidewire Software Full time $100,000 - $200,000 per year

    SummaryWe are searching for a Senior Site Reliability Engineer hungry for a rare chance to transform insurance with the industry's leading cloud platformJob Description*Senior Site Reliability Engineer - Cloud ApplicationThe Opportunity*At Guidewire, growth isn't just financial, it's technical. With Q3 FY2025 revenue up 22% year-over-year and a rapidly...