Site Reliability Engineer
5 days ago
We are looking for a Site Reliability Engineer (SRE) to support and manage our Infrastructure-as-a-Service (IaaS) platform built on the VMware Cloud Foundation (VCF) stack. The role involves maintaining, automating, and optimizing the VCF environment, ensuring high availability, scalability, and operational efficiency
Day-to-Day Responsibilities
- Manage, operate, and optimize VMware Cloud Foundation (VCF) environments.
- Configure and maintain vCenter, ESXi clusters, and vRealize Suite components (vROps, vRA, vRLI, vRNI).
- Develop and maintain automation scripts and playbooks using Ansible and PowerCLI to streamline operations and deployments.
- Monitor system health, capacity, performance, and proactively troubleshoot infrastructure issues.
- Collaborate with Cloud, DevOps, Security, and Infrastructure teams to ensure platform reliability and compliance.
- Implement and maintain configuration management, system upgrades, and patching processes.
- Support Windows and Linux virtual machine environments, including lifecycle management and troubleshooting
Must-Have Skills & Experience
- The ideal candidate will have a strong foundation in IT infrastructure and virtualization, with proven hands-on experience managing enterprise-grade VMware environments.
- Overall IT Experience: 5–7 years of professional experience in IT infrastructure, systems, or virtualization.
- VMware Cloud Foundation (VCF): 2–3 years of experience managing and operating VCF environments, including SDDC Manager, vCenter, NSX, and vSAN integration.
- vCenter & ESXi: 3–5 years of experience in the installation, configuration, and management of vSphere environments, including HA, DRS, resource allocation, and troubleshooting.
- vRealize Suite (vROps, vRA, vRLI, vRNI): 2+ years of experience deploying, monitoring, and automating infrastructure using vRealize tools.
- Automation – Ansible: 2–3 years of experience building and maintaining automation playbooks for provisioning and configuration of infrastructure.
- Automation – PowerCLI / PowerShell: 2+ years of experience developing scripts for VMware automation and task orchestration.
- Operating Systems (Windows & Linux): 2+ years of experience performing administration, configuration, and troubleshooting of virtual machines.
- Monitoring & Troubleshooting: 3+ years of experience identifying, analyzing, and resolving issues in virtualized environments, ensuring high availability and performance.
About Encora
Encora is a global company that offers Software and Digital Engineering solutions. Our practices include Cloud Services, Product Engineering & Application Modernization, Data & Analytics, Digital Experience & Design Services, DevSecOps, Cybersecurity, Quality Engineering, AI & LLM Engineering, among others.
At Encora, we hire professionals based solely on their skills and do not discriminate based on age, disability, religion, gender, sexual orientation, socioeconomic status, or nationality
-
Site Reliability Engineer
3 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Kneat Full timeSite Reliability Engineer – Kuala Lumpur, MalaysiaKneat enables regulated organizations to move from paper-based validation to intelligent, digitized, paperless solutions. And we do it through the ongoing development of a powerful, purpose-built software platform. In 2014, after eight years of intensive software development, we launched Kneat Gx—the...
-
Site Reliability Engineer
3 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Kneat Full timeSite Reliability Engineer – Kuala Lumpur, MalaysiaKneat enables regulated organizations to move from paper-based validation to intelligent, digitized, paperless solutions. And we do it through the ongoing development of a powerful, purpose-built software platform. In 2014, after eight years of intensive software development, we launched Kneat Gx—the...
-
Site Reliability Engineer
5 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia PERSOL APAC Full timeAbout the company:We have partnered with a renowned global leader in information and communications technology (ICT) infrastructure and smart devices. They are providing full-stack, all-scenario solution for products and services carriers, enterprises, governments, and individual consumers worldwide.Our client is looking for enthusiasticSite Reliability...
-
Site Reliability Engineer
2 weeks ago
Kuala Lumpur, Kuala Lumpur, Malaysia VCB Malaysia Berhad Full time 144,000 - 156,000 per yearOverview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...
-
Site Reliability Engineer
2 weeks ago
Kuala Lumpur, Kuala Lumpur, Malaysia PeopleScope Full time 60,000 - 120,000 per yearSite Reliability EngineerJob Description:Ability to debug scripts and automate routine tasks in OS, network, database or application servers. Coding experience beyond simple scripts; Experience in Devops process, programming knowledge in at least one of the following languages: Java, Python, or Go; Scripting skills in at least of the following:...
-
Site Reliability Engineer
5 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia FPT Software Malaysia Sdn. Bhd. Full timeKey Responsibilities:Disaster Recovery Planning (DRP):Design and maintain scalable failover systems, backup strategies, and redundancy mechanisms across cloud and on-prem environments.Develop and update DR documentation, runbooks, and recovery playbooks for infrastructure and application layers.Business Continuity Testing:Plan, coordinate, and execute...
-
Site Reliability Engineer
2 weeks ago
Kuala Lumpur, Kuala Lumpur, Malaysia Abhidi Solution Private Limited Full time 120,000 - 180,000 per yearJob Title: Site Reliability Engineer (SRE)Job Type: Permanent positionWork Location Kuala LumpurResponsibilities:Strong hands-on experience with VMware solutionsStrong experience with patch management for OS & middlewareExperience in VMware server templating/blueprints (RedHat & Windows)Experience with Infrastructure-as-Code, orchestration, configuration...
-
Site Reliability Engineer
2 weeks ago
Kuala Lumpur, Kuala Lumpur, Malaysia Hunters International Full time 19,000 per yearOverview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...
-
Site Reliability Engineer
5 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Encora Full timeAbout the RoleWe are seeking a Site Reliability Engineer (SRE) to support and manage our Infrastructure-as-a-Service (IaaS) platform built on the VMware Cloud Foundation (VCF) stack. You will be responsible for maintaining, automating, and optimizing the VCF environment to ensure high availability, scalability, and operational efficiency. This role is ideal...
-
Senior Site Reliability Engineer
1 week ago
Kuala Lumpur, Kuala Lumpur, Malaysia Guidewire Software Full time $100,000 - $200,000 per yearSummaryWe are searching for a Senior Site Reliability Engineer hungry for a rare chance to transform insurance with the industry's leading cloud platformJob Description*Senior Site Reliability Engineer - Cloud ApplicationThe Opportunity*At Guidewire, growth isn't just financial, it's technical. With Q3 FY2025 revenue up 22% year-over-year and a rapidly...