Specialist, Site Reliability Engineer

2 days ago


Kuala Lumpur, Kuala Lumpur, Malaysia TNG Digital Full time 90,000 - 120,000 per year

We fuel the ideas and ambitions of our people with an environment built on Our DNA of Love, Entrepreneurship, Agility, and Passion – LEAP

We are a culture that empowers everyone to innovate and create solutions that will leave a positive impact on our communities and our nation, Touch 'n Go will always be here to inspire our talents to grow as leaders and innovators giving you the power to make a difference.

What would you do?

Network administration

a) Design, implement and manage network infrastructure

b) Monitor network performance and ensure security compliance

c) Maintain and implement cross/multi cloud networking

d) Implement network segmentation and access control to improve security

e) Implement and maintain monitoring systems to proactively identify performance bottlenecks, security vulnerabilities, and other issues.

Cloud infrastructure management, capacity planning and monitoring

a) Maintain and optimize cloud-based infrastructure

b) Deploy, configure, and manage Linux and Windows servers using automation tools

c) Monitor and troubleshoot infrastructure performance, security, scalability

d) Assess system capacity and performance requirement, implement scalable solutions that meet future growth needs

Cloud FinOps

a) Monitor and analyze cloud spending across multi-cloud environments (AWS, Azure, Alibaba Cloud)

b) Work with finance teams to ensure accurate cloud budgeting, forecasting, and chargeback models

c) Optimize storage, networking, and computing costs without impacting performance

d) Leverage FinOps tools (AWS Cost Explorer, Azure Cost Management, Alibaba Cloud Cost Center, etc.)

Security and compliance

a) Collaborate with other SRE and Security team to manage and optimize infrastructure resources and ensure application of industry standard security

b) Collaborate with development teams to design and implement secure infrastructure architectures, ensuring the confidentiality, integrity, and availability of systems and data

c) Ensuring compliance with security, audits and regulatory requirements

d) Plan and execute disaster recovery plans.

Incident response and troubleshooting

a) Collaborate with different functional teams to address incidents and restore services quickly. Investigate and resolve incidents, apply root cause analysis to prevent recurrence

Who should join us?

  • Bachelor's degree in Computer Science, Network or related field
  • Professional cloud certification
  • Proven 5 experience in a Cloud Network or Cloud Infrastructure role
  • Strong experience in site reliability engineering, infrastructure engineering or a similar role.
  • Strong knowledge on network and protocols, network security and cloud networking
  • Proven strong record of cloud cost optimisation
  • Experience with cloud platforms (e.g., AWS, Azure, GCP, Alibaba Cloud) and infrastructure as code (IaC) tools (e.g., Terraform, CloudFormation). Basic or advanced cloud certification is a plus
  • Experience with containerization technologies like Docker and container orchestration platforms such as Kubernetes is a plus
  • Knowledge of networking principles and protocols
  • Deep knowledge of Linux/Unix systems and administration
  • Strong problem-solving skills and the ability to handle high-pressure situations calmly and effectively
  • Strong attention to detail and a commitment to delivering high-quality results

Our Perks & Benefits:

  • Flexi clock-in hours.
  • Monthly eWallet allowance.
  • Additional 1% employer EPF contribution from your 1st to 3rd year of service, with further increases based on your continued years of service.
  • Unlimited office pantry fruits, snacks and drinks.
  • Mobile and broadband subscription reimbursement.
  • Flexibility to opt dependants coverage (spouse, child, parents or parents-in-law) for outpatient medical benefits.
  • Additional leave including family leave and paid care leave to care for family members.
  • Medical coverage including dental, optometrist, mental care, maternity, registered Traditional Chinese Medicine ("TCM") and Chiropractic.
  • Corporate membership discount and many more to explore.

We believe that you have what it takes to fit into the Touch 'n Go family and help revolutionize the Fintech industry by paving the way to a cashless society. If you're ready to take the next step, apply now

Touch 'n Go is an organization that strives to provide Equal Opportunity Employment, based on merit, qualifications, capabilities, and calibre. It is Touch 'n Go's policy to not discriminate based on age, race, religion, colour or other personal status, identity or characteristics. Fair Opportunity is Our Value and Practice. Please advise us of any accommodations you may need by e-mailing:

Note
: Only shortlisted candidates will be contacted.

Let's keep LEAP-ing forward together



  • Kuala Lumpur, Kuala Lumpur, Malaysia VCB Malaysia Berhad Full time 144,000 - 156,000 per year

    Overview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...


  • Kuala Lumpur, Kuala Lumpur, Malaysia PeopleScope Full time 60,000 - 120,000 per year

    Site Reliability EngineerJob Description:Ability to debug scripts and automate routine tasks in OS, network, database or application servers. Coding experience beyond simple scripts; Experience in Devops process, programming knowledge in at least one of the following languages: Java, Python, or Go; Scripting skills in at least of the following:...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Abhidi Solution Private Limited Full time 120,000 - 180,000 per year

    Job Title: Site Reliability Engineer (SRE)Job Type: Permanent positionWork Location Kuala LumpurResponsibilities:Strong hands-on experience with VMware solutionsStrong experience with patch management for OS & middlewareExperience in VMware server templating/blueprints (RedHat & Windows)Experience with Infrastructure-as-Code, orchestration, configuration...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Hunters International Full time 19,000 per year

    Overview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Unison Consulting Full time 120,000 - 240,000 per year

    As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Career Wise Full time 120,000 - 240,000 per year

    As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Unison Group Full time 120,000 - 240,000 per year

    As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Aisling Group Full time 90,000 - 120,000 per year

    COMPANY PROFILE: Our client is a Tech Ecommerce Scale-Up that provides a single platform for customers to shop for the best price online. Not only that, they also provide data and insights to customers on latest trends and e-commerce sector.   They are looking for a Site Reliability Engineers (SREs) who are responsible for keeping all services and...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Swift Transportation Full time 80,000 - 120,000 per year

    ABOUT USWe're the world's leading provider of secure financial messaging services, headquartered in Belgium. We are the way the world moves value – across borders, through cities and overseas. No other organisation can address the scale, precision, pace and trust that this demands, and we're proud to support the global economy. We're unique too. We were...


  • Kuala Lumpur, Kuala Lumpur, Malaysia Hays Full time 120,000 - 200,000 per year

    Site Reliability Engineer, Kuala LumpurYour new companyThis client enables regulated organisations to move from paper-based validation to intelligent, digitised, paperless solutions. And they do it through the ongoing development of a powerful, purpose-built software platform. In 2014, after eight years of intensive software development, they launched the...