Site Reliability Engineer
2 days ago
COMPANY PROFILE: Our client is a Tech Ecommerce Scale-Up that provides a single platform for customers to shop for the best price online. Not only that, they also provide data and insights to customers on latest trends and e-commerce sector. They are looking for a Site Reliability Engineers (SREs) who are responsible for keeping all services and production systems running smoothly. SREs ensures that services have reliability, uptime appropriate to users' needs and a fast rate of improvement. You'll have the opportunity to work on complex challenges of scale, using your experience in coding, algorithms, and analysis
RESPONSIBILITY:
- Engage in and improve the whole lifecycle of services - from inception and design, through to deployment, operation and refinement.
- Collaborate with engineering teams on their infrastructure needs, and advise them throughout the development lifecycle.
- Maintain services once they are live by measuring and monitoring availability, latency, and overall system health, within our Service Level Objectives.
- Scale systems sustainably through mechanisms like automation; evolve systems by pushing for changes that improve reliability and velocity.
- Practice sustainable incident response and blameless post-mortems.
- Debug production issues across services, databases and levels of the stack.
- Design, develop and manage monitoring tools to provide performance dashboards, alerts, and collect data required to proactively identify issues and/or recommend improvements.
REQUIREMENTS:
- 5-8 years of experience in provisioning environments, deploying applications, and maintaining infrastructures.
- Professional experience using Python, Go, or Ruby.
- Experience with deployment automation/configuration management tools like Chef, Ansible, Puppet, or Terraform.
- Experience in cloud-based environment such as AWS, GCP or Azure.
- Have extensive experience building scalable platforms leveraging containers in a production environment.
- Added bonus if you have experience in operated distributed data storage systems at scale, especially Elasticsearch and SQL Azure.
- Solid knowledge of continuous integration, continuous delivery, automated testing and all phases of the software development lifecycle.
- Experience of working in an agile and multi-cultural environment across many SCRUM teams at the same time.
- A Kaizen mindset and spirit of continuous improvement on a personal level and always up to date with the latest technology trends professionally.
- Ability to identify problems before they happen and implement solutions that detect and prevent outages.
- Expertise in designing, analysing and troubleshooting large-scale distributed systems.
- Ability to debug, optimize code and automate routine tasks.
- Systematic problem-solving approach, coupled with effective communication skills and a sense of drive.
- Understanding of CI/CD principles, Linux fundamentals, networking concepts and IP protocols.
HOW TO APPLY:
- If you're interested, do click apply on the button provided and attach your CV as well. For further information, feel free to speak to Ariff at or email him
-
Site Reliability Engineer
2 weeks ago
Kuala Lumpur, Kuala Lumpur, Malaysia Kneat Full time 80,000 - 120,000 per yearSite Reliability Engineer – Kuala Lumpur, MalaysiaKneat enables regulated organizations to move from paper-based validation to intelligent, digitized, paperless solutions. And we do it through the ongoing development of a powerful, purpose-built software platform. In 2014, after eight years of intensive software development, we launched Kneat Gx—the...
-
Site Reliability Engineer
2 weeks ago
Kuala Lumpur, Kuala Lumpur, Malaysia Kneat Full time 80,000 - 120,000 per yearSite Reliability Engineer – Kuala Lumpur, MalaysiaKneat enables regulated organizations to move from paper-based validation to intelligent, digitized, paperless solutions. And we do it through the ongoing development of a powerful, purpose-built software platform. In 2014, after eight years of intensive software development, we launched Kneat Gx—the...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia VCB Malaysia Berhad Full time 144,000 - 156,000 per yearOverview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia PeopleScope Full time 60,000 - 120,000 per yearSite Reliability EngineerJob Description:Ability to debug scripts and automate routine tasks in OS, network, database or application servers. Coding experience beyond simple scripts; Experience in Devops process, programming knowledge in at least one of the following languages: Java, Python, or Go; Scripting skills in at least of the following:...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Abhidi Solution Private Limited Full time 120,000 - 180,000 per yearJob Title: Site Reliability Engineer (SRE)Job Type: Permanent positionWork Location Kuala LumpurResponsibilities:Strong hands-on experience with VMware solutionsStrong experience with patch management for OS & middlewareExperience in VMware server templating/blueprints (RedHat & Windows)Experience with Infrastructure-as-Code, orchestration, configuration...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Hunters International Full time 19,000 per yearOverview:As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Career Wise Full time 120,000 - 240,000 per yearAs a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Unison Group Full time 120,000 - 240,000 per yearAs a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Swift Transportation Full time 80,000 - 120,000 per yearABOUT USWe're the world's leading provider of secure financial messaging services, headquartered in Belgium. We are the way the world moves value – across borders, through cities and overseas. No other organisation can address the scale, precision, pace and trust that this demands, and we're proud to support the global economy. We're unique too. We were...
-
Site Reliability Engineer
2 days ago
Kuala Lumpur, Kuala Lumpur, Malaysia Guidewire Software Full time 120,000 - 240,000 per yearSummaryAt Guidewire, we deliver the software that Property and Casualty (P&C) insurance companies rely on to protect their customers during crises, natural disasters, accidents, and cyber risks. Our core applications enable insurers to sell and underwrite policies, settle claims, and bill their customers. We also offer a suite of innovative products for data...