Site Reliability Engineer (Sre) | Contract | Penang

Details of the offer

Overview:
As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations, ensuring robust, scalable, and responsive infrastructure. This role emphasizes strong system architecture and design principles, focusing on key SRE practices such as Service Level Objectives (SLOs), Service Level Indicators (SLIs), and the reduction of operational toil. You will collaborate closely with diverse teams to drive reliability improvements and foster a culture of continuous learning and accountability.
Responsibilities:
Design and implement resilient system architecturesthat support high availability and scalability.
Develop automation tools and scripts to enhance operationalefficiency and reduce manual effort.
Define, track, and analyze SLOs and SLIsto ensure reliability and performance meet business needs.
Conduct thorough post-mortem analyses following incidents,driving continuous improvement through root cause identification and solution implementation.
Collaborate with development and operations teams to establish best practices in system reliability and incident management.
Troubleshoot and resolve issues related to database performance, network connectivity, and deployment failures,including diagnosing problems at the underlying platform level (e.g., Kubernetes, virtual machines).
Ensure that issues are resolved within the stipulated Service Level Agreements (SLAs),maintaining high standards of service delivery.
Identify and troubleshoot performance bottlenecks across systems,providing actionable recommendations for enhancements.
Maintain detailed documentation of processes and incident responsesto support knowledge sharing and compliance.
Requirements:
Proficiency in Mandarin is a must, in order to liaise with stakeholders from China.
Proficiency in programming languages such asPython, Golang, Java,or similar, focusing on operational efficiency.
Demonstrated experience insystem architecture and design,prioritizing reliability and scalability.
Strong understanding ofSRE principles , includingSLOs, SLIs, toil reduction, and incident post-mortems.
Experience withcloud environments (e.g., AWS, Azure, Google Cloud)and their operational management.
Strong expertise inLinux system administration.
Proven experience introubleshooting application support issueswith a focus on performance and connectivity.
Familiarity withnetworking conceptsand effective troubleshooting techniques.
Excellent problem-solving abilities and a proactive approach to operational challenges.
Ability to work independently while effectively collaborating within a team environment.
Familiarity with monitoring tools and performance optimization techniques.
Experience inscripting or automation for system administration tasks.
Hands-on knowledge ofcloud platforms (e.g., AWS, Azure, Google Cloud)and their services.
Familiarity withDevOps practices and frameworks, including CI/CD, infrastructure as code, and containerization.
Location:To be based in client's site (Penang)
Remuneration:Up to MYR 15,000 (Based on relevant experience)
Consultant in Charge:
Rodney Chong | ****** | 016 838 2188
This is a contract position with the possibility to be absorbed as a permanent staff.#J-18808-Ljbffr


Nominal Salary: To be agreed

Requirements

Site Reliability Engineer (Sre) | Contract | Penang

Overview: As a Site Reliability Engineer (SRE), you will play a key role in maintaining the reliability and performance of critical services. Your expertise ...


Hunters International Sdn Bhd - Pahang

Published 23 days ago

It Intern

Job Title: IT Intern Years of working experience: 0 to 1 year Job Type: Internship Work-Mode: On-Site Location:5-4-6 Hunza Complex, Jalan Gangsa, Greenlane H...


Sqc Management (Penang) Sdn Bhd - Pahang

Published 23 days ago

Technical Customer Support, Senior (Night Shift)

Remote Work: Hybrid Overview: At Zebra, we are a community of innovators who come together to create new ways of working to make everyday life better. United...


Zebra - Pahang

Published 24 days ago

Programmer It/Analyst Ii

Job Description We are seeking a highly skilled and motivated Data Analyst to join our team and contribute to the development and maintenance of our semicond...


Renesas Electronics - Pahang

Published 22 days ago

Built at: 2024-12-26T10:43:49.442Z