Site Reliability Engineer Jobs in Career Wise in Kuala Lumpur - Malaysia

Site Reliability Engineer

Career Wise

Posted on : 14-12-2024

Employer Active

1 Vacancy

Job Alert

You will be updated with latest job alerts via email

Valid email field required

Send jobs

Send me jobs like this

Job Alert

You will be updated with latest job alerts via email

Valid email field required

Send jobs

Job Location

Kuala Lumpur - Malaysia

Monthly Salary

Not Disclosed

Salary Not Disclosed

Vacancy

1 Vacancy

Posted on : 14-12-2024

Job Description

As a Site Reliability Engineer (SRE) you will play a key role in maintaining the reliability and performance of critical services. Your expertise will help bridge the gap between development and operations ensuring robust scalable and responsive infrastructure. This role emphasizes strong system architecture and design principles focusing on key SRE practices such as Service Level Objectives (SLOs) Service Level Indicators (SLIs) and the reduction of operational toil. You will collaborate closely with diverse teams to drive reliability improvements and foster a culture of continuous learning and accountability.

Key Responsibilities:

Design and implement resilient system architectures that support high availability and scalability.
Develop automation tools and scripts to enhance operational efficiency and reduce manual effort.
Define track and analyze SLOs and SLIs to ensure reliability and performance meet business needs.
Conduct thorough postmortem analyses following incidents driving continuous improvement through root cause identification and solution implementation.
Collaborate with development and operations teams to establish best practices in system reliability and incident management.
Troubleshoot and resolve issues related to database performance network connectivity and deployment failures including diagnosing problems at the underlying platform level (e.g. Kubernetes virtual machines).
Ensure that issues are resolved within the stipulated Service Level Agreements (SLAs) maintaining high standards of service delivery.
Identify and troubleshoot performance bottlenecks across systems providing actionable recommendations for enhancements.
Maintain detailed documentation of processes and incident responses to support knowledge sharing and compliance.

Qualifications:

Proficiency in programming languages such as Python Golang Java or similar focusing on operational efficiency.
Minimum experience of 3 years and above in related field.
Demonstrated experience in system architecture and design prioritizing reliability and scalability.
Strong understanding of SRE principles including SLOs SLIs toil reduction and incident postmortems.
Experience with cloud environments (e.g. AWS Azure Google Cloud) and their operational management.
Strong expertise in Linux system administration.
Proven experience in troubleshooting application support issues with a focus on performance and connectivity.
Familiarity with networking concepts and effective troubleshooting techniques.
Excellent problemsolving abilities and a proactive approach to operational challenges.
Ability to work independently while effectively collaborating within a team environment.

Preferred Skills:

Familiarity with monitoring tools and performance optimization techniques.
Experience in scripting or automation for system administration tasks.
Knowledge of networking concepts and troubleshooting methodologies.
Handson knowledge of cloud platforms (e.g. AWS Azure Google Cloud) and their services.
Familiarity with DevOps practices and frameworks including CI/CD infrastructure as code and containerization.

Employment Type

Full Time

Company Industry

Key Skills

Apply Now

About Company

Career Wise

Report This Job

Disclaimer: Drjobpro.com is only a platform that connects job seekers and employers. Applicants are advised to conduct their own independent research into the credentials of the prospective employer.We always make certain that our clients do not endorse any request for money payments, thus we advise against sharing any personal or bank-related information with any third party. If you suspect fraud or malpractice, please contact us via contact us page.

Free AI Resume Review

Get Hired 3x Faster with free, confidential review from Ai resume review service.

Order Now

Resume, LinkedIn, Cover Letter

Elevate your professional profile with expertly crafted documents including your resume, LinkedIn profile, cover letter.

Start Now

Dr.Job AutoApply

3X your job search with AutoApply's AI for faster dream job results.

Learn More

Reverse Recruiting

Never apply for a job again. We apply and track jobs for you to find your perfect match.

Site Reliability Engineer

Career Wise

Job Description

Employment Type

Company Industry

Key Skills

About Company

Similar Jobs

QA Engineer

QA Engineer

QA Engineer

QA Engineer

Application Engineer

Application Engineer

Application Engineer

QA Engineer