Site Reliability Specialist

4 weeks ago


Singapore GARENA ONLINE PRIVATE LIMITED Full time
Job Overview

We are seeking a skilled Site Reliability Specialist to join our team at GARENA ONLINE PRIVATE LIMITED. The ideal candidate will have a strong background in Linux operating systems, computer networks, and programming languages such as Bash, Python, and Go.



Key Responsibilities:

  • Deep dive into development lines, learning and understanding the mechanism of every application component, and promoting product scalability, stability, and performance.
  • Setup, manage, and maintain product, middleware, big-data applications, and services.
  • Perform regular and ad-hoc server-side deployments, performance fine-tuning, and troubleshooting.
  • Design and develop automations for workflows.
  • Capacity and Resource management.
  • Responsible for the full-chain stress test to enhance performance and remove redundancy of applications.
  • Prepare routine operation documentation.


Requirements:

  • Bachelor's or higher degree in Computer Science, Engineering, Information Systems, or related fields.
  • Minimum 2 years of experience in Site Reliability Engineer roles or equivalent.
  • Extensive and hands-on knowledge with Linux operating systems (Ubuntu, CentOS, etc.).
  • Knowledge of Computer Network (TCP/IP, DNS, etc.) and OS.
  • Hands-on experience with at least one of the programming languages: Bash, Python, Go.
  • Strong analytical and problem-solving skills with the ability to thrive under difficult and stressful situations.
  • Passion and high sense of responsibility for work.
  • Fast learning ability and a good team player.
  • Detailed-oriented, cautious, and prudent.


Preferred Skills:

  • Experience with automation tools like Ansible, Jenkins.
  • Experience with monitoring tools like Prometheus, Zabbix, Grafana, etc.
  • Experience with load balancing tools like LVS, Nginx, Openresty, or HAProxy.
  • Experience with container technology such as Docker, Kubernetes.
  • Experience with Kafka and Codis.
Tell Employers What Skills You Have

Data Analysis
Ubuntu
Business Acumen
Problem Solving
Bash
Computer Networks
Python
CentOS
Operating Systems
Linux

  • Singapore JOHN CRANE SINGAPORE PTE LTD Full time

    Job Summary:As a Reliability Specialist at JOHN CRANE SINGAPORE PTE LTD, you will play a crucial role in ensuring the reliability of mechanical seals for our customers in Singapore. Your primary responsibility will be to provide on-site support and manage reliability contracts for our customers. You will be working closely with our team to identify and...


  • Singapore HW Search & Selection Ltd Full time

    Site Reliability Engineer A new opportunity has arisen for a Site Reliability Engineer for a prestigious investment management firm in Singapore. You will be responsible for providing production support for the trading infrastructure.Your main responsibilities will include:Linux trading infrastructure supportProviding Level II supportUtilizing Python to...


  • Singapore HW Search & Selection Ltd Full time

    Site Reliability Engineer A new opportunity has arisen for a Site Reliability Engineer for a prestigious investment management firm in Singapore. You will be responsible for providing production support for the trading infrastructure. Your main responsibilities will include: Linux trading infrastructure support Providing Level II support Utilizing Python to...


  • Singapore Tencent Full time

    About TencentTencent is a leading Internet-based platform company founded in Shenzhen, China, in 1998. Our mission is to create value for users and drive technological innovation.We are expanding our international operations and seeking top talent to propel us forward. As a Cloud Reliability Specialist, you will have the opportunity to work with a unique...


  • Singapore Aptitude Asia Full time

    At Aptitude Asia, we're seeking a skilled Site Reliability Engineer to join our team. This role is crucial in ensuring the high reliability, availability, and performance of our applications throughout their lifecycle.Key Responsibilities:Develop and implement automation scripts to streamline repetitive tasks and address recurring issues.Collaborate with...


  • Singapore Helius Full time

    Job Title: Site Reliability EngineerJob Summary: Helius is seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for designing, implementing, and operating highly scalable and reliable systems. Your main focus will be on ensuring the smooth operation of our services, resolving technical issues,...


  • Singapore BYTEPLUS PTE. LTD. Full time

    Role OverviewAt ByteDance, we're seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you'll be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing. This includes building automated operation solutions for large-scale systems, cooperating with the...


  • Singapore Aptitude Asia Limited Full time

    Our client, a top-tier hedge fund, is looking to hire a talented Site Reliability Engineer to join their growing SRE team in Singapore. Job Responsibilities: Ensure high reliability, availability, and performance of applications throughout their lifecycle. Automate repetitive tasks and systematically address recurring issues. Generate innovative ideas for...


  • Singapore LANDI INTERNATIONAL (SINGAPORE) PTE. LTD. Full time

    Landi International (Singapore) PTE. LTD.As a Site Reliability Engineer at Landi International (Singapore) PTE. LTD., you will play a crucial role in ensuring the availability, reliability, and scalability of our platforms. Your primary responsibilities will include:· Building, operating, and maintaining our platform infrastructures across various...


  • Singapore BYTEDANCE PTE. LTD. Full time

    About the JobAt ByteDance, we are looking for a talented Site Reliability Engineer to join our team. In this role, you will be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing, while paying attention to system capacity and stability.Key Responsibilities Ensure the reliability and normal...


  • Singapore Skyworks Full time

    Embark on a thrilling career with Skyworks, a pioneer in high-performance analog semiconductors. Our innovative solutions are driving the wireless networking revolution, and we're looking for talented individuals to join our team.About UsSkyworks is a fast-paced environment that values diversity, social responsibility, open communication, mutual trust, and...


  • Singapore ACCESS PEOPLE (SINGAPORE) PTE. LTD. Full time

    Roles & ResponsibilitiesA global energy trading firm is transitioning to a data-centric platform and is seeking a Site Reliability Engineer to support this multi-year program. The role will focus on enhancing the reliability, scalability, and stability of the company's evolving platform. The successful candidate will work on integrating a new event-based,...


  • Singapore Qlik Full time

    Director of Regional Site Reliability EngineeringQlik is seeking an experienced leader to oversee the development and scaling of our regional Site Reliability Engineering (SRE) organization in APAC. This role will be instrumental in ensuring the availability, scalability, and reliability of our services.About QlikWe are a global company that transforms...


  • Singapore APPLE SERVICES PTE. LTD. Full time

    Roles & ResponsibilitiesSummaryThe Apple Services Engineering (ASE) team is one of the most exciting examples of Apple's long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, Fitness+ and Apple Books. And they do it on a massive scale, meeting Apple's high expectations with...


  • Singapore ASIA GULF CLOUD PTE. LTD. Full time

    Roles & ResponsibilitiesAbout SGB:SGB is a new digital bank that will offer a secure and integrated platform to access andmanage conventional and digital assets and financial solutions, including round-the-clock realtime settlement, trading connectivity, custody and asset management. It serves globalinvestors, innovators and institutions looking for a...


  • Singapore Celanese Corporation Full time

    Main Responsibilities• Develop and implement effective reliability strategies to improve the reliability of static equipment.• Provide technical subject matter expertise for static equipment based on engineering codes, SEPs, and RAGAGEPS.• Identify and eliminate recurring problems and bad actors by analyzing failure patterns and implementing corrective...


  • Singapore U3 INFOTECH PTE. LTD. Full time

    Database Reliability SpecialistTo ensure the smooth operation of our databases, we are seeking a skilled Database Reliability Specialist. In this role, you will be responsible for creating and maintaining database standards and policies, supporting database design, creation, and testing activities, and managing database availability and performance.Key...


  • Singapore Tower Research Capital Full time

    Tower Research Capital Job DescriptionJob Title: Site Reliability EngineerJob Summary:We are seeking a highly skilled Site Reliability Engineer to join our team at Tower Research Capital. The successful candidate will be responsible for ensuring the continuous operation of our Linux-based trading infrastructure and addressing day-to-day operational needs.Key...


  • Singapore Riot Games Full time

    Job DescriptionWe are seeking a skilled Game Service Reliability Specialist to join our team at Riot Games. As a key member of our operations team, you will be responsible for ensuring the health and reliability of our game services.


  • Singapore Ripple Labs Singapore Full time

    As a Senior Site Reliability Engineer at Ripple Labs Singapore, you will be responsible for ensuring the high availability and scalability of our systems. Your primary goal will be to design, implement, and maintain a robust and efficient infrastructure that can handle high traffic and complex distributed systems.Key Responsibilities:Design and implement...