Senior Site Reliability Engineer

3 weeks ago


Singapore ACCESS PEOPLE (SINGAPORE) PTE. LTD. Full time
Roles & Responsibilities

A global energy trading firm is transitioning to a data-centric platform and is seeking a Senior Site Reliability Engineer to support this multi-year program. The role will focus on enhancing the reliability, scalability, and stability of the company's evolving platform.


Key Responsibilities:

  • Establish and track SLOs/SLIs, ensuring platform reliability, and lead remediation efforts when breaches occur.
  • Automate production infrastructure tasks, deployment pipelines, and improve overall system stability with the Platform Engineering team.
  • Maintain strong communication with users, providing updates on system performance and remediation plans.
  • Address system bottlenecks and failure points before incidents occur.
  • Enhance platform observability, including tools for user interfaces, analytics, and reporting.
  • Contribute to disaster recovery planning and infrastructure capacity growth.
  • Lead Level 2 production support, ensuring timely resolution of critical incidents.
  • Support the development of web applications, real-time data processors, and data integrations.
  • Implement automated testing, continuous integration, and deployment to boost system efficiency and productivity.
  • Collaborate with IT stakeholders to align with the company’s overall IT strategy.
  • Document system configurations, processes, and troubleshooting guides to ensure knowledge continuity.

Qualifications & Experience:

  • 4 + years of experience in a Site Reliability Engineering role, ideally in trading or a related sector.
  • Expertise in maintaining microservices systems in a production environment.
  • Strong proficiency with .NET, event-driven architecture, and data processing tools (Azure Event Hubs or Apache Kafka).
  • Experience with RESTful APIs, gRPC, and GraphQL.
  • Familiarity with relational and document-based databases.
  • Knowledge of cloud PaaS/IaaS (Microsoft Azure preferred) and containerized microservice architectures (Docker, Kubernetes).
  • Experience with Azure DevOps Pipelines is an advantage.

EA: 14S7084 | Registration No:R1981018


Tell employers what skills you have

Microsoft Azure
Troubleshooting
Remediation
Scalability
Kubernetes
.NET
Azure
Pipelines
Reliability
Reliability Engineering
Apache Kafka
Continuous Integration
Docker
Disaster Recovery
IT Strategy
Databases

  • Singapore Ripple Labs Singapore Full time

    Job SummaryWe are seeking a highly skilled Senior Site Reliability Engineer to join our team at Ripple Labs Singapore. As a key member of our infrastructure team, you will be responsible for ensuring the high availability and reliability of our systems.Key ResponsibilitiesDesign, implement, and maintain high availability systems and infrastructureCollaborate...


  • Singapore Ripple Labs Singapore Full time

    Job Title: Senior Site Reliability EngineerWe are seeking a highly skilled Senior Site Reliability Engineer to join our team at Ripple Labs Singapore. As a key member of our infrastructure team, you will be responsible for ensuring the high availability and scalability of our systems.Key Responsibilities:Design and Implement High Availability Solutions:...


  • Singapore Ripple Labs Singapore Full time

    As a Senior Site Reliability Engineer at Ripple Labs Singapore, you will be responsible for ensuring the high availability and scalability of our systems. Your primary goal will be to design, implement, and maintain a robust and efficient infrastructure that can handle high traffic and complex distributed systems.Key Responsibilities:Design and implement...


  • Singapore HW Search & Selection Ltd Full time

    Site Reliability Engineer A new opportunity has arisen for a Site Reliability Engineer for a prestigious investment management firm in Singapore. You will be responsible for providing production support for the trading infrastructure. Your main responsibilities will include: Linux trading infrastructure support Providing Level II support Utilizing Python to...


  • Singapore Aptitude Asia Full time

    At Aptitude Asia, we're seeking a skilled Site Reliability Engineer to join our team. This role is crucial in ensuring the high reliability, availability, and performance of our applications throughout their lifecycle.Key Responsibilities:Develop and implement automation scripts to streamline repetitive tasks and address recurring issues.Collaborate with...


  • Singapore The Chemical Engineer Full time

    About us At Exxon Mobil, our vision is to lead in energy innovations that advance modern living and a net-zero future. As one of the world’s largest publicly traded energy and chemical companies, we are powered by a unique and diverse workforce fueled by the pride in what we do and what we stand for. The success of our Upstream, Product Solutions and Low...


  • Singapore BYTEPLUS PTE. LTD. Full time

    Role OverviewAt ByteDance, we're seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you'll be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing. This includes building automated operation solutions for large-scale systems, cooperating with the...


  • Singapore Sea Full time

    About SeaSea is a cutting-edge technology company with a hyper-growing business scale, transforming complex problems into technical challenges. Our team of passionate engineers is dedicated to delivering world-class experiences for our users.The Games Site Reliability Engineer (SRE) team at Sea Labs Indonesia plays a crucial role in ensuring the stability...


  • Singapore Aptitude Asia Full time

    Job SummaryAptitude Asia seeks a skilled Site Reliability Engineer to ensure the high reliability, availability, and performance of applications throughout their lifecycle.Key ResponsibilitiesReliability and Performance: Ensure applications operate with high reliability, availability, and performance.Automation and Innovation: Automate repetitive tasks and...


  • Singapore BYTEDANCE PTE. LTD. Full time

    About the JobAt ByteDance, we are looking for a talented Site Reliability Engineer to join our team. In this role, you will be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing, while paying attention to system capacity and stability.Key Responsibilities Ensure the reliability and normal...


  • Singapore LANDI INTERNATIONAL (SINGAPORE) PTE. LTD. Full time

    Landi International (Singapore) PTE. LTD.As a Site Reliability Engineer at Landi International (Singapore) PTE. LTD., you will play a crucial role in ensuring the availability, reliability, and scalability of our platforms. Your primary responsibilities will include:· Building, operating, and maintaining our platform infrastructures across various...


  • Singapore ACCESS PEOPLE (SINGAPORE) PTE. LTD. Full time

    Roles & ResponsibilitiesA global energy trading firm is transitioning to a data-centric platform and is seeking a Site Reliability Engineer to support this multi-year program. The role will focus on enhancing the reliability, scalability, and stability of the company's evolving platform. The successful candidate will work on integrating a new event-based,...


  • Singapore APPLE SERVICES PTE. LTD. Full time

    Roles & ResponsibilitiesSummaryThe Apple Services Engineering (ASE) team is one of the most exciting examples of Apple's long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, Fitness+ and Apple Books. And they do it on a massive scale, meeting Apple's high expectations with...


  • Singapore Ripple Labs Singapore Full time

      WHAT YOU’LL DO: Keeping your assigned site or service up and running or rapid recovery from failures Actively troubleshoot any issues that arise during testing and production, catching and solving issues before launch, Automating work including infrastructure needs, testing, failover solutions, failure mitigation, and much more, Monitor and...


  • Singapore NodeFlair Full time

    Senior Site Reliability EngineerWe are working with one of the leading pioneers in the Cryptocurrency space as one of the largest data platforms, and as part of their continued growth, NodeFlair has been engaged to search for a Senior Site Reliability Engineer to join their Singapore/Remote team.About the RoleOur client, a top player in cryptocurrency data...


  • Singapore Vortexa Full time

    Vortexa is a cutting-edge company that leverages satellite data and AI to provide real-time insights into global energy flows. We're looking for a skilled Site Reliability Engineer to join our Data Services Team, responsible for the developer platform and Amazon AWS estate.The ChallengeOur platform processes massive amounts of data from various sources,...


  • Singapore Tower Research Capital Full time

    Tower Research Capital Job DescriptionJob Title: Site Reliability EngineerJob Summary:We are seeking a highly skilled Site Reliability Engineer to join our team at Tower Research Capital. The successful candidate will be responsible for ensuring the continuous operation of our Linux-based trading infrastructure and addressing day-to-day operational needs.Key...


  • Singapore ASIA GULF CLOUD PTE. LTD. Full time

    Roles & ResponsibilitiesAbout SGB:SGB is a new digital bank that will offer a secure and integrated platform to access andmanage conventional and digital assets and financial solutions, including round-the-clock realtime settlement, trading connectivity, custody and asset management. It serves globalinvestors, innovators and institutions looking for a...


  • Singapore Hireio, Inc. Full time

    About the RoleWe are seeking a highly skilled Senior Site Reliability Engineer to join our Compute Platform team. As a key member of our team, you will be responsible for ensuring the reliability of our major data warehouse products, services, and query engines.Key Responsibilities:Lead a global SRE team for TikTok's Data Platform, distributed across the US...


  • Singapore Helius Full time

    Job Title: Site Reliability EngineerJob Summary: Helius is seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for designing, implementing, and operating highly scalable and reliable systems. Your main focus will be on ensuring the smooth operation of our services, resolving technical issues,...