Site Reliability Engineer
2 days ago
Summary
The Apple Services Engineering (ASE) team is one of the most exciting examples of Apple's long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, Fitness+ and Apple Books. And they do it on a massive scale, meeting Apple's high expectations with dedication to deliver a huge variety of entertainment in over 35 languages to more than 150 countries. These engineers build secure, end-to-end solutions. They develop the custom software used to process all the creative work, the tools that providers use to deliver that media, all the server-side systems, and the APIs for many Apple services. Thanks to Apple's unique integration of hardware, software, and services, engineers here partner to get behind a single unified vision. That vision always includes a deep dedication to strengthening Apple's privacy policy, one of Apple's core values. Although services are a bigger part of Apple's business than ever before, these teams remain small, nimble, and cross-functional, offering greater exposure to the array of opportunities here.
DescriptionAs an SRE at Apple, you will work on world-renowned internet services that serve hundreds of millions of users worldwide. You’ll collaborate closely with development teams and other stakeholders in the entire lifecycle of services from inception through deployment and continuous refinement. Your responsibilities will include monitoring performance and availability, capacity and disaster recovery planning, defining Service Level Objectives (SLO) and ensuring the reliability and resiliency of systems. Join us if you thrive on solving large-scale problems through your expertise and collaborative team work.
- Support and lead cross functional projects across multiple services - Design, analyze and troubleshoot complex mission-critical distributed systems on a global scale.
- Provide OnCall support to 1st level production support teams.
- Participate in incident management and blameless postmortems.
- Detect and resolve performance bottlenecks, enhance service efficiency.
- Write clear technical documentation, production run books.
Apple is an Equal Opportunity Employer that is committed to inclusion and diversity. We also take affirmative action to offer employment and advancement opportunities to all applicants, including minorities, Women, protected veterans, and individuals with disabilities. Apple will not discriminate or retaliate against applicants who inquire about, disclose, or discuss their compensation or that of other applicants. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.
Minimum Qualifications- BS degree in computer science or equivalent field with 5+ years or MS degree with 3+ years experience, or equivalent.
- At least 5 years in a Site Reliability Engineering, DevOps or infrastructure focused role
- Excellent written and verbal communication skills
- Experience in supporting internet-facing production services and distributed systems
- Experience in collaborating effectively across organizations, building strong relationships
- Proficient coding experience with Python or similar scripting languages
- Experience with containers and container orchestration platforms such as Docker and Kubernetes
- Experience with monitoring tools such as Splunk, Grafana, and Prometheus
- Demonstrated ability to deliver results on time with high quality
- Experience in managing and scaling distributed systems in a public, private, or hybrid cloud environment
- Experience in debugging Java code and optimizing JVM performance is plus.
- Understanding of the Linux Operating System, standard networking protocols, and components
- Passion for crafting and building reliable systems
- Strong sense of ownership and integrity proven through clear communication and collaboration
- Automation advocate - you truly believe in removing operation load with software
- Excellent troubleshooting and problem solving skills
Tell employers what skills you have
Excellent Communication Skills
Troubleshooting
Splunk
Kubernetes
Analytical Skills
DevOps
Documentation
Computer Science
Distributed Systems
Python
Site Reliability Engineering
Docker
Building Relationships
Java
Grafana
Debugging
Efficiency Improvement
Linux
Incident Management
-
Site reliability engineer
5 days ago
Singapore HW Search & Selection Ltd Full timeSite Reliability Engineer A new opportunity has arisen for a Site Reliability Engineer for a prestigious investment management firm in Singapore. You will be responsible for providing production support for the trading infrastructure. Your main responsibilities will include: Linux trading infrastructure support Providing Level II support Utilizing Python to...
-
Site Reliability Engineer
4 weeks ago
Singapore Aptitude Asia Full timeAt Aptitude Asia, we're seeking a skilled Site Reliability Engineer to join our team. This role is crucial in ensuring the high reliability, availability, and performance of our applications throughout their lifecycle.Key Responsibilities:Develop and implement automation scripts to streamline repetitive tasks and address recurring issues.Collaborate with...
-
Site Reliability Engineer
2 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRole OverviewAt ByteDance, we're seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you'll be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing. This includes building automated operation solutions for large-scale systems, cooperating with the...
-
Site Reliability Engineer
1 month ago
Singapore Aptitude Asia Full timeJob SummaryAptitude Asia seeks a skilled Site Reliability Engineer to ensure the high reliability, availability, and performance of applications throughout their lifecycle.Key ResponsibilitiesReliability and Performance: Ensure applications operate with high reliability, availability, and performance.Automation and Innovation: Automate repetitive tasks and...
-
Site Reliability Engineer
2 weeks ago
Singapore BYTEDANCE PTE. LTD. Full timeAbout the JobAt ByteDance, we are looking for a talented Site Reliability Engineer to join our team. In this role, you will be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing, while paying attention to system capacity and stability.Key Responsibilities Ensure the reliability and normal...
-
Site Reliability Engineer
2 weeks ago
Singapore LANDI INTERNATIONAL (SINGAPORE) PTE. LTD. Full timeLandi International (Singapore) PTE. LTD.As a Site Reliability Engineer at Landi International (Singapore) PTE. LTD., you will play a crucial role in ensuring the availability, reliability, and scalability of our platforms. Your primary responsibilities will include:· Building, operating, and maintaining our platform infrastructures across various...
-
Senior Site Reliability Engineer
1 month ago
Singapore AIA Singapore Private Limited Full timeAbout the RoleWe are seeking a highly skilled Senior Site Reliability Engineer to join our team at AIA Singapore Private Limited. As a key member of our operations team, you will be responsible for ensuring the reliability and stability of our critical production services and applications.Key ResponsibilitiesLead complex system and champion services...
-
Site Reliability Engineer
2 weeks ago
Singapore ACCESS PEOPLE (SINGAPORE) PTE. LTD. Full timeRoles & ResponsibilitiesA global energy trading firm is transitioning to a data-centric platform and is seeking a Site Reliability Engineer to support this multi-year program. The role will focus on enhancing the reliability, scalability, and stability of the company's evolving platform. The successful candidate will work on integrating a new event-based,...
-
Site Reliability Engineer
2 weeks ago
Singapore Vortexa Full timeVortexa is a cutting-edge company that leverages satellite data and AI to provide real-time insights into global energy flows. We're looking for a skilled Site Reliability Engineer to join our Data Services Team, responsible for the developer platform and Amazon AWS estate.The ChallengeOur platform processes massive amounts of data from various sources,...
-
Site Reliability Engineer
1 week ago
Singapore Tower Research Capital Full timeTower Research Capital Job DescriptionJob Title: Site Reliability EngineerJob Summary:We are seeking a highly skilled Site Reliability Engineer to join our team at Tower Research Capital. The successful candidate will be responsible for ensuring the continuous operation of our Linux-based trading infrastructure and addressing day-to-day operational needs.Key...
-
Site Reliability Engineer
2 weeks ago
Singapore ASIA GULF CLOUD PTE. LTD. Full timeRoles & ResponsibilitiesAbout SGB:SGB is a new digital bank that will offer a secure and integrated platform to access andmanage conventional and digital assets and financial solutions, including round-the-clock realtime settlement, trading connectivity, custody and asset management. It serves globalinvestors, innovators and institutions looking for a...
-
Senior Site Reliability Engineer
3 weeks ago
Singapore Ripple Labs Singapore Full timeAs a Senior Site Reliability Engineer at Ripple Labs Singapore, you will be responsible for ensuring the high availability and scalability of our systems. Your primary goal will be to design, implement, and maintain a robust and efficient infrastructure that can handle high traffic and complex distributed systems.Key Responsibilities:Design and implement...
-
Senior Site Reliability Engineer
1 month ago
Singapore NodeFlair Full timeSenior Site Reliability EngineerWe are working with NodeFlair, a leading pioneer in the Cryptocurrency space, to search for a Senior Site Reliability Engineer to join their Singapore/Remote team.Summary:Our client, a top player in cryptocurrency data monitoring, tracks over 10,000 tokens on 400+ exchanges with 300 million page views from 100+...
-
Senior Site Reliability Engineer
6 days ago
Singapore Ripple Labs Singapore Full timeJob SummaryWe are seeking a highly skilled Senior Site Reliability Engineer to join our team at Ripple Labs Singapore. As a key member of our infrastructure team, you will be responsible for ensuring the high availability and reliability of our systems.Key ResponsibilitiesDesign, implement, and maintain high availability systems and infrastructureCollaborate...
-
Site Reliability Engineer
1 week ago
Singapore Helius Full timeJob Title: Site Reliability EngineerJob Summary: Helius is seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for designing, implementing, and operating highly scalable and reliable systems. Your main focus will be on ensuring the smooth operation of our services, resolving technical issues,...
-
Site Reliability Engineer
2 weeks ago
Singapore ITCAN PTE. LIMITED Full timeRoles & ResponsibilitiesRoles & Responsibilities:The Site Reliability Engineer (SRE) combines software development and system engineering to build and run distributed solutions in a secured multi-tier heterogeneous environment to safeguard, provide and continuously improve the software and systems behind the organization’s cloud platform solutions.The Job:...
-
Senior Site Reliability Engineer
1 week ago
Singapore Sea Full timeAbout SeaSea is a cutting-edge technology company with a hyper-growing business scale, transforming complex problems into technical challenges. Our team of passionate engineers is dedicated to delivering world-class experiences for our users.The Games Site Reliability Engineer (SRE) team at Sea Labs Indonesia plays a crucial role in ensuring the stability...
-
Senior Site Reliability Engineer
1 month ago
Singapore Ripple Labs Singapore Full timeJob Title: Senior Site Reliability EngineerWe are seeking a highly skilled Senior Site Reliability Engineer to join our team at Ripple Labs Singapore. As a key member of our infrastructure team, you will be responsible for ensuring the high availability and scalability of our systems.Key Responsibilities:Design and Implement High Availability Solutions:...
-
Chief Site Reliability Engineer
1 week ago
Singapore Snaphunt Full timeThe OpportunityWe're seeking an experienced Site Reliability Engineer to empower users with a rich feature set, high availability, and stellar performance at First Digital Finance Corp.As we expand customer deployments, the ideal candidate will deliver insights from massive-scale data in real-time, collaborating with a cross-functional team to develop...
-
Reliability Engineer
2 weeks ago
Singapore ADDVALUE INNOVATION PTE LTD Full timeRoles & ResponsibilitiesReliability EngineerResponsibilities Work with product development teams to develop relaibility requirements, establish a reliability / test program and perform appropriate analyse to ensure that new products meet all the relaibility targets. Perform risk / reliabilty analysis (FMEA, FMECA, MTBF) for existing and new products Able...