Site Reliability Engineer
3 weeks ago
TikTok is the leading destination for short-form mobile video. Our mission is to inspire creativity and bring joy. TikTok has global offices including Los Angeles, New York, London, Paris, Berlin, Dubai, Singapore, Jakarta, Seoul and Tokyo.
Why Join Us
Creation is the core of TikTok's purpose. Our platform is built to help imaginations thrive. This is doubly true of the teams that make TikTok possible.
Together, we inspire creativity and bring joy - a mission we all believe in and aim towards achieving every day.
To us, every challenge, no matter how difficult, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always.
At TikTok, we create together and grow together. That's how we drive impact - for ourselves, our company, and the communities we serve.
Join us.
About The Team
Our Recommendation Architecture Team is responsible for building up and optimizing the architecture for our recommendation system to provide the most stable and best experience for our TikTok users.
On the SRE team of Recommendation Architecture, you'll have the opportunity to sharpen your expertise in coding, performance analysis, large-scale system operation, and get heavily involved in the process of hardware/capacity decision-making.
SRE ensures that the recommendation services at ByteDance have the highest level of availability, as well as creating highly automated systems and pipelines.
Responsibilities
- Reliability and operation optimization for large-scale clusters of TikTok Recommendation System.
- Continuous integration and delivery of core services, optimizing the efficiency and automation of operation, and improving service stability and R&D efficiency.
- Cloud platformization, resource optimization and SLA guarantee for large-scale clusters.
- Collaboration with software engineer to design and implement DevOps solutions to Improve the efficiency of the entire R&D process.
- Research, design, and develop computer and network software or specialised utility programs.
- Analyse user needs and develop software solutions, applying principles and techniques of computer science, engineering, and mathematical analysis.
- Update software, enhances existing software capabilities, and develops and direct software testing and validation procedures.
- Work with computer hardware engineers to integrate hardware and software systems and develop specifications and performance requirements.
Qualifications
- Bachelor's degree or above in computer science, software engineering, or a related field
- Operation experience of large-scale systems, familiar with system operation skills on Linux and network.
- Good programming experience with at least one of the following languages: Shell/Python/Perl/Go/C++.
- Expertise in analyzing, and troubleshooting large-scale distributed systems.
- At least 3 years of relevant experience.
TikTok is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At TikTok, our mission is to inspire creativity and bring joy. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.
Tell employers what skills you have
Troubleshooting
Ubuntu
Pipelines
Software Engineering
Computer Hardware
Reliability
Administration Management
Distributed Systems
Infrastructure Architecture
RedHat
Technical Consultation
Technical Engineering
Continuous Integration
Research Design
Linux
-
Site reliability engineer
2 hours ago
Singapore HW Search & Selection Ltd Full timeSite Reliability Engineer A new opportunity has arisen for a Site Reliability Engineer for a prestigious investment management firm in Singapore. You will be responsible for providing production support for the trading infrastructure.Your main responsibilities will include:Linux trading infrastructure supportProviding Level II supportUtilizing Python to...
-
Site reliability engineer
2 weeks ago
Singapore HW Search & Selection Ltd Full timeSite Reliability Engineer A new opportunity has arisen for a Site Reliability Engineer for a prestigious investment management firm in Singapore. You will be responsible for providing production support for the trading infrastructure. Your main responsibilities will include: Linux trading infrastructure support Providing Level II support Utilizing Python to...
-
Site Reliability Engineer
1 month ago
Singapore Aptitude Asia Full timeAt Aptitude Asia, we're seeking a skilled Site Reliability Engineer to join our team. This role is crucial in ensuring the high reliability, availability, and performance of our applications throughout their lifecycle.Key Responsibilities:Develop and implement automation scripts to streamline repetitive tasks and address recurring issues.Collaborate with...
-
Site Reliability Engineer
3 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRole OverviewAt ByteDance, we're seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you'll be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing. This includes building automated operation solutions for large-scale systems, cooperating with the...
-
Site Reliability Engineer
3 weeks ago
Singapore BYTEDANCE PTE. LTD. Full timeAbout the JobAt ByteDance, we are looking for a talented Site Reliability Engineer to join our team. In this role, you will be responsible for ensuring the reliability and normal operation of multiple core systems for big data and online computing, while paying attention to system capacity and stability.Key Responsibilities Ensure the reliability and normal...
-
Site Reliability Engineer
3 weeks ago
Singapore LANDI INTERNATIONAL (SINGAPORE) PTE. LTD. Full timeLandi International (Singapore) PTE. LTD.As a Site Reliability Engineer at Landi International (Singapore) PTE. LTD., you will play a crucial role in ensuring the availability, reliability, and scalability of our platforms. Your primary responsibilities will include:· Building, operating, and maintaining our platform infrastructures across various...
-
Site Reliability Engineer
3 weeks ago
Singapore ACCESS PEOPLE (SINGAPORE) PTE. LTD. Full timeRoles & ResponsibilitiesA global energy trading firm is transitioning to a data-centric platform and is seeking a Site Reliability Engineer to support this multi-year program. The role will focus on enhancing the reliability, scalability, and stability of the company's evolving platform. The successful candidate will work on integrating a new event-based,...
-
Site Reliability Engineer
1 week ago
Singapore APPLE SERVICES PTE. LTD. Full timeRoles & ResponsibilitiesSummaryThe Apple Services Engineering (ASE) team is one of the most exciting examples of Apple's long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, Fitness+ and Apple Books. And they do it on a massive scale, meeting Apple's high expectations with...
-
Process engineer
2 days ago
Singapore The Chemical Engineer Full timeWhy Patients Need You Whether you are involved in the design and development of manufacturing processes for products or supporting maintenance and reliability, engineering is vital to making sure customers and patients have the medicines they need, when they need them. Working with our innovative engineering team, you'll help bring medicines to the...
-
Site Reliability Engineer
3 weeks ago
Singapore Vortexa Full timeVortexa is a cutting-edge company that leverages satellite data and AI to provide real-time insights into global energy flows. We're looking for a skilled Site Reliability Engineer to join our Data Services Team, responsible for the developer platform and Amazon AWS estate.The ChallengeOur platform processes massive amounts of data from various sources,...
-
Site Reliability Engineer
2 weeks ago
Singapore Tower Research Capital Full timeTower Research Capital Job DescriptionJob Title: Site Reliability EngineerJob Summary:We are seeking a highly skilled Site Reliability Engineer to join our team at Tower Research Capital. The successful candidate will be responsible for ensuring the continuous operation of our Linux-based trading infrastructure and addressing day-to-day operational needs.Key...
-
Site Reliability Engineer
3 weeks ago
Singapore ASIA GULF CLOUD PTE. LTD. Full timeRoles & ResponsibilitiesAbout SGB:SGB is a new digital bank that will offer a secure and integrated platform to access andmanage conventional and digital assets and financial solutions, including round-the-clock realtime settlement, trading connectivity, custody and asset management. It serves globalinvestors, innovators and institutions looking for a...
-
Senior Site Reliability Engineer
4 weeks ago
Singapore Ripple Labs Singapore Full timeAs a Senior Site Reliability Engineer at Ripple Labs Singapore, you will be responsible for ensuring the high availability and scalability of our systems. Your primary goal will be to design, implement, and maintain a robust and efficient infrastructure that can handle high traffic and complex distributed systems.Key Responsibilities:Design and implement...
-
Senior Site Reliability Engineer
2 weeks ago
Singapore Ripple Labs Singapore Full timeJob SummaryWe are seeking a highly skilled Senior Site Reliability Engineer to join our team at Ripple Labs Singapore. As a key member of our infrastructure team, you will be responsible for ensuring the high availability and reliability of our systems.Key ResponsibilitiesDesign, implement, and maintain high availability systems and infrastructureCollaborate...
-
Site Reliability Engineer
2 weeks ago
Singapore Helius Full timeJob Title: Site Reliability EngineerJob Summary: Helius is seeking a skilled Site Reliability Engineer to join our team. As a Site Reliability Engineer, you will be responsible for designing, implementing, and operating highly scalable and reliable systems. Your main focus will be on ensuring the smooth operation of our services, resolving technical issues,...
-
Site Reliability Engineer
3 weeks ago
Singapore ITCAN PTE. LIMITED Full timeRoles & ResponsibilitiesRoles & Responsibilities:The Site Reliability Engineer (SRE) combines software development and system engineering to build and run distributed solutions in a secured multi-tier heterogeneous environment to safeguard, provide and continuously improve the software and systems behind the organization’s cloud platform solutions.The Job:...
-
Senior Site Reliability Engineer
2 weeks ago
Singapore Sea Full timeAbout SeaSea is a cutting-edge technology company with a hyper-growing business scale, transforming complex problems into technical challenges. Our team of passionate engineers is dedicated to delivering world-class experiences for our users.The Games Site Reliability Engineer (SRE) team at Sea Labs Indonesia plays a crucial role in ensuring the stability...
-
Site reliability engineer
6 days ago
Singapore Infosight Software And Consulting Services Private Limited Full timeWe can consider EP & Singaporean / PR. Job Role Site Reliability Engineer Experience 5 to 7 Years Work Location Singapore Budget 7800-8000 SGD Duration 6 months, 12 months renewable contract Job Description Strong hands-on experience with using and designing VMware solutions such as NSX-T, v Realize Suite, v Sphere/v Center is mandatory. Strong working...
-
Chief Site Reliability Engineer
2 weeks ago
Singapore Snaphunt Full timeThe OpportunityWe're seeking an experienced Site Reliability Engineer to empower users with a rich feature set, high availability, and stellar performance at First Digital Finance Corp.As we expand customer deployments, the ideal candidate will deliver insights from massive-scale data in real-time, collaborating with a cross-functional team to develop...
-
Reliability Engineer
3 weeks ago
Singapore ADDVALUE INNOVATION PTE LTD Full timeRoles & ResponsibilitiesReliability EngineerResponsibilities Work with product development teams to develop relaibility requirements, establish a reliability / test program and perform appropriate analyse to ensure that new products meet all the relaibility targets. Perform risk / reliabilty analysis (FMEA, FMECA, MTBF) for existing and new products Able...