Cloud Native Engineer, ARK Large Model Platform
3 weeks ago
ByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.
Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create content.
Why Join Us
Creation is the core of ByteDance's purpose. Our products are built to help imaginations thrive. This is doubly true of the teams that make our innovations possible.
Together, we inspire creativity and enrich life - a mission we aim towards achieving every day.
To us, every challenge, no matter how ambiguous, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always.
At ByteDance, we create together and grow together. That's how we drive impact - for ourselves, our company, and the users we serve.
Join us.
About the Team
The Applied Machine Learning (AML) - Enterprise team provides machine learning platform products on VolcanoEngine with cloud native resource scheduling system which intelligently orchestrates different tasks and jobs with minimised costs of every experiment and maximised resource utilisation, rich modelling tools including customised machine learning tasks and web IDE, and multi-framework high performance model inference services.
In 2021, through VolcanoEngine, we released this machine learning infrastructure to the public, to provide more enterprises with reduced costs of computation power, lower barriers to machine learning engineering and deeper developments in AI capabilities.
Responsibilities
Responsible for Ark Large Model Platform development on Volcano Engine, researching systematic solutions on large model solution implementations and applications in various industries, striving to reduce the IT cost of large model applications, meeting the users' ever-growing demand for intelligent interaction and improving the lifestyle and communications of users in the future world.
- Maintain a large-scale AI cluster and develop state-of-the-art machine learning platforms to support a diverse group of stakeholders.
- Tackle extremely challenging tasks which include, but are not limited to, delivering highly efficient training and inference for large language models, managing extremely effective distributed training jobs across clusters with over 10,000 nodes and GPU chips, and constructing highly reliable ML systems with unparalleled scalability.
- The work encompasses various aspects of LLMOps (Large Language Model Operations), such as resource scheduling, task orchestration, model training, model inference, model management, dataset management, and workflow orchestration.
- Investigate cutting-edge technologies related to large language models, AI, and machine learning at large, such as state-of-the-art distributed training systems with heterogeneous hardware, GPU utilization optimization, and the latest in hardware architecture.
- Employ a variety of technological and mathematical analyses to enhance cluster efficiency and performance.
Qualifications
Minimum Qualifications
- B. Sc or higher degree in Computer Science or related fields from accredited and reputable institutions with at least 3 years of R&D experience in the fields of cloud computing or large-scale model systems.
- Experience in Golang/C++/Cuda development with a solid understanding of Linux systems and popular cloud platforms such as Volcano Engine Cloud, AWS, and Azure Cloud.
- Profound knowledge of cloud-native orchestration technologies like Kubernetes, coupled with experience in large-scale cluster maintenance, job scheduling optimization, and cluster efficiency enhancement.
- A strong grasp on various foundational areas of computer science, including computer networking, the Linux file system, object storage services, and SQL as well as NoSQL databases.
- Self-motivated, thirst for innovation, collaborative working aptitude, and consistently uphold high standards in coding and documentation quality.
Preferred Qualifications:
- Experience in developing ML platforms or MLOps platforms. Experience in distributed machine learning model training, ML model fine-tuning, and deployment.
ByteDance is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At ByteDance, our mission is to inspire creativity and enrich life. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.
Tell employers what skills you have
Machine Learning
Hardware Architecture
Scalability
Kubernetes
Azure
Hardware
Cloud Computing
SQL
Networking
Orchestration
Scheduling
Databases
Linux
-
Cloud Native Engineer
3 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRoles & ResponsibilitiesByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the...
-
Cloud Native Engineer for Large Model Platform
3 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRoles and ResponsibilitiesAt ByteDance PTE. LTD., we are seeking an exceptional Cloud Native Engineer for Large Model Platform to join our team. The ideal candidate will have a strong background in cloud computing and large-scale model systems, with expertise in Golang, C++, Cuda, and Linux systems. Key Responsibilities:Develop and maintain large-scale AI...
-
Site Reliability Engineer
3 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRoles & ResponsibilitiesByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the...
-
Singapore BYTEDANCE PTE. LTD. Full timeRoles & ResponsibilitiesByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the...
-
Backend Engineer, ARK Large Model Platform
3 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRoles & ResponsibilitiesByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the...
-
Backend Engineer
3 weeks ago
Singapore BYTEDANCE PTE. LTD. Full timeRoles & ResponsibilitiesByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the...
-
Large Language Model Algorithm Engineer
2 months ago
Singapore BYTEDANCE PTE. LTD. Full timeRoles & ResponsibilitiesFounded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the China market, including Toutiao, Douyin, and Xigua, ByteDance has made it easier and more fun for people to connect with, consume, and create...
-
Singapore BYTEDANCE PTE. LTD. Full timeAbout the RoleAs a Senior Software Engineer, Large Model Development at ByteDance PTE. LTD., you will be responsible for the development of the Ark Large Model Platform on Volcano Engine. This involves researching and implementing systematic solutions for large model applications in various industries, with a focus on reducing the IT cost of large models and...
-
Product Solution Architect
3 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRoles & ResponsibilitiesByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the...
-
Product Solution Architect, Volcano ARK
3 weeks ago
Singapore BYTEPLUS PTE. LTD. Full timeRoles & ResponsibilitiesByteDance will be prioritizing applicants who have a current right to work in Singapore, and do not require ByteDance's sponsorship of a visa.Founded in 2012, ByteDance's mission is to inspire creativity and enrich life. With a suite of more than a dozen products, including TikTok, Helo, and Resso, as well as platforms specific to the...
-
Large Language Model Algorithm Engineer
3 weeks ago
Singapore BYTEDANCE PTE. LTD. Full timeAbout the RoleByteDance PTE. LTD. is seeking a talented AI researcher to join our Machine Learning Platform team. As a key member of our team, you will contribute to the advancement of next-generation artificial intelligence technologies, including large models, multimodal capabilities, text comprehension, generation algorithms, and reinforcement learning...
-
Cloud Native Platform Engineer
5 days ago
Singapore D L RESOURCES PTE LTD Full timeAbout the Role:We are seeking an experienced Cloud Native Platform Engineer to join our team at D L Resources PTE LTD. In this role, you will be responsible for designing, implementing, and configuring OpenShift clusters and containerized applications.Responsibilities:Design and implement scalable and secure containerized environments using Red Hat...
-
Senior Cloud Native Software Engineer
16 hours ago
Singapore Cognizant Full timeJob DescriptionWe are seeking a highly skilled Senior Cloud Native Software Engineer to join our team at Cognizant.About the RoleThe successful candidate will be responsible for designing, developing, and deploying cloud-native web applications on major cloud service providers, with a preference for Amazon Web Services' technology...
-
Cloud Native Software Engineer
5 days ago
Singapore IDC TECHNOLOGIES (SINGAPORE) PTE. LTD. Full timeJob DescriptionWe are seeking a highly skilled Cloud Native Software Engineer to join our team at IDC Technologies (Singapore) Pte. Ltd. As a key member of our engineering team, you will be responsible for designing, developing, and deploying scalable cloud-native applications using cutting-edge technologies.About UsIDC Technologies (Singapore) Pte. Ltd. is...
-
Cloud Engineer, Technology Group
3 weeks ago
Singapore GIC Private Limited Full timeCloud Native Transformation LeadGIC Private Limited is a leading global long-term investor, and we are seeking a Cloud Native Transformation Lead to drive our cloud native transformation efforts. As a key member of our Technology Group, you will play a crucial role in enabling our cloud native strategy and roadmap.Key ResponsibilitiesEnable the next...
-
Cloud Native Software Engineer
2 weeks ago
Singapore Singtel Full timeTransformative Cloud SolutionsWe are seeking an exceptional Cloud Native Software Engineer to join our team in shaping the future of cloud-based technologies. As a key member of our engineering team, you will be responsible for designing, developing, and deploying cutting-edge cloud solutions that drive business growth and innovation.Key...
-
Cloud Native Architect
2 weeks ago
Singapore Dell Global BV Singapore Branch (7032) Full timePrincipal Cloud Native Architect As a Principal Cloud Native Architect at Dell Global BV Singapore Branch (7032), you will be part of a collaborative environment where teams care about the product they create, how it's created, and the impact it has on customers' business objectives. Your primary responsibility will be to provide technical direction...
-
Cloud Native Software Developer
2 weeks ago
Singapore Helius Full timeCloud Native Software Developer We are seeking a Cloud Native Software Developer to join our team at Helius. As a Cloud Native Software Developer, you will be responsible for designing, developing, and deploying cloud native applications and services. Your expertise in Java, HTML, CSS, and JavaScript frameworks such as Angular will be essential in building...
-
Cloud Native Developer
6 days ago
Singapore LUXOFT INFORMATION TECHNOLOGY (SINGAPORE) PTE. LTD. Full timeRoles & ResponsibilitiesProject Description:We are looking for a talented Cloud Native developer proficient with building microservices on the back of Pivotal Cloud Foundry. We are looking at someone with strong experience in this area. The person would be a part of an experienced and established team working on a major rebuild of our...
-
Senior Software Engineer
2 weeks ago
Singapore OCBC Bank Full timeAbout the RoleAs a Cloud Native Software Architect, you will be responsible for designing and implementing cloud-based architectures that meet the needs of our growing organization. This is a fantastic opportunity to work with a talented team and contribute to the development of cutting-edge cloud solutions.Key ResponsibilitiesDesign and implement...