Senior Backend Engineer - AML Engine Orchestration

Byte Dance

Singapore

3-5 Years

Save

Posted 10 days ago
Be among the first 10 applicants

Early Applicant

Job Description

Responsibilities

Team Introduction The mission of our AML team is to push next-generation machine learning algorithms and platforms for the recommendation system, ads ranking and search ranking in our company. We also drive substantial impact on core businesses of the company. Responsibilities: 1. Resource Efficiency Optimization in Distributed Orchestration and Scheduling: - Develop and extend distributed orchestration frameworks within the Kubernetes/Godel ecosystem. Select appropriate frameworks based on different business scenarios, and optimize cluster utilization and load balancing strategies according to the specific characteristics of each scenario - Integrate and expand AutoScaling and automatic parallelization capabilities for various models and tasks. Employ load modeling and analytic methods for different models to automatically optimize resource requests, achieving large-scale improvements in resource usage efficiency and global optimality - Responsible for preemption and re-scheduling mechanisms for services with different prioritties, and manage automatic resource multiplexing across different clusters and resource types handle scheduling and load adaptation across multi-datacenter, multi-region, and multi-cloud environments. 2. Building Training System Architecture for Next-Generation Ultra-Large and Ultra-Deep Recommendation Models: - Develop a flexible, elastic and robust distributed training runtime focused on hyper-scaled embeddings and large-scale GPU training - Design and optimize distributed computing APIs and runtimes geared towards future recommendation and ads model paradigms (e.g., reinforcement learning, fine-tuning and/or distillation) - Collaborate with platform teams to enhance the diagnosability and usability of distributed training systems. 3. Constructing Online Orchestration Architecture for Next-Generation Recommendation Systems: - Build a robust distributed model inference architecture for online learning scenarios involving hyper-scaled embeddings - Optimize the usability of online recommendation and ads model architectures and MLops workflows.

Qualifications

Minimum Qualifications - Bachelor's degree or above, majoring in Computer Science, Engineering or related fields. - Strong programming and coding experience with at least one modern language such as Golang, Python. - Experience contributing to the large scale distributed systems, multi-tenant systems (architecture, reliability and scaling). - Strong analytical abilities and problem solving. - Good communication, self-motivation, engineering practice, documentation, etc. - At least 3 years of relevant experience. Preferred Qualifications - Familiar with large-scale distributed scheduling systems like Kubernetes, Yarn, Flink and/or Spark - Familiar with opensourced orchestration frameworks like VeRL, vLLM, Ray or TFX, etc.

More Info

Job Type:

Permanent Job

Industry:

IT /Computers - Software

Function:

Machine Learning / Software Engineering

Employment Type:

Full time

About Company

Byte DanceJob Source: jobs.bytedance.com

ByteDance is a technology company operating a range of content platforms that inform, educate, entertain and inspire people across languages, cultures, and geographies.
Dedicated to building global platforms of creation and interaction, ByteDance now has a portfolio of applications available in over 150 markets and 75 languages. For example, TikTok, Helo, Vigo Video, Douyin, and Huoshan.
Dedicated to building global platforms of creation and interaction, ByteDance now has a portfolio of applications available in over 150 markets and 75 languages. For example, TikTok, Helo, Vigo Video, Douyin, and Huoshan.

Job ID: 124983279

Jobs by Skill - IT

Jobs by Skill - Non IT

International Jobs

Jobs in Top Cities

Popular Jobs

Last Updated: 29-06-2026 05:45:19 AM

Homejobs in SingaporeSenior Backend Engineer - AML Engine Orchestration