Large Model Application Algorithm Research Scientist-International Content Security Algorithm R[...]

TikTok

Singapore

On-site

SGD 80,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TikTok in Singapore is seeking a machine learning expert for the International Content Safety Algorithm Research Team. You will develop and enhance models that ensure a safe user environment across ByteDance's products.

The ideal candidate holds a PhD in a relevant field and has extensive experience in ML, NLP, and computer vision. Strong programming skills in Python or C++ are a must, along with excellent problem-solving abilities.

Qualifications

  • PhD degree in Computer Science, Electronics, or related fields is required.
  • Extensive experience in ML/CV/NLP/Recommendation Systems.
  • Publications in related conferences preferred.

Responsibilities

  • Develop and iterate on machine learning models for content safety.
  • Monitor potential threats and ensure compliance.
  • Lead development of foundational large models.

Skills

Machine Learning
Natural Language Processing
Computer Vision
Recommendation Systems
Problem-solving
Analytical skills
Python
C++

Education

PhD in Computer Science or related fields

Job description

Responsibilities

The International Content Safety Algorithm Research Team is dedicated to maintaining a safe and trustworthy environment for users of ByteDance's international products. We develop and iterate on machine learning models and information systems to identify risks earlier, respond to incidents faster, and monitor potential threats more effectively. The team also leads the development of foundational large models for products. In the R&D process, we tackle key challenges such as data compliance, model reasoning capability, and multilingual performance optimization. Our goal is to build secure, compliant, and high-performance models that empower various business scenarios across the platform, including content moderation, search, and recommendation.

Project Background

In recent years, Large Language Models (LLMs) have achieved remarkable progress across various domains of natural language processing (NLP) and artificial intelligence. These models have demonstrated impressive capabilities in tasks such as language generation, question answering, and text translation. However, reasoning remains a key area for further improvement. Current approaches to enhancing reasoning abilities often rely on large amounts of Supervised Fine‑Tuning (SFT) data. Acquiring such high‑quality data is expensive and poses a barrier to scalable model development and deployment. To address this, OpenAI's o1 series of models have made progress by increasing the length of the Chain‑of‑Thought (CoT) reasoning process. While this technique has proven effective, how to efficiently scale this approach in practical testing remains an open question. Recent research has explored alternative methods such as Process‑based Reward Model (PRM), Reinforcement Learning (RL), and Monte Carlo Tree Search (MCTS) to improve reasoning. However, these approaches still fall short of the general reasoning performance achieved by OpenAI's o1 series of models. Notably, the recent DeepSeek R1 paper suggests that pure RL methods can enable LLM to autonomously develop reasoning skills without relying on the expensive SFT data, revealing the substantial potential of RL in advancing LLM capabilities.

Project Challenges
  • Design of Reward Models: In the RL process, designing an effective reward model is crucial. It must accurately reflect the effectiveness of the reasoning process and guide the model to iteratively improve its reasoning ability. This involves not only setting appropriate evaluation criteria across different tasks, but also ensuring the reward model to adapt dynamically during training to match the evolving model performance.
  • Stability of the Training Process: In the absence of high‑quality SFT data, ensuring stable training in RL becomes a major challenge. RL often involves extensive exploration and trial‑and‑error, which may lead to unstable training or even performance degradation. Developing robust training strategies is essential to ensure the reliability and effectiveness of the training process for models.
  • Expanding from Mathematics and Code Tasks to Natural Language Tasks: Current RL reasoning methods are primarily applied to mathematics and code tasks, where CoT data is more abundant. However, natural language tasks are more open and complex. Expanding from successful RL strategies to natural language processing tasks requires in-depth research and innovation in both data design and RL methodology to enable cross‑task general reasoning capabilities.
  • Improving Reasoning Efficiency: While maintaining high reasoning quality, improving reasoning efficiency is another critical challenge. Efficient reasoning directly impacts the model's practicality and cost‑effectiveness in real‑world applications. Approaches such as knowledge distillation (transferring knowledge from complex models to smaller models) can be explored to reduce computational resource consumption, or the use of Long Chain‑of‑Thought (Long‑CoT) techniques to improve Short‑CoT models to balance reasoning accuracy with computational efficiency.
Qualifications
  • Got PhD degree in Computer Science, Electronics, or other related fields.
  • Extensive experience in ML/CV/NLP/Recommendation Systems, including but not limited to:
  • Participation in competitions or industry projects in ML, Data Mining, CV, NLP, or Multimodal.
  • Publications in conferences in ML, data mining, AI, or large models (e.g., KDD, WWW, NIPS, ICML, CVPR, ACL, AAAI etc).
  • Plus points:
  • Research experience or innovation in large models or RL.
  • Strong hands‑on skills with contributions to large model projects in the open‑source community.
  • Practical experience in deploying large models in real‑world business scenarios.
  • Strong programming skills and proficient in Python/C++ or other relevant programming languages.
  • Outstanding problem‑solving and analytical skills, with a passion for tackling challenging problems.
  • Strong enthusiasm for technology, with excellent communication skills and collaborative mindset.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Machine Learning Algorithm Research Engineer - Soaring Star Talent Program
Machine Learning Algorithm Research Engineer - Soaring Star Talent Program

ByteDance • Singapore

On-site
SGD 150,000 - 210,000
Research Scientist - LLM Applications for International E-commerce Scenarios - Global Frontier [...]
Research Scientist - LLM Applications for International E-commerce Scenarios - Global Frontier [...]

TikTok • Singapore

On-site
SGD 100,000 - 160,000
Senior Large Language Model Algorithm Engineer/Expert
Senior Large Language Model Algorithm Engineer/Expert

Binance • Singapore

On-site
SGD 100,000 - 150,000
Machine Learning System Engineer - Data AML - Soaring Star Talent Program
Machine Learning System Engineer - Data AML - Soaring Star Talent Program

ByteDance • Singapore

On-site
SGD 80,000 - 120,000
AI Engineer (Managed Services)
AI Engineer (Managed Services)

Avepoint • Singapore

On-site
SGD 90,000 - 120,000
Flexible working hours
Access to high-performance GPU resources
Continued learning and development opportunities
Large Recommendation Model Algorithm Engineer - Global E-Commerce Singapore Regular
Large Recommendation Model Algorithm Engineer - Global E-Commerce Singapore Regular

Pangleglobal • Singapore

On-site
SGD 80,000 - 120,000
Research Scientist - Risk Control Vertical LLM Foundation and Agent - Global Frontier Tech Recr[...]
Research Scientist - Risk Control Vertical LLM Foundation and Agent - Global Frontier Tech Recr[...]

TikTok • Singapore

On-site
SGD 60,000 - 100,000
Senior Data Scientist: Model Risk & Data Analytics, Internal Audit
Senior Data Scientist: Model Risk & Data Analytics, Internal Audit

Visa Hunt • Singapore

On-site
SGD 120,000 - 180,000
Positive team atmosphere
Paid leave
Meals provided
principal engineer
principal engineer

DNA INFOTECH PTE. LTD. • Singapore

On-site
SGD 200,000 - 260,000
Senior AI/Machine Learning Engineer
Senior AI/Machine Learning Engineer

Good co India • Singapore

On-site
SGD 120,000 - 170,000