Research Scientist - Model Capability Boundary Exploration and AI Data Flywheel System Developm[...]

ByteDance

San Jose (CA)

On-site

USD 212,800 - 450,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health/Dental/Vision insurance
401(k) with company match
Parental leave
Wellbeing benefits
Paid time off: holidays, sick days, P‑

Job summary

ByteDance seeks a Research Scientist to explore model capability boundaries and contribute to AI data flywheel system development. Join the Global Frontier Tech Recruitment Program, starting 2027, in San Jose. Doctorate preferred with strong research background in LLMs and agent systems.

The team works on MaaS for LLMs, data-driven feedback loops, and end-to-end AI systems, including logs, prompts, memory, tools, and workflows.

Qualifications

  • Ph.D. in CS/ML/AI/Data Science with strong research exposure.
  • Research in LLM alignment, evaluation, or data-centric AI is valued.
  • Publications or internships demonstrating independent research ability.
  • Ability to formulate problems and run experiments independently.

Responsibilities

  • Build a next-generation big model as a service platform for many LLM-based apps.
  • Develop/offline training, fine-tuning, online inference, and model management.
  • Manage GPU resources and provide efficient computing power.

Skills

LLM post-training & alignment
Model evaluation
Test-time scaling
Agent systems
Data curation & optimization
Publications & research projects

Education

Ph.D. in Computer Science / ML / AI / Data Science

Job description

Research Scientist - Model Capability Boundary Exploration and AI Data Flywheel System Development - Global Frontier Tech Recruitment Program - 2027 Start (PhD)

Location: San Jose

Team: Technology

Employment Type: Regular

Job Code: A04845

We are looking for talented individuals to join our team in 2027. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth. Launch your career where inspiration is infinite at our Company. Successful candidates must be able to commit to an onboarding date by end of year 2027. Please state your availability and graduation date clearly in your resume.

Team Introduction

The Applied Machine Learning Ark team combines system engineering and machine learning to develop and operate Large Language Model (LLM) service platforms that offer businesses Model-as-a-Service (MaaS) solutions, serving both large model providers and downstream users. The US team drives the design, development, and operation of MaaS solutions across the US and international markets outside mainland China. We are building full-stack, end-to-end solutions spanning text and multimodal LLM algorithms, LLM training/fine-tuning/inference frameworks, prompt engineering, model alignment, and intelligent agent systems. Beyond model serving, we operate large-scale log analytics pipelines that process massive volumes of invocation logs from text models, multimodal models, and agent systems — extracting usage patterns, quality signals, and actionable insights to inform model improvement, system optimization, and product decisions through continuous, data-driven feedback loops. We are actively seeking talented engineers and researchers specializing in Large Language Models and AI Agent systems to join our dynamic team.

Project Scope

With foundation models gradually being applied in real ToB scenarios, AI system optimization now extends beyond the foundation model itself to include a complex business system composed of the model, prompt, memory, tools, skills, workflow, and the external environment. Compared to offline benchmarks, real-world cases offer greater potential for optimization but also present challenges such as larger data volumes, higher noise levels, more diverse scenarios, greater structural heterogeneity, and limited user feedback, making them difficult to be directly utilized. Relying on the real-world data accumulated on the Volcano Ark case platform, this project aims to unify logs, cases, feedback, and environmental information into structured objects that are understandable, and attributable, and optimizable. By integrating AI-assisted tools to guide users in providing efficient feedback, it aims to build an AI data flywheel system tailored to real scenarios. This system will both support foundation model iteration and address issues related to environment, memory, tools, and workflows within the business system, focusing on developing agent optimization capabilities that enhance SA/FDE’s efficiency in supporting customers.

Responsibilities

We are looking for talented individuals to join our team in 2027. As a graduate, you will get opportunities to pursue bold ideas, tackle complex challenges, and unlock limitless growth. Launch your career where inspiration is infinite at our Company. Successful candidates must be able to commit to an onboarding date by end of year 2027. Please state your availability and graduation date clearly in your resume.

  • Building a next-generation big model as a service platform to serve hundreds of LLMs based applications
  • To develop and maintain the big model as a service platform, including offline training/finetuning, online inference, model management, and resource orchestration, etc.
  • To manage a huge number of GPU resources and provide computing power efficiently
Qualifications

Minimum Qualifications:

  • Currently pursuing or recently completed a Ph.D. in Computer Science, Machine Learning, Artificial Intelligence, Data Science, or a related technical field.
  • Research experience in one or more of the following areas: LLM post-training and alignment, model evaluation, test-time scaling, agent systems, or large-scale data curation and optimization.
  • Demonstrated research ability through publications, substantial research projects, or internships.
  • Ability to work independently on open-ended research problems, from problem formulation to experimental execution.

Preferred Qualifications:

  • Strong interest in foundation models and data-centric AI, particularly in how large models can improve over time through better data, feedback, and system design. Relevant directions include data flywheels, continual learning, data curation and valuation, and the co-design of algorithms and infrastructure.
  • A strong publication record with multiple first-author papers, in areas of machine learning, NLP, data mining, or related fields.
  • Internship or research experience in similar fields, ideally with experience with scalable ML systems, especially those involving real-world deployment, feedback loops, or human-in-the-loop pipelines.
  • Strong motivation to connect research with practice, and to build end-to-end AI systems spanning modeling, data, evaluation, and infrastructure.
Job Information

The base salary range for this position in the selected city is $212800 - $450000 annually.

Compensation may vary outside of this range depending on a number of factors, including a candidate’s qualifications, skills, competencies and experience, and location. Base pay is one part of the Total Package that is provided to compensate and recognize employees for their work, and this role may be eligible for additional discretionary bonuses/incentives, and restricted stock units.

Benefits may vary depending on the nature of employment and the country work location. Employees have day one access to medical, dental, and vision insurance, a 401(k) savings plan with company match, paid parental leave, short-term and long-term disability coverage, life insurance, wellbeing benefits, among others. Employees also receive 10 paid holidays per year, 10 paid sick days per year and 17 days of Paid Personal Time (prorated upon hire with increasing accruals by tenure).

The Company reserves the right to modify or change these benefits programs at any time, with or without notice.

For Los Angeles County (unincorporated) Candidates:
  • Interacting and occasionally having unsupervised contact with internal/external clients and/or colleagues;
  • Appropriately handling and managing confidential information including proprietary and trade secret information and access to information technology systems;
  • Exercising sound judgment.
Reasonable Accommodation

ByteDance is committed to providing reasonable accommodations in our recruitment processes for candidates with disabilities, pregnancy, sincerely held religious beliefs or other reasons protected by applicable laws. If you need assistance or a reasonable accommodation, please reach out to us at https://tinyurl.com/RA-request

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist - Model Capability Boundary Exploration and AI Data Flywheel System Developm[...]
Research Scientist - Model Capability Boundary Exploration and AI Data Flywheel System Developm[...]

ByteDance • Seattle (WA)

On-site
USD 202,000 - 369,000
Medical, dental, and vision insurance
401(k) plan with company match
Paid parental leave
Applied Scientist - LLM Training System as a Service - Global Frontier Tech Recruitment Program[...]
Applied Scientist - LLM Training System as a Service - Global Frontier Tech Recruitment Program[...]

ByteDance • San Jose (CA)

On-site
USD 212,000 - 450,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+3
Technology - Infrastructure Global Frontier Tech Recruitment Program - 2027 Grad San Jose Regular
Technology - Infrastructure Global Frontier Tech Recruitment Program - 2027 Grad San Jose Regular

ByteDance • San Jose (CA)

On-site
USD 212,000 - 388,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+2
Research Scientist - AI Compute & DPU - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
Research Scientist - AI Compute & DPU - Global Frontier Tech Recruitment Program - 2027 Start (PhD)

ByteDance • Seattle (WA)

On-site
USD 202,000 - 369,000
Software Engineer Intern (Applied Machine Learning-Enterprise) - 2026 Start (PhD)
Software Engineer Intern (Applied Machine Learning-Enterprise) - 2026 Start (PhD)

ByteDance • San Jose (CA)

On-site
Health insurance
Life insurance
Wellbeing benefits
+1
Research Scientist - AI Compute & DPU - Global Frontier Tech Recruitment Program - 2027 Start (PhD)
Research Scientist - AI Compute & DPU - Global Frontier Tech Recruitment Program - 2027 Start (PhD)

ByteDance • San Jose (CA)

On-site
USD 212,000 - 388,000
Medical, dental and vision insurance
401(k) with company match
Paid parental leave
+6
Algorithm Application Scientist - Large Model Applications - Global Frontier Tech Recruitment P[...]
Algorithm Application Scientist - Large Model Applications - Global Frontier Tech Recruitment P[...]

ByteDance • San Jose (CA)

On-site
USD 212,000 - 450,000
Medical, dental, and vision insurance
401(k) with company match
Paid parental leave
+2
Research Scientist in AI Foundation Model Infrastructure - Seed - Graduates - 2027 Start (PhD) [...]
Research Scientist in AI Foundation Model Infrastructure - Seed - Graduates - 2027 Start (PhD) [...]

ByteDance • Seattle (WA)

On-site
USD 232,000 - 428,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+1
Senior Research Scientist - Machine Learning System
Senior Research Scientist - Machine Learning System

ByteDance • San Jose (CA)

On-site
USD 212,800 - 387,600
Medical insurance
Dental insurance
Vision insurance
+5
Data Lake Infrastructure & Data Analytics Research Engineer Graduate (AML-Ark-US) - 2027 Start [...]
Data Lake Infrastructure & Data Analytics Research Engineer Graduate (AML-Ark-US) - 2027 Start [...]

Pangle • San Jose (CA), Northern (KY)

Hybrid
USD 140,000 - 210,000