About Me

I work at ByteDance Seed, as a senior research scientist and a member of TopSeed program. I am working on advancing the reasoning and agentic capabilities of Large Language Models (LLMs) and general Agent foundation models. As one of the algorithm leads for general agent optimization, I contribute to the ByteDance Seed model series (e.g., Seed 2.1 to Seed 1.5) and Doubao products, including building the Agent Task Mode/Expert Mode (豆包办公任务模式和专家模式). We are hiring research interns and looking for academic cooperation, please feel free to email me at wanjun@bytedance.com

Prior to that, I worked at Huawei Noah’s Ark Lab as a research scientist and a member of Huawei TopMind program. I received my Ph.D. degree from the School of Computer Science and Engineering in Sun Yat-sen University (SYSU), as a member of joint Ph.D. program between SYSU and Microsoft Research Asia (MSRA).

As a joint Ph.D. student, I was advised by Dr. Ming Zhou, Prof. Jian Yin and Prof. Jiahai Wang. I was a research intern in the Natural Language Computing Group of MSRA, and was mentored by Dr. Nan Duan

I won the Microsoft Research Fellowship Award (11 outstanding Ph.D. in Asia-Pacific area each year) in 2021, and is selected as a member of Huawei TopMind program in 2023.

I published over 50+ papers in top-tier AI conferences and journals, including NeurIPS, ICLR, ACL, EMNLP, TASLP, NAACL, AAAI, IJCAI, ISSTA, etc.

My Research Interests

  • Large Language Model
  • Agent Foundation Model and Reinforcement Learning
  • Reasoning towards AGI

🔥 News

  • 2026.06 Release Seed 2.1 series, a next-generation agent foundation model for real-world productivity. Grateful to contribute to the general agent optimization as one of the algorithm core contributors.
  • 2026.06 Released the Office Task Mode of 豆包专业版 (Doubao Pro), where I was fortunate to contribute as one of the algorithm leads.
  • 2026.04 Released Agent-World (paper), exploring scalable real-world environment synthesis for evolving general agent intelligence. Check out our demo!
  • 2026.02 Release Seed 2.0 as a core contributor, the leading Agent foundation models!
  • 2026.02 Release 豆包专家模式 (Doubao Expert Mode) — long-CoT reasoning that delivers expert-level answers to complex professional problems — and 豆包超能模式
  • 2025.12 Release Seed 1.8 as a core contributor.
  • 2025.09 Release UI-TARS-2, a multi-modal unified agent models capable of coding, tool using and GUI operation, achieving leading performance.
  • 2025.04: 🎉 Joined ByteDance Seed Edge team as Senior Research Scientist focusing on Large Language Models and Agent foundation models!
  • 2025.04 Released ReTool: A reinforcement learning-based multi-turn tool-use agent training framework!
  • 2025.04: 🎉 Seed-VL-v1.5 technical report released, advancing multi-modal models with great understanding and reasoning capabilities!
  • 2025.04: 🎉 Seed-Thinking-v1.5 technical report released, advancing superb reasoning models with reinforcement learning
  • 2025.01: 🎉 Released UI-TARS: Industry’s open-source GUI+Game Agent foundation model with 6.2K+ GitHub stars!
  • 2024.06: 🎉 Joined ByteDance TopSeed program as a Senior Research Scientist!

📖 Education

  • 2018.09 - 2023.06, Ph.D. in Computer Science and Technology, Sun Yat-sen University (SYSU), Joint Ph.D. program with Microsoft Research Asia (MSRA)
  • 2014.09 - 2018.06, Bachelor in Software Engineering, School of Data Science and Computer Science, Sun Yat-sen University (SYSU)

💼 Work Experience

  • 2024.06 - Present,ByteDance Seed Edge Team - Senior Research Scientist
    • Role: Senior Research Scientist in LLMs and Agents, member of TopSeed program, one of the algorithm leads for general agent optimization
    • Project Experience:
      • Seed 1.8 & Seed 2.0 & Seed 2.1 (Leading next-generation Agent foundation models, core contributor)
      • Office Task Mode of 豆包专业版 (Doubao Pro) — as one of the algorithm leads
      • Doubao Expert Mode (long-CoT reasoning that delivers expert-level answers to complex professional problems, improving quality in STEM, coding, domain expertise and creative writing) & Super Mode (豆包专家模式 & 豆包超能模式)
      • Agent-World (scalable real-world environment synthesis for evolving general agent intelligence)
      • Seed-Thinking long chain-of-thought reasoning model
      • Seed-Agent foundation model:
        • UI-TARS & UI-TARS-2 (Industry-leading open-source GUI+Game Agent foundation models) training
        • ReTool (Agent multi-turn tool calling reinforcement learning training framework)
        • MCP-enhanced general Agent foundation model training
  • 2023.06 - 2024.06, Huawei Noah’s Ark Lab - Speech & Semantic Lab - Research Scientist (TopMind Program)
    • Project Experience: Research scientist in Large Language Models, specializing in PanGu foundation model instruction tuning, data flywheel, Agent super-alignment and complex reasoning research and implementation.
  • 2018.06 - 2023.06, Microsoft Research Asia - Joint Ph.D. Program Long-term Internship
    • Mentor: Dr. Nan Duan and Dr. Ming Zhou

💬 Academic Supervision

  • Ph.D. Advisors: Dr. Ming Zhou (Microsoft Research Asia), Prof. Jian Yin (SYSU), Prof. Jiahai Wang (SYSU)
  • Mentor at MSRA: Dr. Nan Duan (Natural Language Computing Group)

🔬 Research Internship

  • 2018.06 - 2023.06, Research Intern, Natural Language Computing Group, Microsoft Research Asia (MSRA), Beijing
    • Long-term internship as part of joint Ph.D. program
    • Mentor: Dr. Nan Duan

🎖 Honors and Awards

  • 2024 ByteDance TopSeed Program
  • 2023 ACM Outstanding Doctoral Thesis Award (China-Guangzhou Chapter)
  • 2023 Huawei TopMind Program
  • 2021 Microsoft Research Fellowship Award (11 outstanding Ph.D. students in computer science in the Asia-Pacific region each year)
  • 2021 Baidu Scholarship (Global Top 40)
  • 2020 National Scholarship for Doctoral Students (Top 0.2%)
  • 2016 First Prize Scholarship

🏆 Competition Awards

  • 2023 1st Place - CVPR 2023 Ego4D Challenge for Episodic Memory Natural Language Queries
  • 2022 3rd Place - ECCV 2022 Ego4D Challenge for Episodic Memory Natural Language Queries
  • 2018 Outstanding Award - Global (Nanjing) AI Application Competition
  • 2016 National Second Prize - National Mathematical Contest in Modeling
  • 2018 3rd & 7th Place - FASHIONAI Global Challenge (Semi-finals)

📝 Publications

Works in Seed

Large Language Model Reasoning

General Agent Model & System

Multi-modal Agent (GUI etc.)

Tool-Learning Agent

Code Agent

Agent Memory

Agent-driven Training

Benchmark and Evaluation

Self-Learning of LLMs

General LLM Training

Previous Work Before 2023

Multi-Modal

Knowledge-enhanced Language Model Reasoning