Yizhe Yang

杨毅哲

I received my Ph.D. in Computer Science from Beijing Institute of Technology in 2026, advised by Prof. Heyan Huang. From 2023 to 2024 I was a visiting Ph.D. student at Singapore Management University, funded by the China Scholarship Council and hosted by Prof. Ee-Peng Lim and Prof. Jing Jiang.

My research centers on conversational agents and LLM agents, in particular how a language model can sustain a long, goal-directed conversation that stays grounded in knowledge, consistent in persona, and strategically sound. Recent work covers counselor agents for motivational interviewing, faithful client and user simulation, and the preference bias that LLM-based dialogue planners silently acquire from simulated rewards.

My work spans both academia and industry: I have published at top-tier AI venues including ACL, EMNLP, AAAI, and TOIS, with representative works such as CAMI, Consistent Client Simulation, Speaker Verification in Agent-Generated Conversations, and MindLLM, the first bilingual foundation model pre-trained end-to-end at a Chinese university. I regularly serve as a reviewer for top-tier conferences and journals, including ICLR, NeurIPS, ICML, ACL ARR, AAAI, CIKM, and Artificial Intelligence.

Conversational Agent LLM Agents Mental-Health NLP User Simulation

News

2026.06
Defended my Ph.D. thesis, Multi-dimensional Modelling for Domain-specific Dialogue, and graduated from Beijing Institute of Technology.
2026.05
MIThinker and PUPPET are accepted to ACL 2026.
2026.01
Our analysis of preference bias in dialogue planners is accepted to ACM TOIS, with a companion paper as an AAAI 2026 Oral.
2025.11
Received the Wu Wen-Jun AI Science & Technology Progress Award, First Prize (吴文俊人工智能科技进步奖一等奖, CAAI).
2025.05
CAMI and Consistent Client Simulation are accepted to ACL 2025.
2024.03
MindLLM is released, the first bilingual foundation model pre-trained end-to-end at a Chinese university, with 30k+ downloads.

Education

2020.09 – 2026.06
Ph.D. in Computer Science, Beijing Institute of Technology
Advisor: Prof. Heyan Huang. Thesis: Multi-dimensional Modelling for Domain-specific Dialogue
2023.10 – 2024.10
Visiting Ph.D. Student, Singapore Management University
Hosts: Prof. Ee-Peng Lim, Prof. Jing Jiang. Funded by the China Scholarship Council
2016.09 – 2020.06
B.Eng. in Computer Science, Beijing Institute of Technology
GPA 3.8/4.0, ranked 9/218. Admitted to the direct Ph.D. track by recommendation

Selected Publications

  1. 2026
    MIThinker: A Plug-and-Play Policy-Optimized Thinker for Motivational Interviewing Counseling
    Yizhe Yang, Heyan Huang, Palakorn Achananuparp, Jing Jiang, Ee-Peng Lim
    Findings of ACL 2026CCF-A
  2. 2026
    PUPPET: Neural-Symbolic Standardized Patients for Mental Health
    Canyu Xu, Yuhang Ji, Zhiyuan Lv, Yi Yi, Yizhe Yang*, Yuanxing Ji, Chong Chen, et al.
    ACL 2026CCF-A
  3. 2026
    Planner Policies are Bias Learners: A Comprehensive Analysis of Preference Bias in LLM-based Dialogue Planner Policies
    Yizhe Yang, Heyan Huang, Hang Sun, Junyi Li, Yang Gao
    ACM TOISCCF-ASCI Q1doi
  4. 2026
    Simulated Rewards, Skewed Strategies: Tracing the Acquired Preference Bias in LLM-Based Dialogue Planners
    Heyan Huang, Yizhe Yang*, Hang Sun, Junyi Li, Yang Gao
    AAAI 2026CCF-AORAL
  5. 2026
    基于联律规则约束解码的对联生成方法
    杨毅哲, 黄河燕
    CCL 2026
  6. 2026
    Retrieval is NOT Always Needed: Exploring the Timing of Retrieval in Dynamic Retrieval-augmented Generation
    Yuan Liu, Heyan Huang, Yizhe Yang, Zhen Zeng, Yue Zhou, Zhaoming Wu, Yang Gao
    Expert Systems with ApplicationsSCI Q1 TOP
  7. 2025
    CAMI: A Counselor Agent Supporting Motivational Interviewing through State Inference and Topic Exploration
    Yizhe Yang, Heyan Huang, Palakorn Achananuparp, Jing Jiang, et al., Ee-Peng Lim
    ACL 2025CCF-A
  8. 2025
    Consistent Client Simulation for Motivational Interviewing-based Counseling
    Yizhe Yang, Heyan Huang, Palakorn Achananuparp, Jing Jiang, et al., Ee-Peng Lim
    ACL 2025CCF-A
  9. 2025
    Fundamental Capabilities and Applications of Large Language Models: A Survey
    Jiawei Li, Yang Gao, Yizhe Yang, et al.
    ACM Computing SurveysSCI Q1 TOP
  10. 2025
    EvoWiki: Evaluating LLMs on Evolving Knowledge
    Wei Tang, Yixin Cao, Yang Deng, Jiahao Ying, Bo Wang, Yizhe Yang, et al.
    ACL 2025CCF-A
  11. 2024
    Speaker Verification in Agent-Generated Conversations
    Yizhe Yang, Heyan Huang, Palakorn Achananuparp, Jing Jiang, Ee-Peng Lim
    ACL 2024CCF-A
  12. 2024
    Building Knowledge-grounded Dialogue Systems with Graph-based Semantic Modelling
    Yizhe Yang, Heyan Huang, Yang Gao, Junyi Li
    Knowledge-Based SystemsSCI Q1 TOP
  13. 2024
    MindLLM: Lightweight Large Language Model Pre-training, Evaluation and Domain Application
    Yizhe Yang, Hang Sun, Junyi Li, Ruikun Liu, Yulan Li, Yuan Liu, Yang Gao, Heyan Huang
    AI Openmodels
  14. 2024
    Automating Dataset Updates Towards Reliable and Timely Evaluation of Large Language Models
    Jiahao Ying, Yixin Cao, Yushi Bai, Qianru Sun, Bo Wang, Wei Tang, Zhaojun Ding, Yizhe Yang, et al.
    NeurIPS 2024CCF-A
  15. 2024
    Fundamental Capabilities of Large Language Models and Their Applications in Domain Scenarios: A Survey
    Jiawei Li, Yizhe Yang, Yu Bai, et al.
    ACL 2024CCF-A
  16. 2024
    Latent Representation Discretization for Unsupervised Text Style Generation
    Yang Gao, Qianhui Liu, Yizhe Yang, Ke Wang
    Information Processing & ManagementSCI Q1 TOPESI HIGHLY CITED
  17. 2023
    Graph vs. Sequence: An Empirical Study on Knowledge Forms for Knowledge-Grounded Dialogue
    Yizhe Yang, Heyan Huang, Yuan Liu, Yang Gao
    EMNLP 2023
BOOK
2026
大语言模型:技术实践与场景应用 (Large Language Models: Practice and Applications)
Heyan Huang, Zewen Chi, Yu Bai, Yizhe Yang
Publishing House of Electronics Industry

Projects

MindLLM
明德大模型 · bilingual foundation models, 1.3B / 3B

Core contributor to the first foundation model pre-trained end-to-end at a Chinese university. I worked across data curation, tokenizer design, pre-training, instruction tuning, and domain adaptation. The open-source release has been downloaded more than 30,000 times and has been deployed in pilot applications across government, education, and healthcare.

Pre-training Bilingual LLM Instruction Tuning Domain Adaptation Open Source

Awards & Honors

2025
Wu Wen-Jun AI Science & Technology Progress Award, First Prize (吴文俊人工智能科技进步奖一等奖), Chinese Association for Artificial Intelligence
2025
National Scholarship for Doctoral Students (博士研究生国家奖学金), Ministry of Education
2024
Bit_numeval at SemEval-2024 Task 7 (the Mathematical Reasoning task): Enhance Numerical Sensitivity and Reasoning Completeness for Quantitative Understanding