👋 About me
Hi! I’m Junlin Liu, a graduate student majoring in Artificial Intelligence at Institute of Automation Chinese Academy of Sciences (CASIA), and
University of Chinese Academy of Sciences (UCAS).
I’m currently a research intern focusing on Agentic RL and Agent Self-Evolution at Alibaba Qwen MOS Lab. Previously, I held research intern positions at
Meituan LongCat Team,
Baidu Ernie Team and
X-Humanoid LLM Team.
Research Interests:
- 🧬 Self-Evolving Agent — Building adaptive autonomous agents that continuously improve through RL and agent–environment co-evolution.
- 🤖 Agentic Reinforcement Learning — Training general agent intelligence via fundamental RL-based optimization.
- 🔬 LLM Evaluation and Benchmarking — Evaluating and improving the reasoning capabilities of foundation models.
💬 If you are interested in collaborating with me, please feel free to add me on WeChat.
🔥 News
- 2026.08: 🎉🎉 One paper has been accepted by ICONIP 2026!
- 2026.07: 🎤🎤 I was honored to be invited to participate in the Qingyun|UBIQuent|Apex Intelligence Talent Dinner at ICML 2026!
- 2026.05: 🎤🎤 I was honored to be invited to participate in the “AI Talent Corner” at VALSE 2026!
- 2026.04: 🧩🧩 We released General365, advancing LLM reasoning beyond domain-specific tasks toward general real-world scenarios.
- 2026.04: 🎉🎉 One paper has been accepted by ACL 2026! AMO-Bench!
- 2026.03: 🎉🎉 One paper has been accepted by IJCNN 2026! ACE-MAPPO!
- 2026.03: 🎤🎤 I was honored to be invited to participate in the Tencent QingYun Talent Program! “Stars of the Future”
- 2025.10: 📐📐 We released AMO-Bench, a comprehensive benchmark for pushing the boundaries of mathematical reasoning in LLMs.
- 2025.06: 🎉🎉 I received my B.E. degree from Sichuan Agricultural University (SICAU), awarded the Outstanding Graduate and Outstanding Thesis Award, ranking 1st/198 in overall GPA for three years (2022-2025) ! 🌟🌟Student Spotlight🌟🌟
📝 Publications
Full list of publications in Google Scholar.
AMO-Bench: Large Language Models Still Struggle in High School Math Competitions.
🧑💻 Junlin Liu, Shengnan An, Shuang Zhou, Dan Ma, Yehao Lin, Xinxuan Lv, Xuanlin Wang, Xiaoyu Li, Ziwen Wang, Xuezhi Cao, Xunliang Cai.
General365: Benchmarking General Reasoning in LLMs Across Diverse and Challenging Tasks.
🧑💻 Junlin Liu, Shengnan An, Shuang Zhou, Dan Ma, Shixiong Luo, Ying Xie, Yuan Zhang, Wenling Yuan, Yifan Zhou, Xiaoyu Li, Ziwen Wang, Xuezhi Cao, Xunliang Cai.
From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search.
🧑💻 Junlin Liu, Chunji Lv, Jiangwang Chen, Zixin Song, Shuaiyu Zhou, Xingjian Wu, Kailin Jiang,
Jinyang Wu, Bohan yu, Chenxi Zhou, Xiao Yang, Da Zhu, Guanjun Jiang.
DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space.
🧑💻 Jiangwang Chen*, Zixin Song*, Junlin Liu*, Shuaiyu Zhou, Haiyan Wu, Haihan Shi, Chenxi Zhou, Hanqing Li, Xiao Yang, Da Zhu, Guanjun Jiang.
HarnessOpt: Long-Horizon Harness Self-Improvement via Accumulated Optimizer State.
🧑💻 Shuaiyu Zhou*, Jiaying Zhang*, Junlin Liu*, Jiangwang Chen, Zixin Song, Haiyan Wu, Hanqing Li, Jinzhou Song, Xiao Yang, Da Zhu, Guanjun Jiang.
Contrastive Reinforced Policy Optimization via Privileged Self-Distillation.
🧑💻 Xingjian Wu*, Junlin Liu*, Xingchen Liu, Xuhang Zhu, Jianing Wang, Linsen Guo, Xiaoyu Li,
Xuezhi Cao, Xunliang Cai.
PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning.
🧑💻 Chunji Lv*, Yangguang Wei*, Junlin Liu*, Yang Gao, Ming Liu, Xinming Wang, Jinyang Wu,
Guoren Wang, Changsheng Li.
ClawTrack: Towards Trace-Level Evaluation and Improvement of Real-World Autonomous Agents.
🧑💻 Xingjian Wu, Xuhang Zhu, Xingchen Liu, Junlin Liu, Jianing Wang, Linsen Guo, Xiaoyu Li,
Xuezhi Cao, Xunliang Cai.
Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking.
🧑💻 Kailin Jiang, Lei Liu, Jian Xi, Hui Xu, Junlin Liu, Baochen Fu, Shaoqing Ren, Bin Li, Vichwang,
Yu Lu, Haibo Shi.
HalluAgent: Type-Conditioned Visual Evidence for Hallucination Mitigation in Large Vision-Language Models.
🧑💻 Ruipeng Zhang, Zhangtianyi Chen, Zixuan Huang, Tong Ji, Junlin Liu, Yuhao Shen, YiQiLiao,
Bailin Liang, Ruibo Duan.
DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-Agent Reinforcement Learning for Cooperative Air Combat.
🧑💻 Junlin Liu, Yang Gao, Chengwei Li, Hui Chang, Xinchen Zhang, Zhijun Zhao, Hao Zhao.
Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat.
🧑💻 Chengwei Li*, Junlin Liu*, Yang Gao, Hui Chang, Xinchen Zhang, Hao Zhao.
💻 Internship Experience
- 2026.04 – present|
Alibaba,
Qwen Business Group, MOS Lab.
- Research Intern on Agentic RL & Self-Evolving Agent.
- Mentor: Bowen Zhang.
- 2025.09 – 2026.04|
Meituan, LongCat Foundation LLM Team.
- Research Intern on Mathematical & General Reasoning of Large Language Models.
- Mentor: Shengnan An, Xuezhi Cao, Xunliang Cai.
- 2025.07 - 2025.09|
Baidu, ERNIE Foundation LLM Team.
- Research Intern on Multimodal Evaluation & Post-training Data Engineering of Large Language Models.
- Mentor: Dr.Xin Wang.
- 2025.03 - 2025.05|
X-Humanoid (Beijing Innovation Center of Humanoid Robotics), Large Model Department.
- Research Intern on Embodied Agent for Real-world Interaction.
- Mentor: Dr.Xiaoyi Chen.
📖 Educations
- 2025.09 – present|
Institute of Automation Chinese Academy of Sciences (CASIA).
- Master’s Degree in Artificial Intelligence.
- 2025.09 – present|
University of Chinese Academy of Sciences (UCAS).
- Master’s Degree in Artificial Intelligence.
- 2021.09 – 2025.06|
Sichuan Agricultural University (SiCAU).
- Bachelor’s Degree in Computer Science and Technology.
- 2015.09 – 2021.06|
Chengdu Foreign Languages School (CFLS).
🎫 Academic Activities
- 2026.07, Seoul, “Qingyun | UBIQuent | Apex Intelligence Talent Dinner” at ICML 2026.
- 2026.05, Wuhan, “AI Talent Corner” at VALSE 2026.
- 2026.03, BeiJing, Tencent QingYun Program, “Stars of the Future: Technology Exchange Exhibition”.
- 2025.09, BeiJing, Wave Summit 2025, Deep Learning Developers Conference.
🏆 Honors and Awards
- 2025.06, Outstanding Graduates of Sichuan Province, China (Top 2%).
- 2025.06, Outstanding Graduation Thesis (Top 1%).
- 2024.12, Outstanding Student Model (The highest honor of the university, awarded to only 10 students).
- 2024.12, China National Scholarship (Top 0.2%).
- 2023.12, Outstanding Student Scholarship (Top 1%).
📌 Academic Service
- Reviewer for: ACL Rolling (ACL, EMNLP, NAACL, EACL, AACL), NeurIPS, AAAI, IJCNN
- Reviewer for: Applied Sciences











