👋 About me

Hi! I’m Junlin Liu, a graduate student majoring in Artificial Intelligence at iconInstitute of Automation Chinese Academy of Sciences (CASIA), and iconUniversity of Chinese Academy of Sciences (UCAS).

I’m currently a research intern focusing on Agentic RL and Agent Self-Evolution at iconAlibaba Qwen MOS Lab. Previously, I held research intern positions at iconMeituan LongCat Team, iconBaidu Ernie Team and iconX-Humanoid LLM Team.

Research Interests:

  • 🧬 Self-Evolving Agent — Building adaptive autonomous agents that continuously improve through RL and agent–environment co-evolution.
  • 🤖 Agentic Reinforcement Learning — Training general agent intelligence via fundamental RL-based optimization.
  • 🔬 LLM Evaluation and Benchmarking — Evaluating and improving the reasoning capabilities of foundation models.

💬 If you are interested in collaborating with me, please feel free to add me on WeChat.

ICML2026 (首尔·2026.07)
腾讯青云九坤投资超衍智能

中国科学院大学 人工智能学院 (北京·2026.06)

阿里星热爱之旅 (北京·2026.05)

VALSE2026 (武汉·2026.05)

qingyun

腾讯青云计划 (北京·2026.03)

Outstanding Student Model

优秀学生标兵 (成都·2024.12)

⬆ Scrollable

🔥 News

  • 2026.08: 🎉🎉 One paper has been accepted by ICONIP 2026!
  • 2026.07: 🎤🎤 I was honored to be invited to participate in the Qingyun|UBIQuent|Apex Intelligence Talent Dinner at ICML 2026!
  • 2026.05: 🎤🎤 I was honored to be invited to participate in the “AI Talent Corner” at VALSE 2026!
  • 2026.04: 🧩🧩 We released General365, advancing LLM reasoning beyond domain-specific tasks toward general real-world scenarios.
  • 2026.04: 🎉🎉 One paper has been accepted by ACL 2026! AMO-Bench
  • 2026.03: 🎉🎉 One paper has been accepted by IJCNN 2026! ACE-MAPPO
  • 2026.03: 🎤🎤 I was honored to be invited to participate in the Tencent QingYun Talent Program! “Stars of the Future”
  • 2025.10: 📐📐 We released AMO-Bench, a comprehensive benchmark for pushing the boundaries of mathematical reasoning in LLMs.
  • 2025.06: 🎉🎉 I received my B.E. degree from Sichuan Agricultural University (SICAU), awarded the Outstanding Graduate and Outstanding Thesis Award, ranking 1st/198 in overall GPA for three years (2022-2025) ! 🌟🌟Student Spotlight🌟🌟

📝 Publications

Full list of publications in Google Scholar.

ACL 2026 (CCF-A)
sym

icon AMO-Bench: Large Language Models Still Struggle in High School Math Competitions.

🧑‍💻 Junlin Liu, Shengnan An, Shuang Zhou, Dan Ma, Yehao Lin, Xinxuan Lv, Xuanlin Wang, Xiaoyu Li, Ziwen Wang, Xuezhi Cao, Xunliang Cai.

CCF-A Mathematical Reasoning LLM Evaluation

[Paper] [Code] [ProjectPage] [HuggingFace] Stars Citations

EMNLP 2026 (Under Review)
sym

icon General365: Benchmarking General Reasoning in LLMs Across Diverse and Challenging Tasks.

🧑‍💻 Junlin Liu, Shengnan An, Shuang Zhou, Dan Ma, Shixiong Luo, Ying Xie, Yuan Zhang, Wenling Yuan, Yifan Zhou, Xiaoyu Li, Ziwen Wang, Xuezhi Cao, Xunliang Cai.

General Reasoning LLM Evaluation

[Paper] [Code] [ProjectPage] [HuggingFace] Stars Citations

AAAI 2027 (Under Review)
sym

icon From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search.

🧑‍💻 Junlin Liu, Chunji Lv, Jiangwang Chen, Zixin Song, Shuaiyu Zhou, Xingjian Wu, Kailin Jiang,
Jinyang Wu, Bohan yu, Chenxi Zhou, Xiao Yang, Da Zhu, Guanjun Jiang.

Agentic Search On-Policy Self-Distillation Multi-Agent Systems 🤗 #3 Paper of the Day

[Paper] [Code] [HuggingFace] Stars Citations

AAAI 2027 (Under Review)
sym

icon DecoEvo: Score-Decoupled Co-Evolution of Solver and Rubric-Generator Skills in Text Space.

🧑‍💻 Jiangwang Chen*, Zixin Song*, Junlin Liu*, Shuaiyu Zhou, Haiyan Wu, Haihan Shi, Chenxi Zhou, Hanqing Li, Xiao Yang, Da Zhu, Guanjun Jiang.

LLM Agents Self-Evolving Agent Skill and Rubric Learning

[Paper] [HuggingFace] Citations

AAAI 2027 (Under Review)
sym

icon HarnessOpt: Long-Horizon Harness Self-Improvement via Accumulated Optimizer State.

🧑‍💻 Shuaiyu Zhou*, Jiaying Zhang*, Junlin Liu*, Jiangwang Chen, Zixin Song, Haiyan Wu, Hanqing Li, Jinzhou Song, Xiao Yang, Da Zhu, Guanjun Jiang.

LLM Agents Self-Evolving Agent Harness Engineering
AAAI 2027 (Under Review)
sym

icon Contrastive Reinforced Policy Optimization via Privileged Self-Distillation.

🧑‍💻 Xingjian Wu*, Junlin Liu*, Xingchen Liu, Xuhang Zhu, Jianing Wang, Linsen Guo, Xiaoyu Li,
Xuezhi Cao, Xunliang Cai.

Agentic RL On-Policy Self-Distillation Contrastive Policy Optimization

[Paper]

AAAI 2027 (Under Review)
sym

icon PCSD: Persistent Consistency for Self-Distillation in Agentic Reinforcement Learning.

🧑‍💻 Chunji Lv*, Yangguang Wei*, Junlin Liu*, Yang Gao, Ming Liu, Xinming Wang, Jinyang Wu,
Guoren Wang, Changsheng Li.

Agentic RL On-Policy Self-Distillation Adaptive Weighting

[Paper] [HuggingFace]

KDD 2027 (Under Review)
sym

icon ClawTrack: Towards Trace-Level Evaluation and Improvement of Real-World Autonomous Agents.

🧑‍💻 Xingjian Wu, Xuhang Zhu, Xingchen Liu, Junlin Liu, Jianing Wang, Linsen Guo, Xiaoyu Li,
Xuezhi Cao, Xunliang Cai.

Agent Evaluation Long-Horizon Agent Real-World interaction

[Paper] [ProjectPage]

AAAI 2027 (Under Review)
sym

icon Beyond Relevance-Centric Retrieval: Rubric-Oriented Document Set Selection and Ranking.

🧑‍💻 Kailin Jiang, Lei Liu, Jian Xi, Hui Xu, Junlin Liu, Baochen Fu, Shaoqing Ren, Bin Li, Vichwang,
Yu Lu, Haibo Shi.

Retrieval-Augmented Generation Retrieval Evaluation Rubric-Guided

[Paper] [Code] [ProjectPage] [HuggingFace] Stars Citations

AAAI 2027 (Under Review)
sym

icon ATOBench: Tracing How Autonomous Penetration-Testing Agents Verify Vulnerabilities When Target Evidence Lies.

🧑‍💻 Qiyang Chen, Yixi Li, Fengwei Zhang, Junlin Liu.

Agent Evaluation Penetration Testing Agent Security

[Paper] [Code]

AAAI 2027 (Under Review)
sym

icon HalluAgent: Type-Conditioned Visual Evidence for Hallucination Mitigation in Large Vision-Language Models.

🧑‍💻 Ruipeng Zhang, Zhangtianyi Chen, Zixuan Huang, Tong Ji, Junlin Liu, Yuhao Shen, YiQiLiao,
Bailin Liang, Ruibo Duan.

Multi-modal Agents Agent Hallucination Tool-Use
AAAI 2027 (Under Review)
sym

icon PReM: Learning What to Preserve and When to Refresh for Context Compression.

🧑‍💻 Bohan Yu, Lei Shen, Chenxi Zhou, Chen Han, Junlin Liu, Wenbo Su, Yu Cheng, Bo Zheng.

Long-Context LLM Context Compression LLM Memory

[Paper] [Code]

ICONIP 2026 (CCF-C)
sym

icon DRG-MAPPO: Hierarchical Dynamic Role-Graph Multi-Agent Reinforcement Learning for Cooperative Air Combat.

🧑‍💻 Junlin Liu, Yang Gao, Chengwei Li, Hui Chang, Xinchen Zhang, Zhijun Zhao, Hao Zhao.

Reinforce Learning Multi-Agent System Agent Collaboration and Game
IJCNN 2026 (CCF-C)
sym

icon Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat.
🧑‍💻 Chengwei Li*, Junlin Liu*, Yang Gao, Hui Chang, Xinchen Zhang, Hao Zhao.

Reinforce Learning Multi-Agent System Agent Collaboration and Game
Applied Sciences (JCR-Q1)
sym

LCAS-DetNet: A Ship Target Detection Network for Synthetic Aperture Radar Images.
🧑‍💻 Junlin Liu, Dingyi Liao, Xianyao Wang, Jun Li, Bing Yang, Guanyu Chen.
[Paper] [Code]

💻 Internship Experience

  • 2026.04 – present| icon Alibaba, icon Qwen Business Group, MOS Lab.
    • Research Intern on Agentic RL & Self-Evolving Agent.
    • Mentor: Bowen Zhang.
  • 2025.09 – 2026.04| icon Meituan, icon LongCat Foundation LLM Team.
  • 2025.07 - 2025.09| icon Baidu, icon ERNIE Foundation LLM Team.
    • Research Intern on Multimodal Evaluation & Post-training Data Engineering of Large Language Models.
    • Mentor: Dr.Xin Wang.
  • 2025.03 - 2025.05| icon X-Humanoid (Beijing Innovation Center of Humanoid Robotics), Large Model Department.
    • Research Intern on Embodied Agent for Real-world Interaction.
    • Mentor: Dr.Xiaoyi Chen.

📖 Educations

  • 2025.09 – present|icon Institute of Automation Chinese Academy of Sciences (CASIA).
    • Master’s Degree in Artificial Intelligence.
  • 2025.09 – present|icon University of Chinese Academy of Sciences (UCAS).
    • Master’s Degree in Artificial Intelligence.
  • 2021.09 – 2025.06|icon Sichuan Agricultural University (SiCAU).
    • Bachelor’s Degree in Computer Science and Technology.
  • 2015.09 – 2021.06|icon Chengdu Foreign Languages School (CFLS).

🎫 Academic Activities

  • 2026.07, Seoul, “Qingyun | UBIQuent | Apex Intelligence Talent Dinner” at ICML 2026.
  • 2026.05, Wuhan, “AI Talent Corner” at VALSE 2026.
  • 2026.03, BeiJing, Tencent QingYun Program, “Stars of the Future: Technology Exchange Exhibition”.
  • 2025.09, BeiJing, Wave Summit 2025, Deep Learning Developers Conference.

🏆 Honors and Awards

  • 2025.06, Outstanding Graduates of Sichuan Province, China (Top 2%).
  • 2025.06, Outstanding Graduation Thesis (Top 1%).
  • 2024.12, Outstanding Student Model (The highest honor of the university, awarded to only 10 students).
  • 2024.12, China National Scholarship (Top 0.2%).
  • 2023.12, Outstanding Student Scholarship (Top 1%).

📌 Academic Service

  • Reviewer for: ACL Rolling (ACL, EMNLP, NAACL, EACL, AACL), NeurIPS, AAAI, IJCNN
  • Reviewer for: Applied Sciences