I am a senior undergraduate student at Renmin University of China, majoring in Economics with a second major in Artificial Intelligence.
My research interests lie in AI Scientists and Self-Improving AI.
I have been fortunate to work with Prof. Bing Su at RUC, Dr. Lijun Wu at Shanghai AI Laboratory, and Prof. Chuan Cao at Zhongguancun Academy. I also work with Prof. Xuanhe Zhou at Theseus Labs.
News
- 2026.09: Our technical report The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement is now available on arXiv. Thanks for all collaborators!
- 2026.06: BioMatrix technical report is available on arXiv. 👉News
- 2026.06: R3LM was accepted to KDD 2026. 👉News
- 2026.04: Scientific data foundation SciVerse was released! 👉News
- 2026.02: Prism was accepted to ICLR 2026 as an Oral presentation. Thanks for all collaborators! 👉News
Selected Publications
* indicates equal contribution.
2026

The Last AI Built by Humans: Toward Genuine Recursive Self-Improvement
Yi Duan*, Ying Liu*, Zirui Tang*, …, Xuanhe Zhou, Fan Wu
Technical Report 2026
A roadmap toward genuine recursive self-improvement, organizing AI systems by how much responsibility they assume within the improvement loop—from executing predefined improvements to recursively improving the improvement process itself.
Code

Biological Reasoning-Informed Regression for Interpretable Regulatory DNA Activity Prediction
Yi Duan*, Zhao Yang*, Jiwei Zhu, Ying Ba, Chuan Cao, Bing Su
KDD 2026
We design an interface R3LM that enables LLMs to directly understand DNA sequences, introduce CRE-ReasonBench, and develop a biological reasoning-informed regression framework for interpretable regulatory DNA activity prediction.
Code

Zhao Yang*, Yi Duan*, Jiwei Zhu, Ying Ba, Chuan Cao, Bing Su
ICLR 2026, (Oral, 1.15% of submitted papers)
Through systematic experiments, we find that current gene expression predictors have limited long-sequence modeling ability. Prism integrates multimodal epigenomic signals with DNA sequences to achieve state-of-the-art performance in gene expression prediction. Code

Qizhi Pei*, Zhimeng Zhou*, Yi Duan*, …, Lijun Wu
BioMatrix Technical Report, 2026
BioMatrix unifies molecules, proteins, sequences, structures, and language in a shared discrete token space. As a core co-first author, I completed 39 of 80 biological tasks, built the 1D molecule evaluation pipeline, and supported key results and representation analyses. Code

R3: End-to-End Reasoning-based Planning for Multi-step Retrosynthesis via Reinforcement Learning
Yifei Wang*, Qizhi Pei*, …, Yi Duan, …, Lijun Wu, Weiying Ma, Hao Zhou
ACL 2026, Main Conference
An end-to-end retrosynthesis planning framework that uses generative reasoning and reinforcement learning to replace conventional search-heavy multi-step synthesis planning.
Competitions
- CURE-Bench @ NeurIPS 2025: 2nd prize (3rd place among 76 teams) in Internal Reasoning Track.
Educations
- 2023.09 - present, B.A. in Economics; Minor in Artificial Intelligence, Renmin University of China.
Research Experience

Project Lead, School of Computer Science, Shanghai Jiao Tong University / Theseus Labs
Aug. 2026 - Sep. 2026
Lead a collaborative research effort on Recursive Self-Improvement with Prof. Xuanhe Zhou’s team. Proposed the autonomy-centered framework, designed the overall technical narrative and project structure, decomposed and coordinated research/writing tasks, and integrated academic research and industry practices into a unified roadmap toward genuine RSI.

Research Intern, Shanghai Artificial Intelligence Laboratory / OpenDataLab
Remote · Sep. 2025 - Jun. 2026
Worked on multimodal foundation models and large-scale LLM data/evaluation pipelines. Served as a core co-first author of BioMatrix, completing 39 of 80 downstream tasks and building evaluation pipelines; also contributed to SciVerse and Sci-Align for large-scale structured knowledge extraction.

Visiting Student, Zhongguancun Academy
Beijing, China · Mar. 2026 - May 2026
Selected for the “Shenlan” Visiting Program, involved in the research and design of ViraWatch, a viral early-warning AI agent.

Research Assistant, Gaoling School of Artificial Intelligence, Renmin University of China
Beijing, China · Aug. 2024 - Feb. 2026
Worked on LLM reasoning and multimodal learning, leading R3LM and contributing to Prism. Research covered reasoning-informed regression, structured representation learning, multimodal information integration, and rigorous empirical evaluation.