Xinyu Tang 汤昕宇
Ph.D. Student in Artificial Intelligence
人工智能方向博士研究生
Gaoling School of Artificial Intelligence, Renmin University of China
中国人民大学高瓴人工智能学院
I am a Ph.D. student at the Gaoling School of Artificial Intelligence, Renmin University of China, advised by Prof. Wayne Xin Zhao. I am currently also a research intern at Ant Group, working on LLM post-training. I am always open to research collaboration. Feel free to reach out for discussion and collaboration at txy20010310@gmail.com.
我是中国人民大学高瓴人工智能学院的博士研究生,导师为赵鑫教授。目前在蚂蚁集团实习,从事大语言模型后训练相关研究。欢迎交流与合作,联系邮箱:txy20010310@gmail.com。
My research focuses on large language models — in particular LLM post-training, intelligent agents, and reinforcement learning.
我的研究方向聚焦于大语言模型,包括大模型后训练、智能体和强化学习。
🔬 Research Interests🔬 研究方向
🔥 News🔥 最新动态
- 2026.07 Our technical report "Ring-Zero" is released. 🚀
- 2026.06 Three new preprints released: SearchSwarm, GraphPO, and Harness Survey. 📄
- 2026.05 Our paper "Rethinking Sample Polarity in RLVR" was accepted to ACL 2026. 🎉
- 2026.01 Two papers (DEPO & DrugTrail) were accepted to ICLR 2026. 🎉
- 2025.12 Our paper "L2V-CoT" was accepted to AAAI 2026. 🎉
- 2025.10 Our technical report "Ring-1T" is released. 🚀
- 2025.09 Our paper "Incentivizing Dual Process Thinking" was accepted to NeurIPS 2025. 🎉
- 2025.05 Our paper "Unlocking General Long CoT Reasoning via Representation Engineering" was accepted to ACL 2025. 🎉
- 2025.01 Three papers accepted to ICLR 2025 (Competitive-ICL), NAACL 2025 (DAWN-ICL), and AAAI 2025 (GPO). 🎉
- 2024.06 Our technical report "YuLan" is released. 🚀
- 2023.10 Our paper "Rethinking CRS Evaluation" was accepted to EMNLP 2023. 🎉
- 2023.09 I started my Ph.D. journey at the Gaoling School of AI, Renmin University of China. 🎓
- 2023.05 Our paper "CFCRS" was accepted to KDD 2023. 🎉
- 2026.07 技术报告《Ring-Zero》发布。🚀
- 2026.06 三篇新预印本发布:SearchSwarm、GraphPO 和 Harness Survey。📄
- 2026.05 论文《Rethinking Sample Polarity in RLVR》被 ACL 2026 录用。🎉
- 2026.01 两篇论文(DEPO 和 DrugTrail)被 ICLR 2026 录用。🎉
- 2025.12 论文《L2V-CoT》被 AAAI 2026 录用。🎉
- 2025.10 技术报告《Ring-1T》发布。🚀
- 2025.09 论文《Incentivizing Dual Process Thinking》被 NeurIPS 2025 录用。🎉
- 2025.05 论文《Unlocking General Long CoT Reasoning via Representation Engineering》被 ACL 2025 录用。🎉
- 2025.01 三篇论文分别被 ICLR 2025(Competitive-ICL)、NAACL 2025(DAWN-ICL)和 AAAI 2025(GPO)录用。🎉
- 2024.06 技术报告《YuLan》发布。🚀
- 2023.10 论文《Rethinking CRS Evaluation》被 EMNLP 2023 录用。🎉
- 2023.09 进入中国人民大学高瓴人工智能学院攻读博士学位。🎓
- 2023.05 论文《CFCRS》被 KDD 2023 录用。🎉
🎓 Education🎓 教育经历
- 2023.09 – Present Ph.D. in Computer Science Gaoling School of Artificial Intelligence, Renmin University of China Advisor: Prof. Wayne Xin Zhao
- 2023.09 – 至今 计算机科学博士 中国人民大学高瓴人工智能学院 导师:赵鑫 教授
- 2019.09 – 2023.06 B.Sc. in Computer Science Beijing Normal University, Beijing
- 2019.09 – 2023.06 计算机科学学士 北京师范大学
💻 Experience💻 实习经历
- 2025.01 – Present Ant Group Research Intern · LLM Post-training Beijing, China
- 2025.01 – 至今 蚂蚁集团 研究实习生 · 大模型后训练 中国北京
📝 Publications📝 学术论文
* denotes equal contribution; my name is highlighted. A full, up-to-date list is on my Google Scholar and DBLP.
* 表示同等贡献;我的名字已高亮标注。完整论文列表请见 Google Scholar 和 DBLP。
First / Co-first Author第一作者 / 共同第一作者
-
-
ACL 2026CCF-ARethinking Sample Polarity in Reinforcement Learning with Verifiable Rewards -
-
ACL 2025CCF-AUnlocking General Long Chain-of-Thought Reasoning Capabilities of Large Language Models via Representation Engineering -
NAACL 2025CCF-BDAWN-ICL: Strategic Planning of Problem-solving Trajectories for Zero-Shot In-Context Learning -
-
-
arXiv 2025Enhancing Cross-task Transfer of Large Language Models via Fourier Activation Steering -
-
-
EMNLP 2023CCF-BRethinking the Evaluation for Conversational Recommendation in the Era of Large Language Models -
Selected Collaborations合作论文(精选)
-
ICLR 2026CCF-ADrugTrail: Interpretable Drug Discovery via Structured Reasoning and Druggability-Tailored Preference Optimization -
NeurIPS 2025CCF-AIncentivizing Dual Process Thinking for Efficient Large Language Model Reasoning -
-
Tech ReportEvery Step Evolves: Scaling Reinforcement Learning for Trillion-Scale Thinking Model -
-
-