I am an undergraduate student at Yuanpei College, Peking University. My research interests include AI alignment, value alignment, agent safety, reinforcement learning, robotic world models, and natural language processing.
I work with the PKU Alignment Group on building safe, reliable, and human-aligned AI systems.
🌙 Something To Say
I am a science student who loves Chinese, its music, its silences, and the way a single phrase can hold both moonlight and measurement. I hope to be a researcher with an interesting soul, one who does not merely ask whether to be, but tries to become more awake, more useful, and more humane.
There are more things in intelligence than our present theories can name. So I want to do research I love, and research that is worth loving: to take arms against confusion, to seek value amid uncertainty, and to use the rough magic of machines gently, so that the work may leave the world a little safer, clearer, and kinder.
I often, almost naively, imagine myself inside the closing scene of Romain Rolland’s Jean-Christophe: crossing the river between the long night and the rushing current, carrying on my shoulders a child both heavy and bright, step by step toward the farther shore. And when Christophe asks, “Enfant, qui donc es-tu?”, the Child answers, “Je suis le jour qui va naître.” I want to keep walking toward that newborn tomorrow, reborn for the next battle, doing the research I believe in with love, freedom, value, and an unyielding soul.
🔥 News
- 2026.08: Our paper Stable Reasoning, Unstable Responses was accepted to the EMNLP 2026 Main Conference.
- 2026.08: Our paper VISA was accepted to the EMNLP 2026 Main Conference.
- 2026.06: Our paper A Game-Theoretic Negotiation Framework for Cross-Cultural Consensus in LLMs was accepted to the ACL 2026 Main Conference as an Oral Presentation.
- 2026.06: Our paper SafeMCP was accepted to the ACL 2026 Main Conference.
- 2026.05: We released MiraBench, a benchmark for action-conditioned reliability in robotic world models.
- 2026.03: We released VISA and Stable Reasoning, Unstable Responses on arXiv.
📝 Publications
-
SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning
Lichao Wang, Zhaoxing Ren, Tianzhuo Yang, Jiaming Ji, Chi Harold Liu, Yaodong Yang, Juntao Dai. ACL 2026 Main Conference. -
MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models
Tianzhuo Yang, Zihan Shen, Zirui Mi, Zhaoyi Zhang, Jiayi Zhou, Jiaming Ji, Juntao Dai, Jiawei Chen, Boyuan Chen, Yaodong Yang. arXiv 2026. -
Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry
Guoxi Zhang, Jiawei Chen, Tianzhuo Yang, Lang Qin, Juntao Dai, Yaodong Yang, Jingwei Yi. EMNLP 2026 Main Conference. -
VISA: Value Injection via Shielded Adaptation for Personalized LLM Alignment
Jiawei Chen, Tianzhuo Yang, Guoxi Zhang, Jiaming Ji, Yaodong Yang, Juntao Dai. EMNLP 2026 Main Conference. -
AI Deception: Risks, Dynamics, and Controls
Boyuan Chen, Sitong Fang, Jiaming Ji, Yanxu Zhu, Pengcheng Wen, Jinzhou Wu, Yingshui Tan, Boren Zheng, Mengying Yuan, Wenqi Chen, Donghai Hong, Alex Qiu, Xin Chen, Jiayi Zhou, Kaile Wang, Juntao Dai, Borong Zhang, Tianzhuo Yang, et al. arXiv 2025. -
A Game-Theoretic Negotiation Framework for Cross-Cultural Consensus in LLMs
Guoxi Zhang, Jiawei Chen, Tianzhuo Yang, Jiaming Ji, Yaodong Yang, Juntao Dai. ACL 2026 Main Conference, Oral Presentation.
🎖 Honors and Awards
- SenseTime Scholarship (商汤奖学金; 30 recipients nationwide).
- Soong Ching Ling “Future Scholarship” (宋庆龄“未来助学金”).
- Huawei Spark Award (华为火花奖; Team Member).
- Second Prize, National College Student Mathematics Competition (全国大学生数学竞赛二等奖).
- Second Prize, National English Competition for College Students (全国大学生英语竞赛二等奖).
- Peking University Boya Scholarship (北京大学博雅奖学金).
- Dean’s Scholarship, Institute for Artificial Intelligence, Peking University (北京大学人工智能研究院院长奖学金).
- Peking University Social Work Award (北京大学社会工作奖).
- Peking University Freshman Scholarship, First Prize (北京大学新生奖学金一等奖).
- Peking University Second-Class Scholarship (北京大学二等奖学金).
📖 Educations
- 2024 - Present, Undergraduate student, Yuanpei College, Peking University.
💻 Internships
- Research Intern, PKU Alignment and Interaction Research Lab (PAIR Lab), Peking University, Beijing, China.
- Research Intern, Beijing Academy of Artificial Intelligence (BAAI / 智源研究院), Beijing, China.