Yuanlin Duan(段源林)Ph.D. StudentDepartment of Computer Science Rutgers University, New Brunswick Email: yuanlin.duan@rutgers.edu
|
|
I am a fourth-year PhD student in the Department of CS at Rutgers University, advised by Professor
He Zhu. Prior to Rutgers, I got my B.E. degree in Artificial Intelligence
at UESTC. In summer 2026 I was a Machine Learning Scientist Intern at
TikTok in San Jose, CA.
My research asks a single question: how can an agent learn a task quickly when the reward is sparse? In a sparse-reward environment the agent sees almost no signal until the whole task is solved, so exploration, credit assignment and curriculum become the bottleneck rather than optimization. I attack this from two directions that turn out to be the same problem: goal-conditioned and model-based RL for embodied agents, and RL post-training for large models, where a long reasoning trajectory is likewise graded by one outcome-level reward at the very end.
| [08/2026] | Finished my summer internship at TikTok (San Jose, CA) as a Machine Learning Scientist Intern. |
| [05/2026] | Started my summer internship at TikTok (San Jose, CA) as a Machine Learning Scientist Intern. |
| [01/2026] | One paper was accepted by ICLR 2026. |
| [09/2025] | One paper was accepted by NeurIPS 2025. |
| [09/2025] | One paper was accepted by OOPSLA 2025. |
| [09/2024] | Two papers were accepted by NeurIPS 2024. |
| [09/2024] | One paper was accepted by EMNLP 2024. |
| [09/2023] | Begin my journey of PhD study at Rutgers |
| [06/2022] | Begin my research in General Medical Center of West China Hospital, Sichuan University |
Preference-based Policy Optimization from Sparse-reward Offline Dataset
Wenjie Qiu, Guofeng Cui, Shicheng Liu, Yuanlin Duan and He Zhu
International Conference on Learning Representations (ICLR), 2026
[Paper][Poster]
Learning from Demonstrations via Capability-Aware Goal Sampling
Yuanlin Duan, Yuning Wang, Wenjie Qiu and He Zhu
Neural Information Processing Systems (NeurIPS), 2025
[Paper][Poster]
Abstraction Refinement-guided Program Synthesis for Robot Learning from Demonstrations
Guofeng Cui, Yuning Wang, Wensen Mao, Yuanlin Duan and He Zhu
The OOPSLA issue of the Proceedings of the ACM on Programming Languages, OOPSLA 2025
[Paper][Artifact]
Learning World Models for Unconstrained Goal Navigation
Yuanlin Duan, Wensen Mao and He Zhu
Neural Information Processing Systems (NeurIPS), 2024
[Paper][Poster]
Exploring the Edges of Latent State Clusters for Goal-Conditioned Reinforcement Learning
Yuanlin Duan, Guofeng Cui and He Zhu
Neural Information Processing Systems (NeurIPS), 2024
[Paper][Poster]
MoE-I²: Compressing Mixture of Experts Models through Inter-Expert Pruning and Intra-Expert Low-Rank Decomposition
Cheng Yang, Yang Sui, Jinqi Xiao, Lingyi Huang, Yu Gong, Yuanlin Duan, Wenqi Jia, Miao Yin, Yu Cheng, Bo Yuan
Findings of the Association for Computational Linguistics: EMNLP 2024
[Paper][ACL Anthology]
|
TikTok, San Jose, CA, USA        Machine Learning Scientist Intern,   May 2026 - Aug 2026 |
|
General Medical Center of West China Hospital, Chengdu, China        Research Assistant,   Jun 2022 - Aug 2023 |
| Fall | 2024 | Teaching Assistant | CS336: Principles of Information and Data Management |
| Spring | 2024 | Teaching Assistant | CS210: Data Management for Data Science |
| Fall | 2023 | Teaching Assistant | CS210: Data Management for Data Science |