You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
Copy file name to clipboardExpand all lines: research/index.md
+4-4Lines changed: 4 additions & 4 deletions
Display the source diff
Display the rich diff
Original file line number
Diff line number
Diff line change
@@ -27,14 +27,10 @@ For the most up-to-date list of publications, please visit our [Google Scholar p
27
27
28
28
1. Wentse Chen, Jiayu Chen, Hao Zhu, Fahim Tajwar, Ruslan Salakhutdinov, and Jeff Schneider, "Verlog: An Efficient Synchronized Multi-turn RL Framework for LLM Agents", submitted to International Conference on Learning Representations (ICLR), 2026. <spanstyle="background-color: #fff3e0; color: #ef6c00; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">LLM</span> <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span>
29
29
30
-
1. Aravind Venugopal, Jiayu Chen, Xudong Wu, Chongyi Zheng, Benjamin Eysenbach, and Jeff Schneider, "Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning", submitted to International Conference on Learning Representations (ICLR), 2026. <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span>
31
-
32
30
1. Jiayu Chen, Zhekai Wang, and Vaneet Aggarwal, "[Hierarchical Deep Counterfactual Regret Minimization](https://arxiv.org/abs/2305.17327)", submitted to International Conference on Learning Representations (ICLR), 2026. <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span> <spanstyle="background-color: #e3f2fd; color: #1565c0; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RT</span>
33
31
34
32
1. Jiayu Chen, Le Xu, Aravind Venugopal, and Jeff Schneider, "[Policy-Driven World Model Adaptation for Robust Offline Model-based Reinforcement Learning](https://arxiv.org/abs/2505.13709)", submitted to International Conference on Learning Representations (ICLR), 2026. <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span> <spanstyle="background-color: #e0f7fa; color: #00838f; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">FU</span>
35
33
36
-
1. Jiayu Chen, Le Xu, Wentse Chen, and Jeff Schneider, "[Bayes Adaptive Monte Carlo Tree Search for Offline Model-based Reinforcement Learning](https://arxiv.org/abs/2410.11234)", submitted to International Conference on Learning Representations (ICLR), 2026. <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span> <spanstyle="background-color: #e0f7fa; color: #00838f; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">FU</span>
37
-
38
34
1. Chongyu Zhu, Mithun Vanniasinghe, Jiayu Chen, and Chi-Guhn Lee, “Offline Discovery of Interpretable Skills from Multi-Task Trajectories”, submitted to IEEE International Conference on Robotics & Automation (ICRA), 2026. <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span> <spanstyle="background-color: #fce4ec; color: #c2185b; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RO</span>
39
35
40
36
1. Bhargav Ganguly, Abhimanyu Shekhar, Chang-Lin Chen, Jiayu Chen, Vaneet Aggarwal, Shweta Singh, "A Deep Reinforcement Learning Approach for Circular Economy Management", submitted to Nature Sustainability. <spanstyle="background-color: #e8f5e8; color: #2e7d32; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RAPP</span>
@@ -61,6 +57,10 @@ For the most up-to-date list of publications, please visit our [Google Scholar p
61
57
62
58
### Conference Papers (with proceedings):
63
59
60
+
1. Jiayu Chen, Le Xu, Wentse Chen, and Jeff Schneider, "[Bayes Adaptive Monte Carlo Tree Search for Offline Model-based Reinforcement Learning](https://arxiv.org/abs/2410.11234)", accepted in International Conference on Learning Representations (ICLR), 2026. <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span> <spanstyle="background-color: #e0f7fa; color: #00838f; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">FU</span>
61
+
62
+
1. Aravind Venugopal, Jiayu Chen, Xudong Wu, Chongyi Zheng, Benjamin Eysenbach, and Jeff Schneider, "Occupancy Reward Shaping: Improving Credit Assignment for Offline Goal-Conditioned Reinforcement Learning", accepted in International Conference on Learning Representations (ICLR), 2026. <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span>
63
+
64
64
1. Rohit Sonker, Hiro Josep Farre Kaga, Jiayu Chen, Andrew Rothstein, Ian Char, Ricardo Shousha, Egemen Kolemen, and Jeff Schneider, "Offline Reinforcement Learning for Rotation Profile Control in Tokamaks", accepted in Annual Learning for Dynamics and Control Conference (L4DC), 2026. <spanstyle="background-color: #e8f5e8; color: #2e7d32; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RAPP</span> <spanstyle="background-color: #e0f7fa; color: #00838f; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">FU</span>
65
65
66
66
1. Wentse Chen, Yuxuan Li, Shiyu Huang, Jiayu Chen, and Jeff Schneider, "[ME-IGM: Individual-Global-Max in Maximum Entropy Multi-Agent Reinforcement Learning](http://arxiv.org/abs/2406.13930)", accepted in International Conference on Autonomous Agents and Multiagent Systems (AAMAS), 2026 (**Oral Presentation**). <spanstyle="background-color: #f3e5f5; color: #7b1fa2; padding: 2px6px; border-radius: 3px; font-size: 0.8em;">RA</span>
0 commit comments