Search Results for author: Zhengbang Zhu

Found 11 papers, 4 papers with code

Diffusion-based Dynamics Models for Long-Horizon Rollout in Offline Reinforcement Learning

no code implementations • 29 May 2024 • Hanye Zhao, Xiaoshen Han, Zhengbang Zhu, Minghuan Liu, Yong Yu, Weinan Zhang

We propose Dynamics Diffusion, short as DyDiff, which can inject information from the learning policy to DMs iteratively.

Paper
Add Code

Contrastive Diffuser: Planning Towards High Return States via Contrastive Learning

no code implementations • 5 Feb 2024 • Yixiang Shan, Zhengbang Zhu, Ting Long, Qifan Liang, Yi Chang, Weinan Zhang, Liang Yin

Applying diffusion models in reinforcement learning for long-term planning has gained much attention recently.

Contrastive Learning D4RL

Paper
Add Code

DiffStitch: Boosting Offline Reinforcement Learning with Diffusion-based Trajectory Stitching

no code implementations • 4 Feb 2024 • Guanghe Li, Yixiang Shan, Zhengbang Zhu, Ting Long, Weinan Zhang

In offline reinforcement learning (RL), the performance of the learned policy highly depends on the quality of offline datasets.

D4RL Data Augmentation +4

Paper
Add Code

Diffusion Models for Reinforcement Learning: A Survey

1 code implementation • 2 Nov 2023 • Zhengbang Zhu, Hanye Zhao, Haoran He, Yichao Zhong, Shenyu Zhang, Haoquan Guo, Tingting Chen, Weinan Zhang

Diffusion models surpass previous generative models in sample quality and training stability.

reinforcement-learning Reinforcement Learning (RL)

303

Paper
Code

MADiff: Offline Multi-agent Learning with Diffusion Models

1 code implementation • 27 May 2023 • Zhengbang Zhu, Minghuan Liu, Liyuan Mao, Bingyi Kang, Minkai Xu, Yong Yu, Stefano Ermon, Weinan Zhang

MADiff is realized with an attention-based diffusion model to model the complex coordination among behaviors of multiple agents.

Offline RL Trajectory Prediction

Paper
Code

Planning Immediate Landmarks of Targets for Model-Free Skill Transfer across Agents

no code implementations • 18 Dec 2022 • Minghuan Liu, Zhengbang Zhu, Menghui Zhu, Yuzheng Zhuang, Weinan Zhang, Jianye Hao

In reinforcement learning applications like robotics, agents usually need to deal with various input/output features when specified with different state/action spaces by their developers or physical restrictions.

Paper
Add Code

RITA: Boost Driving Simulators with Realistic Interactive Traffic Flow

no code implementations • 7 Nov 2022 • Zhengbang Zhu, Shenyu Zhang, Yuzheng Zhuang, Yuecheng Liu, Minghuan Liu, Liyuan Mao, Ziqin Gong, Shixiong Kai, Qiang Gu, Bin Wang, Siyuan Cheng, Xinyu Wang, Jianye Hao, Yong Yu

High-quality traffic flow generation is the core module in building simulators for autonomous driving.

Autonomous Driving

Paper
Add Code

Understanding or Manipulation: Rethinking Online Performance Gains of Modern Recommender Systems

no code implementations • 11 Oct 2022 • Zhengbang Zhu, Rongjun Qin, JunJie Huang, Xinyi Dai, Yang Yu, Yong Yu, Weinan Zhang

The increase in the measured performance, however, can have two possible attributions: a better understanding of user preferences, and a more proactive ability to utilize human bounded rationality to seduce user over-consumption.

Benchmarking Sequential Recommendation

Paper
Add Code

Plan Your Target and Learn Your Skills: Transferable State-Only Imitation Learning via Decoupled Policy Optimization

2 code implementations • 4 Mar 2022 • Minghuan Liu, Zhengbang Zhu, Yuzheng Zhuang, Weinan Zhang, Jianye Hao, Yong Yu, Jun Wang

Recent progress in state-only imitation learning extends the scope of applicability of imitation learning to real-world settings by relieving the need for observing expert actions.

Imitation Learning Transfer Learning

Paper
Code

Plan Your Target and Learn Your Skills: State-Only Imitation Learning via Decoupled Policy Optimization

no code implementations • NeurIPS 2021 • Minghuan Liu, Zhengbang Zhu, Yuzheng Zhuang, Weinan Zhang, Jian Shen, Jianye Hao, Yong Yu, Jun Wang

State-only imitation learning (SOIL) enables agents to learn from massive demonstrations without explicit action or reward information.

Imitation Learning Reinforcement Learning (RL)

Paper
Add Code

SMARTS: Scalable Multi-Agent Reinforcement Learning Training School for Autonomous Driving

4 code implementations • 19 Oct 2020 • Ming Zhou, Jun Luo, Julian Villella, Yaodong Yang, David Rusu, Jiayu Miao, Weinan Zhang, Montgomery Alban, Iman Fadakar, Zheng Chen, Aurora Chongxi Huang, Ying Wen, Kimia Hassanzadeh, Daniel Graves, Dong Chen, Zhengbang Zhu, Nhat Nguyen, Mohamed Elsayed, Kun Shao, Sanjeevan Ahilan, Baokuan Zhang, Jiannan Wu, Zhengang Fu, Kasra Rezaee, Peyman Yadmellat, Mohsen Rohani, Nicolas Perez Nieves, Yihan Ni, Seyedershad Banijamali, Alexander Cowen Rivers, Zheng Tian, Daniel Palenicek, Haitham Bou Ammar, Hongbo Zhang, Wulong Liu, Jianye Hao, Jun Wang

We open-source the SMARTS platform and the associated benchmark tasks and evaluation metrics to encourage and empower research on multi-agent learning for autonomous driving.

Autonomous Driving Multi-agent Reinforcement Learning +2

894

Paper
Code

Cannot find the paper you are looking for? You can Submit a new open access paper.