Behavior From the Void: Unsupervised Active Pre-Training H Liu, P Abbeel Advances in Neural Information Processing Systems, 2021, 2021 | 226 | 2021 |
Aligning Text-to-Image Models using Human Feedback K Lee, H Liu, M Ryu, O Watkins, Y Du, C Boutilier, P Abbeel, ... arXiv preprint arXiv:2302.12192, 2023 | 202 | 2023 |
Koala: A dialogue model for academic research X Geng*, A Gudibande*, H Liu*, E Wallace*, P Abbeel†, S Levine†, ... Blog post, April 1, 2023 | 198 | 2023 |
Chain of Hindsight Aligns Language Models with Feedback H Liu, C Sferrazza, P Abbeel International Conference on Learning Representations(ICLR), 2024, 2023 | 169* | 2023 |
The false promise of imitating proprietary llms A Gudibande, E Wallace, C Snell, X Geng, H Liu, P Abbeel, S Levine, ... The Twelfth International Conference on Learning Representations (ICLR 2024), 2023 | 164 | 2023 |
URLB: Unsupervised Reinforcement Learning Benchmark M Laskin*, D Yarats*, H Liu, K Lee, A Zhan, K Lu, C Cang, L Pinto, ... arXiv preprint arXiv:2110.15191, 2021 | 150 | 2021 |
Reinforcement learning for fine-tuning text-to-image diffusion models Y Fan, O Watkins, Y Du, H Liu, M Ryu, C Boutilier, P Abbeel, ... Advances in Neural Information Processing Systems 36, 2024 | 147 | 2024 |
Openllama: An open reproduction of llama X Geng*, H Liu* URL: https://github. com/openlm-research/open_llama, 2023 | 147* | 2023 |
APS: Active Pretraining with Successor Features H Liu, P Abbeel International Conference on Machine Learning, 6736-6747, 2021 | 144 | 2021 |
Masked world models for visual control Y Seo, D Hafner, H Liu, F Liu, S James, K Lee, P Abbeel Conference on Robot Learning, 1332-1344, 2023 | 124 | 2023 |
Taming MAML: Efficient unbiased meta-reinforcement learning H Liu, R Socher, C Xiong International Conference on Machine Learning (ICML), 4061-4071, 2019 | 119* | 2019 |
RingAttention with Blockwise Transformers for Near-Infinite Context H Liu, M Zaharia, P Abbeel The Twelfth International Conference on Learning Representations (ICLR 2024), 2024 | 112 | 2024 |
Multimodal Masked Autoencoders Learn Transferable Representations X Geng*, H Liu*, L Lee, D Schuurams, S Levine, P Abbeel arXiv preprint arXiv:2205.14204, 2022 | 101 | 2022 |
Action-depedent Control Variates for Policy Optimization via Stein's Identity H Liu, Y Feng, Y Mao, D Zhou, J Peng, Q Liu International Conference on Learning Representations, 2017 | 101 | 2017 |
Don't Change the Algorithm, Change the Data: Exploratory Data for Offline Reinforcement Learning D Yarats*, D Brandfonbrener*, H Liu, M Laskin, P Abbeel, A Lazaric, ... arXiv preprint arXiv:2201.13425, 2022 | 92 | 2022 |
World model on million-length video and language with ringattention H Liu, W Yan, M Zaharia, P Abbeel arXiv e-prints, arXiv: 2402.08268, 2024 | 86* | 2024 |
CIC: Contrastive Intrinsic Control for Unsupervised Skill Discovery M Laskin, H Liu, XB Peng, D Yarats, A Rajeswaran, P Abbeel Advances in Neural Information Processing Systems, 2022, 2022 | 84* | 2022 |
Competitive Experience Replay H Liu, A Trott, R Socher, C Xiong International Conference on Learning Representations(ICLR) 2019, 2019 | 66 | 2019 |
Variational inference with tail-adaptive f-divergence D Wang, H Liu, Q Liu Advances in Neural Information Processing Systems 31, 2018 | 66 | 2018 |
Instruction-Following Agents with Multimodal Transformer H Liu, L Lee, K Lee, P Abbeel arXiv preprint arXiv:2210.13431, 2022 | 49* | 2022 |