Unsupervised Skill Discovery with Bottleneck Option Learning Paper • 2106.14305 • Published Jun 27, 2021
Time Discretization-Invariant Safe Action Repetition for Policy Gradient Methods Paper • 2111.03941 • Published Jan 27, 2022
HIQL: Offline Goal-Conditioned RL with Latent States as Actions Paper • 2307.11949 • Published Mar 10, 2024
Is Value Learning Really the Main Bottleneck in Offline RL? Paper • 2406.09329 • Published Oct 28, 2024
GHIL-Glue: Hierarchical Control with Filtered Subgoal Images Paper • 2410.20018 • Published Oct 26, 2024
Unsupervised Zero-Shot Reinforcement Learning via Functional Reward Encodings Paper • 2402.17135 • Published Feb 27, 2024
Steering Your Diffusion Policy with Latent Space Reinforcement Learning Paper • 2506.15799 • Published Jun 18, 2025
Diffusion Guidance Is a Controllable Policy Improvement Operator Paper • 2505.23458 • Published May 29, 2025