ShieldVLA: Feasibility-Aware Safety Alignment for Vision-Language-Action Models
Paper • 2609.13231 • Published • 19
None defined yet.
MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning
Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies