End-to-End Content and Plan Selection for Data-to-Text Generation Paper • 1810.04700 • Published Oct 10, 2018
Maximum Expected Hitting Cost of a Markov Decision Process and Informativeness of Rewards Paper • 1907.02114 • Published Nov 4, 2019
Subwords as Skills: Tokenization for Sparse-Reward Reinforcement Learning Paper • 2309.04459 • Published Oct 30, 2024
Running on CPU Upgrade Agents Featured 1.02k Model Memory Utility 🚀 1.02k Calculate GPU memory needed for training Hugging Face models