view article Article Illustrating Reinforcement Learning from Human Feedback (RLHF) +2 natolambert, LouisCastricato, lvwerra, Dahoas โข Dec 9, 2022 โข 419
view article Article Granite 4.0 Nano: Just how small can you go? ibm-granite โข Oct 28, 2025 โข 126