EleutherAI/unsloth-phi-4-Instruct-LORA-Open-R1-Code-GRPO-b2-as4-lr2en5-encouraged
Updated
Large language models, scaling laws, AI Alignment, democratization of DL
Soohak: A Mathematician-Curated Benchmark for Evaluating Research-level Math Capabilities of LLMs
GPT-NeoX-20B: An Open-Source Autoregressive Language Model