ChildVox: A Speech, Audio, and Large Audio-Language Model Benchmark in Understanding and Characterizing Sound across Childhood Paper • 2605.29257 • Published May 28 • 10
dianavdavidson/Vaani-nagamese-majority-lg-English-no-transcript0 Viewer • Updated Jun 2 • 31.7k • 265 • 1
RankE: End-to-End Post-Training for Discrete Text-to-Image Generation with Decoder Co-Evolution Paper • 2605.21195 • Published May 20 • 20
EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation Paper • 2605.23271 • Published May 22 • 82