End-to-end Whispered Speech Recognition with Frequency-weighted Approaches and Pseudo Whisper Pre-training Paper • 2005.01972 • Published May 5, 2020
USAD 2.0: Scaling Representation Distillation for Universal Audio Understanding Paper • 2606.06444 • Published Jun 4 • 3
MARQUIS: A Three-Stage Pipeline for Video Retrieval-Augmented Generation Paper • 2605.17640 • Published May 17
Principled Context Engineering for RAG: Statistical Guarantees via Conformal Prediction Paper • 2511.17908 • Published Jan 19