Delineate Anything v2: A Global Foundation Model for Field Delineation Paper • 2607.19069 • Published 15 days ago • 5
ABot-N1: Toward a General Visual Language Navigation Foundation Model Paper • 2607.10383 • Published 22 days ago • 102
MuSViT: A Foundation Vision Model for Sheet Music Representation Paper • 2606.31811 • Published Jun 30 • 7
SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise Paper • 2602.12783 • Published Feb 13 • 246
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence Paper • 2605.12882 • Published May 13 • 274