Delineate Anything v2: A Global Foundation Model for Field Delineation Paper • 2607.19069 • Published Jul 21 • 6
ABot-N1: Toward a General Visual Language Navigation Foundation Model Paper • 2607.10383 • Published Jul 14 • 102
mlboydaisuke/Streaming-Sortformer-Diar-CoreAI Voice Activity Detection • Updated 13 days ago • 925 • 2
MuSViT: A Foundation Vision Model for Sheet Music Representation Paper • 2606.31811 • Published Jun 30 • 8
electricsheepasia/asia-owid-deaths-from-substance-use-disorders-who Viewer • Updated Jun 3 • 1.03k • 16 • 1
SQuTR: A Robustness Benchmark for Spoken Query to Text Retrieval under Acoustic Noise Paper • 2602.12783 • Published Feb 13 • 246
CiteVQA: Benchmarking Evidence Attribution for Trustworthy Document Intelligence Paper • 2605.12882 • Published May 13 • 274