Stale but Stable: Staleness-Adaptive Trust Regions for Stabilizing Asynchronous Reinforcement Learning Paper • 2607.18722 • Published 12 days ago • 35
anorim/bertimbau-fusion-2-daretask-p0.9-l0.3-bestcross-hatebr-hspt-olidbr-toldbr-tupy-v4 0.1B • Updated 17 days ago • 23 • 1
sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2 Sentence Similarity • 0.1B • Updated Jan 28 • 60M • • 1.34k
AGVBench: A Reliability-Oriented Benchmark of Data Augmentation for Vein Recognition Paper • 2607.02271 • Published Jul 2 • 17
Graph-Native Reinforcement Learning Enables Traceable Scientific Hypothesis Generation through Conceptual Recombination Paper • 2607.00924 • Published Jul 1 • 11
QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents Paper • 2605.27068 • Published May 26 • 24
Perception or Prejudice: Can MLLMs Go Beyond First Impressions of Personality? Paper • 2605.22109 • Published May 21 • 171