Health-ORSC-Bench: A Benchmark for Measuring Over-Refusal and Safety Completion in Health Context Paper • 2601.17642 • Published Jan 25
CheMM-R1: Benchmark, Dataset and Model Collection CheMM-R1: Enhancing Chemical Structure Recognition and Elucidation with Reasoning Multimodal Large Language Models • 3 items • Updated 18 days ago
RU-AI: Dataset and Model Collection RU-AI: A Large Multimodal Dataset for Machine-Generated Content Detection • 4 items • Updated 19 days ago