OCR-MetaReasoning Benchmark: Evaluating the Meta-Reasoning Ability of MLLMs in Text-Rich Image Understanding Paper • 2608.30678 • Published Aug 31 • 1
OCR-MetaReasoning Benchmark: Evaluating the Meta-Reasoning Ability of MLLMs in Text-Rich Image Understanding Paper • 2608.30678 • Published Aug 31 • 1
Don't Take the Premise for Granted: Evaluating the Premise Critique Ability of Large Language Models Paper • 2505.23715 • Published May 29, 2025 • 2
Don't Take the Premise for Granted: Evaluating the Premise Critique Ability of Large Language Models Paper • 2505.23715 • Published May 29, 2025 • 2
Can Large Multimodal Models Actively Recognize Faulty Inputs? A Systematic Evaluation Framework of Their Input Scrutiny Ability Paper • 2508.04017 • Published Aug 6, 2025 • 12
Large Language Model Evaluation via Matrix Nuclear-Norm Paper • 2410.10672 • Published Oct 14, 2024 • 19