2024
tgy2024
AI & ML interests
None yet
Organizations
None yet
Github
-
MMMR: Benchmarking Massive Multi-Modal Reasoning Tasks
Paper • 2505.16459 • Published • 45 -
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
Paper • 2510.16062 • Published • 1 -
AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery
Paper • 2605.23204 • Published • 29
Website
Github
-
MMMR: Benchmarking Massive Multi-Modal Reasoning Tasks
Paper • 2505.16459 • Published • 45 -
Can LLMs Correct Themselves? A Benchmark of Self-Correction in LLMs
Paper • 2510.16062 • Published • 1 -
AutoResearch AI: Towards AI-Powered Research Automation for Scientific Discovery
Paper • 2605.23204 • Published • 29
models 0
None public yet
datasets 0
None public yet