chenhao
ch-outcomes-ai
AI & ML interests
None yet
Organizations
None yet
MATH
VLM-VLA
-
EmbRACE-3K: Embodied Reasoning and Action in Complex Environments
Paper • 2507.10548 • Published • 37 -
OmniEAR: Benchmarking Agent Reasoning in Embodied Tasks
Paper • 2508.05614 • Published • 20 -
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models
Paper • 2507.12806 • Published • 21 -
DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning
Paper • 2508.05405 • Published • 65
TOOL
MATH
AGENT
VLM-VLA
-
EmbRACE-3K: Embodied Reasoning and Action in Complex Environments
Paper • 2507.10548 • Published • 37 -
OmniEAR: Benchmarking Agent Reasoning in Embodied Tasks
Paper • 2508.05614 • Published • 20 -
MCPEval: Automatic MCP-based Deep Evaluation for AI Agent Models
Paper • 2507.12806 • Published • 21 -
DeepPHY: Benchmarking Agentic VLMs on Physical Reasoning
Paper • 2508.05405 • Published • 65