Spaces:
Running
Running
|
Download README.md from SeerRay-Lab/DAREBench: direct link, hf CLI and curl.
- Browser
- Download file 808 Bytes
-
https://huggingface.co/spaces/SeerRay-Lab/DAREBench/resolve/main/README.md
- Command line
-
hf download hf://spaces/SeerRay-Lab/DAREBench/README.md
-
curl -L -o README.md https://huggingface.co/spaces/SeerRay-Lab/DAREBench/resolve/main/README.md
808 Bytes
| title: DAREBench | |
| emoji: π¦ | |
| colorFrom: red | |
| colorTo: blue | |
| sdk: static | |
| app_file: index.html | |
| pinned: true | |
| short_description: Deployment-Aware and Reliable Evaluation of Models as Agents | |
| tags: | |
| - benchmark | |
| - agents | |
| - openclaw | |
| - evaluation | |
| - leaderboard | |
| datasets: | |
| - SeerRay-Lab/DAREBench | |
| thumbnail: https://seerray-lab-darebench.static.hf.space/assets/img/social-card.png | |
| # DAREBench: Deployment-Aware and Reliable Evaluation of Models as Agents | |
| Project page for DAREBench. A workload- and deployment-aware benchmark for evaluating models as agents. | |
| - π Paper: [arXiv 2609.06059](https://arxiv.org/abs/2609.06059) | |
| - π» Code: [github.com/SeerRay-Lab/DareBench](https://github.com/SeerRay-Lab/DareBench) | |
| - π€ Dataset: [SeerRay-Lab/DAREBench](https://huggingface.co/datasets/SeerRay-Lab/DAREBench) | |