Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Quantummed
's Collections
coding-benchmarking dataset
coding-benchmarking dataset
updated
Oct 11, 2025
data-sets for benchmarking LLM for software devt
Upvote
-
Sort: Collection
livebench/liveswebench
Viewer
•
Updated
Mar 31, 2025
•
53
•
288
•
3
livebench/liveswebench-patches
Viewer
•
Updated
Mar 31, 2025
•
1
•
359
•
1
livebench/reasoning
Viewer
•
Updated
Apr 7, 2025
•
200
•
11.5k
•
20
livebench/data_analysis
Viewer
•
Updated
Apr 7, 2025
•
150
•
9.43k
•
8
livebench/coding
Viewer
•
Updated
Apr 7, 2025
•
128
•
9.86k
•
12
livebench/instruction_following
Viewer
•
Updated
Apr 7, 2025
•
400
•
10.7k
•
6
livebench/math
Viewer
•
Updated
Apr 7, 2025
•
368
•
13.6k
•
3
livebench/language
Viewer
•
Updated
Apr 7, 2025
•
190
•
9.55k
•
1
livebench/model_judgment
Viewer
•
Updated
Apr 7, 2025
•
60.4k
•
2.14k
•
2
livebench/model_answer
Viewer
•
Updated
Oct 22, 2024
•
93.7k
•
333
•
1
Upvote
-
Sort: Collection
Share collection
View history
Collection guide
Browse collections