Roadmap
Development Roadmap
Text-to-3D β’ 40.2M β’ Updated β’ 4Note VoxelModel v1: tiny text-to-3D voxel diffusion
Field
π3Pat animated cows on a hill for fun
Note These little grazing cows used to live on the bench-labs blog. Now they have their own hill β click a cow to pat it!
bench-labs/PixelModel-v4
Text-to-Image β’ 40.1M β’ Updated β’ 16 β’ 5Note a tiny latent diffusion transformer, and the result is a roughly 10x jump in FID.
bench-labs/PixelModel-v3
Text-to-Image β’ 919k β’ Updated β’ 90 β’ 7Note New architecture, not a scale-up. SIREN decoder + FiLM + learned embeddings. Beats v1 FID.
bench-labs/PixelModel-v2
Text-to-Image β’ 200k β’ Updated β’ 54 β’ 9Note scaled up
bench-labs/bench-mid-7-2026
Viewer β’ Updated β’ 300 β’ 89 β’ 3Note 300 dual-mode evaluation items, balanced across all 17 Bench Labs categories
bench-labs/bench-easy-7-2026
Viewer β’ Updated β’ 300 β’ 90 β’ 3Note 300 dual-mode evaluation items, balanced across all 17 Bench Labs categories
bench-labs/bench-effortless-7-2026
Preview β’ Updated β’ 90 β’ 2Note 300 dual-mode evaluation items, balanced across all 17 Bench Labs categories
Leaderboard
π6Generate a live leaderboard of AI model benchmarks
Note The leaderboard is here, let's evaluate
bench-labs/Conversations-Human-AI-100k
Viewer β’ Updated β’ 100k β’ 99 β’ 2Note This artifact is part of a larger research, we do not recommmend the usage of this data corpus.
bench-labs/pixelmodel-v1
Text-to-Image β’ 23.7k β’ Updated β’ 106 β’ 10Note A continuation of previous pixelmodel
bench-labs/bench-AGI
Viewer β’ Updated β’ 1 β’ 247 β’ 2Note recursive self improvement for AGI, still updating
bench-labs/bench-mid-6-2026
Viewer β’ Updated β’ 143 β’ 176 β’ 2Note A mid quality benchmark.
bench-labs/bench-easy-6-2026
Viewer β’ Updated β’ 238 β’ 124 β’ 3Note Our new benchmark, comes with a custoom script.
Members
π5Explore partners and members of Benchβlabs community
Note We seek colaborators to utilize our benchmarks and work with us
bench-labs/bench-effortless-6-2026
Viewer β’ Updated β’ 240 β’ 120 β’ 3Note 240 rows simple benchmark
bench-labs/pixelmodel
Text-to-Image β’ 203k β’ Updated β’ 66 β’ 6Note A toy, fun to play with. text to image model
Blog
π6Search and filter Bench Labs blog posts
Note Our blog.