Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
π
In a Training Loop
66657.2
TFLOPS
VIDRAFT_LAB
SeaWolf-AI
118
43
244
Follow
takuM23's profile picture
lrlily2823's profile picture
wwwcxc520's profile picture
321 followers
Β·
274 following
https://www.vidraft.net
AI & ML interests
Contact: arxivgpt@gmail.com
Recent Activity
updated
a Space
about 2 hours ago
FINAL-Bench/AX-RAY
updated
a Space
about 6 hours ago
FINAL-Bench/gate-tetris
replied
to
their
post
about 12 hours ago
The cost of a judging gate is usually quoted as a number. This puts it on a Tetris board. Three boards get the same piece order, and on every move the same proposal and the same noise β a paired comparison. The gate decides one thing: keep this move, or draw again. Each board gets the same 60 seconds of gate time. The text-writing gates get through 15β22 moves. The generation-free gate gets through 40β50. The boards that stop simply run out of clock. It does not win on accuracy: on the same 2,018-question LODO set, JEV scores AUC 0.7350 against ZTC-Judge-27B's 0.7289. The separation is elsewhere. Clock β 2.1 s vs 0.0615 s per call, and on a 200-candidate agent screen one judging call measured 3.206 s generative vs 0.033 s readout, same server. Calibration β a gate is a threshold, and at ECE 0.4985 (vs ZTC 0.0245) a threshold stops carrying information. Mechanism β a text judge can name option 42 when there is no option 42; a scoring readout cannot. Not a lower error rate. No path. The curve in the ZTC panel is real online fitting, scored prequentially β predict first, learn after β with base weights untouched. Not recursive self-improvement. Limits, also stated on the page: Laya's AUC and latency are not our measurements and are set equal to JEV's, so calibration is the only measured axis it differs on. The page is a simulation driven by measured constants. KO / EN / ZH. https://huggingface.co/spaces/FINAL-Bench/Tetris-JEV-LAYA-ZTC https://huggingface.co/FINAL-Bench/ZTC-Judge-27B
View all activity
Organizations
SeaWolf-AI
's models
4
Sort:Β Recently updated
SeaWolf-AI/Darwin-36B-KR
Updated
Jun 23
SeaWolf-AI/Darwin-Qwen3.5-27B-x-Qwen3.5-27B-Claude-4-08162
28B
β’
Updated
Apr 12
β’
12
SeaWolf-AI/Darwin-Darwin-4B-Opus-x-gemma-4-E4B-it-The-D-08412
8B
β’
Updated
Apr 10
β’
6
β’
8
SeaWolf-AI/Darwin-gemma-4-E4B-it-x-Gemma-4-E4B-Claude-4-08292
Updated
Apr 8
β’
7