Dean Byrne PRO
Quazim0t0
AI & ML interests
DaisyChainAI🌼 / SmallLM's / San Francisco / Open Source
Recent Activity
liked a model about 19 hours ago
SmallScale/Simple-Stories-Hindi-20M reacted to geekwrestler's post with 👍 about 20 hours ago
Her is live as a standalone project! Try it here at : https://her.zerodegrees.tech
Built for the Build Small Hackathon, Her enables you to analyze your Claude Code & Pi Code sessions, skills and Posture. Try it out - https://her.zerodegrees.tech !
Read more about the original Hackathon entry:
https://huggingface.co/blog/build-small-hackathon/her-blog reacted to kanaria007's post with 👍 about 20 hours ago
✅ Article highlight: Benchmark Publication Without Governance Inflation (art-60-274, v0.1)
TL;DR:
This article argues that a benchmark result is not a governance maturity claim.
A score may be real, reproducible, and worth publishing—and still say nothing by itself about safety, deployability, assurance, institutional quality, or platform maturity. 274 treats benchmark publication as a discipline of comparability, disclosure, lifecycle limits, and anti-inflation.
Read:
https://huggingface.co/datasets/kanaria007/agi-structural-intelligence-protocols/blob/main/article/60-supplements/art-60-274-benchmark-publication-without-governance-inflation.md
Why it matters:
• prevents measured results from being inflated into safety or maturity claims
• separates historical results from current comparability
• makes scope, freshness, omissions, and unsupported readings visible
• allows honest publication without requiring full platform assurance
• treats narrower wording as trust discipline, not underselling
What’s inside:
• the publication triad: comparability, disclosure, and anti-inflation
• bounded publication outcomes such as PUBLISHABLE, PUBLISHABLE_WITH_LIMITS, NOT_COMPARABLE, and NOT_PUBLISHABLE
• benchmark publication profiles
• comparability disclosure notes
• public non-claims registers
• inflation checklists for result-to-maturity, comparison-to-assurance, historical-to-current, and wording inflation
Key idea:
Do not say:
“this system scored well, therefore it is mature, safe, or ready to deploy.”
Say:
“this result was observed under this benchmark and comparability frame, remains valid within these lifecycle and disclosure limits, and does not support these broader governance claims.”
Better benchmark publication is not a louder score.
It is a result that is harder to overread.