Dated evals, honest limitations, and the failure cases we didn't sand off before publishing.
Rhet
AIIT-Threshold
AI & ML interests
AI Safety, LLM Evaluation, Sycophancy, Conversational AI, Local AI, Companion AI, Speech, Memory, Human Feedback, Open Models
Organizations
None yet
Companion Safety Benchmarks
Scripted spirals that test whether a companion stays warm without capitulating to risky demands.
The Buddy Stack — a fully local AI companion, open sourced
Every layer open. Tessera-1B is a seed, NOT yet the companion — no tested model safely is. Honest model guide + mission: companion-spiral-bench.
Reports, Receipts, and Failures
Dated evals, honest limitations, and the failure cases we didn't sand off before publishing.
Human-Authored Training Data
No synthetic text, no model-transcript training. Every example here was written by a person.
Companion Safety Benchmarks
Scripted spirals that test whether a companion stays warm without capitulating to risky demands.
Start Here: Truth Over Engagement
The thesis, the receipts, and where to look first if you have 30 seconds.
The Buddy Stack — a fully local AI companion, open sourced
Every layer open. Tessera-1B is a seed, NOT yet the companion — no tested model safely is. Honest model guide + mission: companion-spiral-bench.