Dated evals, honest limitations, and the failure cases we didn't sand off before publishing.
Rhet
AIIT-Threshold
AI & ML interests
AI Safety, LLM Evaluation, Sycophancy, Conversational AI, Local AI, Companion AI, Speech, Memory, Human Feedback, Open Models
Organizations
None yet