Where Does Social Reasoning Come From? Capability Provenance in Language Models Paper • 2606.19625 • Published Jun 17
On the Quantization Robustness of Diffusion Language Models in Coding Benchmarks Paper • 2604.20079 • Published Apr 22
Do Language Models Agree with Human Perceptions of Suspense in Stories? Paper • 2508.15794 • Published Aug 13, 2025
Running Agents Capabilibara - Capability Provenance in Language Models 🦫 Capability provenance in language models (COLM 2026).
Trackstar Benchmark Datasets Collection Datasets of edited benchmarks and their OLMo3 completions for TranckStar to calculate query/completion influence on OLMo3 • 1 item • Updated Jun 10
Dolma3 — Query Data Collection OLMES evaluation queries as attribution targets for OLMo-3-7B variants. • 3 items • Updated May 25 • 1
Trackstar Benchmark Datasets Collection Datasets of edited benchmarks and their OLMo3 completions for TranckStar to calculate query/completion influence on OLMo3 • 1 item • Updated Jun 10