AI & ML interests

Explainability, assurance, alignment

Recent Activity

Articles

peterAmberTrace 
published an article 13 days ago
view article
Article

Quantisation and the Safety Direction of Decisions

AmberTraceLabs
•
peterAmberTrace 
published an article 14 days ago
view article
Article

Faithfulness of Stated Reasoning Under Verifiable-Reward RL

AmberTraceLabs
•
peterAmberTrace 
published an article 14 days ago
view article
Article

The Direction of Error in Open-Weight Decision Models

AmberTraceLabs
•
peterAmberTrace 
published an article 14 days ago
view article
Article

Verifiable Rewards Beyond Maths and Code

AmberTraceLabs
•