·
AI & ML interests
Explainability, Assurance, Alignment
Recent Activity
Organizations
view article Quantisation and the Safety Direction of Decisions
AmberTraceLabs
• view article Faithfulness of Stated Reasoning Under Verifiable-Reward RL
AmberTraceLabs
• view article The Direction of Error in Open-Weight Decision Models
AmberTraceLabs
• view article Verifiable Rewards Beyond Maths and Code
AmberTraceLabs
•