Recent posts

Who said what, and was it right

7 minute read

A speculative project on AI assurance for court interpreting, and two problems underneath it: putting each turn against the right speaker, and keeping an eag...

Terraform from my phone

7 minute read

Two days directing Claude from my phone while it stood up an AWS organisation, a permissions-bounded identity for itself, and a working AI demo on top. Mostl...

An escalation ladder for LLM judgement

6 minute read

Batched mid-tier models make the calls and adversarial verifiers try to break them. I read the two places they disagreed – notes on a cheap architecture for ...