Request a scoping call Contact

Research & insights

Follow the
evidence.

Independent thinking on AI architecture, governance and risk. Explore the methods, trade-offs and evidence behind better decisions.

Practical research for better AI decisions.

Sources. Questions. Connections.
A connected view of the evidence.

83 published articles · Architecture / Governance / Method / Data

The lead feature

A closer look at the decision and the evidence behind it.

Recent editions

The latest thinking.

Recent editions, with sources and limitations made explicit.

04 Oct 2026

Architecture · 6 min

Give long-running agents a working account of the task

New research on explicit belief states gives agents a way to track what remains unknown and recognise stalled progress. The opportunity is a more useful investigation workflow, with a measurable cost for maintaining that state.

Read the edition →
02 Oct 2026

Data · 6 min

Hold out the customer before funding the rollout

Fresh tickets from familiar accounts do not establish performance on unseen customers. Match the evaluation boundary to the proposed rollout and report the evidence for each population separately.

Read the edition →
01 Oct 2026

Architecture · 7 min

Restore the AI service from a compatible record

Approve an AI recovery plan against a restored service that completes useful work. Test document and index compatibility, reconcile current access permissions and state the recovery scope that the evidence supports.

Read the edition →

The research library

Explore the archive.

Browse by subject or explore earlier editions.

30 Sept 2026 Test the questions your documents cannot answer Test how an assistant responds when its permitted evidence does not contain the answer. Separate unsupported answers from unnecessary refusals, and measure the human handoff before approving the service scope. Method 29 Sept 2026 Randomise the capacity pool before testing the AI When trial users compete for shared appointments, stock or staff, treatment can change the control group’s opportunities. Design the comparison around that resource before inferring the effect of a full rollout. Method 28 Sept 2026 Trace the shared dependency behind every AI fallback Accept an AI fallback against the work it can complete during a defined disruption. Map shared identity, retrieval, gateway and supplier dependencies, then test the minimum service and its operating limits. Governance 27 Sept 2026 Test where an agent token stops working Measure when every receiving service rejects an agent’s credential after authority is withdrawn. Include cached checks, active sessions and queued work before accepting the access design. Architecture 26 Sept 2026 Give outcome labels time to mature Recent records may contain unresolved outcomes rather than negative results. Define the target window and reporting allowance before using conversion labels for retraining or investment decisions. Data 25 Sept 2026 Build reassessment into AI procurement An AI supplier can change the model, data handling or workflow after contract award. Define material-change triggers, evidence rights and review authority so continued use remains an explicit buyer decision. Method 24 Sept 2026 Put a decision gate on live AI testing A live AI trial needs a defined customer outcome, exposure boundary and executable stop decision. Use the proof of concept to design that trial, then use the trial to decide whether the service should expand. Governance 23 Sept 2026 Constrain the inventory policy before training the agent An inventory agent acts on cash, stock and service levels. Define its permitted actions, comparison policy and fallback before asking it to optimise replenishment. Architecture 22 Sept 2026 Time saved at a desk is not capacity released A shorter individual task does not establish extra organisational capacity. Follow the work through queues, reviews and bottlenecks, and value working-pattern improvements separately from increased output. Method