Writing · Topic

Research

Measured studies of the system — evaluation protocol, citation behaviour, and the paper behind them.

2 pieces

Common questions

Why do AI assistants name some brands and cite others?
Because those are two different phenomena. Being named from training data and being retrieved as a source for one specific answer have different causes and different fixes, and almost nobody separates them.
How do you measure whether an agent harness actually works?
On two planes, never averaged: judgment and architecture, graded separately. Averaging them produces one number that hides which half is failing, which is the half you needed to know.

Ongoing measurement

The AI Visibility IndexHow often agencies are returned when a buyer asks an AI assistant for a recommendation — measured across real audits, updated as the sample grows.