Why this page exists first
A comparison site that publishes rankings without publishing its method is asking for trust it has not earned.
Most “we tested every AI tool” articles were not tested. There is no prompt set, no rubric, no date, and no way for a reader to repeat the work and get the same answer. We would rather show the instrument than the verdict, and we would rather publish an empty results table with a date on it than a full one we cannot defend.
Everything below is fixed before any model is run. The prompts do not change between quarters; if one has to change, the change is logged on this page and the affected scores are marked as not comparable.