How we work
The experiments are the evidence.
Our method decides what that evidence can support. We record the exact question, the evaluator, and the conditions, we keep the original output, and we keep what happened separate from what we think it means.
The instruments are standardized and still being validated. That means we report what we observed. It does not mean we certify it.
How a run is conducted
Six steps, the same way every time.
- 01
Start with a commercial question
The run has to connect to a real decision, a company being ruled out, a requirement, or an evidence problem.
- 02
Define the conditions
We record the exact wording, the evaluator, its version, access mode, date, and sampling conditions.
- 03
Run it fresh, more than once
Repeated runs reduce the chance that one favourable or unfavourable answer gets mistaken for a pattern.
- 04
Keep the original output
The raw answer stays available for inspection. A summary is never treated as the source.
- 05
Record only what occurred
Inclusion, recommendation, exclusion, evidence use, and next-step behaviour are recorded separately.
- 06
State the limitation
A system's stated reason is evidence of what it said, not proof of how it actually works.
How strong is the evidence
Six levels, weakest to strongest.
- 01Narrated
- 02Observed
- 03Replicated
- 04Causally Supported
- 05Cross Evaluator
- 06Real World Corroborated
The order is settled. What each level requires is not. Everything on this site currently sits at the weakest level.
Published methods
The procedures behind the runs.
- M-1Evidence tier definitions: the working constitution
- M-4Ingestion and consolidation: how raw evidence becomes a record
Instruments · 0
What does the measuring.
None registered yet. As observation scales, each evaluator becomes its own record carrying model, version, and access mode, and every observation carries that stamp. AI systems are the current instrument class, not the only possible one.
Vocabulary · 14
The terms we use precisely.
- Client Zero
- Commercial evaluation
- Commercial Evaluation Intelligence
- Competitor displacement
- Evidence tier
- Frontrunner movement
- Machine representation
- Recommendation set formation
- Recommendation stability
- Recommendation survivability
- Requirement interpretation
- Requirements
- Validation and evidence
- Vendor elimination