L-02 · Concept

Observable evidence in AI answers

Primary intent
Distinguish direct public-output evidence from inference and unknown system behavior.
Evidence state
Source-grounded reference
Review owner
Matthias Ramahi · independent review not claimed
Last reviewed
2026-08-22
Direct answer

Observable evidence

Observable AI-answer evidence is information a researcher can directly capture from the public interaction: the submitted question, visible answer, visible links or citations, declared settings, route, locale and timestamp. Explanations of hidden retrieval, ranking or reasoning remain inference unless the provider publishes corresponding evidence.

Use this boundary before designing a dataset, writing a claim, or reviewing a visual that appears to show internal system behavior.

Keep three evidence layers separate

A usable observation record distinguishes the submitted input, the returned public output, and the observer’s interpretation. Mixing the layers makes later review almost impossible: an analyst cannot tell whether a label came from the interface, the provider documentation or the researcher.

The first two layers can be recorded directly when terms and rights allow. The third layer is legitimate analysis, but it needs its own field, method and confidence. It should never overwrite the raw observation.

  • Observed: text and controls visible in the public interaction.
  • Supported: a mechanism described in named provider documentation.
  • Inferred: an interpretation derived from observed patterns.
  • Unknown: information neither observed nor documented for the case.

Minimum capture for a reviewable observation

Record a stable observation ID, timestamp with timezone, surface and route, locale, stable question ID, protocol version, visible source URLs, and a missing-data reason when an expected field is unavailable. If raw answer retention is restricted, record the permitted representation and the rule that produced it.

A screenshot alone is weak provenance. It may show appearance, but often omits route, locale, timing, configuration and machine-readable source normalization. Pair visual evidence with a structured record whenever possible.

Do not upgrade an observation into a system claim

A source link appearing in an answer proves that the link was visible in that captured output. It does not prove how the system found, ranked, read or weighted the page. Similar answers across two runs show similarity across those runs, not permanent model stability.

This is the most common reporting error in AI-search studies: the observation is real, but the sentence travels beyond it. A reviewer should be able to trace every claim back to the exact field or documentation source that supports it.

A compact claim test

Before publication, ask four questions: What exactly was captured? Which part of the sentence comes from documentation? Which part is analysis? What plausible alternative explanation remains? If any answer is missing, reduce the claim or collect better evidence.

The result may sound less dramatic, but it becomes more useful. Readers can repeat the observation, challenge the interpretation and understand what would change the conclusion.

S

Source notes

These sources support the definitions, standards or project boundaries named in this reference. They do not prove that a public observation dataset exists.

  1. portfolio-dossier
    Canonical ai-fanout.com domain dossier

    Confirmed ownership, accepted public Evidence Lab purpose, named Research Owner, indexable website launch and separately gated provider research.

    Owner record
  2. google-query-fanout
    AI features and your website

    Google describes query fan-out publicly without exposing a general private-query inspection interface.

    Open
  3. nist-ai-rmf-genai
    NIST AI RMF Generative AI Profile

    Supports explicit measurement, documentation, monitoring and limitations for generative-AI evaluations.

    Open
  4. w3c-prov-o
    PROV-O: The PROV Ontology

    Provides provenance concepts for entities, activities, agents, derivations, sources and versions.

    Open