What models say
We extract structured claims from model outputs so that observations can be compared consistently rather than judged from summaries alone.
SemanticRisk observes how AI systems interpret public website content, compares those interpretations across models and over time, and separates meaningful semantic differences from simple wording variation.
We extract structured claims from model outputs so that observations can be compared consistently rather than judged from summaries alone.
Claims may be equivalent, possibly equivalent, narrower, broader, contradictory or unrelated. Shared vocabulary by itself is not treated as proof of equivalence.
Repeated observations identify interpretation drift, including cases where model interpretation changes even when normalized source content appears unchanged.
SemanticRisk reports how AI systems interpret available public content. It does not independently certify every extracted statement as legally, commercially or factually complete. The distinction is central to the product.
Calibration results are dated and based on the reviewed sample available at that time. New reviewed examples may change thresholds, classifications and published performance statistics.
Current public methodology snapshot: 30 July 2026.