Pranav Bhave
Claim AF-001 — 12 of 21 in the registry Supported within scope

AF-001Warrant reviewed 2026-08-28 · expires 2026-11-26

AI Fluency Index prevalences for 11 of 24 behaviours

Anthropic's AI Fluency Index reports overall prevalences for the 11 of its framework's 24 behaviours that are directly observable in conversations, and percentage-point differences for conversations that produced an artifact, but not the artifact share and not which comparison group the differences are against. Those published numbers therefore do not identify the subgroup rates — but they do bound them. Under every feasible artifact share and under either reading of the comparison, fact-checking is between 30% and 43% less prevalent in artifact conversations, a far larger effect than the stated 3.7 percentage points suggests on a base rate of 8.7%. The three behaviours that rise in artifact conversations are Description behaviours; the three that fall are Discernment behaviours.

01Falsifier — what changes this claim

The report states an artifact share or comparison group under which the relative reduction falls outside [30%, 43%]; or a published figure differs from the transcription; or the closed-form bound disagrees with a direct sweep over feasible shares.

Consequence NARROW

This is the condition and consequence recorded in the registry. The vocabulary this is published in defines what each consequence commits the author to.

02Scope

Arithmetic on the figures published in the Index summary, nothing more. A pooled rate is a mixture of its subgroups, so a stated overall rate and a stated between-group difference bound each subgroup rate and their ratio; scripts/mixture_bounds.py computes those bounds in closed form and cross-checks them against a sweep. This is a re-expression of published numbers on a relative scale, not a re-analysis of the underlying conversations, to which this record has no access. It inherits every limitation the report declares: the findings are correlational, and the report states that users may perform fluency behaviours mentally without expressing them conversationally — so a fall in transcript-visible checking cannot be separated from checking that moved off-platform, which is where checking an artifact ordinarily goes. The measured evaluation gap is confounded with a measurement gap and these numbers cannot separate them. The framework is CC BY-NC-SA 4.0, copyright 2025 Rick Dakan, Joseph Feller, and Anthropic PBC; it is cited and analysed here, never redistributed or productised.

03Forbidden rescues

Repairs declared unavailable in advance; using one after a failure would breach the recorded commitment.

04Non-claims — what this does not license
05Binding and freshness
Binding
the-ai-fluency-index
The published summary. Figures transcribed into scripts/mixture_bounds.py and re-asserted against this claim's expected block on every CI run; a change upstream fires the manual trigger below rather than silently altering a bound.
Reviewed
2026-08-28 · window 90 days
Expires
2026-11-26 — after this date the recorded review is overdue; this does not make the claim false
Triggers
  • executable fires when the transcribed figures or the bound change
  • manual the report publishes the artifact share, states which comparison group the differences are against, corrects a figure, or expands the analysis to a different sample
Dimensions
visibilityPublicprovenanceMachine-generated, owner-executedsupport roleExecuted outputmaturityExperimental