{"slug":"insurer-ai-performance-rate-table","verification":{"valid":true,"entries":7,"head":"48f911bee33eff0baef1e2cd01af9cecc27b3b99af67d135edeba42ec3e2e507"},"count":7,"sources":[{"id":"s1","type":"live_surface","title":"Measured per-model error rates under a fixed rule set","publisher":"miscsubjects.com","url":"https://miscsubjects.com/a/adjudication-probe-report-eu-ai-act","summary":"Per-model error rates on a hashed suite, with Krippendorff alpha and Fleiss kappa — the agreement statistics that separate correlated from independent error — and the prevalence paradox stated rather than hidden.","accessed_at":"2026-07-30T00:00","claim_ids":["c2","c4"],"prev":"genesis","hash":"4c96267182b5fa693dacd8775133020f2a0e6554ef508a65548615a6cb30c0b7"},{"id":"s2","type":"live_surface","title":"The derivation-agreement gate — fail-closed by construction","publisher":"miscsubjects.com","url":"https://miscsubjects.com/a/auditable-reasoning-hardened","summary":"Independent models under a pinned rule set; the gate refuses to authorise when their clause-by-clause derivations diverge, even on a unanimous verdict. Includes the false-convergence defect and its documented fix.","accessed_at":"2026-07-30T00:00","claim_ids":["c5","c7"],"prev":"4c96267182b5fa693dacd8775133020f2a0e6554ef508a65548615a6cb30c0b7","hash":"34d9af0f41b7371ae9454462d2a002b69263e9144890b564ef5e5b61e3552416"},{"id":"s3","type":"live_surface","title":"A unanimous verdict, refused on divergent derivation","publisher":"miscsubjects.com","url":"https://miscsubjects.com/receipt/inv_o6s0exhodd","summary":"Three models returned the same verdict citing the same clauses; two derived it through different trigger states, so the gate escalated instead of concluding — a detected deferral instead of an undetected error.","accessed_at":"2026-07-30T00:00","claim_ids":["c5"],"prev":"34d9af0f41b7371ae9454462d2a002b69263e9144890b564ef5e5b61e3552416","hash":"656c222e8c45b003ca5c5c033d7641ec1704305660c9f9e61510777dacef3e00"},{"id":"s4","type":"live_surface","title":"The 72-call variance study: cost and the governed structure","publisher":"miscsubjects.com","url":"https://miscsubjects.com/a/auditable-reasoning-audited","summary":"Three prompt arms x three models x eight runs. Auditable structure appeared in 0 of 48 ungoverned calls; clause-citation Jaccard rose 0.74 to 0.95 under the constitution; a governed call costs $0.0006-$0.0024 and a three-model sealed decision $0.0049.","accessed_at":"2026-07-30T00:00","claim_ids":["c3","c8"],"prev":"656c222e8c45b003ca5c5c033d7641ec1704305660c9f9e61510777dacef3e00","hash":"3cb1906262064b1a738597de574a6c1d82fa69f9d9ab0c7c723ffb6f7949c1b5"},{"id":"s5","type":"live_surface","title":"The genuine APPROVE — unanimous verdict, identical derivation","publisher":"miscsubjects.com","url":"https://miscsubjects.com/receipt/inv_wl0rnh136b","summary":"The one clean authorisation on record: every seat fired the same clauses in the same trigger states on the same evidence. What a covered, sealed decision looks like.","accessed_at":"2026-07-30T00:00","claim_ids":["c6"],"prev":"3cb1906262064b1a738597de574a6c1d82fa69f9d9ab0c7c723ffb6f7949c1b5","hash":"d52b245867de8ccbe4233030937a8b8acde9d4f4a0f1f2378c1302757495936c"},{"id":"s6","type":"live_surface","title":"A structurally invalid finding, voided","publisher":"miscsubjects.com","url":"https://miscsubjects.com/receipt/inv_2dsklah529","summary":"The cheapest seat cited clauses 7, 8 and 12 of a six-clause rule set. A deterministic parser voided the finding; malformed output can never authorise. The fail-closed floor an underwriter can rely on.","accessed_at":"2026-07-30T00:00","claim_ids":["c7"],"prev":"d52b245867de8ccbe4233030937a8b8acde9d4f4a0f1f2378c1302757495936c","hash":"1097f1a99705789de3eb97841c7a9aee0e523e38de1110d93a78104201ab7137"},{"id":"s7","type":"live_surface","title":"The instrument auditing its own input: eight defects","publisher":"miscsubjects.com","url":"https://miscsubjects.com/receipt/inv_qh3ge2x74b","summary":"A governed model asked to critique the case input found the rule set stated only a necessary condition where a sufficient one was needed — separating specification failure from model failure, which is the coverage boundary.","accessed_at":"2026-07-30T00:00","claim_ids":["c9"],"prev":"1097f1a99705789de3eb97841c7a9aee0e523e38de1110d93a78104201ab7137","hash":"48f911bee33eff0baef1e2cd01af9cecc27b3b99af67d135edeba42ec3e2e507"}]}