Native PDF text before OCR
Why a text layer should be the first extraction path for modern Gazette documents.
/ BLOG
Product notes about the evidence layer, search interface, and the choices behind GazetteIntel. Specifics over slogans.
We started by looking at the archive. The faster idea was not another portal. It was a question, a source, and a watch that keeps working after you leave.
Read the note ↗Why a text layer should be the first extraction path for modern Gazette documents.
Search, page boundaries, and the small details that keep a source usable.
TOOLS
Explore the planned question pin, alert rule, collection, and API surfaces.
View tools ↗