Capture
Preserve the original page and record the best source available.
Document reconstruction
The goal is a genuine digital document: selectable text, clean lines, meaningful structure, editable output and a clear link back to the source.
The original remains available. Reconstruction should never quietly replace evidence with a guess.
The reconstruction path
Preserve the original page and record the best source available.
Identify text, layout, page regions and the kind of document.
Create real digital text, lines and structure where confidence is sufficient.
Produce a useful PDF or bounded editable output with provenance and review.
What works today
It uses on-device text recognition, reconstructs supported text and layout into PDFs, retains confidence and processing metadata, and preserves a page when safe digital replacement cannot be proven.
Current outputs include searchable PDFs and bounded Word export modes. Results still require review, especially for handwriting, complex tables, unusual type and specialist notation.
What is not public on the web yet
The existing development web pipeline requires a local Node and Python worker. It is not compatible with this static Cloudflare Pages release, so the public website does not claim that online Rebuild processing is available.
A future web release needs a deliberate privacy, retention, deletion and specialist-engine architecture before document upload is enabled.
Learn how scanned-text editing worksUse what is available now
Merge, organise, compress and convert without sending your file away.