The actual processing path

From a crowded bill to sourced facts and checked arithmetic

ExpenseReader combines a correctable text preview, bounded AI extraction, exact evidence alignment, and deterministic calculations. Each layer has a different job and a visible limit.

Six stages

What happens to a document

  1. 01

    Choose pasted text or a text-layer PDF

    PDF extraction reads selectable text in bounded request memory. Scans, photographs, encrypted files, and PDFs without a usable text layer are rejected. The original PDF is not sent to the AI provider.

  2. 02

    Correct and mask the extracted text

    You see every extracted page before analysis. Correct columns, signs, and decimal points; remove names and identifiers; then confirm the review. Pattern detection can help, but it is not complete.

  3. 03

    Classify and extract structured facts

    Only the corrected text is sent to OpenAI for bounded classification and category-specific structured extraction. Requests set store:false and use no provider files, web search, tools, threads, or persistent conversation IDs. Provider and hosting retention are described in the privacy notice.

  4. 04

    Check the math outside the model

    ExpenseReader uses signed integer minor units to reproduce totals, quantity-by-rate calculations, and category-specific relationships. A passing calculation establishes consistency, not that a charge or rule is valid.

  5. 05

    Show evidence, unknowns, and questions

    Document facts must point to exact submitted text. Calculations identify their inputs. Missing or ambiguous information stays unknown, and the result proposes questions instead of filling gaps with invented facts.

  6. 06

    Clear or leave the page

    Document, result, and follow-up state live in page memory only and disappear on refresh, Clear, or session expiry. Clearing the page cannot retract processing already performed by hosting or AI providers.

Published limits

What the workflow accepts

PDF size
4,000,000 bytes maximum
Pages
40 per PDF
Extracted text
120,000 characters per document
Supported geography
U.S. documents; local rules remain unverified unless sourced
Daily allowance
Five analyses per anonymous browser token, resetting at UTC day boundaries and subject to the shared budget
Medical PDFs
Require an additional provider-processing acknowledgement

Start with the right vocabulary

Choose a bill guide

Each category has a worked fictional example, terminology, questions, and the same real analyzer preselected for that document type.

Browse the seven guides

A provider boundary

Temporary here is not zero retention

ExpenseReader does not intentionally save bill or conversation content. OpenAI requests use store:false, but that does not by itself establish Zero Data Retention. Provider processing and retention are explained in the privacy notice.

Read the privacy notice, current usage terms, and FAQ for the precise boundaries.