Chooses the right reading method
Text PDFs are read directly, while scanned PDFs and images use the image-processing path.
marvinjb.devLIVE AI SYSTEM · EXTRACTION AGENT
Upload a PDF or image of an invoice. The AI extracts the vendor, dates, totals, and line items, validates the result, and lets you ask questions about the document.
01 · LIVE DEMO
Upload a PDF, scanned PDF, JPG, or PNG. The system reads the document, extracts the invoice fields, validates the result, and lets you inspect the structured data or ask questions about the invoice.
PDF, scanned PDF, JPG, or PNG. One invoice at a time.
Upload → AI Extracts → Explore → Ask
02 · HOW IT WORKS
Text-based PDFs are read directly. Scanned PDFs and images are processed visually. Both paths produce the same validated invoice data.
03 · ENGINEERING DECISIONS
The system can read different invoice formats, but every document is converted into the same validated invoice structure before it is returned.
Text PDFs are read directly, while scanned PDFs and images use the image-processing path.
Every supported document ends as the same structured invoice result.
The model returns specific invoice fields instead of free-form text.
Extracting the invoice and asking questions about it are handled as separate requests.
04 · RELIABILITY
The system checks files before processing them, validates the AI result before showing it, and gives the user a clear error when something fails.
Empty, unsupported, or oversized files are rejected.
Invalid extraction or Q&A responses are not shown as valid data.
Upload, extraction, timeout, network, and question failures stay visible to the user.
File handling, API behavior, document routing, and structured outputs are covered by tests.
05 · TECHNOLOGY STACK
INSPECT THE IMPLEMENTATION
Explore the API flow, validation, tests, and deployment setup behind the Extraction Agent.