Reading receipts, scanned deposit slips, handwritten cash-count sheets and PDF invoices directly, extracting structured fields, and loading them as source data — removing the manual keying step entirely.
Vision-capable language model with a defined output schema; supersedes template-based OCR by handling unseen layouts without per-format configuration
Vision-capable LLM API; classic OCR (Tesseract) for the text-only baseline