Invoice Parsing Demo
Open the invoice playground.
vlm-1 extracts structured data from invoice PDFs and images.
The response includes visual grounding.
Only schema-requested values are retrieved and visualized.
OCR returns all text in the document, with no context.

Parsing an invoice with visual grounding enabled
Parsing Invoices in 2 Steps
1
Submit an Invoice Parsing Job
2
Wait for the Job to Complete
High-Accuracy Parsing with Grounding
Setconfig=GenerationConfig(grounding=True) to enable Visual Grounding.