Volerai / Document processing
Read documents and extract the data you need
You do not have to process every document type the same way. Volerai combines reading, classification, and extraction according to the document and the work. You can label fields and start training from a small set of examples.
Label fields and train on a few examples
You mark fields on the document and turn them into training examples. Model training can start from a small set of examples. How many examples are enough is decided with the document type.
Decide the document type
Define classes for invoices, forms, contracts, or application documents. PDFs, scanned pages, tables, and image-based files can be processed in the scenario. Route documents with text or image models to the right processing path.
Define the fields the work needs
Turning a document into plain text may not be enough. Define fields such as number, date, amount, person, or organization so the process receives a usable result.
Handle tables and options
Use different extraction methods for line tables, form fields, and marked boxes. Structure the information with a profile and checks that fit the layout.
Use rules and models together
Use rules for stable layouts and trainable models for variable documents. Pass extracted values through format and required-field checks before the next step.
Review the result with its source
Check extracted fields next to the document view. Carry the user’s corrections into the result the rest of the process uses.