Changelog

24 Mar 2025

Fine-tuning generative AI for extraction

You can now fine-tune generative AI models specifically for your document types, delivering strongly improved extraction accuracy with minimal training data.

This isn't just an incremental update, it's a next step for document understanding. Here’s what it means for you:

  • Up to 10x less training data: where you previously needed hundreds of examples, you can now achieve superior results with as few as 10-20 documents per type, allowing you to go live in hours, not weeks. 
  • Visual understanding: the new models don't just read; they see the whole document. This visual understanding allows them to interpret information that was previously out of reach. The AI can now recognize logos to identify a supplier, understand technical drawings, or use the placement of a signature to validate a form. It comprehends the document's structure visually, just like a human would.
  • Deep contextual awareness: go beyond simple key-value extraction. They understand the document context fully, allowing them to make logical inferences about data relationships in e.g. complex table structures while maintaining accuracy.
  • Superior handwriting recognition: we've made a big leap in processing handwritten text. The new models are significantly more adept at deciphering even messy, cursive, or rushed handwriting. This unlocks reliable automation for a whole new class of documents, like handwritten intake forms, field service reports, and annotated logs.

We use cookies to provide you with the best user experience. For more information, please read our Cookie Policy.