Sort incoming documents
Invoices, quotes, letters: the Sort files pipeline reads each document, labels it and files it into the right folder
Automatic sorting, folder by folder
OCR, classification and summary pipelines, inside your perimeter
Assemble a pipeline in the console: a trigger, a source, steps, a destination. Six pre-wired templates read your PDFs and attachments, classify them, summarize them and file them into Iceberg tables or folders. OCR and classification run on a model served in your own cluster, or on a third-party API if you prefer.
Invoices, quotes, letters: the Sort files pipeline reads each document, labels it and files it into the right folder
Automatic sorting, folder by folder
Extract documents runs a bucket of PDFs through OCR and lands the result in an Iceberg table, queryable like any other table
From PDF to Iceberg table
Memory distillation summarizes a table into L0 and L1 levels (Hippocampe): your AI agents search the synthesis before diving into the detail
Progressive search
A property management team receives invoices, quotes and letters by email every day
No more manual sorting: every document lands classified in the right place
A data team needs to make years of accumulated PDF archives usable
Years of archives become a queryable knowledge base
Discover how document pipelines build your knowledge base.