Skip to main content
Document sets are processed differently from both individual documents and documents with multiple sections. To process a document set, ABBYY FlexiCapture uses special completeness rules to check that no documents are missing from the set. Completeness rules range from simple document listings to complex rules that require a set to contain certain documents if they are referred to by other documents or included on the inventory list.

Processing stages

A document set goes through the following stages:
  1. Check that the set contains all required documents, verify the number of documents of each type, and optionally check the order of the documents.
  2. Capture data from one main document in the set, or from several documents while detecting contradictions (for example, confirming that all documents relate to the same person or organization).
  3. Visually check the documents for signatures and seals.
  4. Create a searchable PDF from all the documents in the set.
  5. Export the captured data to a database, together with links to the original document images.
A document set may contain documents from which no data should be captured but whose images must still be included in the results. These documents do not require optical recognition, but their type still needs to be detected so that no documents are missing from the set. Examples include hand-written applications, certificates, and receipts.

Recognition of document sets

Document set recognition has two distinct features. You do not have to list child documents. Instead, specify only the document sets to be recognized. Open the batch type properties and, on the Recognition tab, select the sets that correspond to the required definitions. FlexiCapture then fully recognizes those sets. If a child document is moved to the top level of a set, an assembly error occurs because the matched definition does not comply with the set structure. To avoid such errors, add the child document definitions to the general recognition list.