Document Eligibility Filtering

Document Eligibility Filtering

Document Eligibility Filtering indicates whether a document is eligible for training, based on internal checks in the application and our machine learning logic. It provides additional information about documents that were excluded from the training set. With Document Eligibility Filtering, you can see which documents are incompatible with training and why, allowing you to address any issues accordingly and achieve better model performance.

Using Document Eligibility Filtering

Before using Document Eligibility Filtering, make sure that you’ve:

  1. Go to Library > Models and click on the name of your layout to access the Training Data Management tools.

  2. Find the model you want to manage in the respective tab and scroll to the Training Data Health card.

The Training Data Health card shows information about the quality of your training data. The bar indicates how many documents you need to meet the minimum required for training.

All documents with the Training Status Ready to Annotate or Never appear as ineligible for training until you change the status and reanalyze the data.

.png?sv=2026-02-06&spr=https&st=2026-07-27T09%3A02%3A23Z&se=2026-07-27T09%3A15%3A23Z&sr=c&sp=r&sig=BhMv2DDf6pCIYlUtI%2Bgk0wJe0RaL5p3B7MMz%2FF8j2Ko%3D)

  1. To see ineligibility details for the training set, click See Ineligibility details >> in the Field Identification Model or Table Identification Model card.

Ineligibility information is displayed in the right-hand sidebar, showing the reasons documents are ineligible for training and how many documents are ineligible for each reason.

Always make sure to reanalyze your data to see updated information on your training dataset. If documents have been added, removed, or modified since the last analysis, the ineligibility details may be outdated.

Depending on the results of the analysis, a yellow indicator may appear on the left-hand side of a document’s record in the Training Data card. Hover over it to see whether an anomaly was detected in the document or the document is ineligible for training.

.png?sv=2026-02-06&spr=https&st=2026-07-27T09%3A02%3A23Z&se=2026-07-27T09%3A15%3A23Z&sr=c&sp=r&sig=BhMv2DDf6pCIYlUtI%2Bgk0wJe0RaL5p3B7MMz%2FF8j2Ko%3D)

You can filter the documents by training-ineligibility reason by clicking Filters and selecting a reason in the Training Eligibility drop-down list. Click Apply Filters to view the results.

  1. To view ineligibility details for a particular document, click its ID in the Training Data table.

Ineligibility information is displayed in the right-hand sidebar, showing the reasons the document is ineligible for training.

Ineligibility reasons

Reason Description
Ineligible status A document will always be ineligible for training if its Training Status is Ready to Annotate or Never.
Incompatible layout version The information about the layout is incompatible with the documents provided for training.
Overlapping bounding boxes If the annotated bounding boxes of two or more fields overlap, the document is ineligible for training.
Consecutive page breaks The annotated bounding boxes for a field with multiple bounding boxes span across more than two consecutive pages.
Example:

A tool used to annotate, manage, import, and export training documents. It is also used to train models by working directly with the training data (“ground truth”) obtained from each document in the training set.