TDM for Identification Models

TDM for Identification Models

Training Data Management (formerly Keyer Data Management) includes tools for controlling and managing Identification models performance. In this article, you will learn how to navigate TDM for Identification Models and its features. Learn how to use them together in Training a Semi-structured Model.

TDM features

Training Data Management includes:

Learn how to use Training Data Management in Training a Semi-Structured Model.

Navigating TDM for ID models

Each locator model trained for the specific semi-structured layout has its tab on the Model Management page (i.e., Field Identification or Table Identification). Find the type of model you want to manage by clicking on its respective tab.

You can upload or download training documents by clicking the buttons located in the upper-right corner of the page.

Learn how to navigate through TDM for ID models below.

Model Summary Card

Navigating Model Summary card

The Model summary card displays the Live and the Candidate models for this layout. You can see the following information:

If the requirements for training a model are not met, then the model summary card will display the status “Reqs not met,” and a View training data button will appear on the card. It will redirect you to the Training Data table.

If your model is ready for training, the status “Ready to train” will appear on the model summary card. You’ll be able to start training by clicking the Train button located on the right-hand side of the card.

If you want to cancel your training, you can click the Cancel training job button next to the status of your model in the model summary card.

Projected Automation Chart

The Projected Automation chart displays the performance of the model that’s currently live.

The chart displays how the target accuracy affects the automation. The lower the accuracy, the higher the automation, and vice versa.

Note that projected model performance (i.e., accuracy and automation) can increase by adding more QA documents. You can also see the margin of error (MoE) for this model.

The Margin of Error (MoE) indicates the allowable range of inaccuracy in the system's results. It shows you how much the output can differ from the true value while still being acceptable. A smaller margin of error means the system is more accurate.

The chart will display the projected automation of your model with the target accuracy you specify. You can determine the target accuracy value that best meets your needs by entering test values.

Identification Report

The Identification Report displays the number of identified fields (whether the machine or a human identified them), their accuracy, and the field-level automation (i.e. the automation of the fields the model was trained on).

The Identification Report is available only for Field Identification models.

Select a specific date range for the report to see charts for the total number of identified fields (machine-identified and human-identified) and their respective accuracy values.

Calculation points are the total fields used to calculate accuracy. They represent the number of evaluated fields. For instance, an accuracy of 50% could come from 1/2 or 400/800 evaluated fields. Learn more in our Accuracy article.

Fields Identified Chart

The Fields Identified chart displays the number of machine- and manually-identified entries for a specific period.

Field Identification Accuracy Chart

The Field Identification Accuracy chart displays the percent accuracy for the selected time. You can see:

Field / Table Level Automation

The Field / Table Level Automation card displays the automation percentage of the fields or columns your model was trained on:

Training Data Health Card

The Training Data Health card displays a breakdown of your dataset. It shows the following insights on the uploaded documents:

Learn more about eligibility in Document Eligibility Filtering.

Learn how to optimize your data and achieve better model performance by using Training Data Analysis, as described in Step 4 of Training a Semi-structured Model.

Training Data Table

The Training Data table shows all documents available for use as training data for your model. You can filter the contents of the table by training document ID, group ID, number of pages, submission date, training status, and scheduled deletion. You can also search for documents by their IDs.

Selecting at least one training document allows you to use the Actions drop-down menu. This drop-down menu has the following buttons:

The Training Data table contains the following columns:

Learn how to annotate your documents to achieve a top-performing model in Step 5 in Training a Semi-structured Model.

Model History Table

The Model History table is located at the bottom of the Model Management page. You can upload a model from another instance, run training, and adjust your Test Target Accuracy. The table displays the following columns:

For example, if you train a model with 5 fields but remove a field from the current layout version, the numbers in the column will be as follows: 4 / 5 (i.e., the layout has 4 fields, but the model was trained on 5).

To train a Semi-structured model using TDM see Training a Semi-structured Model.