Document Classification Task

Document Classification Task

Classification models in Hyperscience

Classification models are a crucial part of document processing as they help the system determine which layout should be used to process each page you upload. In Hyperscience, we have two types of document classification:

In this article, you’ll learn how to work with Document Classification Task. Learn more about Identification in TDM for Identification Models. To learn about Transcription and Flexible Extraction, see our Transcription and Flexible Extraction articles.

Document Classification task

Document Classification is a Supervision task used to group pages into documents and assign the correct layout. These tasks are created when automatic classification isn’t possible (e.g., when no model is available or the model’s confidence is low).

.png?sv=2026-02-06&spr=https&st=2026-07-27T09%3A12%3A53Z&se=2026-07-27T09%3A28%3A53Z&sr=c&sp=r&sig=GnlVklgFM1j1eJEFBqNS%2BWX51Fl4vWY7iahFb%2BKg7bQ%3D)

Expand the section below to learn how to navigate the interface and complete a Document Classification task.

Document Classification Interface

To open a Document Classification task, go to the Tasks section and click Perform Tasks under the Supervision task type table.

Document Classification allows you to:

.jpg?sv=2026-02-06&spr=https&st=2026-07-27T09%3A12%3A53Z&se=2026-07-27T09%3A28%3A53Z&sr=c&sp=r&sig=GnlVklgFM1j1eJEFBqNS%2BWX51Fl4vWY7iahFb%2BKg7bQ%3D)

Left Panel - Uncategorized

The left panel displays all uncategorized pages - pages that haven’t yet been grouped into a document.

Order of Uncategorized pages

Pages in this panel appear in the order they were submitted to the system, with a thumbnail preview for quick scanning.

.gif?sv=2026-02-06&spr=https&st=2026-07-27T09%3A12%3A53Z&se=2026-07-27T09%3A28%3A53Z&sr=c&sp=r&sig=GnlVklgFM1j1eJEFBqNS%2BWX51Fl4vWY7iahFb%2BKg7bQ%3D)

Middle Panel - Grouped Documents

The middle panel shows all documents you’ve grouped.

In this panel, you can:

Searching for a specific layout

You can search for a specific layout using the Layout Group drop-down list.

Searching for layout variations

You can search for and select a layout variation directly from the Layout Variation drop-down list. Selecting a layout variation automatically assigns the corresponding layout group.

Page Numbers

All pages in the left and middle panels show their submission page number based on the order in which they were uploaded.

Right Panel - Document Preview

.jpg?sv=2026-02-06&spr=https&st=2026-07-27T09%3A12%3A53Z&se=2026-07-27T09%3A28%3A53Z&sr=c&sp=r&sig=GnlVklgFM1j1eJEFBqNS%2BWX51Fl4vWY7iahFb%2BKg7bQ%3D)

The right panel shows a zoomable full-page view of the selected page.

Machine-classified grouped documents rotation

Pages in machine-classified grouped documents cannot be modified in place. You need to remove the page from the document and once it falls under Uncategorized, you can rotate it.

Reprocessing Misclassified documents

If a document is marked as Layout Incorrect during Flexible Extraction, Identification,or Transcription Supervision, you can reprocess it and send it back to Document Classification for rework.

Instead of being marked as Complete, the submission remains in progress until all documents are properly classified.

Using Reprocessing

You can trigger reprocessing from any Identification or Transcription Supervision or Flexible Extraction task. To learn more, see Supervision and QA Introduction.

  1. Go to Submissions.

  2. Find the mismatched submission by its ID.

  3. Click on the Perform Tasks link in the Tasks column for the submission.

  4. In the right sidebar of the task, expand the document information and click Mark Layout Variation Incorrect.

    • Doing so removes the layout from all pages of the current document and creates a Document Classification task.
  5. Click Continue on the warning message.

    • The document status changes to Reprocessing and is sent back to Document Classification.
      • On the Document Classification task, you’ll see the following message:
  1. Reclassify your documents by following the guidance in Document Classification Interface section.

After you’ve classified your Structured documents, the next step is to extract the data you need. Learn how to work with Flexible Extraction by reading our Flexible Extraction article.

A manual task in Hyperscience where you confirm or correct the location of fields in Semi-structured documents. You can adjust or draw bounding boxes around field values to help your model learn where to look for the data you want to extract. You may need to perform a Field ID task when the machine is not confident enough in its prediction for a field, based on the target accuracy.

A Supervision task that allows you to review or enter text the system couldn’t confidently read from a document. This task enables you to ensure accurate final data when the system’s confidence is low.

A task in Hyperscience that involves human intervention to validate or correct data extraction for Structured documents. This task is used when automatic extraction isn’t fully reliable, allowing you to transcribe or adjust specific fields to ensure accuracy.

The first step in Supervision. It is used to categorize and combine pages that were not classified by the machine.

A manual task that is created when the system’s confidence in a prediction is below the confidence threshold. Supervision allows a human to review and correct the output, ensuring data accuracy through human-in-the-loop input.