The best AI document classification software in 2026: 14 tools for sorting and splitting documents

For operations teams that receive mixed files, and the engineers who build their pipelines. How 14 tools split and label documents, how each one learns a new document type, and what classification costs per page.

A binder-clipped packet of pages coming apart into four stacks, each with a differently shaped tab, and one loose page set apart in a dashed outline

Key takeaways

  • Document classification software splits a mixed file into its documents and labels each one by type, such as bank statement or W-2. Then it sends each document to the right extraction model, queue or person.
  • The 14 tools here fall into five groups. Intelligent document processing (IDP) platforms are Docsumo, ABBYY Vantage, Hyperscience, Tungsten TotalAgility and UiPath IXP. Capture and content platforms, which take in and store scanned documents, are Doxis, Grooper and ParaScript. Rossum is accounts payable (AP) software. The cloud classifiers come from Google, Microsoft and AWS. Reducto and LlamaParse are LLM document APIs that developers call from their own code.
  • Tools learn a new document type from rules, from labeled samples or from a short description that a language model reads. Docsumo, Google, Microsoft, AWS, UiPath and Grooper now offer this route too, alongside Reducto and LlamaParse.
  • Where prices are published, classification alone is cheap, from $1.25 per 1,000 pages in LlamaParse's fast mode to $15 for Reducto's Deep Classify. The time people spend fixing misfiled documents can cost more than the classifier.
  • Score a pilot on your own mixed files. Count splits and labels per document type, and watch what each tool does with a page that fits no type.
On this page
  1. What document classification software does
  2. How tools learn your document types
  3. The 14 best document classification tools
  4. Document classification software compared
  5. What classification costs per page
  6. How to test a classifier on your own files
  7. Classification for lending, insurance and healthcare files
  8. Classify only, or classify and extract?
  9. Frequently asked questions

Document classification software works out what type each incoming document is, such as a bank statement or a claim form. It also splits files that hold several documents. Then each document goes to the right extraction model, queue or person. The right tool depends on who will run it. Most operations teams want one tool that sorts and reads each document. Docsumo, ABBYY Vantage and Hyperscience do both. Engineers can build on a cloud classifier from Google, Microsoft or AWS. They can also use an API such as Reducto.

We compared 14 tools. Every fact comes from the vendor's own website or documentation, checked on September 29, 2026, and the links are in the sources at the end. We didn't run our own accuracy test, so the testing section explains how to run yours.

Which team are you?

For lending, insurance and healthcare operations
Docsumo sorts each file into the document types you name, with no training set. It extracts the fields and sends the ones it's unsure about to your reviewers. On the Enterprise plan it also checks the documents against each other.

What document classification software does#

Classification is the first step in document processing. A file arrives by email, upload, scanner or API, often as one PDF holding several documents. The software finds where each document starts and ends, and decides what each one is. Then it hands each document to an extraction model for that type, a work queue, or a person when it isn't sure.

  • Loan files
  • Claim submissions
  • Scanned mail
  • Email attachments
Classifier
  1. 01Split into documents
  2. 02Label each type
  3. 03Extract its fields
  4. 04Hold unsure ones for review
Queues, case files and your systems
Where classification sits in document processing

Splitting and labeling are separate skills, and tools differ on both. Take a loan file scanned as one 60-page PDF. It has to be cut at the right pages before anything can be labeled. If the tool cuts in the wrong place, the last page of a pay stub becomes part of the bank statement after it.

One naming note before the list. Security and records teams also use the word "classification" for labels such as public or confidential. Those labels decide who can open a file and how long it's kept. That's a different kind of software, and it isn't covered here.

How tools learn your document types#

Rules are the oldest method. A page with a form's title or a known header gets that form's label. Anything else goes to a person. Rules are cheap to run and easy to explain. But they break quietly when a layout changes or a new form arrives. So most tools now offer rules as an extra layer on top of a model.

Most platforms and the older cloud classifiers learn from labeled examples instead. Azure's custom classifier needs at least two document types, which Azure calls classes, and five samples of each. UiPath's trained classifier has the same minimum, but UiPath recommends 150 documents per type. Hyperscience wants at least 10 pages per layout, with 120 recommended. Accuracy on your own layouts is usually good, but every new type means collecting and labeling more samples, then training again.

The newest classifiers use a large language model (LLM) that reads each type's name and a short description. So adding a type can mean writing one sentence instead of labeling a folder of samples. Reducto and LlamaParse work only this way. Google's classifier runs on its Gemini AI models. It works from just your label names and descriptions, or after fine-tuning, which means extra training on your own samples. Microsoft's Content Understanding takes up to 200 described categories. AWS's Bedrock Data Automation matches documents to blueprints, which are AWS's name for document types you describe. UiPath and Grooper have added the same option to their older methods, and Tungsten has added an LLM-based classification copilot. Docsumo sorts files into the document types you name, with no training set. Its AI Split takes a plain-English definition of each type.

Description-based classifiers are quick to set up. Their weak point is look-alike document types, which they don't always label the same way. Put look-alike types in your pilot. Whatever the method, ask what the tool does when it isn't sure. Most tools give each label a confidence score, which says how sure the tool is. Ask what happens to a page when that score falls below the level you set. Reducto always returns its best match and advises adding an "other" category, and Microsoft gives the same advice for Azure. Docsumo holds documents it can't classify as Unclassified until a person assigns them. A classifier that files every page under its best guess makes mistakes nobody sees until later. A classifier that holds uncertain pages for a person makes mistakes you can count.

The 14 best document classification tools#

Docsumo is our product, so we've put it first and said plainly what it doesn't do. The others are grouped by what they're built for.

IDP platforms

1. Docsumo

IDP platformOur product
Docsumo is an intelligent document processing platform. Its AI Auto-Classifier splits a multi-page upload into the documents inside it. It labels each one and sends it to the right extraction model. Its LLM-based models do this without a template, a fixed layout set up for each form. It runs in the cloud only, with no on-premises option.
Splits
Mixed files arriving by email, upload or API. AI Split takes a plain-English definition of each document type, for packets such as loan binders and claim packets. Auto-classification and splitting are on the Business plan
Learns types
Name the document types you want, and it sorts files into them with no training set. Pre-trained models cover 250+ document types
Review
Documents it can't classify wait as Unclassified for a person to assign, and a reviewer can move any document to another type. Fields it's unsure about go to your own reviewers, with thresholds set per field
Output
API and webhooks; cross-document checks, case management and automated workflows on the Enterprise plan
Custom steps
You can add a workflow step that calls an LLM or runs your own Python code, for checks and calculations the built-in models don't cover
Pricing
A free 14-day trial for up to 1,000 pages; Business and Enterprise plans are priced on request
Best fitOperations teams whose loan, claim or patient files mix many document types
  • 99%field-level accuracy across 250+ document types
  • 95%+of documents processed straight through, without manual review
  • <5 minper document, down from 2+ hours

2. ABBYY Vantage

IDP platform
ABBYY Vantage is a low-code IDP platform built from pre-trained and trainable skills. A classification step sorts each document by type and routes it to the skill that extracts it, and an Assemble step groups pages into documents. It runs in ABBYY's cloud or on-premises.
Splits
Yes, by page classification, or with splitter skills for invoices, purchase orders and brokerage statements
Learns types
A pre-trained Vantage Classifier covers 70+ types, from ACORD and tax forms to bank statements. A trainable skill needs a few examples per type, or 10 to 100 when types differ only slightly
Review
Documents outside its list are labeled Unknown. A review step can send people only the documents with an unknown type, uncertain fields or rule errors
Pricing
Not published; trial on request
Best fitEnterprises with many document types, including companies that run software on their own servers

3. Hyperscience

IDP platform
Hyperscience's Hypercell platform builds document automation from processing blocks. Classification and splitting come early, and people handle the exceptions. It runs as SaaS, on-premises or air-gapped, and it has a FedRAMP High authorized cloud.
Splits
Yes, with a rule per layout, either a fixed page count or a text pattern on the first page, the last page or across pages
Learns types
Fixed forms are matched by layout; semi-structured documents train a model on at least 10 pages per layout, with 120 recommended
Review
Low-confidence pages go to a person in a classification task, or are marked No Layout Found
Pricing
Not published; volume-based
Best fitLarge enterprises and government agencies with high page volumes

4. Tungsten TotalAgility

IDP platform
Tungsten Automation was called Kofax until January 2024. It sells TotalAgility, which sorts and separates documents as part of the business processes it runs. Version 2026.1 added an LLM-based Copilot for Classification for highly variable documents.
Splits
Yes. Trainable Document Separation learns from correctly separated samples
Learns types
Sample documents per type, optionally with rules; a clustering tool groups unknown documents into candidate types
Review
Document review and validation steps in the workflow, and Copilot confidence scores (new in 2026.3) that let high-confidence results pass on their own
Pricing
Not published
Best fitEnterprises that want capture and process automation from one vendor, in the cloud or on-premises

5. UiPath IXP

Automation suite
UiPath IXP is the document processing product inside the UiPath platform, and its page now leads with classification and extraction. Classified documents go on to UiPath's software robots and AI agents, and Document Understanding is still available from within IXP.
Splits
Yes, with a trainable splitter that finds where each document starts and ends in a packet and labels each part. It's an early preview release, for customers whose UiPath cloud is hosted in the US or Europe
Learns types
Pre-trained types, from W-2s and ACORD forms to CMS-1500s. A trained classifier needs at least 5 documents per type. A generative classifier reads a name and description for each type
Review
Unrecognized documents come back as Unknown, and validation by a person is built in
Pricing
Community is free, and Basic starts at $25 a month. Classification comes with the Standard and Enterprise plans, which are priced on request. On the Flex plan, classification uses 0.2 of UiPath's AI units a page. Standard has a 60-day trial
Best fitTeams already running UiPath automations

Capture and content platforms

6. Doxis AI.dp (formerly Klippa DocHorizon)

Capture and content
Doxis AI.dp is the IDP module of the Doxis content platform, from the company called SER Group until January 2026. It captures, classifies, extracts and routes business documents, and it took over Klippa's DocHorizon after Doxis bought Klippa in March 2025.
Splits
By page ranges you set; automatic boundary detection isn't described
Learns types
50+ document types are ready to use, and custom types train from a few examples
Review
Confidence thresholds per workflow, with documents below them held for a person
Pricing
Not published; priced on request, by volume and use case
Best fitEnterprises that already store and manage documents in Doxis

7. Grooper

Capture and content
Grooper, from BIS, is an IDP platform built for hard, messy documents, and it runs on-premises, in the cloud or both. It lists six classify methods, from rules and trained examples to computer vision and an LLM classifier.
Splits
Yes. ESP Auto Separation classifies and separates in one pass, and an LLM-based AI Separate needs no training
Learns types
Trained examples, rules or visual matching, while the LLM Classifier needs only each type's name and description
Review
Confidence scores, with low-confidence documents flagged for an operator
Pricing
Not published; licensed by page volume
Best fitTeams with difficult paper, such as EOBs, leases and auto loan deal jackets, that want control over the models

8. ParaScript

Capture and content
ParaScript's FormXtra.AI learns to classify and split documents from tagged samples of your own documents. ParaScript's website highlights handwriting and poor-quality images. Stakk announced on September 24, 2026 that it had completed its purchase of ParaScript.
Splits
Yes, without separator pages or barcodes; version 8.0 added deep-learning separation that can split two documents on one page
Learns types
Examples of your documents (no count stated), with optional rules; Auto-Discovery clusters unknown documents into groups you then name
Review
FormXtra.AI Capture has validation workflows at the document or field level, including double-blind checks; confidence scores aren't described
Pricing
Not published
Watch for
The newest FormXtra.AI release announced on its site is version 8.4, from January 2023, so ask what has shipped since
Best fitMortgage, claims and records teams with handwriting-heavy paper

AP software

9. Rossum

AP software
Rossum sells AI document processing for transactional work such as AP and order processing, and Coupa announced its acquisition on May 12, 2026. Its built-in document type field covers a fixed list of AP document types.
Splits
Yes. It suggests splits in bundled files for a person to confirm, and QR-code separator pages split files at scan time
Learns types
A built-in document type field with six values, from tax invoice and credit note to receipt and other. A sorting extension routes each type to its own queue
Review
Split suggestions and uncertain fields are checked on a validation screen
Pricing
Starter from $18,000 a year; higher plans priced on request; 14-day trial
Watch for
Split suggestions cover the first 32 pages of a file by default
Best fitAP teams, especially those on Coupa

Cloud classifiers

10. Google Document AI

Cloud classifier
Google Cloud's Document AI has a custom classifier and a custom splitter. Both run on Google's Gemini AI models. They work from just your label names and descriptions, or after extra training on your samples.
Splits
Yes. The splitter returns page ranges and a type for each document, and you cut the file with Google's SDK
Learns types
Label names and descriptions with no training, a fine-tuned model, or a trained model. For training, Google recommends at least 10 documents per label
Review
Confidence scores, and a catch-all type for documents that fit no other. Google deprecated its human-review feature in January 2024, so you build the review step yourself
Pricing
$5 per 1,000 pages for the first million pages a month, then $3, for the classifier or the splitter. Google lists $0.05 an hour for hosting a custom processor
Watch for
The older lending and procurement splitters were discontinued on June 30, 2026, and the processor list gives English as the language
Best fitEngineering teams on Google Cloud

11. Azure Document Intelligence

Cloud classifier
Microsoft's Document Intelligence is now part of Azure Content Understanding in Foundry Tools. Its custom classification model labels documents and splits files. Content Understanding adds classification from category descriptions, with no training set.
Splits
Yes, once you set the split mode to auto, since the v4.0 default is none
Learns types
A trained classifier with at least two classes and five samples of each, or up to 200 described categories in Content Understanding
Review
A confidence score per result; Microsoft suggests a threshold or an "other" type, and no review screen is included
Pricing
Custom classification $3 per 1,000 pages (East US); 500 free pages a month on the free tier
Best fitEngineering teams on Azure

12. Amazon Textract and Bedrock Data Automation

Cloud classifier
AWS spreads classification across three services. Textract's Analyze Lending splits and classifies mortgage loan packages. Bedrock Data Automation matches each document to a blueprint, a document type you describe, and it can split files. Comprehend trains custom classifiers on labeled examples.
Splits
Yes in Analyze Lending and in Bedrock Data Automation projects; Comprehend doesn't split
Learns types
Analyze Lending's fixed list of mortgage document types, up to 40 described blueprints per Bedrock project, or labeled training data in Comprehend
Review
Confidence scores; AWS's human review service, A2I, no longer takes new customers
Pricing
Analyze Lending $0.07 a page for the first million pages a month (US West, Oregon); Bedrock custom output from $0.040 a page
Best fitEngineering teams on AWS, especially mortgage lenders

LLM document APIs

13. Reducto

LLM document API
Reducto is a document API for developers. It pitches Classify as a cheap first step that sends each document to the right parsing and extraction settings.
Splits
Yes. Split returns the page ranges for each section you describe, and Deep Split handles harder files
Learns types
Categories written as plain-language criteria, with no training; Classify reads the first five pages by default
Review
A confidence score and reasoning per category; it always returns a best match, so Reducto advises adding an "other" category
Pricing
Classify $7.50 and Split $20 per 1,000 pages, with $150 in free credits; Growth and Enterprise priced on request
Best fitEngineering teams that want classification as one API call, with VPC or on-premises options on Enterprise

14. LlamaParse

LLM document API
LlamaParse is LlamaIndex's hosted platform, called LlamaCloud until February 2026. It parses documents for AI apps and agents, and its Classify and Split APIs sort and cut files before extraction.
Splits
Yes. The Split API, in beta, finds where each document ends and labels each segment
Learns types
A type name plus a plain-language description, with no training
Review
A type, a confidence score and step-by-step reasoning for each file
Pricing
Classify 1 credit a page (2 in multimodal mode) and Split 4, at $1.25 per 1,000 credits; a free plan includes 10,000 credits a month
Best fitEngineering teams building AI agents or retrieval apps

Document classification software compared#

The table shows how each tool splits files, learns a type and handles uncertain documents. It follows each vendor's own site, checked September 29, 2026.

ToolSplitsLearns fromReviewPricing
IDP platforms5 tools
DocsumoOur productLoan, claim and patient filesYesType names you choose; 250+ pre-trained typesUnclassified pile; your reviewersFree trial; priced on request
ABBYY VantageMany document types, on-premises or cloudYesPre-trained classifier (70+ types) or a few samples per typeUnknown type; manual reviewNot published
HyperscienceLarge enterprises and governmentYesLayouts; 10+ pages per layoutClassification tasks for peopleNot published
Tungsten TotalAgilityCapture plus process automationYesSamples and rules; LLM CopilotReview and validation stepsNot published
UiPath IXPTeams on UiPathYes (preview)Pre-trained types, 5+ samples or descriptionsUnknown type; built-in validationPriced on request; 0.2 AI units a page
Capture and content platforms3 tools
Doxis AI.dpDoxis content platform usersPage ranges you set50+ types; a handful of samplesConfidence thresholdsNot published
GrooperHard paper, on-premisesYesSamples, rules, visual or descriptionsOperator flags on low confidenceNot published
ParaScriptHandwriting-heavy paperYesSamples plus rulesValidation workflowsNot published
AP software1 tool
RossumAP teams on CoupaYes (suggestions)Built-in invoice typesValidation screenFrom $18,000 a year
Cloud classifiers3 tools
Google Document AIEngineering teams on Google CloudYesDescriptions or 10+ samples per labelBuild your own$5 per 1,000 pages
Azure Document IntelligenceEngineering teams on AzureYes5+ samples per type, or descriptionsBuild your own$3 per 1,000 pages
Amazon Textract and BedrockEngineering teams on AWSYesMortgage types, or blueprint descriptionsBuild your own$70 per 1,000 pages (Lending)
LLM document APIs2 tools
ReductoAPI-first teamsYesCategory descriptionsBuild your own$7.50 per 1,000 pages
LlamaParseAI app and agent buildersYes (beta)Type descriptionsBuild your own$1.25 per 1,000 pages

Side-by-side pages with Docsumo are on our compare hub, including Google Document AI, Azure Document Intelligence, Grooper and Reducto.

What classification costs per page#

Where vendors publish prices, classification alone is the cheap part. Per 1,000 pages, LlamaParse's Classify costs $1.25 in its fast mode, and Azure's custom classifier costs $3. Google's classifier or splitter costs $5 for the first million pages a month. Reducto's Classify costs $7.50, or $15 for its Deep Classify. Services that classify and extract in one step cost more, such as AWS's Analyze Lending at $0.07 a page, or $70 per 1,000 pages. Most platforms, Docsumo included, price their plans on request instead, and Rossum's Starter plan begins at $18,000 a year.

The per-page price rarely decides the budget. A misfiled or badly split document costs a person minutes to find and fix. Those minutes are worth far more than a fraction of a cent. Put your own volumes into the calculator, with the tool's price per page and the share of documents your team still touches.

Cost per document: the classifier plus review time

Per-page prices look small until you add the time people spend fixing what the tool gets wrong.

From the tool's pricing page or your quote
Tool cost per month
$200
Review time per month
$1,167
Total per month
$1,367
Total cost per document
$0.27
How it's worked out
  • Tool cost = documents × pages × price per page.
  • Review time = documents × the share reviewed × minutes per review ÷ 60 × hourly cost.
  • Compare tools on the total per document, not the price per page: a cheaper tool that sends more documents to review can cost more.

How to test a classifier on your own files#

  1. Pull a real month of filesTake whole packets as they arrived, not hand-picked single documents, and keep the rare types and the pages that fit no type.
  2. Write the answer key firstMark where each document starts and what it is before any tool sees the files, so every vendor is scored against the same truth.
  3. Score splitting and labeling apartCount wrong page boundaries separately from wrong labels, because one bad cut can spoil two documents.
  4. Read accuracy per typeAn overall average can hide a type the tool keeps confusing with its look-alike, such as a bank statement and a printed transaction history.
  5. Check what happens below the thresholdSee whether uncertain documents go to a person with a reason, or get filed under the tool's best guess.
  6. Add one new type during the pilotCount the samples, rules or descriptions it took, and how long before the tool labeled it reliably.

Our take. Buy the classifier that admits doubt. In a pilot, a tool that files every page somewhere looks better on day one than one that sets aside the pages it can't place. But a misfiled pay stub comes back weeks later as a wrong income figure, while a page set aside costs a reviewer a few seconds. We'd rather see a short pile of exceptions, each with the reason it was held, than trust a queue that is quietly wrong.

Classification for lending, insurance and healthcare files#

Lenders work on the whole loan file, not one document at a time. One borrower upload can hold bank statements, pay stubs, W-2s, tax returns and an ID. The classifier decides which checks run on which pages. Docsumo's lending platform splits the file, reads each document and, on the Enterprise plan, checks the documents against each other. Lenders also shortlist Ocrolus. Its Classify step sorts a loan file into more than 2,000 pre-built document types before deeper capture. We compare the two in Docsumo vs Ocrolus.

Insurance submissions arrive as broker emails with an application, supplements, loss runs and a schedule of values attached. Each attachment needs its own label before anyone can read it. On the Enterprise plan, Docsumo also checks the loss runs against the losses reported on the ACORD 125. See Docsumo for insurance.

Healthcare back offices get claim forms, EOBs, referrals and records, often faxed as one long file. Docsumo for healthcare splits and reads them, and it signs a BAA with customers who process protected health information. If your volume is mostly scanned mail rather than case files, our guide to the digital mailroom covers intake and routing.

Classify only, or classify and extract?#

Pick by what happens after the label. Some documents only need to reach the right folder or person, as in a mailroom or an archive. For those, a capture tool such as Grooper or ParaScript can be enough, or the classifier on your own cloud. Other documents have data that must be read and checked before anyone decides. For those, choose a platform that classifies and extracts in one place. If it has to run on your own servers, look at ABBYY Vantage, Hyperscience or Tungsten TotalAgility. Engineers building their own pipeline can start with Reducto, LlamaParse or their cloud's classifier.

If your team processes loan, claim or healthcare files and wants only the uncertain fields in front of its reviewers, that's where Docsumo fits. For extraction tools, see our comparison of IDP software. For how classifiers work, see our guide to document classification.

Book a demo and bring one of your messiest mixed files, or start a free trial to test extraction on your own documents.

Frequently asked questions#

What is document classification software?

It's software that works out what each incoming document is, such as an invoice or a claim form. It also splits files that hold several documents. Then each document can be routed and have its data extracted. Most tools also score their confidence, so uncertain documents can go to a person.

What is the best document classification software?

It depends on who will run it. Operations teams that process loan, claim or healthcare files usually want one platform that classifies and extracts. Docsumo, ABBYY Vantage and Hyperscience are examples. Engineers building their own pipeline can start with their cloud's classifier from Google, Microsoft or AWS, or an API such as Reducto or LlamaParse.

Can LLMs be used for document classification?

Yes. Several tools now classify with a large language model that reads each type's name and description, so you don't need a labeled training set. Docsumo, Google's custom classifier, Reducto, LlamaParse and Grooper's LLM Classifier work this way. Test them on look-alike document types, and keep a confidence threshold that sends uncertain pages to a person.

Is document classification the same as data classification?

No. Data classification labels files as public, confidential and so on. Security teams use those labels to decide who can open a file and how long to keep it. Document classification, as covered here, sorts documents by type so each one can be read and routed.

How much does document classification software cost?

Most cloud and API classifiers charge per page. Per 1,000 pages, LlamaParse's Classify costs $1.25 in its fast mode and Azure's custom classifier costs $3. Google's custom classifier costs $5 for the first million pages a month, and Reducto's Classify costs $7.50. Most IDP platforms, Docsumo included, price their plans on request. Add the time your team spends fixing misfiled documents.

How accurate is AI document classification?

Vendors publish figures measured on their own documents, so they don't transfer to yours. Run a month of your own files through each tool. Then measure accuracy for each document type. An overall average can hide one type the tool keeps confusing with another.

Sources

  1. Docsumo: pricing
  2. ABBYY: document classification and splitting
  3. ABBYY: Vantage
  4. ABBYY: Vantage trial
  5. ABBYY: Vantage Classifier (docs)
  6. ABBYY: training a classification skill (docs)
  7. ABBYY: Assemble activity (docs)
  8. ABBYY: Invoice Splitter skill (docs)
  9. ABBYY: Purchase Order Splitter skill (docs)
  10. ABBYY: Brokerage Statement Splitter skill (docs)
  11. ABBYY: Manual Review activity (docs)
  12. Hyperscience: Hypercell platform
  13. Hyperscience: semi-structured document classification (help center)
  14. Hyperscience: training a classification model (help center)
  15. Hyperscience: automatically organize large file submissions (Auto-Splitting)
  16. Tungsten Automation: TotalAgility
  17. Tungsten Automation: TotalAgility release highlights
  18. Tungsten Automation: TotalAgility 2026.1 (March 5, 2026)
  19. Tungsten Automation: TotalAgility Features Guide 2026.2 (PDF)
  20. Tungsten Automation: classification in DocAI Studio (help)
  21. Tungsten Automation: about (formerly Kofax)
  22. UiPath: IXP
  23. UiPath: pricing
  24. UiPath: trainable splitter (docs)
  25. UiPath: classify documents automatically (docs)
  26. UiPath: Document Understanding migration to IXP (docs)
  27. UiPath: train a classifier (docs)
  28. UiPath: generative classifier (docs)
  29. UiPath: metering and charging logic (docs)
  30. Doxis: document classification
  31. Doxis: Doxis AI.dp
  32. Klippa: Klippa is now Doxis
  33. Doxis: SER Group rebrands to Doxis (January 19, 2026)
  34. Doxis AI.dp: split endpoint (docs)
  35. Grooper: document classification
  36. Grooper: LLM Classifier (wiki)
  37. Grooper: AI Separate (wiki)
  38. Grooper: classify methods (wiki)
  39. Grooper: Grooper and AI (wiki)
  40. Grooper: IDP vendors 2026 (deployment and licensing)
  41. ParaScript: document classification software
  42. ParaScript: FormXtra.AI
  43. ParaScript: FormXtra.AI SDK
  44. ParaScript: FormXtra.AI Capture
  45. ParaScript: FormXtra.AI 8.0 (February 16, 2021)
  46. ParaScript: FormXtra.AI 8.4 (January 30, 2023)
  47. ParaScript: Stakk completes acquisition of ParaScript (September 24, 2026)
  48. Rossum: pricing
  49. Rossum: Aurora
  50. Rossum: splitting documents (knowledge base)
  51. Rossum: document sorting extension (knowledge base)
  52. Rossum: Coupa acquires Rossum (May 12, 2026)
  53. Google Cloud: Document AI custom classifier
  54. Google Cloud: Document AI custom splitter
  55. Google Cloud: Document AI processor list
  56. Google Cloud: Document AI release notes
  57. Google Cloud: Document AI deprecations
  58. Google Cloud: Document AI pricing
  59. Microsoft: Document Intelligence custom classification model
  60. Microsoft: Document Intelligence pricing
  61. Microsoft: Document Intelligence in Foundry Tools
  62. Microsoft: Content Understanding classification
  63. AWS: Textract Analyze Lending (docs)
  64. AWS: Amazon Textract pricing
  65. AWS: Amazon Bedrock Data Automation
  66. AWS: Bedrock Data Automation document splitting (docs)
  67. AWS: Bedrock Data Automation projects (docs)
  68. AWS: Amazon Bedrock pricing
  69. AWS: Amazon Comprehend custom classification (docs)
  70. AWS: Amazon A2I human review loops (no longer open to new customers)
  71. Reducto: pricing
  72. Reducto: Classify (docs)
  73. Reducto: Split (docs)
  74. Reducto: credit usage (docs)
  75. LlamaIndex: LlamaParse pricing
  76. LlamaIndex: Classify (docs)
  77. LlamaIndex: Split (docs)
  78. LlamaIndex: newsletter, LlamaCloud renamed LlamaParse (February 24, 2026)
  79. Ocrolus: Classify (docs)
  80. Ocrolus: mortgage

First published .

See Docsumo read your own documents

Bring a few real samples. We'll show the fields extracted, the checks that ran and what a reviewer would see.