Classification before document processing
The starting scope is a known set of document families and destinations. We retain the original file and make the routing decision traceable.
- Tools
- Azure AI Document Intelligence, n8n, Airtable
- Trigger
- A document enters the agreed intake with a source ID and access to the pages needed for classification.
- Outcome
- A document category, relevant page ranges and routing decision are recorded. Eligible files reach the selected queue or extractor; uncertain cases enter manual review.
Categories your team can apply consistently
We define each class with examples and exclusions, including similar-looking documents that need different handling. Categories correspond to a business destination or next step. We test how the model handles unfamiliar material rather than assuming every unknown file will identify itself as unknown.
Page ranges for mixed document packets
When an upload contains several documents, we assess whether classification should operate on the whole file, individual pages or detected document sections. Splitting behavior is configured deliberately and tested against packet examples. The original upload remains available while routing records point to the relevant pages.
Routing with an exception queue
Airtable can record the source, predicted class, checks and destination. Agreed thresholds and rule conflicts send a file to review. The processing record identifies which handoff completed, so a retry does not repeatedly notify an owner or submit the same pages to an extractor.