Document processing has been a solved problem for structured documents. PDF extraction, OCR for clean print, template-based field extraction — these work well when documents are consistent, clearly typed, and properly formatted. Most documents that enterprises actually process are not like that. They are scanned at angles. They have handwritten annotations alongside printed text. They […]