Abstract
Capturing of new documents into the enterprise document repository involves some kind of (usually labour intensive) task of metadata attributing. Attributes include document types, folders, cases and others. This is, in fact, a classification process and may be addressed by Machine Learning based classification methods. The research presented in this paper aims to introduce a flexible framework for document capturing automation employing Machine Learning based document classification bots and providing an intuitive user interface featuring analysis of the document classification performance, creation of the domain specific rules, and configuration or the process parameters. The paper outlines important aspects of the framework - channels, rules, handling imbalanced data and continuous training. The prototype DOBO of the framework is presented and its performance evaluation results are discussed.
Author supplied keywords
Cite
CITATION STYLE
Rats, J., & Pede, I. (2020). Document capturing automation: Case study. Baltic Journal of Modern Computing, 8(2), 259–274. https://doi.org/10.22364/BJMC.2020.8.2.04
Register to see more suggestions
Mendeley helps you to discover research relevant for your work.