Improving Document Image Understanding with Reinforcement Finetuning

Nguyen, Bao-Sinh; Le, Dung Tien; Vu, Hieu M.; Nguyen, Tuan Anh D.; Nguyen, Minh-Tien; Le, Hung

Computer Science > Information Retrieval

arXiv:2209.12561 (cs)

[Submitted on 26 Sep 2022]

Title:Improving Document Image Understanding with Reinforcement Finetuning

Authors:Bao-Sinh Nguyen, Dung Tien Le, Hieu M. Vu, Tuan Anh D. Nguyen, Minh-Tien Nguyen, Hung Le

View PDF

Abstract:Successful Artificial Intelligence systems often require numerous labeled data to extract information from document images. In this paper, we investigate the problem of improving the performance of Artificial Intelligence systems in understanding document images, especially in cases where training data is limited. We address the problem by proposing a novel finetuning method using reinforcement learning. Our approach treats the Information Extraction model as a policy network and uses policy gradient training to update the model to maximize combined reward functions that complement the traditional cross-entropy losses. Our experiments on four datasets using labels and expert feedback demonstrate that our finetuning mechanism consistently improves the performance of a state-of-the-art information extractor, especially in the small training data regime.

Comments:	Accepted to ICONIP 2022
Subjects:	Information Retrieval (cs.IR); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
Cite as:	arXiv:2209.12561 [cs.IR]
	(or arXiv:2209.12561v1 [cs.IR] for this version)
	https://doi.org/10.48550/arXiv.2209.12561

Submission history

From: Hieu Vu [view email]
[v1] Mon, 26 Sep 2022 10:27:29 UTC (467 KB)

Computer Science > Information Retrieval

Title:Improving Document Image Understanding with Reinforcement Finetuning

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Information Retrieval

Title:Improving Document Image Understanding with Reinforcement Finetuning

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators