Financial table extraction in image documents

Proceedings of the First ACM International Conference on AI in Finance Pub Date : 2020-10-15 DOI:10.1145/3383455.3422520

W. Watson, Bo Liu

引用次数: 0

Abstract

Table extraction has long been a pervasive problem in financial services. This is more challenging in the image domain, where content is locked behind cumbersome pixel format. Luckily, advances in deep learning for image segmentation, OCR, and sequence modeling provides the necessary heavy lifting to achieve impressive results. This paper presents an end-to-end pipeline for identifying, extracting and transcribing tabular content in image documents, while retaining the original spatial relations with high fidelity.

查看原文

微信好友朋友圈 QQ好友复制链接

本刊更多论文

财务表格提取图像文档

长期以来，表提取一直是金融服务中普遍存在的问题。这在图像领域更具挑战性，因为内容被锁定在繁琐的像素格式后面。幸运的是，深度学习在图像分割、OCR和序列建模方面的进步为实现令人印象深刻的结果提供了必要的提升。本文提出了一种端到端的管道，用于识别、提取和转录图像文档中的表格内容，同时高保真地保留原始空间关系。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文去求助

来源期刊

Proceedings of the First ACM International Conference on AI in Finance

自引率

0.00%

发文量

期刊最新文献

Recommending missing and suspicious links in multiplex financial networks A hybrid learning approach to detecting regime switches in financial markets Financial table extraction in image documents Index tracking with differentiable asset selection Deep Q-network-based adaptive alert threshold selection policy for payment fraud systems in retail banking