Towards Analysing Invoices and Receipts with Amazon Textract
Sneha Oommen
Gabby Sanchez
Cassandra T. Britto
Di Wang
Jordan Chiou
Maria Spichkova
- LMTDXAIMLAUWaLMVOT3DPCSLRISegPERMQViTUQCVCoGePILMBDLDMLReCodMedImUDHAIAIFinELMAI4TSMIALMRALMOnRL3DVUQLMMILMSSegGPAAMLSupRMLTAI4MHMDEHILMOSLMWSODLM&MASSLGNNFAttMUALMMoENAIAILawMGenSILMOTPICV3DGSCLLKELMWSOLCLIPVGenAI4ClSyDaVLMFedMLTTAOffRLCMLReLMLLMSVAI4CELRMLM&RoOCLPINNAuLLMAI4Ed3DH
Main:9 Pages
2 Figures
Bibliography:3 Pages
6 Tables
Abstract
This paper presents an evaluation of the AWS Textract in the context of extracting data from receipts. We analyse Textract functionalities using a dataset that includes receipts of varied formats and conditions. Our analysis provided a qualitative view of Textract strengths and limitations. While the receipts totals were consistently detected, we also observed typical issues and irregularities that were often influenced by image quality and layout. Based on the analysis of the observations, we propose mitigation strategies.
View on arXivComments on this paper
