0% Complete
Home
/
13th International Conference on Computer and Knowledge Engineering
SUT: a new multi-purpose synthetic dataset for Farsi document image analysis
Authors :
Elham Shabaninia
1
Fatemeh sadat Eslami
2
Ali Afkari Fahandari
3
Hossein Nezamabadi-pour
4
1- Department of Applied mathematics, Faculty of Sciences and Modern Technologies, Graduate University of Advanced Technology, Kerman, Iran
2- Department of Computer Engineering, Sirjan University of Technology, Sirjan, Iran
3- Department of Electrical Engineering, Shahid Bahonar University of Kerman, Kerman, Iran
4- Department of Electrical Engineering, Shahid Bahonar University of Kerman, Kerman, Iran
Keywords :
Document Image Analysis،Farsi database،Document Classification،Optical Character Recognition
Abstract :
This paper introduces a new large-scale dataset for Farsi document images, named SUT, which aims to tackle the challenges associated with obtaining diverse and substantial ground-truth data for supervised models in document image analysis (DIA) tasks, like document image classification, text detection and recognition, and information retrieval. The dataset comprises 62,453 images that have been categorized into 21 distinct classes, including identity documents featuring synthetically generated personal information superimposed on various backgrounds. The dataset also includes corresponding files with labeling information for the images. The ground-truth data is organized in CSV files containing image file paths and associated information about the embedded data. To demonstrate the efficacy of the SUT dataset in DIA tasks, it was utilized for document classification (achieving an accuracy of 86% using a convolutional neural network) and OCR (achieving a CER of 0.083 and 0.072 using Tesseract and EasyOCR engines, respectively). The SUT dataset serves as an esteemed asset for scholars engaged in the development and assessment of supervised models in Farsi document image analysis.
Papers List
List of archived papers
IR-LPR: Large Scale of Iranian License Plate Recognition Dataset
Mahdi Rahmani - Melika Sabaghian - Seyyedeh Mahila Moghadami - Mohammad Mohsen Talaie - Mahdi Naghibi - Mohammad Ali Keyvanrad
Graph Representation Learning Towards Patents Network Analysis
Mohammad Heydari - Babak Teimourpour
Automated Person Identification from Hand Images\\using Hierarchical Vision Transformer Network
Zahra Ebrahimian - Seyed Ali Mirsharji - Ramin Toosi - Mohammad Ali Akhaee
Enhanced Autoencoder-based Clustering for Message Analysis in Binary Protocols
Mohaddese Nemati - Shiva Mahmoudzadeh - Mehdi Teimouri
Artificial Intelligence applications addressing different aspects of the Covid-19 crisis and key technological solutions for future epidemics control
Nadia Khalili - Hojatollah Hamidi
Vaccine Distribution Modelling in Pandemics through Multi-Agent Systems: COVID-19 Case
Hossein Yarahmadi - Mohammad Ebrahim Shiri - Hamid Reza Navidi - Arash Sharifi - Moharram Challenger - Hassan Piriaei
Sotfware defined content popularity estimation for wireless D2D caching networks
Maede Rezaei - AhmadReza Montazerolghaem
Enhanced Principal-curve based Classifiers for Time-series Label Prediction
Seyed Aref Hakimzadeh - Koorush Ziarati
TCAR: Thermal and Congestion-Aware Routing Algorithm in a Partially Connected 3D Network on Chip
Majid Nezarat - Masoomeh Momeni
SUT: a new multi-purpose synthetic dataset for Farsi document image analysis
Elham Shabaninia - Fatemeh sadat Eslami - Ali Afkari Fahandari - Hossein Nezamabadi-pour
more
Samin Hamayesh - Version 42.2.1