0% Complete
Home
/
15th International Conference on Computer and Knowledge Engineering
Persian Legal Text Simplification Leveraging Transformer-Based Models
Authors :
Mohammadreza Joneidi Jafari
1
Saedeh Tahery
2
Amirhossein Nikoofard
3
1- Department of Electronics, Faculty of Electrical Engineering
2- Department of Artificial Intelligence, Faculty of Computer Engineering
3- Department of Systems and Control, Faculty of Electrical Engineering
Keywords :
Text Simplification،Persian Legal Documents،ChatGPT،Large Language Models،Transformer-Based Models
Abstract :
Legal documents often use complex and domain-specific language, which limits their accessibility to the general public. Despite growing interest in text simplification within natural language processing, the legal domain in Persian remains largely unexplored due to the lack of annotated corpora and domain-adapted models. This study presents a practical approach to Persian legal text simplification by leveraging synthetic supervision alongside fine-tuned transformer-based models. This paper creates a new labeled dataset by generating simplified versions of legal rulings using ChatGPT, and validate a subset of the outputs through expert review to ensure data quality. To address resource constraints common in real-world applications, we fine-tune lightweight encoder-decoder models, enabling efficient deployment without requiring large-scale annotation or extensive inference infrastructure. Our results show that a compact model such as ParsT5 outperforms zero-shot large language models like PersianLLaMA. The model is also enhanced with an existing attention extension, which enables efficient processing of long inputs without truncation. As a result, this work introduces the first benchmark for Persian legal text simplification, demonstrating that well-adapted, efficient models can achieve high performance in low-resource and domain-specific scenarios. By taking this solid first step, it paved the way for future research and development in natural language processing for Persian legal texts. The released dataset and code are publicly available at https://github.com/mrjoneidi/Simplification-Legal-Texts.
Papers List
List of archived papers
Segmentation of Hard Exudates in Retinal Fundus Images Using BCDU-Net
Nafise Ameri - Nasser Shoeibi - Mojtaba Abrishami
An Automated Visual Defect Segmentation for Flat Steel Surface Using Deep Neural Networks
Dorna Nourbakhsh Sabet - Mohammad Reza Zarifi - Javad Khoramdel - Yasamin Borhani - Esmaeil Najafi
Automatic Infrared-Based Volume and Mass Estimation System for Agricultural Products
Seyed Muhammad Hossein Mousavi - S. Muhammad Hassan Mosavi
Computational Microscopy Based on Fourier Ptychography using Embedded Architecture
Rezvan Mir - Abedin Vahedian
A Robust Network for Embedded Traffic Sign Recognation.
Omid Nejati Manzari - Shahriar Baradaran Shokouhi
A Synergistic Hybrid Architecture with Residual Attention and Mixture-of-Experts for Robust Hour-Ahead Forex Forecasting
Alireza Abbaszadeh - Seyyed Abed Hosseini - Mohammad Reza Akbarzadeh Totonchi
Machine Learning-Driven Prediction of Anti-Alzheimer Drug Efficacy Using PubChem Molecular Fingerprints
Mohammad Javad Sadeghi - Mohammad Javad Nemati - AliAsghar Zare - Mohammadreza Shams
Enhancing Vehicle Make and Model Recognition with 3D Attention Modules
Narges Semiromizadeh - Omid Nejati Manzari - Shahriar B. Shokouhi - Sattar Mirzakuchaki
Optimization Resource Allocation in NOMA-based Fog Computing with a Hybrid Algorithm
Zohreh Torki - S.Mojtaba Matinkhah
A New Application of Machine Learning Based Methods for Disk Space Variation Fault Diagnosis in Transformer Windings
Reza Behkam - Amir Lotfi - Gevork B. Gharehpetian
more
Samin Hamayesh - Version 43.7.0