Legal Document Classification: An Application to Law Area Prediction of Petitions to Public Prosecution Service

Noguti, Mariana Y.; Vellasques, Eduardo; Oliveira, Luiz S.

doi:10.1109/IJCNN48605.2020.9207211

Abstract:In recent years, there has been an increased interest in the application of Natural Language Processing (NLP) to legal documents. The use of convolutional and recurrent neural networks along with word embedding techniques have presented promising results when applied to textual classification problems, such as sentiment analysis and topic segmentation of documents. This paper proposes the use of NLP techniques for textual classification, with the purpose of categorizing the descriptions of the services provided by the Public Prosecutor's Office of the State of Paraná to the population in one of the areas of law covered by the institution. Our main goal is to automate the process of assigning petitions to their respective areas of law, with a consequent reduction in costs and time associated with such process while allowing the allocation of human resources to more complex tasks. In this paper, we compare different approaches to word representations in the aforementioned task: including document-term matrices and a few different word embeddings. With regards to the classification models, we evaluated three different families: linear models, boosted trees and neural networks. The best results were obtained with a combination of Word2Vec trained on a domain-specific corpus and a Recurrent Neural Network (RNN) architecture (more specifically, LSTM), leading to an accuracy of 90\% and F1-Score of 85\% in the classification of eighteen categories (law areas).

Subjects:	Information Retrieval (cs.IR); Machine Learning (cs.LG)
Cite as:	arXiv:2010.12533 [cs.IR]
	(or arXiv:2010.12533v1 [cs.IR] for this version)
	https://doi.org/10.48550/arXiv.2010.12533
Related DOI:	https://doi.org/10.1109/IJCNN48605.2020.9207211

Computer Science > Information Retrieval

Title:Legal Document Classification: An Application to Law Area Prediction of Petitions to Public Prosecution Service

Submission history

Access Paper:

References & Citations

DBLP - CS Bibliography

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators