Przeskocz do nawigacji głównej Przeskocz do wyszukiwania Przeskocz do głównej treści

A comparative evaluation of the effectiveness of document splitters for large language models in legal contexts

  • Silesian University of Technology

Wyniki badań: Wkład do czasopismaArtykułrecenzja

13 Cytowania z bazy Scopus

Abstrakt

The study explores the development and application of an advanced artificial intelligence-based system aimed at improving the efficiency and accuracy of legal document processing. Due to the specific nature and complexity of legal texts, traditional document management techniques often prove inadequate and error-prone, creating significant challenges for legal practitioners. The proposed method leverages natural language processing and machine learning algorithms to automate key processes such as analysis, search, and classification. By utilizing vector embedding techniques, the system enables precise information retrieval from large legal document collections, while advanced splitting methods generate concise and relevant chunks of extensive texts. The study employs a Retrieval-Augmented Generation approach, combining Large Language Models (LLMs) with external knowledge bases to enhance the accuracy and contextual relevance of generated responses, addressing common issues such as hallucinations and outdated information in traditional LLMs. The research provides an in-depth analysis of the application of various text-splitting algorithms in the context of legal document databases. The findings highlight the characteristics of appropriate algorithms and offer recommendations on the conditions under which specific mechanisms should be employed.

Język oryginałuangielski
Numer artykułu126711
CzasopismoExpert Systems with Applications
Tom272
Identyfikatory DOI
Status publikacjiOpublikowano - 5 maj 2025

Obszary tematyczne ASJC Scopus

  • Inżynieria ogólna
  • Zastosowania informatyki
  • Sztuczna inteligencja

Fingerprint

Zanurz się w tematy badawcze publikacji „A comparative evaluation of the effectiveness of document splitters for large language models in legal contexts”. Razem tworzą niepowtarzalny odcisk palca.

Cytowanie