2025/09/16 by Wenjun Ke, Yifan Zheng, Youlan Li +5
Computer Science · #Handwritten Text Recognition Techniques #Natural Language Processing Techniques #Topic Modeling
paper · doi:10.1145/3768156
openalex publication_date 2025/09/16 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/17
The rapid proliferation of documents has made document intelligence increasingly critical across various industries. In recent years, Large Language Models (LLMs) have dramatically transformed the field of document intelligence, allowing for more advanced and accurate document processing solutions. Despite these advancements, most existing surveys have failed to focus on these breakthroughs, instead concentrating on traditional methods and earlier machine learning techniques. This survey seeks to fill that gap by offering an in-depth analysis of approximately 300 papers published between 2021 and mid-2025, thus providing a comprehensive overview of the impact of LLMs in document intelligence. The key topics explored include Retrieval-Augmented Generation (RAG), long-context processing, and fine-tuning LLMs for document comprehension. Furthermore, the survey highlights essential datasets, practical applications, current challenges, and future research directions, offering critical insights for both researchers and industry practitioners looking to advance the field.