Preprint
Machine Learning

Artificial intelligence for literature reviews: opportunities and challenges

F. J. Bolaños(The Open University), Angelo A. Salatino(The Open University), Francesco Osborne(The Open University), Enrico Motta(The Open University)
August 17, 2024Artificial Intelligence Review218 citations

218

Citations

12

Influential Citations

Artificial Intelligence Review

Venue

2024

Year

Abstract

Abstract This paper presents a comprehensive review of the use of Artificial Intelligence (AI) in Systematic Literature Reviews (SLRs). A SLR is a rigorous and organised methodology that assesses and integrates prior research on a given topic. Numerous tools have been developed to assist and partially automate the SLR process. The increasing role of AI in this field shows great potential in providing more effective support for researchers, moving towards the semi-automatic creation of literature reviews. Our study focuses on how AI techniques are applied in the semi-automation of SLRs, specifically in the screening and extraction phases. We examine 21 leading SLR tools using a framework that combines 23 traditional features with 11 AI features. We also analyse 11 recent tools that leverage large language models for searching the literature and assisting academic writing. Finally, the paper discusses current trends in the field, outlines key research challenges, and suggests directions for future research. We highlight three primary research challenges: integrating advanced AI solutions, such as large language models and knowledge graphs, improving usability, and developing a standardised evaluation framework. We also propose best practices to ensure more robust evaluations in terms of performance, usability, and transparency. Overall, this review offers a detailed overview of AI-enhanced SLR tools for researchers and practitioners, providing a foundation for the development of next-generation AI solutions in this field.

Analysis

Why This Paper Matters

Systematic Literature Reviews (SLRs) are a cornerstone of evidence-based research but are notoriously time-consuming and labor-intensive. As the volume of scientific publications grows exponentially, the need for automated or semi-automated tools becomes critical. This paper addresses that need by providing a structured, up-to-date overview of how Artificial Intelligence (AI) is being applied to streamline SLR processes, particularly in the screening and extraction phases. The timing is especially relevant given the recent surge in large language models (LLMs) like GPT-4, which are now being integrated into research workflows. By systematically analyzing both traditional and AI-enhanced tools, the authors offer a practical guide for researchers seeking to adopt these technologies, while also highlighting gaps that future work must fill.

Technical Contributions

The paper's main technical contribution is its dual-framework analysis of SLR tools. First, it evaluates 21 leading tools using a feature set that combines 23 traditional capabilities (e.g., duplicate detection, citation management) with 11 AI-specific features (e.g., active learning, natural language processing for screening). Second, it examines 11 recent tools that leverage LLMs for tasks like literature search and academic writing. This structured comparison allows the authors to identify which phases of the SLR process are most amenable to AI automation and where current tools fall short. Key innovations highlighted include:

  • AI for screening: Tools using active learning and text classification to prioritize relevant studies.
  • AI for extraction: Named entity recognition and relation extraction to automatically populate data tables.
  • LLM integration: Chat-based interfaces for query refinement and summarization.
  • Knowledge graphs: Representing relationships between concepts to support literature synthesis.

The paper also proposes a standardized evaluation framework, addressing a critical gap in the field where tools are often assessed on ad-hoc metrics.

Results

The review does not present new experimental results but synthesizes findings from the literature. Key observations include:

  • Most AI-enhanced tools focus on the screening phase, with reported reductions in manual effort of 30-50% in some studies.
  • LLM-based tools show promise for generating literature summaries but lack rigorous evaluation of accuracy and bias.
  • Only a minority of tools (about 20%) incorporate knowledge graphs or advanced reasoning capabilities.
  • The authors note that no single tool covers all SLR phases effectively, and interoperability between tools remains a challenge.

Significance

This paper serves as a valuable reference for both researchers developing new SLR tools and practitioners looking to adopt existing ones. By clearly delineating the state of the art and the remaining challenges, it sets a research agenda for the next generation of AI-assisted literature review systems. The emphasis on standardized evaluation is particularly important, as it could lead to more reproducible and comparable benchmarks. As AI continues to permeate academic workflows, this work provides a roadmap for ensuring that these tools are not only powerful but also trustworthy and user-friendly.