Anything you can do, A-I can do better... Or can it? Comparing ChatGPT's Search Strategy Outputs with Cochrane Review Searches
Rebecca Carlson, Katherine Howell, Elizabeth Moreton, Emily Jones · UNC Libraries · 2026
AI-generated evidence extraction, verified across multiple analytical personas. Not a substitute for the peer-reviewed original.
This is an AI analysis. Read the peer-reviewed original at the publisher: https://doi.org/10.17615/3drp-qx72
Methodology & findings
Study design
Comparative measurement study comparing ChatGPT-generated search strategy outputs with Cochrane Review searches
Main result
The abstract indicates that the authors "measured to what extent ChatGPT could help develop comprehensive literature search strategies," investigating ChatGPT's capabilities for completing specific literature search tasks such as term generation, database syntax, and search hedge formatting, though the abstract notes that "generalizability is lacking" in previous studies.
Reports effect sizes.
Research paradigm
Empiricist/Positivist
Author conclusions
The abstract suggests the authors are investigating "improving efficiency with GenAI" in systematic review workflows, with the stated motivation that "designing literature searches for systematic reviews is time-consuming, even for experienced librarians, so improving efficiency with GenAI is a possibility worth investigating."
Risk of bias
Potential selection bias in choice of Cochrane reviews for comparison; Possible technology bias given rapid evolution of GenAI tools; No mention of blinding in methodology; Lack of blinded assessment mentioned in abstract
Limitations
- The abstract notes that "generalizability is lacking" in previous studies that "measured ChatGPT's capabilities for completing specific literature search tasks, such as term generation, database syntax, and search hedge formatting," suggesting this study aims to address but may still have scope limitations.
Open questions raised
- The authors identify that while "previous studies have measured ChatGPT's capabilities for completing specific literature search tasks," there is a gap in understanding how ChatGPT performs on comprehensive search strategy development and generalizability across different review contexts.
- The authors identified that "Previous studies have measured ChatGPT's capabilities for completing specific literature search tasks, such as term generation, database syntax, and search hedge formatting, but generalizability is lacking," indicating a gap in comprehensive evaluation of ChatGPT for full search strategy development.
- The authors identify that "previous studies have measured ChatGPT's capabilities for completing specific literature search tasks, such as term generation, database syntax, and search hedge formatting, but generalizability is lacking," indicating a gap in understanding ChatGPT's broader applicability to comprehensive literature search strategy development.
Explore related topics
Related papers
- PRISMA Extension for Scoping Reviews (PRISMA-ScR): Checklist and ExplanationAndrea C. Tricco · 2018 · 40,391 citations
- Rayyan—a web and mobile app for systematic reviewsMourad Ouzzani · 2016 · 24,664 citations
- Cochrane Handbook for Systematic Reviews of Interventions2019 · 14,420 citations
- Guidance for conducting systematic scoping reviewsMicah D.J. Peters · 2015 · 7,472 citations
- Updated methodological guidance for the conduct of scoping reviewsMicah D.J. Peters · 2020 · 6,688 citations
- Which academic search systems are suitable for systematic reviews or meta‐analyses? Evaluating retrieval qualities of Google Scholar, PubMed, and 26 other resourcesMichael Gusenbauer · 2019 · 2,116 citations