Human-in-the-Loop AI Reviewing: Feasibility, Opportunities, and Risks
Iddo Drori, Dov Te’eni · Journal of the Association for Information Systems · 2024
AI-generated evidence extraction, verified across multiple analytical personas. Not a substitute for the peer-reviewed original.
This is an AI analysis. Read the peer-reviewed original at the publisher: https://doi.org/10.17705/1jais.00867
Methodology & findings
Study design
Experimental case study with comparative analysis.
Primary method
Comparative evaluation between LLM-generated reviews and human reviews; no specific statistical software mentioned in abstract.
Main result
The study found that "current AI-augmented reviewing is sufficiently accurate to alleviate the burden of reviewing but not completely and not for all cases." The authors demonstrated feasibility by evaluating and comparing LLM reviews with human reviews, and identified key opportunities and risks including bias, value misalignment, and misuse in AI-augmented academic peer review.
Reports effect sizes.
Research paradigm
Pragmatist/Mixed-methods (combines empirical experimentation with interpretive analysis)
Author conclusions
The authors conclude that "we explore the feasibility, opportunities, and risks of using large language models (LLMs) for reviewing academic submissions, while keeping the human in the loop" and that they "conclude with recommendations for managing these risks." They demonstrate that AI-augmented reviewing can alleviate reviewer burden while requiring human oversight to mitigate identified risks.
Risk of bias
Potential selection bias in choosing which submissions to review; Risk of algorithmic bias in GPT-4 reviews; Value misalignment between LLM outputs and reviewer standards; Representativeness of sample LLM reviews for broader applicability; Potential for LLM bias in academic reviewing; Value misalignment between AI systems and academic standards; Limited generalizability of GPT-4 performance across different submission types and domains; Bias in AI reviews (explicitly identified as a risk); Value misalignment between AI systems and human reviewers; Potential misuse of AI-augmented reviewing systems
Open questions raised
- The authors present 'open questions' regarding opportunities of AI-augmented reviewing and recommend further work on risk management strategies for bias, value misalignment, and misuse in AI-assisted academic reviewing.
- The authors present 'open questions' regarding AI-augmented reviewing and identify future directions related to managing risks of bias, value misalignment, and potential misuse of AI in academic peer review processes.
- The authors present "open questions" regarding opportunities of AI-augmented reviewing and identify gaps in understanding how to manage risks of bias, value misalignment, and misuse in academic peer review contexts.
Explore related topics
Related papers
- Rayyan—a web and mobile app for systematic reviewsMourad Ouzzani · 2016 · 24,664 citations
- What if the devil is my guardian angel: ChatGPT as a case study of using chatbots in educationAhmed Tlili · 2023 · 1,587 citations
- Embracing the future of Artificial Intelligence in the classroom: the relevance of AI literacy, prompt engineering, and critical thinking in modern educationYoshija Walter · 2024 · 805 citations
- Generative AI tools and assessment: Guidelines of the world's top-ranking universitiesBenjamin Luke Moorhouse · 2023 · 343 citations
- AI-assisted peer reviewAlessandro Checco · 2021 · 261 citations
- Leveraging ChatGPT for Enhancing Critical Thinking SkillsYing Guo · 2023 · 223 citations