• Skip to primary navigation
  • Skip to main content
  • Skip to primary sidebar

information for practice

news, new scholarship & more from around the world


advanced search
  • gary.holden@nyu.edu
  • @ Info4Practice
  • Archive
  • About
  • Help
  • Browse Key Journals
  • RSS Feeds

Tackling challenges in large language model–based data extraction via context engineering: A commentary on Jansen et al. (2025).

Psychological Bulletin, Vol 152(4), Apr 2026, 404-419; doi:10.1037/bul0000520

Systematic reviews, particularly meta-analyses, involve crucial yet labor-intensive and error-prone stages of data extraction. Recent advances in large language models (LLMs) have unlocked new avenues for automating this process, potentially enhancing both efficiency and reliability. Recently, Jansen et al. (2025) systematically evaluated the accuracy and error patterns of LLM-assisted data extraction across 22 reviews published in Psychological Bulletin. Their findings indicated that while achieving acceptable-to-good accuracy for some variables describing study characteristics, LLMs struggled with numerical variables, especially those related to effect sizes. In this commentary, we discuss the current challenges of automated data extraction and potential pathways to improve the work reported in Jansen et al.’s study. We situate our discussion within the framework of context engineering, aiming to refine the information provided to LLMs through dynamic optimization strategies tailored to specific tasks. We identify five key challenges that reflect either LLMs’ unique patterns or standard practices in research synthesis: parsing semistructured data, understanding long contexts, performing arithmetic induction, engaging in complex reasoning, and ensuring the reproducibility of coding protocols. We then outline potential solutions inspired by context engineering implementations such as retrieval-augmented generation and tool-integrated reasoning. For illustration, we present four examples: extracting semistructured data via optical character recognition, reliably computing effect sizes through function calls, performing adaptive retrieval with LLM-based agents, and iteratively improving outputs through self-refinement. We conclude by calling for future research in automated data extraction to advance beyond simple instruction-following paradigms toward more reliable forms of context engineering. (PsycInfo Database Record (c) 2026 APA, all rights reserved)

Read the full article ›

Posted in: Journal Article Abstracts on 07/10/2026 | Link to this post on IFP |
Share

Primary Sidebar

Categories

Category RSS Feeds

  • Calls & Consultations
  • Clinical Trials
  • Funding
  • Grey Literature
  • Guidelines Plus
  • History
  • Infographics
  • Journal Article Abstracts
  • Meta-analyses - Systematic Reviews
  • Monographs & Edited Collections
  • News
  • Open Access Journal Articles
  • Podcasts
  • Video

© 1993-2026 Dr. Gary Holden. All rights reserved.

gary.holden@nyu.edu
@Info4Practice