// NATURE NEWS — SPAZIO & SCIENZA
My four-step workflow for reading scientific papers (without trying to read them all)
Shea Andrews is an assistant professor of psychiatry and behavioral sciences at the University of California, San Francisco.
Search author on:
PubMed
Google Scholar
Careful use of a reference manager can help scientists to keep control of their to-read pile. Credit: Impact Photography/Shutterstock
In the past 13 years, I have added 10,000 papers to the reference manager that I first used when I was a PhD student. The topics span Alzheimer’s disease, genetics, biostatistics, epidemiology, the philosophy of science and more. Some of these works taught me about my own field; others pulled me away from it, informing my understanding of other areas of science, alternative methods and fresh ways of seeing a problem.
That range reflects how I learnt to read the literature as a scientist. Early on in my PhD, during which I studied the genetic and environmental factors that contribute to cognitive decline, I thought I needed to read every paper closely and critically, especially if I planned to cite it. That belief turned my reference manager into an ever-expanding to-read list. I quickly realized, while writing my thesis proposal, that I could not read every paper closely. Part of my development as a scientist has been learning to read selectively, to the depth required for the task at hand.
Even selective reading, however, cannot fully relieve the pressure of keeping up with the exponentially growing scientific literature. Artificial-intelligence tools built on large language models (LLMs) promise fresh ways to search, summarize and sort the literature. When I first used ChatGPT, I was impressed by how well it could summarize scientific topics. But I soon learnt that it could also fabricate citations, and hallucinated references have become a visible sign of uncritical AI use in manuscripts.
Although the most recent LLMs, with built-in search, return real papers and summarize the literature more reliably, they do not eliminate the risk of uncritical citations. This is a modern version of an old problem, in which scientists cite papers they have not read, or have misunderstood or inherited from other reference lists. A key part of developing as a scientist today is learning how to use AI tools for literature management without letting them replace scientific judgement. This is also how I frame these tools for my students and trainees.
In my own work, and when conducting training, I treat reading as a set of decisions about purpose, relevance and depth. AI can assist with those decisions, but it cannot eliminate them. Those decisions usually fall into four broad categories.
Discovery. Many papers that I read reach me before I go looking for them. This is passive discovery. Each morning, I check an RSS feed that pulls works from 35 journals, including specialist publications such as Alzheimer’s & Dementia and multidisciplinary ones such as Nature and Science. This gives me a regular view of what is being published across my main areas of interest. Newsletters, citation alerts and recommendations from colleagues serve a similar function, and social media can also reveal which papers are gaining attention across the scientific community.
Active discovery, by contrast, uses a search strategy to address a specific research need, such as identifying papers that inform a particular idea I want to develop. In the past, this meant using keywords or search terms in bibliographic databases such as PubMed or Google Scholar, and following citations from relevant articles. AI tools now add another route. When I do not know the right search terms, I increasingly use chatbots to do a first pass, asking them to identify potentially relevant papers across databases and the web. For example, for a grant application about biological ageing in Alzheimer’s disease, ChatGPT identified 52 potentially relevant papers across several searches; I added around half of these to my