What is it about?

When two experts score the same grant application, their scores often differ a lot. We wanted to know why. Using almost 135,000 reviews from four Norwegian research funders, we tested whether disagreement depends on the reviewers (such as their gender, age, experience and expertise), on the application, or on how the review process is organised. We fixed our questions, methods and analysis code before analysing the data, and had them peer reviewed in advance (a Registered Report). Most factors made only a small difference. Only about 3 percent of the disagreement came from stable habits of individual reviewers. About a quarter was tied to the specific application, and more than half was left unexplained, which points to ordinary variability in human judgement.

Featured Image

Why is it important?

Research funders distribute large sums through peer review, and a common response to disagreement is to fix the reviewers: train them, select them differently or replace them. Our results suggest this alone is unlikely to make reviews much more reliable, because so little of the disagreement is tied to the reviewers themselves. Gains are more likely to come from clearer criteria, more structured review formats and applications that leave less room for different readings, although even these should be expected to bring gradual rather than dramatic improvements. To our knowledge, this is the first study of the question to combine data from several funders with questions, methods and code fixed in advance.

Perspectives

I work as a programme director at one of the four funders in this study, so these results are not only academic to me. For years, the natural instinct has been to look for the harsh or overly generous reviewer. This study tells me that is mostly the wrong place to look. The more useful question is how we ask reviewers to judge: what the criteria say, how the form is structured, and how clearly an application can be read. Fixing the analysis in advance also made it easier to trust the findings, including the ones that went against what earlier research led us to expect.

Jan-Ole Hesselberg
Universitetet i Oslo

Read the Original

This page is a summary of: Why do reviewers disagree? Evidence from four funders and 134,000 reviews, PLOS One, September 2026, PLOS,
DOI: 10.1371/journal.pone.0356108.
You can read the full text:

Read
Open access logo

Contributors

The following have contributed to this page