Systematic Reviews

Bias in Individual Studies vs. Bias in the Review Process Itself

July 24, 2026·Dr. Elena Kowalski·5 min read
On this page

Discussions of bias in systematic reviews often focus almost entirely on risk-of-bias assessment of individual included studies, but this addresses only one of two conceptually distinct sources of bias -- the review process itself, independent of any individual study's quality, can introduce its own genuinely consequential distortions.

Bias within individual studies

This is the more familiar category, addressed directly by tools like RoB 2 and ROBINS-I: did a specific study's own design, conduct, or reporting introduce a systematic distortion in its measured effect -- inadequate randomization, unblinded outcome assessment, selective outcome reporting within that individual trial. This is bias existing within the primary evidence itself, before your review even begins synthesizing it.

Bias introduced by the review process

A separate, equally real category concerns whether your systematic review's own conduct -- how you searched, screened, extracted, and synthesized -- introduced distortion independent of the quality of any individual included study. A review with a comprehensive search and perfectly appraised, genuinely high-quality included studies can still produce a biased overall conclusion if the review process itself was flawed.

Selection bias in the review process

If your search strategy systematically misses a category of relevant studies -- non-English publications, given the language bias discussed elsewhere, or grey literature containing null results -- your review's conclusion can be skewed even though every individual included study is itself high quality and well-conducted. This is bias in what your review chose to include, not bias within what it did include.

Reviewer-introduced bias during screening and extraction

If screening or extraction is conducted by a single reviewer with a specific expectation about the likely finding, subtle, unconscious bias can influence borderline eligibility judgments or extraction decisions in a way dual independent screening and extraction are specifically designed to catch and correct. This is a review-process bias risk distinct from anything present in the underlying primary studies.

Synthesis-stage bias

Choosing a statistical model, a subgroup analysis, or a sensitivity analysis after seeing preliminary results, rather than pre-specifying these choices, introduces a form of review-process bias sometimes called outcome-switching or analytical flexibility at the review level, conceptually parallel to outcome switching within an individual primary study, but occurring instead at the level of how the review itself chose to analyze and present its pooled evidence.

Why this distinction matters practically

A systematic review can score well on every individual included study's risk-of-bias assessment while still producing an unreliable overall conclusion if the review's own search, screening, or analytical process was flawed. Conversely, a review of individually flawed studies can still produce a genuinely useful, honestly caveated conclusion if the review process itself transparently accounts for and communicates those individual study limitations. These are separate quality dimensions, both worth checking independently.

How PRISMA addresses both dimensions

PRISMA 2020's reporting items address both categories, though not always using this exact framing -- items covering search comprehensiveness and screening process address review-process bias risk, while items covering risk-of-bias assessment of included studies address bias within the primary evidence itself. Reading PRISMA with this two-category distinction in mind clarifies why the checklist covers such a broad range of seemingly disparate items.

GRADE and the two bias categories

GRADE's risk-of-bias domain specifically addresses bias within included studies. Its other domains -- inconsistency, indirectness, imprecision, publication bias -- more directly reflect concerns about the overall evidence base and how your review process has assembled and interpreted it, effectively addressing review-process-level concerns even though GRADE doesn't explicitly frame them using this specific two-category language.

A practical checklist covering both dimensions

When appraising or conducting a systematic review, separately ask two distinct questions: are the individual included studies themselves methodologically sound, and did the review process -- search, screening, extraction, synthesis -- introduce its own independent distortion regardless of individual study quality? A rigorous systematic review needs a defensible answer to both questions, not just careful attention to one while overlooking the other.

Why this framing matters for how reviews get critiqued

When a systematic review's conclusions are later challenged or debated, the critique often implicitly targets one of these two bias categories without making the distinction explicit -- a critique arguing the review missed key evidence is a review-process critique, while a critique arguing the underlying trials themselves were poorly conducted is a primary-study critique. Recognizing which category a specific criticism actually falls into helps a review team respond to it precisely and appropriately, rather than defending against a different kind of concern than the one actually being raised.

Building both considerations into your protocol from the start

A well-designed protocol should anticipate both bias categories explicitly -- planning a comprehensive, well-documented search and dual independent screening to guard against review-process bias, alongside a clearly matched risk-of-bias tool to assess bias within your eventual included studies. Treating these as two separate, equally important protocol design considerations, rather than focusing primarily on the more familiar risk-of-bias assessment step alone, produces a genuinely more defensible systematic review overall. Holding both considerations in mind simultaneously, rather than treating one as the primary concern and the other as secondary, reflects the fuller, more complete understanding of bias that rigorous evidence synthesis genuinely requires. This dual awareness, once genuinely internalized, tends to make a researcher a noticeably more careful reader of other published systematic reviews as well, not just a more careful conductor of their own. This dual sensitivity to both bias categories, cultivated deliberately over time, is ultimately what separates a genuinely sophisticated understanding of evidence synthesis from a more superficial, checklist-only approach to systematic review methodology. Cultivating this dual awareness deliberately, rather than assuming it develops automatically through experience alone, is a worthwhile and genuinely achievable goal for any researcher committed to conducting truly rigorous evidence synthesis work. Ultimately, this dual awareness is what allows a researcher to move beyond mechanically following a checklist toward genuinely understanding why each specific methodological safeguard exists and what it is actually protecting against, which is ultimately the deeper understanding good systematic review training aims to build.

#bias#systematic reviews#methodology