Manuel B. Garcia

Manuel B. Garcia serves as the Senior Director for Educational Technology and Digital Learning at FEU Institute of Technology, Manila, Philippines. Read More

Contact Info

1607, FEU Tech Building,
P. Paredes St, Sampaloc,
Manila, Philippines
mbgarcia@feutech.edu.ph

Follow Me

How Do You Handle Multiple Papers From the Same Study or Dataset?

One study can produce several papers, and those papers should not automatically be counted as separate studies. Learn how to link related reports, combine their information, and prevent the same participants or data from influencing your review more than once.

125
Multiple Papers From the Same Study Guide 125 of 247
01 · The Question

What If Several Papers Actually Come From the Same Research?

You find three papers that all appear eligible for your review. Their titles are different. They were published in different journals. One examines the primary outcome, another reports a follow-up, and a third analyzes a secondary outcome.

Then you notice something familiar: several authors overlap, the sample sizes are similar, participants were recruited from the same institutions, and the study dates seem to match.

Are these three studies, or three papers reporting one study?

This distinction matters because systematic reviews generally treat the study, not the individual publication, as the primary unit of interest. A single study may generate a protocol, conference abstract, primary results paper, follow-up article, secondary analysis, qualitative companion paper, or other report. If those publications are mistakenly counted as independent studies, the same participants or underlying evidence can influence the review more than once.

The correct approach is therefore not to discard all but one paper. It is to determine which reports belong to the same underlying study, link them together, and use the relevant information across those reports without double-counting the study.

02 · The Short Answer

Treat the Study as the Unit, and the Papers as Reports About It

In Brief

If multiple papers report the same underlying study, cohort, trial, or participant sample, link those papers together and treat the study as one unit rather than counting every publication as an independent study.

Do not automatically discard the additional papers. Different reports may contain complementary information about methods, participants, outcomes, follow-up periods, or analyses. The task is to combine information across reports carefully while preventing the same participants, outcomes, or data from being counted more than once.

03 · What You Need to Know

One Study Can Produce Many Publications

Start by Distinguishing a Study From a Report

In everyday academic conversation, researchers often use “paper” and “study” almost interchangeably. During evidence synthesis, that shortcut can become problematic.

A study is the underlying research investigation. A report is one source through which information about that study becomes available.

Study The underlying investigation involving particular participants, data, interventions, observations, or research procedures.
Report A publication or other record communicating information about that study, such as a journal article, conference abstract, protocol, registry entry, thesis, or follow-up paper.

Cochrane explicitly identifies studies rather than reports as the primary units of interest in systematic reviews. Searches, however, usually retrieve reports. This creates an additional task during study selection: identifying which retrieved reports belong to the same underlying study.

That distinction also explains why deciding which papers belong in a review eventually becomes more complicated than simply counting eligible articles.

Why Would One Study Produce Several Papers?

There are many legitimate reasons.

A large research project may answer several questions that cannot reasonably be reported in one article. Different publications may address:

  • the study protocol or design;
  • primary outcomes;
  • secondary outcomes;
  • different participant subgroups;
  • different measurement instruments;
  • short-term and long-term follow-up;
  • qualitative findings embedded within a larger project;
  • process or implementation evaluation;
  • economic analysis;
  • adverse events;
  • secondary analyses; or
  • additional research questions using the same dataset.

Multiple publications are therefore not inherently suspicious. A rich longitudinal study may legitimately support a substantial publication program.

The methodological problem arises when the relationship among those papers is not recognized and the same underlying evidence is treated as though it came from independent studies.

Multiple Reports Are Not the Same as Duplicate Database Records

This distinction is particularly important during screening.

If you search several databases, the exact same journal article may appear multiple times. Those are duplicate records of one report and can usually be removed during deduplication.

Multiple reports from the same study are different. They may have different titles, years, journals, authors, outcomes, and even sample sizes.

What you found What it represents Typical action
The identical journal article retrieved from Scopus and Web of Science Duplicate records of the same report Deduplicate the records
A conference abstract and later full journal article reporting the same trial Multiple reports of one study Link the reports to the same study
A primary results article and a five-year follow-up from the same cohort Multiple reports of one underlying study or cohort Link them and use the relevant time-point information appropriately
Two independent samples collected by the same research team Potentially separate studies Treat separately if they are genuinely independent
Two papers analyzing overlapping participants from one dataset Related analyses with non-independent data Identify the overlap and avoid treating the evidence as independent

The neighboring problem of duplicate publications and the bias they can create therefore overlaps with, but is not identical to, the broader problem of multiple legitimate reports from one study.

How Can You Tell Whether Two Papers Come From the Same Study?

Sometimes the relationship is obvious. Both papers state the trial registration number or explicitly identify themselves as analyses from the same project.

Other cases require detective work.

Cochrane recommends comparing characteristics such as:

  • trial or study registration numbers;
  • author names;
  • institutions and study locations;
  • intervention details;
  • sample sizes and baseline characteristics;
  • recruitment dates;
  • study duration; and
  • other distinctive study characteristics.

No single clue is always decisive.

Author lists can change across publications. Sample size can decrease at follow-up. Recruitment may occur at several sites. A secondary analysis may use only a subset of the original participants. Conversely, two studies conducted by the same authors at the same institution may genuinely be independent.

Look for a pattern of matching characteristics rather than relying on one superficial similarity.

Registration Numbers and Study Identifiers Are Especially Useful

When available, trial registration numbers and other unique study identifiers can make report linkage much easier.

For example, two papers carrying the same ClinicalTrials.gov identifier strongly indicate that they relate to the same registered study, even if their titles, authorship order, or outcomes differ.

Registries can also help identify reports you have not yet found. A registry entry may link to publications, identify the study's primary and secondary outcomes, or reveal that what appears to be a standalone paper is actually a secondary analysis from a larger trial.

Identifiers are not universally available, particularly for older research and study designs that are not routinely registered. They should therefore be used as strong evidence when present rather than as a requirement for establishing that reports are related.

Author Overlap Is Helpful but Not Conclusive

Shared authorship is another useful clue.

If two papers involve similar samples, the same institution, matching recruitment dates, and several common authors, they may well originate from the same study.

But author lists can vary considerably. A statistician may appear only on one analysis. A doctoral student may lead a secondary paper. A multicenter study may generate publications with different subsets of investigators.

The reverse problem also occurs. A research group may conduct several similar studies over successive years, producing papers with almost identical author lists and settings.

Shared authors therefore raise the possibility of overlap. They do not prove it.

Sample Size Differences Do Not Necessarily Mean Different Studies

This can be especially confusing.

One paper reports 500 participants. Another apparently related article reports 436. A third analyzes 212 participants.

Are these separate studies?

Possibly, but several other explanations exist. The first report may describe everyone enrolled, the second only participants with complete follow-up data, and the third a subgroup eligible for a particular analysis.

Cochrane specifically notes that participant numbers may differ across publications from the same study. Recruitment dates and baseline characteristics can help determine whether changing sample sizes reflect different stages or subsets of one study rather than independent research.

Do Not Choose One Paper and Throw the Others Away

Once you determine that several papers describe one study, the tempting solution is to keep the “main paper” and delete the rest.

That can lose valuable information.

Cochrane requires multiple reports from the same study to be collated and specifically cautions against discarding secondary reports because they may contain useful information about study design, conduct, outcomes, or other characteristics.

One article may provide the clearest description of participant recruitment. Another may contain the outcome needed for your review. A follow-up paper may provide the only long-term results. A protocol may explain allocation procedures more clearly than the results article.

The study should therefore have one identity in your review, but that identity may draw information from several sources.

Identify potentially related reports Flag papers that appear to describe the same study, cohort, trial, or dataset.
Compare study characteristics Check identifiers, authors, sites, sample characteristics, intervention details, and dates.
Link confirmed reports Assign the related papers to one study-level record or study identifier.
Extract complementary information Use relevant details from all reports rather than automatically relying on only one article.
Count the study once Ensure that multiple publications do not become multiple independent contributions of the same underlying participants or data.

You May Still Need to Identify a Primary Report

Although information can come from several reports, it is often useful to identify one publication as the primary or main report for the study.

Cochrane requires review authors to choose and justify which report is used as the principal source for study results when multiple reports exist. The most appropriate report is not necessarily the earliest publication or the paper in the highest-ranked journal.

The primary report might be the one that:

  • provides the most complete description of the main study;
  • reports the outcomes most relevant to the review;
  • contains the most complete participant information;
  • corresponds to the prespecified primary analysis; or
  • provides the clearest and most complete results for the relevant time point.

Secondary reports should remain linked to that study because they may supplement or clarify the primary report.

What If the Reports Disagree?

Multiple reports from the same study do not always tell the same story.

Sample sizes may differ. Outcome values may be reported differently. A later paper may use a different analysis. One report may label an outcome as primary while another emphasizes something else.

These discrepancies should not be resolved by automatically choosing whichever result is most favorable or easiest to extract.

First determine whether the apparent conflict has a legitimate explanation:

  • different follow-up periods;
  • different analytic populations;
  • updated or corrected data;
  • adjusted versus unadjusted analyses;
  • subgroup versus full-sample results;
  • different outcome definitions; or
  • different stages of recruitment.

If the discrepancy remains unresolved, check protocols, registrations, corrections, supplementary materials, or other reports. Contacting investigators may also be appropriate.

PRISMA 2020 asks systematic reviewers to report decision rules used when selecting data from multiple reports of the same study and any steps taken to resolve inconsistencies across reports.

One Dataset Can Produce Several Genuine Research Questions

The issue becomes more subtle when researchers publish multiple analyses from the same dataset.

Suppose a national student survey contains responses from 20,000 participants. One paper examines academic stress, another investigates digital literacy, and a third analyzes generative AI use. All three use the same underlying dataset but address different questions.

These are distinct papers and may contain genuinely different analyses. Whether they should be treated as one “study” for every purpose depends on the structure of your review and the specific evidence being synthesized.

The critical concern is statistical and evidential independence. If two papers contribute different outcomes from the same participants, that may be entirely appropriate. If they contribute estimates to the same synthesis as though they came from independent samples, the shared participants can create dependency and effectively give that dataset more influence than intended.

Watch Out

“Different paper” does not necessarily mean “independent evidence.” When publications use the same or overlapping participants, determine whether their contributions to your synthesis are statistically or conceptually dependent before treating them as separate observations.

Overlapping Samples Are Harder Than Identical Samples

Sometimes two papers do not use exactly the same participants, but the samples overlap.

For example, one article may analyze the first three waves of a longitudinal cohort while another analyzes waves two through five. Or two papers may draw participants from the same large registry during overlapping recruitment periods.

This creates a dependency problem rather than a simple duplicate-publication problem.

You may need to determine:

  • how much participant overlap exists;
  • whether the same outcomes and time points are being analyzed;
  • whether one report contains a subset of another;
  • whether estimates can legitimately enter the same synthesis; and
  • whether statistical methods are needed to account for dependence.

The appropriate solution depends on the synthesis method. The central principle remains that the same observations should not be treated as independent simply because they appear in separate publications.

Multiple Time Points Do Not Automatically Create Multiple Studies

A longitudinal study may produce separate publications for six-month, one-year, and five-year outcomes.

These are still reports from the same underlying study, although each may provide evidence relevant to different time frames in your review.

Cochrane guidance on outcome multiplicity recommends prespecifying how multiple eligible time points or measures will be handled. Selecting among them after seeing which result is most favorable can introduce bias.

If your review distinguishes short-, medium-, and long-term outcomes, different reports from the same study may legitimately contribute to different time frames. What you should not do is count the same study as three independent studies merely because three publications exist.

Multiple Reports Can Improve Critical Appraisal

Linking reports is not only about preventing double-counting.

Additional publications can reveal information that changes your understanding of methodological quality or risk of bias.

A brief primary article may provide little detail about allocation procedures, while the protocol describes them clearly. A follow-up paper may reveal attrition that was not obvious in an earlier report. A registry entry may show that outcomes were prespecified differently from how they were presented in the published article.

For this reason, critical appraisal should consider relevant information across reports of the same study rather than treating each article as an isolated methodological object.

Study-Level Organization Makes the Review Easier to Manage

Once you recognize multiple reports, create a study-level identifier.

For example:

Study ID: Santos 2024 AI Feedback Trial

Under that study record, you might link:

  • Santos et al. 2023 protocol;
  • Santos et al. 2024 primary outcomes;
  • Reyes et al. 2024 qualitative process evaluation; and
  • Santos et al. 2025 twelve-month follow-up.

The exact naming convention does not matter as much as maintaining the relationship. Your screening, extraction, appraisal, and synthesis records should make clear that these publications originate from one research project.

For large reviews, systematic review software may support study-level grouping. A spreadsheet can also work if it contains separate identifiers for reports and underlying studies.

PRISMA Distinguishes Studies From Reports for a Reason

PRISMA 2020 explicitly distinguishes the number of studies included from the number of reports describing those studies.

The flow diagram can therefore show, for example, that 42 studies were included but those studies were represented by 57 reports.

That is not a contradiction. It communicates that some studies generated more than one relevant report.

This distinction becomes particularly important when preparing a PRISMA flow diagram for study selection. Counting every report as a separate study would misrepresent the evidence base.

Do Not Confuse Multiple Reports With Multiple Independent Studies in One Paper

The reverse situation can also occur: one paper may report more than one independent study.

An article might contain “Study 1” and “Study 2,” each with different participants and procedures. A paper might report two separate experiments. A multi-cohort article may include independent samples that satisfy your eligibility criteria separately.

In such cases, the publication count is smaller than the study count.

This is another reason paper-level counting is unreliable. The relationship between reports and studies can be one-to-one, many-to-one, or occasionally one-to-many.

The correct unit depends on what was actually done, not on how many PDF files you downloaded.

04 · A Practical Example

Four Papers, but Only One Underlying Study

Hypothetical Example

A University Trial Produces Several Publications

Suppose you are reviewing studies of generative AI-generated feedback among university students. Your search identifies four apparently eligible papers from the same research group.

Paper What it reports Key clues
Paper A Protocol for an AI-feedback trial Same trial registration number and planned sample
Paper B Primary learning and student-experience outcomes Same institutions, intervention, recruitment period, and registration number
Paper C Qualitative interviews with a subset of participants Participants drawn explicitly from the trial sample
Paper D Twelve-month follow-up Same original cohort with fewer participants remaining at follow-up
Recognize the relationship The registration number, institutions, intervention, dates, and participant descriptions establish that the four publications relate to the same underlying research project.
Create one study-level record All four papers are linked under a single study identifier rather than entered as four independent studies.
Use each report where it contributes The protocol helps clarify planned methods, the primary article provides the main outcomes, the qualitative paper provides participant experiences, and the follow-up provides long-term results.
Avoid double-counting If Papers B and D report the same outcome at different time points, the review follows its prespecified time-point rules rather than treating them as independent samples.
Report transparently The review can identify one included study represented by four relevant reports.

Now suppose Paper C uses only 25 participants from the original sample but reports qualitative experiences not available anywhere else. It should not be discarded merely because the participants already appear in the larger trial. Its qualitative findings may provide genuinely different evidence. The important point is to recognize the relationship and handle the resulting dependence and study structure appropriately.

05 · What Researchers Often Get Wrong

Common Mistakes When the Same Study Produces Several Papers

Misconception

Every Eligible Paper Counts as a Separate Study

No. A single study may produce multiple publications. Counting every report as an independent study can exaggerate the size of the evidence base and may double-count participants or outcomes.

Misconception

You Should Keep the Main Paper and Delete the Others

Secondary reports may contain important methodological details, additional outcomes, subgroup analyses, or follow-up data. Link the reports and use their complementary information rather than automatically discarding them.

Misconception

Different Sample Sizes Mean the Papers Must Describe Different Studies

Not necessarily. Attrition, subgroup analyses, missing data, different analytic populations, and follow-up periods can produce different sample sizes across reports from the same underlying study.

Misconception

Different Authors Mean the Studies Are Independent

Authorship can change across reports from one project. Compare identifiers, sites, recruitment periods, interventions, participant characteristics, and other study details rather than relying on author lists alone.

Misconception

Using the Same Dataset Is Never a Problem if the Outcomes Are Different

Different outcomes may legitimately be reported from the same participants, but the relationship still matters. Depending on the synthesis, estimates derived from the same participants may be statistically dependent, and the dataset should not be represented as several independent samples.

Misconception

The Newest Paper Automatically Replaces All Earlier Reports

A later publication may provide updated results without containing every methodological detail or outcome reported previously. Review all relevant reports and determine what each contributes rather than assuming publication date creates a simple replacement hierarchy.

06 · What This Means for You

Build Your Review Around Studies, Then Attach the Papers to Them

When two or more papers look suspiciously similar, do not immediately decide that one is a duplicate. First determine the relationship among them.

A simple decision framework

If two records are copies of the exact same publication
Deduplicate the records so the same report is screened only once.
If different publications clearly describe the same underlying study
Link them under one study-level record and retain the reports that provide useful information.
If publications use the same dataset but analyze different questions or subsets
Record the shared data source and determine whether their contributions to your synthesis are independent or dependent.
If two papers appear related but the relationship is uncertain
Compare identifiers, authors, institutions, recruitment dates, participant characteristics, interventions, and sample details; seek clarification from investigators if necessary.
If multiple reports contain conflicting information
Investigate the reason, use prespecified decision rules where available, document how the discrepancy was resolved, and avoid choosing results merely because they are favorable.

A practical data-management structure is to assign both a study ID and a report ID. The study ID groups all material from the same underlying research, while each report ID identifies the individual publication or source.

For example:

Study ID: AI-FEEDBACK-01

Reports: AI-FEEDBACK-01-A, AI-FEEDBACK-01-B, AI-FEEDBACK-01-C

This simple distinction can prevent a surprisingly large amount of confusion once screening, extraction, critical appraisal, and synthesis begin.

It also makes later documentation easier because inclusion and exclusion decisions ultimately need to be tracked at the appropriate study level, even though searches initially identify records and reports.

07 · A Quick Checklist

Before Treating Two Papers as Independent Studies

Compare the reports for:
Matching trial registration numbers, project identifiers, cohort names, or other unique study identifiers.
Overlapping authors or research groups, while remembering that authorship can change between publications.
The same institutions, recruitment sites, countries, or settings.
Matching or overlapping recruitment dates and study periods.
Similar participant numbers, demographic characteristics, baseline values, or distinctive sample descriptions.
Identical or highly similar interventions, exposures, procedures, or comparison groups.
Statements that participants, data, or analyses came from a previously reported trial, cohort, survey, registry, or project.
Whether counting both papers independently would cause the same participants, outcomes, or dataset to contribute more than once to the same synthesis.
08 · Frequently Asked Questions

Questions About Multiple Papers From the Same Study or Dataset

Can one study have more than one paper?

Yes. A single study may generate protocols, conference abstracts, primary results papers, secondary analyses, subgroup reports, qualitative companion studies, and follow-up publications. Systematic reviews therefore need to distinguish the underlying study from the reports describing it.

Should I include both papers if they come from the same study?

You may need information from both, but do not count them automatically as two independent studies. Link the reports to the same underlying study and use each source for the information it contributes while preventing duplicate use of the same evidence.

What if two papers use the same dataset but investigate different outcomes?

Both may be relevant, particularly if they provide different outcomes required by your review. Record that they share the same dataset and consider whether their contributions are statistically dependent. The same participants should not be represented as independent samples merely because the outcomes were published separately.

What if two papers use different subsets of the same dataset?

Treat the relationship explicitly. Determine how much the samples overlap, which participants contribute to each analysis, and whether both estimates can enter the same synthesis without violating assumptions of independence. The appropriate solution depends on your analysis and review design.

How can I tell whether two papers report the same study?

Compare study identifiers, authors, institutions, recruitment dates, intervention details, sample sizes, baseline characteristics, and study duration. No single characteristic is always decisive. If substantial uncertainty remains and the distinction matters, contacting the investigators may be appropriate.

Which paper should I cite as the main study?

Choose the report that most appropriately represents the study for the purpose of your review, often the main or most complete results publication, and justify that choice when necessary. Continue using secondary reports for relevant supplementary information rather than discarding them.

What if different papers from the same study report conflicting numbers?

Investigate whether the difference reflects follow-up timing, analytic populations, corrections, adjusted analyses, subgroups, or another legitimate explanation. Consult protocols, registrations, supplementary materials, and other reports where useful. If the discrepancy remains unresolved, document the decision rule used to select the information and consider contacting the study authors.

Does PRISMA count papers or studies?

PRISMA 2020 distinguishes records, reports, and studies. A systematic review may therefore include more reports than studies because several reports can correspond to one underlying study. The study-selection reporting should preserve that distinction.

09 · The Bottom Line

Count the Research Once, but Use All the Relevant Reports

The Bottom Line

When several papers come from the same underlying study or dataset, link them together and treat the study as the primary unit rather than counting every publication as independent evidence.

Do not simply discard secondary papers, because they may contain methods, outcomes, follow-up data, or other information missing from the main report. Compare identifiers and study characteristics carefully, document overlapping samples and conflicting information, and make sure the same participants or data do not acquire extra influence merely because researchers published them more than once.

10 · Sources and Further Reading

Authoritative Guidance on Multiple Reports From the Same Study

11 · Cite this Guide

How to Cite This Guide

This guide is intended to be read, shared, and used in research, teaching, and academic work. If you draw on its ideas, explanations, or other content, please acknowledge the source by citing the guide. Doing so gives appropriate credit and helps your readers locate the original resource.

Has the Field Guide helped your research?

If a guide helped clarify a question, inform a research decision, or move your work forward, I would love to hear about your experience. Your story may also help other researchers discover the Field Guide.

Share Your Experience
Takes only a few minutes