Manuel B. Garcia

Manuel B. Garcia serves as the Senior Director for Educational Technology and Digital Learning at FEU Institute of Technology, Manila, Philippines. Read More

Contact Info

1607, FEU Tech Building,
P. Paredes St, Sampaloc,
Manila, Philippines
mbgarcia@feutech.edu.ph

Follow Me

Can You Include Pilot Study Participants or Data in the Main Study?

Pilot participants or data can sometimes contribute to the main study, particularly when the pilot was prospectively designed as an internal pilot. Pooling is more difficult when the pilot was separate or when substantial changes make its data incompatible with the definitive study.

174
Including Pilot Data in the Main Study Guide 174 of 217
01 · The Question

Can the People and Data From Your Pilot Count Toward the Main Study?

You complete a pilot and discover something encouraging: the procedures work reasonably well, and you already have usable data from participants who resemble those you intend to recruit for the main study.

Discarding those observations may seem wasteful. Why not simply add the pilot participants to the main sample?

Sometimes that is entirely defensible. In other situations, combining pilot and main-study data can create methodological problems. The answer depends partly on whether the pilot was designed from the outset to form part of the definitive study, whether participants were recruited under compatible protocols, what researchers learned from the pilot data, and how much the study changed afterward.

The crucial distinction is therefore not merely whether the pilot data look usable. It is whether incorporating them preserves the scientific and ethical integrity of the main study.

02 · The Short Answer

Sometimes, but the Decision Should Ideally Be Planned in Advance

In Brief

Pilot participants or data can sometimes be included in the main study, especially when the pilot was prospectively designed as an internal pilot and the relevant protocol, eligibility criteria, interventions, outcomes, data collection, and analysis remain sufficiently compatible.

Data from a separate external pilot should not simply be pooled into the definitive study because they happen to be available. Consider what changed after piloting, whether pilot outcomes influenced design or analysis decisions, whether participants were enrolled under appropriate approvals and consent, and whether combining the data would preserve a valid interpretation.

03 · What You Need to Know

The Answer Depends on How the Pilot Relates to the Main Study

Start With the Difference Between an Internal and External Pilot

The clearest starting point is whether the pilot is internal or external to the definitive study.

An external pilot is conducted separately before the main study. Its purpose is to investigate feasibility and refine the future research. Pilot data ordinarily remain separate from the definitive study, giving researchers greater freedom to modify procedures before the main study begins.

An internal pilot, by contrast, is designed as an initial phase of the definitive study. Participants recruited during that phase may contribute to the final analysis, provided the study proceeds and the conditions for retaining those data are satisfied.

External pilot A separate preliminary study conducted before the definitive study; its data are ordinarily analyzed separately from the main-study data.
Internal pilot A planned initial phase of the definitive study in which eligible participants and their data may contribute to the final analysis if progression occurs appropriately.

This distinction should ideally be established prospectively rather than invented after researchers see that the pilot data would be convenient to keep.

Why Internal Pilots Can Retain Participants

An internal pilot can reduce the inefficiency of recruiting participants solely for feasibility assessment and then starting the definitive sample from zero. The initial participants form part of the planned study, while early data provide an opportunity to examine specified feasibility parameters or design assumptions.

For example, an internal pilot within a randomized controlled trial might assess recruitment, retention, or other prespecified parameters during an initial phase. If progression criteria are met and no incompatible changes are required, recruitment continues and the internal-pilot participants remain part of the final trial dataset.

This approach requires careful planning because the feasibility assessment and definitive analysis are not completely independent. The protocol should explain the internal-pilot design, progression criteria, possible adaptations, and treatment of the initial participants.

External Pilot Data Are Different

An external pilot is intentionally separated from the definitive study so researchers can learn from it and make changes without necessarily preserving comparability with the later research.

The pilot may use an earlier version of an intervention, different questionnaire wording, different eligibility criteria, an incomplete follow-up procedure, a different recruitment strategy, or a data system that is subsequently redesigned.

Those differences may be exactly why the pilot was useful. They can also make its observations inappropriate for direct pooling with the main study.

If you originally designed the pilot as external, do not assume afterward that scientifically usable data should automatically become definitive data.

Ask What Changed After the Pilot

The most important practical question is often simple: what did you change?

Minor administrative modifications may have little effect on comparability. Correcting typographical errors, clarifying staff instructions without altering the participant procedure, or fixing a data-entry interface may not necessarily change what was measured or how participants were treated.

Other changes are more consequential. Researchers might modify the intervention, eligibility criteria, outcome measure, follow-up schedule, randomization process, recruitment population, assessment mode, or analytical strategy.

The more a change affects what participants experienced, what was measured, when it was measured, or who could enter the study, the harder it becomes to treat pilot and main-study observations as though they arose under one unchanged protocol.

If piloting leads to substantial modifications, consider whether the revised protocol remains the same study in a meaningful methodological sense.

Outcome Changes Require Particular Attention

Suppose the pilot shows that the planned primary outcome is difficult to collect, so the researchers replace it before the main study. Pilot participants may not have data on the new primary outcome at all.

Even subtler changes can matter. Revising questionnaire items, changing measurement timing, switching administration modes, or altering scoring can affect comparability.

Before pooling, ask whether the variable has the same meaning and measurement process across both phases. If not, simply placing observations in the same dataset does not make them methodologically equivalent.

Intervention Changes Can Make Pilot Participants Non-Comparable

A pilot may reveal that an intervention needs modification. Perhaps its duration changes, content is revised, staff training is strengthened, or delivery moves from face-to-face to online.

If pilot participants received a materially different intervention from main-study participants, combining their outcome data without accounting for that difference can blur the interpretation of the treatment being evaluated.

The appropriate response depends on the extent of the modification and the study design. Minor refinements may sometimes be compatible with a planned adaptive or internal-pilot framework. Major intervention changes may make pooling inappropriate.

Eligibility and Recruitment Changes Can Alter the Study Population

Suppose pilot recruitment is poor, so eligibility criteria are broadened before the main study. The pilot participants remain eligible under the new criteria, but the population from which later participants are recruited has changed.

That does not automatically prohibit pooling, but it requires methodological consideration. Changes to inclusion or exclusion criteria can alter the target population, event rates, baseline characteristics, intervention response, or generalizability of the study.

Similarly, adding new sites or recruitment channels may introduce differences that need to be understood rather than ignored.

Be Careful When Pilot Outcomes Influenced Main-Study Decisions

A particularly important issue arises when researchers inspect substantive pilot outcomes before deciding how to design or analyze the definitive study.

Suppose you examine which outcome favors the intervention and then designate that outcome as primary for the main study. Or you inspect subgroup effects and change the analysis accordingly. If those same pilot observations are then included in the definitive analysis, the data have influenced both the question and the answer.

This can introduce bias and undermine the independence of the planned analysis.

Testing whether the planned analytical workflow operates correctly is different from examining substantive pilot results and choosing a final analysis because it produces a favorable pattern.

Consent and Ethics Approval Must Cover the Intended Use

Methodological compatibility is not the only consideration. Participant data must be used consistently with the approved protocol, consent process, applicable regulations, and institutional requirements.

If pilot participants consented to a separate preliminary study, researchers should not assume that their data can automatically be repurposed as part of a definitive study simply because the variables are similar.

The exact requirements depend on the jurisdiction, institution, study type, consent language, and ethics or regulatory framework. When reuse was not prospectively planned, researchers should consult the relevant ethics review body or institutional authority rather than infer permission from methodological convenience.

Combining Data Is Not the Same as Reusing Participants

Two related questions should be separated.

You might ask whether data already collected from pilot participants can enter the definitive analysis. Alternatively, you might ask whether people who participated in the pilot can enroll again in the main study.

Allowing the same people to participate again can introduce additional concerns. Prior exposure may make them more familiar with study procedures, questionnaires, interventions, experimental tasks, or study hypotheses. That experience could change their behavior in the main study.

Whether repeat participation is acceptable depends on the design. It should not be assumed merely because existing pilot data will remain separate.

Plan the Decision Before the Pilot Whenever Possible

If you anticipate wanting pilot participants to contribute to the definitive study, address this during design rather than after seeing the pilot results.

Specify whether the pilot is internal, what progression criteria apply, what modifications are permissible while retaining the initial data, how adaptations will be documented, and what would require treating the pilot as external.

This prospective approach is methodologically cleaner because the decision to retain data is not made opportunistically after researchers know what those data contain.

Watch Out

Do not decide to pool pilot and main-study data merely because excluding the pilot would reduce your sample size. Statistical efficiency does not override differences in protocol, outcome measurement, intervention exposure, eligibility, consent, or data-dependent design decisions.

04 · A Practical Example

When Keeping Pilot Data Is Reasonable and When It Becomes Problematic

Hypothetical Example

An Internal Pilot of a Multisite Intervention Study

A research team plans a definitive randomized study with an internal pilot involving the first participating sites. The protocol states that early recruitment and follow-up will be assessed against prespecified progression criteria.

Initial phase Participants are recruited using the definitive eligibility criteria, randomized using the intended procedure, receive the intended interventions, and complete the planned outcome measures.
Feasibility review Recruitment is slightly slower than expected, but the progression criteria support continuation with a modest extension of the recruitment period.
Minor modification The research team improves administrative scheduling but does not change eligibility, intervention content, outcome measurement, or the planned substantive analysis.
Decision Because the pilot was prospectively internal and the initial participants remain compatible with the definitive protocol, their data continue as part of the main study.
Different scenario Imagine instead that the pilot leads the researchers to replace the intervention, change the primary outcome, and substantially broaden eligibility. Those pilot observations would no longer represent the same study conditions in the same straightforward way.

The key issue is not whether the pilot dataset is technically mergeable. It is whether combining the observations preserves the meaning of the definitive study.

05 · What Researchers Often Get Wrong

Common Mistakes When Reusing Pilot Participants or Data

Misconception

Pilot Data Should Never Be Included in the Main Study

That is too absolute. Internal pilots are specifically designed so that eligible initial participants may contribute to the definitive study when progression occurs under the planned conditions.

Misconception

If the Variables Are the Same, the Data Can Be Combined

Variable names alone do not establish comparability. Consider eligibility, intervention exposure, measurement procedures, timing, recruitment conditions, protocol versions, and whether pilot results influenced subsequent design or analysis decisions.

Misconception

Minor Changes Automatically Mean Pilot Data Must Be Discarded

Not every modification destroys comparability. The question is whether the change materially alters the participants, intervention, measurement, procedures, or interpretation relevant to the definitive analysis. Planned internal pilots may permit certain modifications while retaining initial data.

Misconception

If the Pilot Sample Is Small, Including It Cannot Affect the Study Much

The methodological issue is not only numerical influence. Even a small set of observations can create concerns if those data informed outcome selection, analytical decisions, eligibility changes, or other design adaptations and are subsequently reused in the analysis they helped shape.

Misconception

Participants Can Simply Enroll Again Because the Pilot Was Separate

Previous participation can change familiarity with procedures, interventions, tasks, or study aims. Whether repeat participation is appropriate depends on the design and should be considered explicitly rather than assumed.

Misconception

If Participants Consented to Research, Their Pilot Data Can Automatically Be Used in the Main Study

Consent and ethics requirements depend on what participants were told, what was approved, and the applicable institutional and regulatory framework. Researchers should verify that the intended use is covered rather than treating general participation in research as unlimited permission for subsequent use.

06 · What This Means for You

Decide Whether the Pilot and Main Study Still Produce Comparable Evidence

Before pooling anything, reconstruct what happened between the beginning of the pilot and the final main-study protocol. Compare the two phases rather than focusing only on the datasets.

A simple decision framework

If the pilot was prospectively designed as an internal pilot
Follow the prespecified progression, adaptation, analysis, and data-retention plan when deciding whether initial participants remain in the definitive study.
If the pilot was designed as a separate external study
Do not assume retrospective pooling is appropriate merely because the observations appear compatible.
If eligibility, intervention, outcomes, timing, or important procedures changed
Assess whether the changes materially alter comparability before considering combined analysis.
If substantive pilot outcomes influenced the final hypothesis, outcome, or analytical strategy
Consider the risk of data-dependent decision-making before allowing those same observations to contribute to the definitive inference.
If consent or ethics coverage for the intended use is uncertain
Verify the applicable requirements with the responsible ethics or institutional authority before reusing the data.

If you have not yet begun the pilot and believe retaining participants may be valuable, design that possibility prospectively. It is considerably easier to justify a planned internal pilot than to convert an external pilot into one after seeing the results.

07 · A Quick Checklist

Before Including Pilot Participants or Data in the Main Study, Check

Before pooling pilot and main-study data, check:
Was the pilot prospectively designed as an internal part of the definitive study or as a separate external pilot?
Were pilot and main-study participants recruited under sufficiently compatible eligibility and enrollment procedures?
Did participants receive the same relevant intervention, exposure, comparison condition, or study procedures?
Were important outcomes measured in comparable ways and at compatible time points?
Did pilot results influence selection of the primary outcome, hypothesis, subgroup, model, or another substantive analytical decision?
Were any changes after the pilot substantial enough to alter the meaning of the study or comparability of the observations?
Does the approved protocol and consent process permit the intended use of the pilot participants or data?
If participants would enroll again, could previous exposure alter their behavior, responses, or intervention experience?
Have the decision and rationale for including or excluding pilot data been documented transparently?
08 · Frequently Asked Questions

Frequently Asked Questions About Using Pilot Data in the Main Study

What is an internal pilot study?

An internal pilot is a planned initial phase of the definitive study. Feasibility or design parameters are assessed during that phase, and participants may remain part of the final analysis if the study progresses according to the planned conditions.

What is an external pilot study?

An external pilot is conducted separately from the definitive study. It allows researchers to test feasibility and modify the future protocol without necessarily preserving the pilot observations for the final analysis.

Can I combine pilot and main-study data if nothing changed?

Possibly, but unchanged procedures alone do not settle the issue. Consider whether pooling was planned, whether the pilot and definitive study share compatible design and analysis, whether pilot outcomes influenced subsequent decisions, and whether the intended use is covered ethically and procedurally.

What if I changed only the questionnaire wording after the pilot?

It depends on what changed and whether the revision affects measurement comparability. Correcting a minor typographical problem is different from changing the meaning of an item, response options, scoring, or construct measurement. Assess the methodological consequence rather than categorizing every wording change identically.

Can pilot participants join the main study again?

Sometimes, but previous exposure may affect responses or behavior. The implications depend on the study design, intervention, tasks, outcomes, washout or learning effects where relevant, and ethics requirements. Repeat participation should be considered prospectively when possible.

Can I keep pilot data if I change the sample size of the main study?

A sample-size modification does not automatically make pilot data incompatible, particularly within a properly planned internal-pilot design. The justification depends on how the adaptation was made, what information informed it, and whether the study's statistical integrity is preserved.

What if major changes are needed after the pilot?

Major changes may make the pilot observations inappropriate for direct inclusion in the definitive analysis. The next question is whether the original design should be revised, retested, or reconsidered before the main study proceeds.

09 · The Bottom Line

Pilot Data Can Count, but Only When Their Role in the Main Study Is Defensible

The Bottom Line

You can sometimes include pilot participants or data in the main study, particularly when the pilot was prospectively designed as an internal pilot and the initial observations remain compatible with the definitive protocol and analysis.

Do not pool data simply to avoid losing participants. Examine what changed, whether pilot outcomes influenced later decisions, whether the observations still represent the same research conditions, and whether the intended use is ethically and procedurally covered. When retaining pilot data matters, planning for that possibility before the pilot begins is usually the cleaner approach.

10 · Sources and Further Reading

Sources and Further Reading

11 · Cite this Guide

How to Cite This Guide

This guide is intended to be read, shared, and used in research, teaching, and academic work. If you draw on its ideas, explanations, or other content, please acknowledge the source by citing the guide. Doing so gives appropriate credit and helps your readers locate the original resource.

Has the Field Guide helped your research?

If a guide helped clarify a question, inform a research decision, or move your work forward, I would love to hear about your experience. Your story may also help other researchers discover the Field Guide.

Share Your Experience
Takes only a few minutes