03 · What You Need to Know
Ethical Limits Can Create Genuine Limits on Empirical Knowledge
An important question and an ethically permissible study are not the same thing
Research questions are usually evaluated for significance, clarity, originality, answerability, and feasibility. Ethical feasibility adds another constraint: even if a study could technically produce the desired evidence, researchers must still be justified in conducting it.
The Belmont Report makes this tension explicit in its discussion of beneficence. Research can require exposure to some risk because learning what causes harm or benefit sometimes cannot occur without uncertainty and risk. At the same time, investigators are obligated to maximize possible benefits and minimize possible harms, and there are circumstances in which anticipated benefits do not justify the risks involved.
That means empirical possibility does not establish ethical permissibility.
Worth asking
The question addresses a meaningful uncertainty whose answer could contribute important scientific, theoretical, clinical, educational, policy, or social understanding.
Directly researchable
A study capable of producing sufficiently direct evidence can be conducted within the ethical, legal, scientific, and practical constraints that apply.
A question can satisfy the first condition without satisfying the second.
The clearest examples involve harmful exposures researchers should not create
Suppose researchers want to know the long-term causal effect of severe childhood neglect on cognitive development.
An experiment that randomly assigned children to experience severe neglect or adequate care would provide an extraordinarily direct test of the causal question. It would also be ethically indefensible. Researchers cannot deliberately create serious neglect merely because randomization would strengthen causal inference.
The causal question nevertheless matters. Understanding the consequences of childhood neglect can inform prevention, treatment, social services, and policy.
Researchers therefore study naturally occurring exposure using observational cohorts, administrative data, natural experiments, longitudinal designs, sibling comparisons, and other ethically defensible approaches. None should be described as equivalent to deliberately randomized harmful exposure. Instead, researchers build an evidentiary case while acknowledging the limitations of each design.
The inability to perform the hypothetical ideal experiment does not make the causal question meaningless. It determines which forms of evidence can responsibly be used to investigate it.
Ethics can rule out withholding as well as exposure
The problem can also arise when the decisive study would require withholding something known or strongly expected to be beneficial.
Once an effective intervention has become established, a research design that deliberately denies it to participants may become ethically problematic depending on the clinical or research context. In medical research, the Declaration of Helsinki contains specific requirements governing the use of placebo or no intervention when a proven intervention exists.
This illustrates a broader point. Ethical research does not operate in a vacuum. What is permissible can change as knowledge changes.
A comparison that was ethically defensible when genuine uncertainty existed may no longer be defensible after sufficiently strong evidence accumulates. The scientific question might remain interesting, but the range of permissible methods for investigating it can narrow.
Some direct evidence would require unacceptable privacy intrusion
Ethical limits do not arise only from physical or psychological interventions.
A researcher might want to know how an extremely sensitive characteristic affects employment decisions, political behavior, healthcare use, or another outcome. The most direct study might require linking highly identifiable records across institutions or observing private behavior without participants' knowledge.
If obtaining the necessary information would violate applicable privacy protections, confidentiality commitments, consent conditions, or legal requirements, researchers cannot simply declare the question important enough to justify access.
In some cases, privacy and confidentiality protections fundamentally limit the evidence available. The question may remain meaningful while the preferred dataset or linkage remains unavailable.
Scientific value does not provide unlimited permission to create risk
The value of a research question matters ethically. Asking people to accept research burdens is harder to justify when a study is trivial, redundant, poorly designed, or foreseeably unable to produce useful knowledge.
Under the U.S. Common Rule, for example, risks to participants in covered research must be reasonable in relation to anticipated benefits, if any, and the importance of the knowledge that may reasonably be expected to result. The Belmont Report likewise places risk-benefit assessment within the application of beneficence.
But this relationship should not be reversed into a claim that sufficiently important knowledge justifies any procedure.
The Belmont Report describes the ethical problem as determining when it is justifiable to seek benefits despite research risks and when the benefits should instead be foregone because of those risks. In other words, there are circumstances in which researchers must accept that a desirable form of knowledge cannot be pursued through the proposed means.
The 2024 Declaration of Helsinki makes the boundary particularly clear for medical research involving human participants: the purposes of generating knowledge and improving health must not take precedence over the rights and interests of individual research participants.
Do not confuse "cannot answer directly" with "cannot investigate at all"
The loss of the ideal study does not necessarily leave researchers empty-handed.
Many important questions are investigated through evidence that is less direct than the hypothetical perfect experiment. Depending on the question, researchers may use observational studies, longitudinal cohorts, natural experiments, quasi-experimental designs, existing records, historical evidence, qualitative research, modelling, animal or laboratory research where appropriate, or combinations of evidence from different methods.
The relevant alternative depends on what the question asks. There is no universal substitute for an unethical experiment.
The objective is to determine what parts of the question remain empirically accessible and what inferential assumptions each alternative requires.
Natural experiments can sometimes provide unusually informative alternatives
Researchers sometimes encounter situations in which an exposure or intervention occurs because of circumstances outside their control. Policy changes, administrative thresholds, environmental events, implementation schedules, or institutional differences may create variation that can be studied without researchers deliberately producing the ethically problematic condition.
When the assumptions of the design are credible, such natural or quasi-experimental situations can sometimes support stronger causal inference than ordinary observational comparisons.
But they do not become randomized experiments merely because researchers did not assign the exposure. Selection processes, confounding, concurrent events, measurement problems, and other threats to inference still require careful analysis.
Researchers should therefore ask what the natural variation actually identifies rather than treating "natural experiment" as an ethical and methodological magic wand.
Triangulation can strengthen an answer without manufacturing certainty
When no single ethical design can provide decisive evidence, several different forms of evidence may collectively become informative.
Suppose observational cohorts show an association, a natural experiment produces a similar pattern, qualitative research supports a plausible mechanism, and results are consistent across different populations. The convergence may strengthen the interpretation even though none of those studies independently reproduces the hypothetical unethical experiment.
This is one reason cumulative research matters. Scientific understanding often develops from multiple imperfect studies rather than one definitive design.
Yet convergence should not be exaggerated. Several studies sharing the same bias do not necessarily correct one another. Triangulation is most informative when different methods have meaningfully different assumptions, weaknesses, or sources of error.
Sometimes the ethical alternative answers a different question
This is where researchers need particular discipline.
Suppose the original question is:
Does prolonged exposure to severe social isolation cause depression in adolescents?
Because deliberately assigning adolescents to prolonged severe isolation would be ethically unacceptable, researchers instead conduct a cross-sectional survey comparing reported isolation with depressive symptoms.
The survey can provide evidence about an association. It does not automatically answer the original causal question. Depression might contribute to social isolation, unmeasured factors might influence both, or measurement error might affect the observed relationship.
The ethical alternative has therefore changed the evidentiary basis.
You can respond in two defensible ways. You can reformulate the immediate study question around association, or you can retain the broader causal question as the scientific problem while describing the survey as one source of indirect evidence relevant to it.
What you should not do is keep the causal wording and behave as though the ethical constraint disappeared when the method changed.
Not every hypothetical experiment is actually the "gold standard"
There is another subtle problem. Researchers sometimes imagine an unethical randomized experiment and then assume it would provide perfect knowledge if only ethics would get out of the way.
That is often unrealistic.
The artificial exposure might not resemble the real-world phenomenon. Participants willing to enter the study might differ substantially from the population of interest. Outcomes measured over a short experimental period might not represent long-term consequences. Compliance, attrition, measurement error, and limited external validity could remain.
Ethics may prohibit a study without that prohibited study being epistemically perfect.
This matters because researchers should not romanticize the forbidden experiment. The appropriate comparison is between realistic designs, including their actual strengths and weaknesses, rather than between an ethical observational study and an imaginary experiment with flawless validity.
An unanswered question is different from an unanswerable question
These terms should also be distinguished.
Currently unanswered
The evidence needed to resolve the question has not yet been produced, but an ethically and practically plausible route may exist.
Not directly answerable through an ethical study
The decisive or sufficiently direct evidence would require procedures that researchers cannot ethically conduct under the relevant conditions.
Even the second category should be used cautiously. Methods, data availability, natural events, regulations, and scientific knowledge can change. A question that cannot be investigated directly today may become approachable through a new ethical method tomorrow.
It is therefore often more accurate to specify the limitation: no currently available ethically acceptable design can directly establish the particular claim.
Indirect evidence should not be downgraded simply because the perfect experiment is impossible
Researchers can make the opposite mistake too. Once they realize that the idealized direct experiment cannot be conducted, they may conclude that nothing useful can ever be learned.
That standard would discard enormous areas of legitimate scholarship.
Questions about harmful exposures, historical events, social structures, long-term environmental conditions, rare disasters, and many population-level phenomena frequently cannot be studied through direct experimental manipulation. Researchers nevertheless develop credible knowledge using designs appropriate to those phenomena.
The relevant question is not whether the evidence reproduces an impossible experiment. It is how strongly the available evidence supports particular interpretations given the assumptions and limitations of the designs used.
Sometimes the correct conclusion really is "we do not know"
Research culture can make uncertainty uncomfortable. A paper is expected to end with findings. A dissertation is expected to answer something. A grant proposal is expected to promise progress.
But ethical limits occasionally leave important uncertainty unresolved.
If every ethical source of evidence is too indirect, too confounded, too incomplete, or otherwise insufficient to distinguish competing explanations, then the responsible conclusion may be that the question remains unanswered.
That is preferable to converting an association into a causal conclusion, treating inaccessible data as though they were unnecessary, or lowering ethical protections until a cleaner analysis becomes possible.
Uncertainty is not a methodological embarrassment when the evidence genuinely warrants uncertainty. Peer reviewers have survived seeing the phrase "cannot determine" before.
Do not expose participants to risk for a study that cannot meaningfully answer its question
The fact that direct evidence is unavailable does not justify conducting any indirect study merely because some evidence is better than none.
The alternative study still needs sufficient scientific value and methodological adequacy.
OHRP's Secretary's Advisory Committee on Human Research Protections has emphasized that foreseeably uninformative research can waste resources and devalue participants' contributions. Its recommendations connect scientific validity with the ethical justification for exposing people to research risk.
If an ethical substitute is so weak that it cannot meaningfully inform the question, asking participants to accept burdens or risks may itself become difficult to justify.
The goal is therefore not simply to find an ethical method. It is to find an ethical method capable of producing worthwhile evidence.
The research question may need to become more modest
Sometimes the best response is to narrow the claim rather than abandon the topic.
A causal question may become an associational question. A question about individual effects may become one about population patterns. A question about a hidden behavior may become one about self-reported behavior. A question about long-term harm may become one about short-term indicators.
Each change loses something. That loss should be visible in the wording.
If the strongest method would be unethical, redesigning the method is usually the first response. When every acceptable redesign changes the substantive target, however, the ethical constraint may need to change the question itself.
Watch Out
Do not use the phrase "this question cannot be studied ethically" simply because your preferred design was rejected or inconvenient. First determine whether alternative methods, existing evidence, natural variation, different data, or a narrower formulation can investigate the underlying problem responsibly.