Skip to main content
ExplainerClinical TrialsEvidence Explainer· 5 min read· in Science

Phase I, II, III, and IV: The Four Metrics That Define Drug Safety, Efficacy, and Approval

A drug earns regulatory approval by passing four escalating statistical tests, scaling from dozens of healthy volunteers to thousands of patients. This sequential expansion is a mathematical necessity designed to detect rare adverse effects before and after a medication reaches the global market.

By Nicolas Laurent

Medical Regulators 40%Clinical Researchers 35%Patient Advocates 25%
Medical Regulators
Prioritize absolute safety and statistical certainty before allowing widespread public exposure to a new molecule.
Clinical Researchers
Focus on rigorous methodology, double-blind architectures, and accurate data collection to prove biological efficacy.
Patient Advocates
Argue that the decade-long timeline of traditional clinical trials is too slow for patients facing terminal or rapidly progressing diseases.

Perspectives this story doesn't cover

  • Pharmaceutical Executives
  • Health Insurance Providers
20–100
Phase I healthy volunteers
300–3,000
Phase III patient cohort
33%
Phase II success rate
25–30%
Phase III success rate

Fast facts

  • Phase I trials test a drug's basic safety and dosage limits in 20 to 100 healthy volunteers.
  • Phase II evaluates preliminary efficacy in several hundred patients, with only 33% of drugs passing this stage.
  • Phase III confirms the drug's benefit against a placebo in a large cohort of 300 to 3,000 patients.
  • Phase IV involves post-market surveillance to detect rare, long-term side effects in the general population.

How we got here

  1. Phase I

    Researchers test the drug in 20 to 100 healthy volunteers to determine the maximum tolerated dose and basic safety.

  2. Phase II

    The drug is administered to several hundred patients with the target disease to look for preliminary efficacy signals.

  3. Phase III

    A large-scale, randomized, double-blind trial of 300 to 3,000 patients confirms efficacy against a placebo or standard care.

  4. Phase IV

    Following regulatory approval, ongoing post-market surveillance monitors the general population for rare, long-term side effects.

A drug earns regulatory approval by passing four escalating statistical tests: Phase I establishes basic safety in dozens of people, Phase II tests efficacy in hundreds, Phase III confirms the benefit against a placebo in thousands, and Phase IV monitors for rare risks in the general population indefinitely. This sequential expansion is not merely a bureaucratic hurdle; it is a mathematical necessity driven by the limits of sample size. Before a chemical compound can be prescribed at a pharmacy, researchers must prove that its therapeutic benefits outweigh its physiological risks, a calculation that requires exposing progressively larger populations to the molecule.[7]

The architecture of clinical research is designed to minimize human risk while maximizing data collection. According to the World Health Organization's 2020 framework, "Clinical trials are a type of research that studies new tests and treatments and evaluates their effects on human health outcomes." This evaluation begins in the laboratory, but in vitro and animal models can only predict human responses up to a point.[1]

The transition into human testing begins with Phase I, which is entirely focused on safety and pharmacokinetics—how the body absorbs, metabolizes, and excretes the drug. The U.S. Food and Drug Administration notes in its 2018 guidance that these initial trials typically enroll 20 to 100 healthy volunteers or people with the disease. The primary metric here is not whether the drug cures the condition, but identifying the maximum tolerated dose before unacceptable toxicity occurs.[3]

Because the sample size in Phase I is so small, the statistical power is limited to detecting only the most common and acute side effects. The FDA reports that approximately 70 percent of experimental drugs pass this initial safety threshold. Those that fail usually do so because they cause severe liver toxicity, cardiovascular disruption, or other immediate adverse events that were not apparent in animal models.[3]

The clinical trial pipeline requires exponentially larger patient cohorts at each phase to prove safety and efficacy.

Drugs that clear Phase I advance to Phase II, where the metric shifts from pure safety to preliminary efficacy. This phase expands the cohort to several hundred patients who actually have the disease or condition the drug is intended to treat. Here, researchers are looking for a biological signal that the intervention works, while continuing to monitor short-term side effects.[2][3]

Phase II is historically the most lethal stage for experimental medications. The FDA estimates that only 33 percent of drugs successfully navigate this phase. The high attrition rate occurs because a molecule that is perfectly safe in healthy volunteers may prove entirely ineffective at altering the disease pathway in actual patients. This phase establishes the optimal dosage that balances efficacy with tolerability, setting the parameters for the definitive tests to follow.[3]

Phase II is historically the most lethal stage for experimental medications.

The definitive test is Phase III, the large-scale randomized, double-blind, placebo-controlled trial. Enrollment scales up massively, typically requiring 300 to 3,000 patients across multiple clinical centers and often multiple countries. The Danish Medicines Agency emphasizes that Phase III trials are designed to confirm the drug's efficacy, monitor side effects, and compare it to commonly used treatments.[3][6]

In a Phase III trial, neither the patients nor the administering physicians know who is receiving the experimental drug and who is receiving a placebo or standard-of-care treatment. This double-blind architecture eliminates observation bias. The primary endpoint is a predefined clinical outcome—such as tumor shrinkage, reduced blood pressure, or symptom resolution—measured against the control group with strict statistical significance.[1][2]

The financial and logistical burden of Phase III is immense, often costing hundreds of millions of dollars and spanning one to four years. Despite the extensive preparation, the FDA notes that only 25 to 30 percent of drugs that enter Phase III successfully complete it and prove their intended benefit. Success at this stage provides the comprehensive data package required for regulatory agencies to grant market approval.[3]

Detecting a 1-in-10,000 adverse event requires the massive scale of Phase IV post-market surveillance.

However, the mathematical reality of sample sizes means that even a successful 3,000-person Phase III trial cannot guarantee absolute safety. If a severe adverse reaction occurs in one out of every 10,000 patients, a 3,000-patient trial is statistically unlikely to detect it. This limitation necessitates Phase IV, also known as post-market surveillance.[4][7]

Phase IV begins after the drug is approved and available to the general public. As the patient population scales from thousands to millions, the statistical power to detect rare, long-term side effects reaches its maximum. The National Center for Biotechnology Information outlines that these observational studies track the drug's performance in real-world settings, outside the tightly controlled environment of a clinical trial.[4]

During Phase IV, researchers can identify drug interactions with other medications, long-term risks, and efficacy in specific subpopulations, such as pregnant women or patients with complex comorbidities. The MD Anderson Cancer Center highlights that this phase is crucial for understanding how a treatment affects quality of life over an extended period. If severe, previously undetected side effects emerge during this phase, regulatory agencies can update the drug's warning labels or, in extreme cases, withdraw the medication from the market entirely.[5]

The four-phase system represents a deliberate trade-off between speed and certainty. While patient advocacy groups often push for faster access to experimental treatments, the sequential scaling of clinical trials ensures that the public is not exposed to widespread harm. The metrics that define each phase—safety, efficacy, confirmation, and surveillance—form the foundation of modern evidence-based medicine, ensuring that a drug's true profile is revealed through the uncompromising lens of statistical power.[7]

What we don’t know

  • How artificial intelligence and in silico modeling might eventually reduce the number of human participants required in early-phase trials.
  • The exact long-term effects of many recently approved accelerated-pathway drugs, which rely heavily on Phase IV surveillance to confirm their initial promise.

Sources

Source coverage

7 outlets

3 viewpoints surfaced

Medical Regulators 40%Clinical Researchers 35%Patient Advocates 25%
  1. [1]World Health Organization (WHO)Medical Regulators

    Clinical trials

    Read on World Health Organization (WHO)
  2. [2]National Institutes of Health (NIH)Clinical Researchers

    The Basics

    Read on National Institutes of Health (NIH)
  3. [3]FDAMedical Regulators

    Step 3: Clinical Research

    Read on FDA
  4. [4]NCBI BookshelfClinical Researchers

    Drug Trials - StatPearls - NCBI Bookshelf

    Read on NCBI Bookshelf
  5. [5]MD Anderson Cancer CenterClinical Researchers

    Phases of Clinical Trials

    Read on MD Anderson Cancer Center
  6. [6]Danish Medicines AgencyMedical Regulators

    The clinical trial phases

    Read on Danish Medicines Agency
  7. [7]Factlen Editorial TeamPatient Advocates

    Synthesis by Factlen editorial team

    Read on Factlen Editorial Team

Comments

Stay informed

Every angle. Every day.

Get Science stories with full source coverage and perspective breakdowns delivered to your inbox.