What is a phase-3 trial and why does it matter?
Phase-3 trials are the large (often 1000–5000 participant), randomised, controlled, blinded trials that establish efficacy and safety for regulatory approval. They are the highest evidence tier in clinical medicine.
Last reviewed 2026-07-13
Read the answer
Clinical trial development follows a standard four-phase structure. Phase 1 tests safety and pharmacokinetics in a small number of healthy volunteers (typically 20–100). Phase 2 tests efficacy and dose-finding in a moderate patient population (typically 100–500). Phase 3 tests efficacy and safety at scale in a large patient population (typically 1000–5000) using randomised, controlled, blinded design. Phase 4 is post-market surveillance after approval.
The phase-3 trial is what regulatory agencies typically require for approval because its design controls for the largest number of confounders. Randomisation controls for baseline differences between arms. Placebo control controls for placebo effects and natural history. Double-blinding controls for observer bias and expectation effects. Large sample size gives statistical power to detect clinically meaningful differences and to observe rare adverse events. Pre-specified endpoints prevent post-hoc analysis inflation of false positives.
When a compound page cites 'STEP-1 (Wilding 2021 NEJM)' as pivotal evidence, that reference identifies a specific phase-3 trial: sample size in the thousands, defined primary endpoint (percent weight loss at 68 weeks), randomised placebo-controlled double-blind design, published in a top-tier peer-reviewed journal. That combination is the highest standard of evidence for demonstrating that a compound produces a specific clinical effect at scale. Compounds without this evidence tier can still be scientifically interesting, but claims about clinical effect require weaker evidence to be interpreted more cautiously.
Related on Healthy Mango
Related compounds
Related research categories
References
- Once-Weekly Semaglutide in Adults with Overweight or Obesity (STEP 1)· Wilding JPH, Batterham RL, Calanna S, et al. · 2021
- Tirzepatide versus semaglutide once weekly in patients with type 2 diabetes (SURPASS-2)· Frías JP, Davies MJ, Rosenstock J, et al. · 2021
Question
Related questions
Other questions that touch the same biology, evidence, or laboratory concepts.
What do the evidence levels A, B, C, D mean?
Healthy Mango's evidence tiers describe how much human clinical evidence supports a compound. A = strong human RCTs; B = moderate human trials; C = early human data; D = preclinical only.
Why doesn't strong preclinical evidence guarantee a human clinical effect?
Mouse and human biology are similar but not identical, and rodent injury models differ from human clinical contexts in important ways. Preclinical strength predicts human effect in some biology (receptor pharmacology) and not others (tissue repair, cognition).
Why do some peptides become approved medicines while others don't?
Approval depends on biology plus patent protectability plus indication clarity plus market size — not simply on whether a compound works. Naturally-occurring sequences with broad, diffuse effects are often the least investible even when their biology is interesting.
