Evidence Literacy
How to read an evidence grade
Every grade on this site is shorthand for the quality of the human evidence behind a claim. The Study Design Snapshot keeps the practical takeaway prominent and tucks the “why” — trial design and limitations — into an optional, expandable panel. The same component is embeddable inside our structured education content. For the full methodology guide, see Evidence Hierarchy.
What each grade looks like
Study design snapshot
Evidence grade: StrongWhen several large, well-run human trials agree, a claim earns the strongest grade — you can act on it with reasonable confidence.
Why this grade — design & limitations
Why this grade
Multiple large randomized controlled trials and meta-analyses converge on the same direction and rough size of effect.
Pivotal study design
- Study type
- Multiple RCTs + meta-analysis
- Population
- Large, varied adult samples
- Participants
- Hundreds to thousands pooled
- Duration
- Weeks to months
- Comparator
- Placebo and/or active control
Limitations & context
- Even strong evidence describes averages, not individual response.
- Effect sizes can still be modest in practical terms.
A strong grade means the effect is real and replicated — not that it will be large for everyone.
Study design snapshot
Evidence grade: ModerateA handful of small human trials point the same way, but the evidence is thinner — promising, not settled.
Why this grade — design & limitations
Why this grade
A few randomized trials show a consistent signal, but small samples, short durations, or funding concerns limit confidence.
Pivotal study design
- Study type
- Several small RCTs
- Population
- Modest adult samples
- Participants
- Dozens per trial
- Duration
- 4–8 weeks
- Comparator
- Placebo
Limitations & context
- Small samples widen the uncertainty around the true effect.
- Short trials cannot speak to long-term use.
- Some trials are industry-funded.
Study design snapshot
Evidence grade: PreliminaryThe rationale is mostly mechanistic or traditional. Treat it as a hypothesis worth watching, not a recommendation.
Why this grade — design & limitations
Why this grade
Support comes largely from lab, animal, or traditional-use evidence with little or no controlled human data.
Pivotal study design
- Study type
- Mechanistic / animal / traditional use
- Population
- Pre-clinical or anecdotal
Limitations & context
- Mechanistic plausibility frequently fails to translate to human benefit.
- No controlled human trials means effect and safety are unestablished.
A low grade is not a verdict that something does not work — it means the human evidence is not there yet.
Why design factors matter
Blinding, a placebo comparator, sample size, and duration are what separate a persuasive trial from a misleading one. The snapshot surfaces these so a grade is never a black box — and so you can judge the evidence for yourself.
References
- [1] Concato J, et al. (2000). RCTs vs observational studies. N Engl J Med, 342(25): 1887-1892. PubMed →
Learning context
How this concept connects to supplement decisions
How evidence grades are assigned and which clinical trial design factors matter — shown through embeddable Study Design Snapshots that keep the practical takeaway prominent. Learning pages explain the reasoning layer behind the herb and compound library. They are designed to make mechanisms, evidence quality, safety tradeoffs, and product claims easier to interpret.
Use How to read an evidence grade to build better questions before choosing a supplement: what outcome is being targeted, what mechanism is claimed, what human evidence exists, what dose was studied, and what risks could change the answer for a specific person?
Mechanistic plausibility is useful, but it should be weighed against trial design, safety history, product quality, and the possibility that a simpler intervention may be more appropriate.