Evidence Literacy

How to read an evidence grade

Every grade on this site is shorthand for the quality of the human evidence behind a claim. The Study Design Snapshot keeps the practical takeaway prominent and tucks the “why” — trial design and limitations — into an optional, expandable panel. The same component is embeddable inside our structured education content. For the full methodology guide, see Evidence Hierarchy.

What each grade looks like

Study design snapshot

Evidence grade: Strong

When several large, well-run human trials agree, a claim earns the strongest grade — you can act on it with reasonable confidence.

Why this grade — design & limitations

Why this grade

Multiple large randomized controlled trials and meta-analyses converge on the same direction and rough size of effect.

Pivotal study design

Study type
Multiple RCTs + meta-analysis
Population
Large, varied adult samples
Participants
Hundreds to thousands pooled
Duration
Weeks to months
Comparator
Placebo and/or active control

Limitations & context

  • Even strong evidence describes averages, not individual response.
  • Effect sizes can still be modest in practical terms.

A strong grade means the effect is real and replicated — not that it will be large for everyone.

Study design snapshot

Evidence grade: Moderate

A handful of small human trials point the same way, but the evidence is thinner — promising, not settled.

Why this grade — design & limitations

Why this grade

A few randomized trials show a consistent signal, but small samples, short durations, or funding concerns limit confidence.

Pivotal study design

Study type
Several small RCTs
Population
Modest adult samples
Participants
Dozens per trial
Duration
4–8 weeks
Comparator
Placebo

Limitations & context

  • Small samples widen the uncertainty around the true effect.
  • Short trials cannot speak to long-term use.
  • Some trials are industry-funded.

Study design snapshot

Evidence grade: Preliminary

The rationale is mostly mechanistic or traditional. Treat it as a hypothesis worth watching, not a recommendation.

Why this grade — design & limitations

Why this grade

Support comes largely from lab, animal, or traditional-use evidence with little or no controlled human data.

Pivotal study design

Study type
Mechanistic / animal / traditional use
Population
Pre-clinical or anecdotal

Limitations & context

  • Mechanistic plausibility frequently fails to translate to human benefit.
  • No controlled human trials means effect and safety are unestablished.

A low grade is not a verdict that something does not work — it means the human evidence is not there yet.

Why design factors matter

Blinding, a placebo comparator, sample size, and duration are what separate a persuasive trial from a misleading one. The snapshot surfaces these so a grade is never a black box — and so you can judge the evidence for yourself.

References

  1. [1] Concato J, et al. (2000). RCTs vs observational studies. N Engl J Med, 342(25): 1887-1892.

Learning context

How this concept connects to supplement decisions

How evidence grades are assigned and which clinical trial design factors matter — shown through embeddable Study Design Snapshots that keep the practical takeaway prominent. Learning pages explain the reasoning layer behind the herb and compound library. They are designed to make mechanisms, evidence quality, safety tradeoffs, and product claims easier to interpret.

Use How to read an evidence grade to build better questions before choosing a supplement: what outcome is being targeted, what mechanism is claimed, what human evidence exists, what dose was studied, and what risks could change the answer for a specific person?

Mechanistic plausibility is useful, but it should be weighed against trial design, safety history, product quality, and the possibility that a simpler intervention may be more appropriate.