GRADE assessment rates the certainty of evidence for each outcome in a systematic review, judging how much confidence a reader can place in the estimated effect. Our service rates every prioritised outcome across the five domains that can lower certainty, applies the criteria that can raise it, and builds a Summary of Findings table in GRADEpro with the certainty narrative drafted for your results and abstract.
GRADE assessment and Summary of Findings tables
We rate the certainty of your evidence outcome by outcome across every GRADE domain and hand you a Summary of Findings table your reviewers and guideline panel will accept.
4.9 / 5 across 1,194+ delivered projects
- Cochrane and PRISMA 2020 methods
- PhD methodologists
- 100% human-written, no generative AI
- Reproducible R and Stata code
- Free quote within 24 hours
- Mutual NDA on request
Why certainty is rated per outcome, not per study
The single idea that trips up most authors is that GRADE is applied per outcome, not per study. A review of the same trials can be high certainty for mortality and low certainty for quality of life, because each outcome carries its own risk of bias, its own consistency across studies, and its own precision. Rating the studies once and stamping a single grade across the whole review misreads the method and is the error reviewers catch first. We build the assessment around your prioritised outcomes so each certainty rating reflects the evidence that actually informs it.
The rating also starts from the design. Randomised trials begin at high certainty and observational studies begin at low, and from there the certainty can move down through five domains or up through three. This is where GRADE goes beyond a quality assessment of individual studies, and why it feeds on, but is not the same as, your risk of bias appraisal.
The five domains that can downgrade certainty
Risk of bias
Serious limitations in the studies behind an outcome, carried over from the RoB 2 or ROBINS-I appraisal, lower the certainty for that outcome.
Inconsistency
Unexplained heterogeneity in the effect across studies, seen in a wide I-squared or non-overlapping confidence intervals, reduces confidence in a single pooled number.
Indirectness
Evidence from different populations, interventions, comparators, or surrogate outcomes than your question asks about weakens how directly it answers that question.
Imprecision
A confidence interval wide enough to include both meaningful benefit and no effect, or too few events, lowers certainty regardless of how clean the studies were.
The fifth downgrading domain is publication bias. Where the smaller studies point one way and the picture from a funnel plot or a small-study-effects test suggests missing negative results, we document the concern and downgrade honestly rather than assuming the published record is complete. Each domain can cost one or two levels, so an outcome can fall from high to very low certainty when several concerns stack up.
When certainty can be rated up
Downgrading is only half of GRADE. For observational evidence that has not already been marked down, three criteria can raise the certainty rating. A large effect that is unlikely to be explained by bias, a clear dose-response gradient where more exposure tracks more effect, and plausible residual confounding that would only have worked against the observed effect all argue that the true effect is at least as strong as the estimate. We apply these criteria explicitly and record the reasoning, so an upgrade is defensible rather than optimistic.
What the GRADE service covers
- 1
Prioritise the outcomes
We agree the critical and important outcomes to rate, ideally from your protocol, so the assessment stays focused on the outcomes that drive the decision.
- 2
Rate every domain per outcome
For each outcome we assess risk of bias, inconsistency, indirectness, imprecision, and publication bias, then apply the upgrading criteria where they fit, with a footnote justifying every move.
- 3
Build the Summary of Findings table
We assemble the table in GRADEpro GDT with participant and study counts, absolute and relative effects, the certainty symbols, and explanatory footnotes linked to each rating.
- 4
Draft the certainty narrative
We write the certainty statements for your results and the plain-language summary for your abstract, phrased in the standard GRADE wording readers expect.
What you receive
You receive a completed certainty rating for every prioritised outcome, a Summary of Findings table built in GRADEpro with linked footnotes, and the drafted narrative for your results and abstract. The assessment reads directly off your pooled estimates and heterogeneity statistics and off your risk of bias judgements, so the certainty ratings and the GRADE certainty of evidence table tell one consistent story with the rest of the review. Where an outcome cannot be pooled, we carry the certainty logic into the narrative synthesis instead of forcing a rating the data will not support.
- A certainty rating for each prioritised outcome, from high to very low
- A Summary of Findings table built in GRADEpro GDT ready for your manuscript
- An explanatory footnote justifying every downgrade and every upgrade
- Absolute and relative effects reported alongside the certainty for each outcome
- The certainty narrative drafted for your results section
- A plain-language evidence statement for your abstract in standard GRADE wording
Who commissions a GRADE assessment
Guideline developers
Panels who need defensible certainty ratings and Summary of Findings tables to move from evidence to a graded recommendation.
Doctoral and clinical reviewers
Authors whose systematic review needs a formal certainty assessment to meet supervisory or journal standards.
Authors answering reviewers
Researchers asked in revision to add GRADE ratings or a Summary of Findings table before a manuscript can be accepted.
Teams without a methodologist
Groups with the clinical expertise but not the time to rate every domain and assemble the table correctly.
How a GRADE assessment is quoted
We quote a fixed fee shaped by the number of outcomes you need rated, whether the evidence is randomised, non-randomised, or a mix that changes the starting certainty, and whether you want the Summary of Findings table and the certainty narrative drafted as well as the ratings. Send your outcomes, your pooled results, and your risk of bias appraisal, and we will return a fixed quote for a fully documented assessment.
Common mistakes we correct
| Consideration | The common mistake | Working with us |
|---|---|---|
| Unit of rating | One certainty grade stamped across the whole review | A separate rating for each prioritised outcome, as GRADE requires |
| Domain judgements | Downgrades asserted without a stated reason | Every downgrade and upgrade justified in a linked footnote |
| Imprecision | Clean studies read as high certainty despite a wide interval | Precision judged on the confidence interval and the event count |
| Observational evidence | Rated up informally or dismissed outright | Started at low certainty, then raised only on the explicit criteria |
Frequently asked questions
- What is the difference between GRADE and risk of bias?
- Risk of bias appraises a single study, judging how its design and conduct could distort its result. GRADE rates the certainty of the whole body of evidence for one outcome, and risk of bias is only the first of its five domains. Two studies can each be low risk of bias, yet the outcome they inform can still be low certainty once inconsistency, indirectness, imprecision, or publication bias is considered. Risk of bias is an input; GRADE is the overall verdict.
- How many outcomes should be GRADEd?
- GRADE is applied to the outcomes that matter to the decision, not to every outcome you measured. Most reviews rate the certainty for a focused set of critical and important outcomes, typically around five to seven, chosen when the protocol is written. Rating everything dilutes the Summary of Findings table and buries the outcomes readers actually use, so we help you prioritise rather than pad.
- What is a Summary of Findings table?
- A Summary of Findings table is the standard one-page output of GRADE. For each prioritised outcome it reports the number of participants and studies, the pooled or absolute effect, the certainty rating with its symbols, and a plain-language statement of what the evidence shows. It is the table editors and guideline panels look for first, and it is usually built in GRADEpro GDT so the ratings, footnotes, and effect estimates stay linked.
- Can GRADE be applied to observational studies?
- Yes. GRADE starts non-randomised evidence at low certainty rather than excluding it, then allows that rating to move up where the signal is strong. A large effect, a dose-response gradient, or plausible confounding that would only have diluted a real effect can all raise the certainty of observational evidence. Randomised trials start at high certainty, so the two designs enter the same framework from different baselines.
- Do you need GRADE for a scoping review?
- No. A scoping review maps what evidence exists and does not rate the certainty of an effect, so a formal GRADE assessment and a Summary of Findings table are not expected. GRADE belongs with systematic reviews that synthesise effects on defined outcomes. If your question is really about how confident you can be in an effect, a systematic review rather than a scoping review is the right design.
Send us your outcomes and your pooled results, and we will rate the certainty of your evidence and deliver a Summary of Findings table for a fixed quote.
Free quote within 24 hours. 100% human-written by PhD methodologists.
Methodology reviewed by
Lead Review Methodologist
Leads protocol design, screening, and reporting across the practice, with a focus on reviews that survive editorial scrutiny.
Why researchers bring in a PhD methodologist
80+
systematic reviews are published every day (Hoffmann et al., 2021)
67.3 weeks
average time to complete a review in-house (Borah et al., BMJ Open 2017)
~70%
of published reviews rate critically low on AMSTAR 2 quality appraisal
4.9 / 5
our client rating across 1,194+ delivered projects
