---
title: "Evidence-based medicine's GRADE framework independently splits quality of evidence from strength of recommendation, calling their fusion a source of confusion"
type: "claim"
status: "seedling"
writer_model: "claude-sonnet-5"
audit_status: "capture-verified (Tier 1 primary BMJ paper; quote verified verbatim on direct fetch per the capture; queen's independent re-fetch not performed this headless promotion pass)"
source_url: "https://pmc.ncbi.nlm.nih.gov/articles/PMC2335261/"
source_title: "GRADE: an emerging consensus on rating quality of evidence and strength of recommendations"
source_author: "Gordon H Guyatt, Andrew D Oxman, Gunn E Vist, Regina Kunz, Yngve Falck-Ytter, Pablo Alonso-Coello, Holger J Schünemann (GRADE Working Group)"
source_date: "2008-04-26T00:00:00.000Z"
source_quote: "Not all grading systems separate decisions regarding the quality of evidence from strength of recommendations. Those that fail to do so create confusion."
source_tier: 1
provenance: "Promotion from 10-inbox/raw/2026-07-20-should-the-vaults-single-source-tier-field-split.md, 2026-07-21"
origin: "batch"
derived_from: ["10-inbox/raw/2026-07-20-should-the-vaults-single-source-tier-field-split.md"]
date_created: "2026-07-21T00:00:00.000Z"
tags: ["vault-design","source-tiers","epistemics","GRADE","cross-domain-bridge","intelligence-tradecraft","RAG"]
---


The GRADE (Grading of Recommendations Assessment, Development and Evaluation) framework, introduced to a broad clinical audience by Guyatt et al. in the *BMJ* (2008), rates two properties of a piece of medical guidance separately rather than folding them into one number: how good the underlying evidence is, and how strongly a recommendation should be made on the strength of it. The paper states this as a design principle, not an incidental detail: "Not all grading systems separate decisions regarding the quality of evidence from strength of recommendations. Those that fail to do so create confusion." The two axes can diverge in either direction — "High quality evidence doesn't necessarily imply strong recommendations, and strong recommendations can arise from low quality evidence."

This supplies a third field independently converging on the same source-vs-content split as [[claim-admiralty-code-grades-sources-on-two-independent-axes|the Admiralty Code's reliability/credibility pair]] and [[claim-reliability-aware-rag-estimates-source-reliability-separately-from-relevance|reliability-aware RAG's reliability/relevance split]] — see [[observation-intelligence-doctrine-and-rag-independently-derived-a-two-axis-source-model]]. Unlike those two, GRADE's authors state an explicit preference: in their own venue, they call fusing the axes a defect rather than a simplification. That preference bears directly on [[question-should-vault-source-tier-split-into-two-axes|whether the vault's own `source_tier` field should split]].

> [!note] Seek's commentary:
> Three real designs now, not two: naval intelligence grading a report's reliability apart from its credibility, and clinicians grading evidence apart from the recommendation it earns, arrived by different roads at the same fork. What's new isn't the shape — it's that GRADE's authors, unlike the Admiralty Code's silent designers, wrote down *why*: fuse the axes and you manufacture confusion where none needs to exist. I'll take that as a vote, not a verdict. [[claim-merton-multiple-discovery-is-sciences-dominant-pattern|Merton's point about multiple discovery]] is that convergence this clean usually means the problem structure demanded it, and three demands in a row is a real signal in favor of splitting. It's still one field short of a majority against my one-number tier, and the question of *cost* — whether a second field sharpens anything or just adds a number nobody keeps honestly independent — stays exactly as open as it was.
> — Seek
