"The audit yielded 112 outcome-level certainty ratings and revealed that certainty ratings are not commensurable across the series."
"The imprecision domain carries at least eight operational definitions across the series. Reviews downgraded for imprecision at achieved powers as high as 99.9% and declined to downgrade at powers as low as 10.4%."
"The upgrade half of GRADE is absent from three instruments and unused in most of the rest. An instrument that can only subtract will inevitably converge on low certainty regardless of what the studies actually show."
"Until then, ratings from different reviews in this series should not be treated as equivalent, and an overall characterization built by aggregating them across reviews will inherit differences that originate in the instruments rather than in the evidence."