Accepted answer
Because two papers on SURPASS-3 are usually reporting two different estimands from the same randomisation. The treatment-policy estimand asks what happened to everyone assigned, including those who stopped; the trial-product estimand asks what happens if you keep taking it. The second is always the larger number, and both are legitimate answers to different questions. Then there is the analysis population — randomised, treated, or completers — and the handling of missing data, where a last-observation-carried-forward and a multiple imputation can differ by a point or more. Neither paper is wrong. Read the statistical methods section and you will find both figures defined in it.
The honest answer here is that the published evidence supports part of the claim and is silent on the rest, and it is worth being precise about which part is which.
Trial populations are selected. Exclusion criteria in this class routinely remove people with significant renal impairment, prior pancreatitis and unstable psychiatric illness, which is exactly the population the results are then quoted for.
Relative to absolute, worked
| Quantity | Value | Derivation |
|---|
| Control-arm event rate | 8.0 % | From the trial table, not the abstract |
| Hazard ratio | 0.80 | Reported |
| Treated event rate | 6.4 % | 8.0 × 0.80 |
| Absolute risk reduction | 1.6 pp | 8.0 − 6.4 |
| Number needed to treat | 63 | 1 ÷ 0.016 |
| Relative risk reduction | 20 % | 1 − 0.80 |
The last two rows describe the same finding. Only one of them is used in headlines.
Confidence intervals matter more than point estimates when two trials disagree. Two studies reporting fifteen and twenty per cent whose intervals overlap heavily have not disagreed about anything.
Registry entries at ClinicalTrials.gov carry the pre-specified primary endpoint with a timestamp, which is the cheapest available check on whether an endpoint was changed after the data were seen.
If a claim cannot be traced to a named trial with a named endpoint, treat it as a claim rather than as evidence.
6Good answer, but the confidence interval in the cited trial is wider than implied. – gunnar_isaksen 44 days ago add a comment