Fair depends on the comparator arm, and in SUSTAIN-6 that means asking whether the comparator was titrated to the same ambition as the experimental one. A head-to-head that runs its comparator to a dose below the one it is licensed at is not measuring the two agents, it is measuring one agent against a handicapped version of the other. Check three things: the maximum comparator dose reached, the proportion of the comparator arm that reached it, and whether the titration schedules had the same duration. If those match, the comparison is fair on dosing and the argument moves to the endpoint. If they do not, the effect size is partly an artefact of the protocol.
A trial establishes what happened to a defined group under a defined protocol. Extending it beyond that group is inference, and inference is allowed as long as it is labelled.
Trial populations are selected. Exclusion criteria in this class routinely remove people with significant renal impairment, prior pancreatitis and unstable psychiatric illness, which is exactly the population the results are then quoted for.
Non-inferiority and superiority designs are not interchangeable. A non-inferiority result says the new agent is not meaningfully worse against a pre-specified margin — it does not say it is as good, and it certainly does not say it is better.
The cardiovascular outcome programme in this class runs to several large randomised trials — LEADER for liraglutide, SUSTAIN-6 and SELECT for semaglutide, REWIND for dulaglutide — and they are the reason the class is discussed as more than a weight intervention.
Read the protocol and the statistical analysis plan if the result matters to you. Both are usually published alongside.
edited 2 Jun 2026 by Dr_Ilse_Vandenberg — tightened the wording; no substantive change
2The number needed to treat is the framing that finally made this concrete for me. – marcus_thorbjorn 7 days ago add a comment