Accepted answer
Fair depends on the comparator arm, and in PIONEER-1 that means asking whether the comparator was titrated to the same ambition as the experimental one. A head-to-head that runs its comparator to a dose below the one it is licensed at is not measuring the two agents, it is measuring one agent against a handicapped version of the other. Check three things: the maximum comparator dose reached, the proportion of the comparator arm that reached it, and whether the titration schedules had the same duration. If those match, the comparison is fair on dosing and the argument moves to the endpoint. If they do not, the effect size is partly an artefact of the protocol.
It helps to be literal here: a trial establishes what happened to a defined group under a defined protocol. Extending it beyond that group is inference, and inference is allowed as long as it is labelled.
Confidence intervals matter more than point estimates when two trials disagree. Two studies reporting fifteen and twenty per cent whose intervals overlap heavily have not disagreed about anything.
More usefully, non-inferiority and superiority designs are not interchangeable. A non-inferiority result says the new agent is not meaningfully worse against a pre-specified margin — it does not say it is as good, and it certainly does not say it is better.
The cardiovascular outcome programme in this class runs to several large randomised trials — LEADER for liraglutide, SUSTAIN-6 and SELECT for semaglutide, REWIND for dulaglutide — and they are the reason the class is discussed as more than a weight intervention.
Read the protocol and the statistical analysis plan if the result matters to you. Both are usually published alongside.
edited 28 Jul 2024 by Dr_Tomas_Kral — clarified the distinction between purity and content
Good answer, but the confidence interval in the cited trial is wider than implied. – Dr_Tomas_Kral 30 days ago add a comment