Accepted answer
The STEP 2 placebo arm is the only thing that makes its treatment arm interpretable, and it is the row nobody quotes. Symptoms reported under placebo in these programmes are not rare, because the population is being asked about them weekly and would have had some of them regardless. The attributable figure is the treated rate minus the placebo rate, and that difference is routinely a fraction of the headline. Two cautions on the subtraction: the arms must have been assessed the same way, and a discontinuation for an event removes that participant from later time points in both arms, which flatters whichever arm loses more people.
Stated carefully, a trial establishes what happened to a defined group under a defined protocol. Extending it beyond that group is inference, and inference is allowed as long as it is labelled.
Placebo arms in this class are not nothing. Lifestyle-intervention placebo arms in the major obesity trials commonly lose two to three per cent of body weight, so an active-arm figure quoted without its comparator overstates the drug effect by roughly that much.
Non-inferiority and superiority designs are not interchangeable. A non-inferiority result says the new agent is not meaningfully worse against a pre-specified margin — it does not say it is as good, and it certainly does not say it is better.
Registry entries at ClinicalTrials.gov carry the pre-specified primary endpoint with a timestamp, which is the cheapest available check on whether an endpoint was changed after the data were seen.
If a claim cannot be traced to a named trial with a named endpoint, treat it as a claim rather than as evidence.
edited 4 Nov 2024 by wren_calloway — added a caveat about sampling
5Same experience here, different supplier. – tamsin_wray 5 months ago 6Adding that the endpoint definition differs between the two trials being compared here. – Dr_Fatima_Belkacem 6 months ago add a comment