Numbers first: SURMOUNT-4 · nausea.
I can parse the result. I am less sure what it licenses me to conclude.
I have deliberately not looked at anyone else’s interpretation yet.
What is the correct interpretation, and what is the common misreading?
Numbers first: SURMOUNT-4 · nausea.
I can parse the result. I am less sure what it licenses me to conclude.
I have deliberately not looked at anyone else’s interpretation yet.
What is the correct interpretation, and what is the common misreading?
Start with the population. The inclusion criteria of the trial determine what its result can be extrapolated to, and the extrapolation people want is usually to a population the trial excluded.
Absolute risk reduction, worked: if the control-arm event rate is 8.0 per cent over the follow-up period and the hazard ratio is 0.80, the treated rate is approximately 6.4 per cent, the absolute risk reduction is 1.6 percentage points, and the number needed to treat is 1 ÷ 0.016 ≈ 63 over that period. A 20 per cent relative reduction and a number needed to treat of 63 are the same finding stated two ways, and only one of them sounds impressive.
| Quantity | Value | Derivation |
|---|---|---|
| Control-arm event rate | 8.0 % | From the trial table, not the abstract |
| Hazard ratio | 0.80 | Reported |
| Treated event rate | 6.4 % | 8.0 × 0.80 |
| Absolute risk reduction | 1.6 pp | 8.0 − 6.4 |
| Number needed to treat | 63 | 1 ÷ 0.016 |
| Relative risk reduction | 20 % | 1 − 0.80 |
The last two rows describe the same finding. Only one of them is used in headlines.
The relevant detail is that apoB and LDL-C disagree because they measure different things: LDL-C is the cholesterol mass carried in the LDL fraction, ApoB is a count of atherogenic particles. Small dense particles carry less cholesterol each, so a person with many small particles has a concordantly higher ApoB than their LDL-C suggests. When they disagree, ApoB is the better risk marker.
STEP 1 reported a mean weight change of approximately −14.9 per cent with semaglutide 2.4 mg versus −2.4 per cent with placebo at 68 weeks[1]; the difference between the figures quoted from this trial in different places is an estimand difference.
The caveat is the population. Trial participants were screened, monitored and supported; the effect size in an unmonitored setting is not the trial effect size, and it is not obvious in which direction the difference runs.
Read the confidence interval, read the estimand, and compute the absolute effect yourself. It takes two minutes and it changes how the result feels.
edited 5 Mar 2026 by tobias_maartens — removed a claim I could not source
Aggregated, published test results and vendor ratings built from submitted batches. Methodology stated, dataset browsable, no listing fees.
Browse resultsAsk PeptideStack is a static archive. Posting is closed, but the norms are worth stating: answer the question that was asked, show your working, cite the trial or the certificate, and say plainly where the evidence runs out.