A single laboratory value is a point on a noisy curve. What you want is a trend across at least three draws under comparable conditions, and "comparable" is doing a lot of work in that sentence.
Liver enzymes are a poor surrogate for hepatic histology in both directions: substantial steatohepatitis with normal transaminases is common, and modest enzyme elevation with minimal fibrosis is common. If the question is fibrosis, the answer comes from a non-invasive score such as FIB-4 or a stiffness measurement, not from ALT.
Headline results, principal programmes
| Trial | Agent | n | Duration | Primary result |
|---|
| STEP 1 | Semaglutide 2.4 mg | 1,961 | 68 wk | −14.9 % vs −2.4 % weight |
| STEP 2 | Semaglutide 2.4 mg, T2DM | 1,210 | 68 wk | −9.6 % vs −3.4 % weight |
| SURMOUNT-1 | Tirzepatide 5/10/15 mg | 2,539 | 72 wk | −15 / −19 / −21 % weight |
| SURMOUNT-4 | Tirzepatide, withdrawal | 670 | 88 wk | Continued loss vs substantial regain |
| SELECT | Semaglutide 2.4 mg | 17,604 | ~40 mo | MACE HR 0.80 (0.72–0.90) |
| FLOW | Semaglutide 1.0 mg, CKD | 3,533 | ~3.4 yr | Renal composite reduced; stopped early |
| SURMOUNT-OSA | Tirzepatide, OSA | 469 | 52 wk | AHI reduced with and without PAP |
Absolute risk reduction, worked: if the control-arm event rate is 8.0 per cent over the follow-up period and the hazard ratio is 0.80, the treated rate is approximately 6.4 per cent, the absolute risk reduction is 1.6 percentage points, and the number needed to treat is 1 ÷ 0.016 ≈ 63 over that period. A 20 per cent relative reduction and a number needed to treat of 63 are the same finding stated two ways, and only one of them sounds impressive.
FLOW tested a composite renal endpoint — kidney failure, sustained 50 per cent eGFR decline, or renal or cardiovascular death — in type 2 diabetes with chronic kidney disease, and was stopped early for efficacy[1].
I would resist reading a subgroup finding as a result. Subgroups in these trials were not powered, and a striking subgroup in a large trial is the expected consequence of multiplicity.
If the trend across three draws is flat, the difference between draws one and two was noise. Most of what people react to is noise.
Confirming from the other direction: I did the wrong thing and got exactly the predicted outcome. – kirsi_lahtinen 6 days ago Is there a reason to prefer the second method over the first, other than cost? – marta_okonkwo 2 months ago add a comment