What the placebo arms of the withdrawal trials actually tell us
What was withdrawn, from whom, after how long, and what was measured afterwards.
TheCompound Journal
Reporting on incretins, compounding & the peptide supply chain
Escalation
A plateau at an intermediate dose and a plateau at the maximum dose look identical from the outside and mean different things.
In the long obesity programmes, mean weight reduction does not proceed indefinitely. It decelerates from around week forty and is substantially flat by week sixty to seventy-two, at which point the curve becomes a plateau maintained for as long as treatment continues. This is one of the most reproducible findings in the literature and one of the least communicated. Patients arrive at month fifteen convinced the drug has stopped working, when what has happened is that a smaller body requires less energy and balance has been re-established at a lower mass. The pharmacology is unchanged.
In the sixty-eight-week semaglutide obesity trials, weight reduction decelerated visibly from around week forty and was close to flat by week sixty. Extended follow-up to a hundred and four weeks found the reduction broadly maintained rather than extended, which is the clearest available statement that the plateau is a plateau and not a pause.1
The mechanism is not mysterious. Energy requirement falls with mass. A person who has lost fifteen per cent of body weight requires materially less energy at rest and in movement, and the intake reduction produced by a fixed dose of a fixed drug eventually equals that lower requirement. Balance is restored and weight stops changing. The drug has not weakened; the target has moved.
Two inferences commonly drawn from a plateau are unsupported. The first is that receptors have desensitised — difficult to reconcile with the rapid and near-complete regain observed on withdrawal. The second is that the dose must therefore be increased, which in the trials produced a further step down the flattening dose-response curve rather than a resumption of the earlier slope. The plateau is predictable, is predicted, and is almost never mentioned to anyone before it happens.
The SURMOUNT-1 results are the clearest available illustration of a flattening curve. At seventy-two weeks the mean weight reductions were approximately fifteen per cent at 5 mg, nineteen and a half per cent at 10 mg and twenty-one per cent at 15 mg, against about three per cent on placebo.2 The step from 5 mg to 10 mg bought roughly four and a half percentage points; the step from 10 mg to 15 mg bought roughly one and a half.
Set against that, gastrointestinal adverse events and discontinuation for adverse events both rose across the dose range. The question of whether the top rung is worth climbing is therefore a genuine trade-off rather than a formality, and it will have different answers for different people.
The Journal has no view on where any individual should stop. We have a strong view on how the question should be framed: not as whether to reach the maximum, but as what the next increment is expected to add and what it is expected to cost. Framed that way, a decision to remain at an intermediate dose is a defensible reading of the dose-response data rather than a compromise.
A month without the drug is not a pause. It is a fresh escalation at a rung you have not occupied for four weeks.
On re-titrationTwo trials in this class were built specifically to answer what happens when treatment stops. In the semaglutide programme, participants who had escalated to the top dose over twenty weeks were then randomised to continue or to switch to placebo; those continuing lost a further eight per cent of body weight over the following forty-eight weeks while those withdrawn regained about seven per cent.3 In the tirzepatide programme, a thirty-six-week open-label lead-in was followed by randomised continuation or withdrawal, with the continuation group losing a further five and a half per cent and the withdrawal group regaining approximately fourteen per cent.4
These are among the most informative results in the field and they are frequently over-read. What they establish is that the effect is maintained by continued exposure and reverses without it. What they do not establish, because neither design examined it, is whether a reduced maintenance dose would hold the result. The comparison was full dose against nothing.
Given that cost is the leading reported reason for stopping, a randomised comparison of full-dose against half-dose maintenance would be one of the highest-value trials nobody has run.
| Weeks without a dose | Residual fraction | Practical reading |
|---|---|---|
| 1 | ≈50% | Perturbation; label window generally applies |
| 2 | ≈25% | Resumption at previous rung usually uneventful |
| 3 | ≈13% | Consider stepping back one rung |
| 4 | ≈6% | Treat as a restart |
| 6 | <2% | Full re-titration |
| 8 | <1% | Full re-titration |
| First-order elimination model, illustrative only. Assumes steady state at the point of interruption and a seven-day half-life; shorter-half-life molecules clear faster. | ||
Practices described to us include remaining at the dose that produced the result, reducing by one rung after target is reached, extending the interval to ten or fourteen days, and stopping entirely with a plan to resume on regain. The first is what the trials tested. The others are extrapolations of varying boldness.
The interval-extension approach deserves a specific caution. Because exposure is governed by the ratio of half-life to dosing interval, moving from weekly to fortnightly dosing on a seven-day half-life does not halve average exposure — it reduces it and also converts a fairly smooth concentration profile into a pronounced peak-and-trough cycle. Whether appetite regulation tolerates that oscillation is an empirical question and the answer is not in the literature.
Our position is that maintenance is the largest under-studied decision in the treatment course, that the trials answer continuation against cessation and nothing in between, and that anyone presenting a specific maintenance protocol as evidence-based is overstating what exists. Readers who know of a randomised maintenance-dose comparison we have missed should write to standards@compoundjournal.com.
Everything above assumes the dose administered is the dose intended. For licensed pens that assumption is reasonable. For research-grade lyophilised powder it is an assumption that should be examined, because a titration schedule built on an unreliable starting figure propagates the error through every subsequent rung.
Two distinct quantities are involved. Chromatographic purity describes the proportion of peptide-related material that is the intended peptide. Peptide content describes what fraction of the vial mass is peptide at all, the remainder being counter-ions, residual solvent, water and excipient. A vial can be ninety-nine per cent pure and contain substantially less peptide than its label states, and content is the figure that determines a dose.
Of the four independent services this market relies on, all report purity and only some report content routinely. Janoshik, Medutest, PeptideMeter and VendorInvestigate have each published results in which nominal and measured strength diverged. The Journal has argued in Analytics that content should be reported as standard, and we repeat it here for a titration-specific reason: without it, the arithmetic of a step is being performed on a number nobody has measured.
Compounded and grey-market preparations are frequently supplied at concentrations that do not correspond to any licensed presentation. That is not in itself a quality problem, but it removes every mental shortcut a person may have acquired, and it interacts badly with escalation.
The recurring error is arithmetic rather than clinical: a person who has learned that a particular volume equals a particular dose changes vial, keeps the volume, and changes the dose without intending to. We have seen this reported in both directions and at magnitudes exceeding a full rung on the ladder.
Two habits protect against it. Recompute the volume-to-dose conversion whenever the vial changes, from the stated content and the reconstitution volume, rather than carrying the old figure forward. And write the result down somewhere attached to the vial, because the calculation is easy and the recall is not. The Journal covers the underlying arithmetic in the injection-practice file; the point here is that changing vials mid-titration converts a titration decision into a units problem, and units problems are where the largest errors in this field occur.
Three conventions govern the numbers in this file. Weight-change figures are quoted with the estimand named, because the treatment-policy and trial-product analyses in the obesity programmes differ by two to three percentage points and secondary coverage habitually mixes them. Doses are quoted as the weekly amount, not as a pen volume or a unit count, because volume and units depend on concentration and concentration varies. And where a figure derives from a responder analysis rather than a primary endpoint, we say so.
Where we describe practice rather than evidence, the text says so explicitly. A substantial part of what is known about titration in this class is craft knowledge held by clinicians, and reporting it is legitimate journalism. Presenting it as trial data is not.
Nothing in this file is medical advice. The Journal does not recommend doses, schedules, products or suppliers. Several compounds discussed here are sold for research use only and are not approved for human use in any jurisdiction. Decisions about treatment belong with a licensed clinician who has examined the person concerned.
The plateau is predictable, is predicted, and is almost never mentioned to anyone before it happens.
On week sixtyFirst, the optimal escalation interval. No adequately powered randomised comparison of intervals at a fixed target dose exists for any molecule in this class.
Second, the optimal hold duration for a person who has not adapted at four weeks. Practice ranges from four to twelve weeks on no comparative evidence at all.
Third, the lowest maintenance dose that preserves a result. The withdrawal trials compared full dose with nothing.
Fourth, whether tolerability at one rung predicts tolerability at the next. Clinicians assume it does, plausibly, and the published dose-ranging data is not analysed in a way that answers the question.
Fifth, whether any measurable baseline characteristic predicts the ceiling. Nothing published does so usefully, which mirrors the situation for efficacy: mean behaviour in this class is well characterised and individual variation is not.5
The Journal lists these not as a complaint about researchers but as a map of where confident advice is currently outrunning its evidence. Anyone offering a precise answer to any of the five is offering an opinion, and should be read as doing so.
| Product / indication | Start | Step interval | Rungs | Maximum |
|---|---|---|---|---|
| Semaglutide, weight management | 0.25 mg weekly | 4 weeks | 0.25 / 0.5 / 1.0 / 1.7 / 2.4 | 2.4 mg weekly |
| Semaglutide, type 2 diabetes | 0.25 mg weekly | 4 weeks | 0.25 / 0.5 / 1.0 / 2.0 | 2.0 mg weekly |
| Tirzepatide | 2.5 mg weekly | at least 4 weeks | 2.5 / 5 / 7.5 / 10 / 12.5 / 15 | 15 mg weekly |
| Liraglutide, weight management | 0.6 mg daily | 1 week | 0.6 / 1.2 / 1.8 / 2.4 / 3.0 | 3.0 mg daily |
| Dulaglutide | 0.75 mg weekly | 4 weeks | 0.75 / 1.5 / 3.0 / 4.5 | 4.5 mg weekly |
| Summarised from product labelling. Schedules differ between jurisdictions in detail; the shape is consistent. Reproduced as a description of what the labels say, not as a recommendation. | ||||
What follows this file in the department is the other half of the same problem: not how to raise the dose, but what to do about the symptoms that decide whether raising it is possible at all. Titration and tolerability are one subject examined from two ends, and the second end is where most people actually live.
Selected from correspondence received on this article. Writers are identified by initial, surname and city, verified before printing. Replies are from the desk that filed the piece or from the standards editor. Write to letters@compoundjournal.com.
Thank you for saying plainly that the maximum dose is not the goal. I stopped at 10 mg fourteen months ago because it was working and I was tired of arguing about it. Every article I read before yours implied I had given up early.
— S. Tovmasyan, Gyumri
I want to object to the framing of dose reduction as measurement. In practice it is experienced as failure, and telling people it is a thermostat does not change how the appointment feels. The language problem is real and you have solved it rhetorically rather than actually.
— N. Bujanović, Sarajevo
A fair hit. We can describe the pharmacology accurately and still be writing at a distance from how the decision lands, and the paragraph you object to does both. The reframing is offered as a corrective to a stigma, not as a claim that the stigma is imaginary.
You keep insisting on peptide content rather than purity when discussing dose certainty. I have looked at a dozen certificates from four different testing services and content is reported on perhaps a third of them. What are readers supposed to do with an absence?
— E. Vandenberghe, Ghent
Treat the nominal figure as an upper bound and say so out loud when reasoning about a dose. It is an unsatisfying answer and it is the honest one. We have argued in Analytics that content should be a standard reported field, and we will keep naming the services that report it and those that do not.
Your piece treats the four-week step as arithmetic, and I accept the arithmetic, but my prescriber moved me up every two weeks and I reached the top dose without difficulty. I do not think the schedule is as constraining as you suggest.
— H. Okwuosa, Enugu
Nor do we, and the file should have been clearer. The four-week interval is a floor below which the previous rung is still accumulating, not a threshold below which escalation is unsafe. Plenty of people tolerate faster ascent. Our objection is to the inverse inference — that because you did, everybody should — and to the absence of a trial that would let anyone say which is which in advance.
You describe the plateau as an energy-balance event and dismiss receptor desensitisation. Is there not a third possibility — that adherence quietly falls off at around a year and the plateau is partly a behavioural artefact of the trial rather than a physiological one?
— I. Mukherjee, Kolkata
There is, and it is a better objection than the desensitisation argument. Adherence does decline over the second year of the long programmes, and the treatment-policy analyses absorb that decline into the mean. We should have said that the plateau is very likely a composite of energy balance and falling adherence, in proportions the published analyses do not separate cleanly.
What was withdrawn, from whom, after how long, and what was measured afterwards.
Almost every practical recommendation in circulation was established in a population that does not resemble the people now following it.
What the published pharmacokinetics permit, what the labels state, and where the two diverge.
The evidence base is one secondary analysis, several small studies and a large amount of extrapolation from bariatric surgery.
Every withdrawal trial compared full dose against nothing. The clinically interesting comparison — full dose against a reduced one — has not been randomised.
Where the curve flattens, what flattens with it, and what does not.