Continued treatment against continued placebo: the comparison that settled it
The evidence on stopping is better than the evidence on almost anything else in this field, because somebody deliberately randomised it.
TheCompound Journal
Reporting on incretins, compounding & the peptide supply chain
Titration
The four-week step exists because four to five weeks is approximately how long a once-weekly drug takes to stop rising at a fixed dose. That is a good reason, and it is not a trial result.
Consider two steps on the same ladder. Moving from 0.25 mg to 0.5 mg of semaglutide is a doubling. Moving from 1.7 mg to 2.4 mg is an increase of about forty-one per cent. The absolute increments are 0.25 mg and 0.7 mg respectively, so the second looks larger and is in fact the gentler event. Receptor occupancy responds to ratio, not to difference, and every titration ladder in this class is built on tapering ratios for precisely that reason. Almost nobody explains this to the person holding the pen, who reasonably concludes that the steps are getting bigger.
Tirzepatide begins at 2.5 mg weekly for four weeks, moves to 5 mg, and thereafter increases in 2.5 mg increments at intervals of not less than four weeks, to a maximum of 15 mg. The structural difference from semaglutide is important: after the first doubling the increments are fixed in absolute terms, which means the ratio falls steadily — 1.5-fold, then 1.33, then 1.25, then 1.20.
The practical consequence is that the upper half of the tirzepatide ladder is unusually gentle in proportional terms, and the first step from 2.5 mg to 5 mg is by some distance the most demanding thing the schedule asks. Clinicians we spoke to described the 2.5-to-5 transition as the point at which most early attrition occurs, which is what the ratios predict.
The label also states, in language that repays attention, that 5 mg is a therapeutic dose in its own right and that escalation beyond it should reflect response and tolerability. That is a materially different instruction from a ladder with a fixed destination, and it is closer to how the drug is actually used.1
The elimination rate constant of a drug is 0.693 divided by its half-life. Fractional approach to steady state after time t is 1 minus e to the power of minus k times t. For a seven-day half-life this yields about seventy-five per cent of steady state at two weeks, eighty-eight per cent at three, ninety-four per cent at four and ninety-seven per cent at five.
Four weeks is therefore the point at which a once-weekly dose has essentially finished getting stronger. Escalate at two weeks and the person receives the increment written on the pen plus roughly a further quarter of the previous rung still accumulating underneath it. That is not dangerous in any dramatic sense, but it does mean the symptom burden attributed to the new dose is partly the tail of the old one, and it makes the escalation harder to interpret.
For tirzepatide, with a half-life closer to five days, four weeks corresponds to more than five half-lives and the previous rung is fully settled. The same interval is therefore slightly conservative for one molecule and exactly adequate for the other, which is a small illustration of how a shared convention can be right for different reasons.2
Four weeks is the point at which a dose has finished getting stronger on its own. That is a kinetic fact, not a clinical result.
On the escalation intervalA dose increase should be described as a ratio, because receptor occupancy and the exposure-response relationship are governed by proportional change rather than by absolute milligrams. On the semaglutide weight-management ladder the ratios are 2.00, 2.00, 1.70 and 1.41. On the tirzepatide ladder they are 2.00, 1.50, 1.33, 1.25 and 1.20.
Two consequences follow. The first is that the early rungs are the hard ones, in both ladders, and the widespread expectation that titration gets progressively more difficult is backwards. The second is that a person who has tolerated the doubling at the bottom of the ladder has already survived the largest proportional insult the schedule contains.
There is a third, less obvious consequence for anyone dosing from a multi-dose vial rather than a fixed pen. Fixed-pen users move in the ratios above. Vial users can move in any ratio they like, including ratios small enough to be pharmacologically meaningless and large enough to be foolish. Freedom of increment is the single largest practical difference between pen and vial administration, and the arithmetic is the only guardrail.
| Programme | Molecule | Dose | Duration | Mean weight change |
|---|---|---|---|---|
| STEP 1 | Semaglutide | 2.4 mg weekly | 68 weeks | −14.9% |
| STEP 1 | Placebo | — | 68 weeks | −2.4% |
| STEP 5 | Semaglutide | 2.4 mg weekly | 104 weeks | −15.2% |
| SURMOUNT-1 | Tirzepatide | 5 mg weekly | 72 weeks | −15.0% |
| SURMOUNT-1 | Tirzepatide | 10 mg weekly | 72 weeks | −19.5% |
| SURMOUNT-1 | Tirzepatide | 15 mg weekly | 72 weeks | −20.9% |
| SURMOUNT-1 | Placebo | — | 72 weeks | −3.1% |
| Treatment-policy estimand where reported. Figures are means from the primary publications and are not comparable across programmes, which differed in population, duration and analysis. | ||||
The evidence on deviating from four-week steps is observational and one-sided. Slower escalation — five, six or eight weeks per rung — is reported by clinicians to reduce early discontinuation, is consistent with the tachyphylaxis data, and has never been randomised against the standard interval in a trial of adequate size. Faster escalation has no supporting rationale we can identify and a clear kinetic argument against it.
What can be said with confidence is that the cost of going slower is bounded and calculable: a longer time to target exposure, and therefore a later arrival at the efficacy plateau. Because the plateau itself sits at sixty weeks or beyond, adding four or eight weeks to the escalation phase is a small fraction of the treatment course. The cost of going faster is a higher probability of discontinuation, and discontinuation costs the entire effect.
That asymmetry is the strongest thing the Journal is willing to say on the subject. It is an argument from consequence rather than from trial data, and we flag it as such rather than dressing it as a finding.3
Initiation dose: the first rung, chosen for tolerability and generally sub-therapeutic. Not a low treatment dose. Target dose: the dose a protocol or prescriber intends to reach. Maintenance dose: the dose continued once the intended effect is achieved. Maximum approved dose: the highest dose in the label, set by the studied range and the tolerability ceiling.
Escalation interval: the time between increments. Hold: deliberately remaining at a rung beyond the standard interval. Re-titration: re-ascending after exposure has been substantially cleared. Dose-limiting: describing an effect severe enough to prevent escalation, which is a property of the person and the dose jointly, not of the drug alone.
Steady state: the condition in which drug entering the body equals drug leaving it. Accumulation ratio: steady-state average concentration divided by first-dose average concentration. Precision here matters more than it sounds: a large share of the correspondence this desk receives about titration turns out on inspection to be a disagreement about which of these words the writer meant.
First, the optimal escalation interval. No adequately powered randomised comparison of intervals at a fixed target dose exists for any molecule in this class.
Second, the optimal hold duration for a person who has not adapted at four weeks. Practice ranges from four to twelve weeks on no comparative evidence at all.
Third, the lowest maintenance dose that preserves a result. The withdrawal trials compared full dose with nothing.
Fourth, whether tolerability at one rung predicts tolerability at the next. Clinicians assume it does, plausibly, and the published dose-ranging data is not analysed in a way that answers the question.
Fifth, whether any measurable baseline characteristic predicts the ceiling. Nothing published does so usefully, which mirrors the situation for efficacy: mean behaviour in this class is well characterised and individual variation is not.4
The Journal lists these not as a complaint about researchers but as a map of where confident advice is currently outrunning its evidence. Anyone offering a precise answer to any of the five is offering an opinion, and should be read as doing so.
Readers who think a paragraph above has outrun its evidence should write to the standards desk. Titration is a subject on which practically everybody has an opinion and practically nobody has a trial, and our correction log for this file is longer than we would like. That is the correct outcome of publishing numbers in a field where the numbers keep being checked.
The evidence on stopping is better than the evidence on almost anything else in this field, because somebody deliberately randomised it.
The features that should prompt urgent assessment, stated once and plainly.
The evidence base is thin and the document says so, which is to its credit.
We set out the questions that distinguish a symptom to manage from a dose to change.
Mass and function are different endpoints and training affects them differently. Most coverage treats them as one.
The graduation interval differs between barrel sizes, and a 1 mL barrel is frequently marked in two-unit steps. Reading one as though it were marked in single units halves…