STEP 4 extension data: what happens after the trial stops
A design note rather than a result: what the comparator was, and what that permits you to conclude.
TheCompound Journal
Reporting on incretins, compounding & the peptide supply chain
Maintenance
A supply interruption is a discontinuation with no notice, no taper and no plan. That is a distinct clinical situation.
The shortage years produced an unintentional natural experiment in interrupted treatment, and its main lesson was about escalation rather than about regain. People who had been stable at a maintenance dose for months, lost access for six or eight weeks, and then resumed at the dose they had been on, frequently found the resumption far less tolerable than the original escalation had been. That is exactly what the pharmacokinetics predicts, it was widely experienced, and it was almost never explained in advance.
This is the most practically consequential item in the whole subject and the one least often stated in advance. Gastrointestinal tolerability to these agents develops over weeks of continued exposure and decays when exposure is removed. After four weeks without the drug, plasma concentrations are a small fraction of steady state and the tolerability accommodation has substantially reset. Resuming at the previous maintenance dose therefore presents the system with an exposure step it has not experienced for a month.
The clinical convention — resume at a lower dose and re-escalate — follows from the pharmacokinetics rather than from caution.1 Product labelling for several agents in the class advises consideration of re-initiation at a lower dose after an extended interruption, and the threshold at which this applies differs between products, which is a detail worth checking against the specific label rather than a general rule.
The shortage period demonstrated the consequence of ignoring this at scale. Large numbers of people lost access for six to ten weeks, resumed where they had left off, and experienced nausea and vomiting considerably worse than during their original escalation. It was predictable, it was predicted by anybody who had read the label carefully, and it was almost never communicated.
Maintenance has a thinner evidence base than escalation, which is the opposite of where people spend their time. That imbalance is the largest gap between what has been studied and what is lived, and it is not closing quickly.
Analyses of pharmacy claims consistently find that persistence with these agents for weight management is poor relative to their efficacy, with a large minority of people no longer filling prescriptions within a year of starting and discontinuation concentrated in the first three months.2 The pattern tracks coverage, deductible reset timing and cash price far more closely than it tracks clinical response, which is the signature of an economic rather than a therapeutic discontinuation.
Almost none of this appears in the clinical literature on withdrawal. The trials studied people who stopped because a protocol told them to, with the drug supplied free, in a population willing to be randomised. That is close to the opposite of the situation in which most discontinuation actually occurs: unplanned, unsupervised, at a time set by an insurer or a price rise rather than by a clinical assessment, and frequently without anybody being told it has happened.
The Journal reports discontinuation in both this department and The Ledger for that reason. The clinical trajectory after stopping is a Patient Notes question; why people stop is an economics question; and the two literatures currently do not speak to one another at all.
A seven-day half-life tapers itself. What a taper buys is behavioural, and it should be argued for on those terms.
On coming offA supply gap is a discontinuation with no notice, no plan and no taper. It differs from every other route to stopping in that it is imposed on both the patient and the prescriber, its duration is unknown at the outset, and it frequently ends as abruptly as it began. The shortage listings of recent years produced these events at population scale, and they have not been studied as a clinical exposure.
Three features make them distinctive. The patient cannot plan a maintenance strategy around an interruption of unknown length. Substitution — to a different agent, a different dose, or a compounded preparation — happens under time pressure and often without a dose-equivalence basis, since no head-to-head equivalence data exists between agents in this class. And the resumption problem described above applies in full, because the gaps were typically long enough to reset tolerability.
The Journal reported these events as they occurred and continues to think they represent the largest uncontrolled interruption experiment in the history of the class. What nobody collected was outcome data: how much weight was regained during the gaps, how many people never resumed, and what happened to the glycaemic control of those taking the drugs for diabetes rather than for weight.
| Reason | Randomised evidence on outcome | Typical notice | Resumption likely? |
|---|---|---|---|
| Protocol-driven withdrawal | Three designs | Planned | Not applicable |
| Reached target weight | None | Planned | Sometimes |
| Intolerable side effects | Discontinuation rates only | Days | Sometimes, lower dose |
| Cost or coverage loss | None | Weeks or none | Often, when coverage returns |
| Supply interruption | None | None | Usually, at reset tolerability |
| Discontinuation rates for adverse events are reported in every pivotal trial; outcomes after discontinuation for the other reasons are not, because the trials did not enrol people who stopped for them. | |||
The withdrawal question changes shape when the drug was prescribed for something other than weight. In the cardiovascular outcome trial of semaglutide in overweight and obesity without diabetes, the reduction in major adverse cardiovascular events emerged over years of continued treatment, and the trial provides no information about what happens to that benefit on cessation.3 The same applies to the renal outcome data in chronic kidney disease with type 2 diabetes, where the effect on kidney disease progression was measured over a median of several years of treatment.4
There is no reason to expect an outcome benefit that accrues over years to persist after the exposure ends, and no trial has tested it. For a person taking the drug for glycaemic control, stopping has an immediate and measurable consequence in HbA1c over the following three months. For a person taking it for cardiovascular or renal risk, stopping has no measurable short-term consequence at all, which makes the decision harder rather than easier.
This is the situation in which the Journal thinks the withdrawal-trial coverage has done the most damage. Framing discontinuation as a weight question invites a person taking the drug for kidney disease to reason about it in the wrong currency entirely.
Extending the interval between doses is a dose reduction expressed in time. Framed that way it connects to the maintenance literature instead of sitting apart from it as an unstudied practice.
Restarting after months away is well tolerated in general and the response is broadly reproducible: people who lost weight on an agent and stopped generally lose weight again on resuming, at a similar rate. There is no established phenomenon of a diminished second response in this class, and the withdrawal trials that re-offered treatment after their observation periods did not report one.
Three practical features recur. Escalation has to start again from a low dose for tolerability reasons, which means several weeks before the previous maintenance exposure is re-established. The nausea of a second escalation is frequently reported as worse than the first, for which the Journal has seen no mechanistic explanation and would not rule out reporting bias. And the weight trajectory on restarting begins from wherever the person now is, so a second course is a longer project than the first if regain was substantial.
None of this constitutes advice about whether to restart, which is a clinical decision. It is offered as a description of what the trial reports and the correspondence describe, and readers should note that no trial has been designed to study re-initiation as its primary question.
This is reporting on a body of trial evidence and it is not advice about whether or how to stop taking a medicine. The decision to discontinue an agent prescribed for glycaemic control, cardiovascular risk or kidney disease is materially different from the decision to discontinue one prescribed for weight, and in every case it belongs with a clinician who has seen the person and knows why the drug was started.
Two further notes. Compounds sold for research use only are not approved for human use in any jurisdiction, and nothing here should be read as guidance about using them or about stopping their use. And where this piece describes what clinicians report doing about maintenance dosing, that is description of practice and not a schedule anybody should adopt from a magazine.
The Journal takes correspondence on this subject at letters@compoundjournal.com and factual challenges at standards@compoundjournal.com. Letters describing a personal experience of stopping are read with attention and are published, where they are published, as accounts rather than as evidence — a distinction this department tries hard to preserve in both directions.
The correspondence this department receives on stopping divides almost evenly between people frightened by regain figures they have seen quoted without denominators and people who stopped without difficulty and cannot understand the alarm. Both groups are reading the same trials. The difference is almost entirely a matter of which number was quoted to them and whether anybody explained what it was a proportion of.
Selected from correspondence received on this article. Writers are identified by initial, surname and city, verified before printing. Replies are from the desk that filed the piece or from the standards editor. Write to letters@compoundjournal.com.
Stopping without telling anybody is common and it makes the observational data worse than it looks, because somebody recorded as continuing may have stopped months earlier. Persistence measured from prescription records and persistence in fact are different quantities.
— H. Barreto, Recife
The reporting convention that irritates me most is the pie chart. Reasons are not mutually exclusive, they are not equally weighted, and a shape that implies both is doing damage to the reader’s understanding before a single number is read.
— Y. Sasaki, Sapporo
Intermittent use is studied almost nowhere and practised widely, which is an uncomfortable combination for anybody trying to write about it responsibly. The honest position is that the evidence base is close to empty, and saying so is more useful than assembling anecdotes.
— L. Marulanda, Medellín
That is our position and we restate it in each piece. Where there is no evidence, the finding is the absence, and filling the space with accounts would misrepresent what is known.
A last point about this publication’s own framing. Research-use material is not approved for human use in any jurisdiction, and a maintenance discussion conducted around such material is a discussion about a supply chain rather than about a therapy. Keeping that distinction visible is the useful thing you do.
— R. Mothibi, Gaborone
It is the distinction the department is built on and we restate it deliberately rather than as a formality. What we cover is a market and a body of published evidence, not a course of treatment.
Cost is a maintenance variable and it is treated as though it were outside the clinical question. For most people the sustainable dose is the affordable one, and a plan that ignores that will be abandoned rather than followed.
— H. Baptiste, Fort-de-France
Which is why the ledger department and this one keep landing on the same story from opposite ends.
A design note rather than a result: what the comparator was, and what that permits you to conclude.
What the labels permit, what clinicians do, and the size of the gap between them.
Every withdrawal trial compared full dose against nothing. The clinically interesting comparison — full dose against a reduced one — has not been randomised.
A design note rather than a result: what the comparator was, and what that permits you to conclude.
A survey of the maintenance evidence, which is shorter than the survey of the withdrawal evidence.
Efficacy was never the question in this appraisal. Duration of treatment was.