Forty-four per cent: reading the nausea figure properly
Trial adverse-event tables count episodes reported to a study nurse. They are the best data we have and they systematically under-record the mundane.
TheCompound Journal
Reporting on incretins, compounding & the peptide supply chain
Adverse events
Severity in these tables is graded by interference with activity, not by how unpleasant the experience was. Those are different measurements.
The number most often quoted about this drug class is that around forty-four per cent of participants on the highest semaglutide dose in the pivotal obesity trial reported nausea. It is a real figure and it is routinely misused. In the same trial, seventeen per cent of the placebo group reported nausea, which tells you something about how much ordinary gastrointestinal discomfort a population reports when asked weekly and given a form. The drug-attributable excess is the difference between the two, and it is meaningful without being the number in the headline.
The pivotal semaglutide obesity trial randomised 1,961 adults to 2.4 mg weekly or placebo for sixty-eight weeks. Gastrointestinal disorders were reported by around seventy-four per cent of the active arm and about forty-eight per cent of placebo. Within that, nausea was reported by roughly forty-four per cent against seventeen per cent, diarrhoea by about thirty-two per cent against sixteen, vomiting by about twenty-five per cent against seven, and constipation by roughly twenty-three per cent against ten.1
Three features of that table are routinely lost. The placebo rates are high, which is what happens when a large population is asked systematically about gut symptoms every few weeks. The events were predominantly graded mild or moderate. And discontinuation attributable to gastrointestinal events ran to about four and a half per cent of the active arm, against under one per cent on placebo.
The gap between three-quarters of participants reporting a gastrointestinal event and four and a half per cent stopping because of one is the most informative thing in the table. Most of this effect profile is endured rather than disabling, and any account that quotes the first figure without the second is describing something other than what happened.
A number in an adverse-event table counts participants who reported at least one episode of a coded term at any point during the treatment period. It says nothing about how many episodes, how long they lasted, or how bad they were beyond a three-level severity grade defined by interference with usual activity.
This construction has predictable consequences. A cumulative figure over sixty-eight weeks is the union of many short episodes and cannot be read as a prevalence. Two populations with identical percentages can have entirely different lived experiences. And severity grading captures function rather than distress, so an episode of severe nausea that did not stop somebody working is graded moderate.
None of this is a criticism of the trials, which followed standard practice and reported it transparently. It is a caution about a specific and common misreading: that a forty-four per cent nausea figure describes a state rather than an event count. The published tolerability analyses that break events down by timing and duration are considerably more informative than the summary tables, and are cited far less often.2
A forty-four per cent nausea figure is the union of many short episodes, not a description of a state.
On what an adverse-event percentage countsGastrointestinal events in this class are concentrated in the escalation phase. Reported incidence rises in the days following a dose increase, declines over the subsequent weeks at an unchanged dose, and rises again at the next increment. Analyses that plot event onset against week show a series of peaks aligned to the escalation schedule rather than a flat burden across the trial.2
Two things follow. The first is that the escalation phase is where discontinuation risk lives, which means the tolerability problem in this class is largely a titration problem. The second is that a symptom appearing eight months into stable dosing should not be attributed to the drug by default, because that is not where the drug-attributable events cluster.
There is a corollary that patients find useful and are rarely told. The worst week of a given dose is usually the first one. A person who has been unwell for four days after an increase is, on the published pattern, at the point where things typically begin to improve rather than at the beginning of a permanent state. That is a statement about a population and not a promise about an individual, and we put it that way deliberately.
| Measure | Target | Evidence in this population | Basis |
|---|---|---|---|
| Hold dose / extend escalation interval | Nausea, vomiting, satiety | Protocol-permitted; supported by tolerability analyses | Dose- and time-dependence of the effect |
| Step back one rung | Any dose-limiting effect | Observational and protocol practice | Same |
| Smaller, more frequent meals | Early satiety, nausea | None randomised | Delayed gastric emptying |
| Reduced dietary fat | Nausea, fullness | None randomised | Fat further slows emptying |
| Osmotic laxative | Constipation | Strong in general populations; none specific | Transfer from general constipation evidence |
| Deliberate fluid intake | Volume depletion | None randomised; mechanism clear | Thirst is appetite-linked and suppressed |
| Ondansetron or similar | Nausea, vomiting | None adequately powered here | Transfer from other emetic settings |
| Ginger | Nausea | None here | Small trials in pregnancy and chemotherapy |
| Graded by the Journal on the published literature as of this issue. Inclusion is not endorsement and this table is not a treatment protocol. | |||
Discontinuation for adverse events ran to roughly four and a half per cent on top-dose semaglutide and between four and seven per cent across the tirzepatide dose range, against one to three per cent on placebo. The great majority of those discontinuations were gastrointestinal and the great majority occurred during escalation.13
Those figures should be read as a floor. Trial participants receive weekly contact, free product, a nurse who can be telephoned, and an investigator with a strong interest in retention, and they are pre-selected by their willingness to enter a trial. Real-world persistence data for this class is markedly worse, with a substantial proportion of people no longer filling prescriptions at twelve months, for reasons that combine tolerability with cost and supply.
The Journal draws one inference. If most intolerance-driven discontinuation happens during escalation, and escalation practice is the least evidence-based part of the treatment course, then the largest available improvement in outcomes in this class is probably not a new molecule. It is a better answer to the titration question, which nobody has run a trial to obtain.4
Nausea: the sensation preceding or in place of vomiting; a symptom. Vomiting: forceful expulsion of gastric contents; a sign. Retching: the effort without the expulsion. Early satiety: fullness disproportionate to volume consumed. Dyspepsia: upper abdominal discomfort, often used loosely to include all of the above.
Gastroparesis: a clinical diagnosis of delayed gastric emptying with characteristic symptoms and no mechanical obstruction. It is not a synonym for drug-induced emptying delay, and the two are conflated constantly. Ileus: failure of propulsion without mechanical obstruction. Obstruction: mechanical blockage.
Incidence: proportion of a population experiencing at least one event in a period. Prevalence: proportion affected at a point in time. Adverse-event tables report the first and are read as the second. Adjudicated: reviewed against predefined criteria by a committee blinded to treatment, which is a materially stronger standard than a reported term.
If one paragraph of this file survives, we would prefer it to be the one about fluid. The dramatic harms in this area are rare and the mundane one is common: appetite suppression removes the signal that drives drinking, and volume depletion follows quietly. It is prevented by drinking on a schedule rather than on a sensation, and it accounts for the great majority of renal events reported in association with these drugs.
Selected from correspondence received on this article. Writers are identified by initial, surname and city, verified before printing. Replies are from the desk that filed the piece or from the standards editor. Write to letters@compoundjournal.com.
I want to push back on the ginger paragraph. You describe the evidence as transferred from pregnancy and chemotherapy, which is accurate, and then include it in the table anyway. Either it belongs or it does not.
— P. Ekundayo, Akure
It belongs, labelled. The table is a map of what is recommended and on what basis, not a list of endorsements, and excluding widely used low-risk measures because their evidence is transferred would make the map less useful rather than more honest. We have made the column heading clearer.
Your figures show diarrhoea at thirty-two per cent and constipation at twenty-three per cent in the same trial arm. I assumed one of these was an error until your mechanism section. It would be worth putting that explanation before the table rather than after it.
— J. Wenninger, Graz
The gallbladder section says some of the excess is attributable to weight loss rather than the drug. If the drug causes the weight loss, is that not a distinction without a difference for the person who ends up in theatre?
— A. Fournier, Nantes
For the individual, largely yes. For the question of whether one molecule is safer than another, or whether the risk would fall on slower loss, the distinction is the whole question. We should have made clear that it is a mechanistic distinction rather than a consoling one.
As an anaesthetist I read the perioperative section with interest and one objection. You frame disclosure as the patient obligation. In my experience the failure is more often ours: the pre-assessment questionnaire in my own institution did not include these drugs until eighteen months after the first guidance appeared.
— L. Fontaine, Brussels
A fair correction and we have amended the text. If the question is not on the form, the absence of an answer is not a patient failure. We would be interested to hear from readers in other institutions about whether their pre-assessment documentation has caught up.
Your incidence tables are from the licensed products. I use compounded material at a concentration that does not match any pen. Are the figures transferable at all?
— F. Duquesne, Lyon
The mechanism transfers; the incidence figures transfer only to the extent that your actual exposure matches the trial exposure, which is unknown unless the content has been measured. That is not evasion. It is the reason we argue for peptide content as a standard reported field rather than purity alone.
Trial adverse-event tables count episodes reported to a study nurse. They are the best data we have and they systematically under-record the mundane.
A tour of the tissues where the receptor is expressed, and what happens in each.
We give background rates alongside trial rates, because an event occurring during treatment is not thereby caused by it.
The evidence base is thin and the document says so, which is to its credit.
The four-week step exists because four to five weeks is approximately how long a once-weekly drug takes to stop rising at a fixed dose. That is a good reason, and it is not…
The mechanism is well described. The variance is not.