Load, not cardio: the distinction the general advice keeps losing
The older-adult diet-and-exercise trials are the closest analogue to rapid pharmacological weight loss, and they are twenty years old.
TheCompound Journal
Reporting on incretins, compounding & the peptide supply chain
Panels
Which markers are informative, which are confounded, and which move for reasons unrelated to nutrition.
Several of the markers used are also confounded in ways that matter here. Ferritin is an acute-phase reactant, and inflammation falls markedly during weight loss, so a falling ferritin during successful treatment may represent resolving inflammation, depleting iron stores, or both, and cannot distinguish them without a transferrin saturation and a C-reactive protein alongside. Vitamin D concentrations rise during weight loss partly because the volume of adipose tissue in which the vitamin distributes has shrunk. These are not obscure technicalities; they determine whether a result is acted on.
Two results in the same person differ for three reasons: the analyte genuinely changed, the assay is imprecise, and the analyte varies within the person from day to day. The last two are quantified in the biological variation literature as the analytical coefficient of variation and the within-subject coefficient of variation, and databases of the latter have been maintained for decades.1
The reference change value combines them: approximately 2.77 times the square root of the sum of their squares, for a two-sided ninety-five per cent probability that a difference is real. The results are instructive. Sodium, with tiny biological variation, has a reference change value of around three per cent. Creatinine is about fourteen per cent. Alanine aminotransferase, with within-subject variation above twenty per cent, requires something like a sixty per cent change. Triglycerides, more variable still, require more.
Apply that to a routine monitoring situation. An ALT moving from 28 to 41 units per litre — a rise of forty-six per cent that crosses no threshold and is unlikely to be flagged — sits inside the reference change value and may be nothing at all. An ALT moving from 28 to 62 has moved. Nothing on the report distinguishes the two cases, and the distinction is the entire question.
The renal outcome programme in type 2 diabetes with chronic kidney disease is the only trial in this class designed with kidney endpoints as its primary purpose. It randomised participants with established chronic kidney disease and reported a reduction in a composite of kidney disease progression, kidney death and cardiovascular death, together with a slower annual decline in estimated glomerular filtration rate, over a median follow-up of several years.2
Two features of the eGFR data matter for anybody reading a panel. There is an initial dip in estimated filtration rate on starting treatment, of the order of one millilitre per minute per 1.73 square metres, which resolves and is followed by a slower long-term decline than in the comparator arm. That pattern — an acute dip followed by long-term preservation — is familiar from other renoprotective drug classes and is generally understood as a haemodynamic effect rather than injury.
The practical implication is that a small fall in eGFR in the first months of treatment is expected and is not evidence of harm, while a large fall is not expected and is. Distinguishing them requires knowing the reference change value for creatinine, which is around fourteen per cent, and knowing whether the person has been vomiting, which changes everything.
The earlier cardiovascular outcome trials in the class carried renal composites as secondary endpoints and reported reductions in new or worsening nephropathy driven largely by albuminuria, which is a weaker endpoint than the eGFR-based composites of the dedicated renal trial.34 Anybody quoting renal benefit from those programmes should say which component of which composite they mean.
Everything required to interpret a laboratory result correctly is omitted from the document that reports it.
Perpetua Nwachukwu, Contributing Writer, Laboratory MedicineThe fall in alanine aminotransferase during successful treatment is one of the few laboratory movements in this field with a directly demonstrated mechanism, because liver fat was measured by imaging in several programmes rather than inferred from enzymes. A trial of semaglutide in biopsy-confirmed steatohepatitis reported resolution of steatohepatitis without worsening of fibrosis in a substantially greater proportion of treated participants than placebo, with corresponding falls in transaminases.5 The larger phase 3 programme in the same indication subsequently reported histological improvement on both resolution and fibrosis endpoints.6
Alongside that sits the imaging evidence from the diabetes programme, where liver fat content measured by magnetic resonance fell considerably more on a dual agonist than on insulin at broadly comparable glycaemic control, which separates the hepatic effect from the glycaemic one.
What this establishes is that the falling ALT is tracking a real change in the liver rather than reflecting reduced enzyme release for some incidental reason. What it does not establish is how much of the change is attributable to the weight loss and how much to a direct hepatic effect, since the two are not separable in a trial where the treated arm also lost more weight.
| Analyte | Direction during rapid loss | Principal reason | Finding or artefact? |
|---|---|---|---|
| Serum creatinine | Falls | Reduced muscle mass | Artefact of composition |
| eGFR (creatinine-based) | Rises | Follows creatinine | Artefact of composition |
| Alanine aminotransferase | Falls | Reduced hepatic fat | Finding |
| Triglycerides | Fall | Improved insulin sensitivity | Finding |
| LDL cholesterol | Falls slightly | Weight loss | Finding, small |
| Lipoprotein(a) | Little change | Largely genetic | Neither |
| Free triiodothyronine | Falls | Energy restriction adaptation | Artefact of deficit |
| C-reactive protein | Falls | Reduced adipose inflammation | Finding |
| Ferritin | Falls | Both inflammation and iron stores | Ambiguous |
| 25-hydroxyvitamin D | Rises | Smaller distribution volume | Artefact of composition |
| Lipase, amylase | Rise modestly | Drug class effect | Finding of unclear significance |
| Directions are typical rather than universal. The classification is the Journal’s own and is offered as an interpretive aid, not as a clinical rule. | |||
Amylase and lipase rise modestly on treatment with this drug class, by something in the region of ten to twenty per cent on average, and elevations above the upper reference limit are more common on drug than on placebo. This has been characterised most thoroughly in the liraglutide cardiovascular outcome programme, which followed more than nine thousand participants for a median of 3.8 years and therefore had the events to adjudicate.7 A dedicated analysis within it found higher mean enzyme concentrations on treatment with no corresponding excess of adjudicated acute pancreatitis, and concluded that the elevations had no useful predictive value for the clinical event.8
The diagnostic threshold for acute pancreatitis is a lipase above three times the upper reference limit in the presence of characteristic abdominal pain, or imaging evidence. Both limbs are required. A lipase of twice the upper limit in an asymptomatic person on treatment is a common finding with no established significance, and investigating it as though it were the first limb of a diagnosis produces imaging, anxiety and no information.
The Journal notes that this is one of the few places in this subject where the trial evidence is genuinely clarifying: somebody asked the question directly, measured the enzymes systematically, adjudicated the clinical events independently, and reported that the two did not track. That is what a useful safety analysis looks like.
Sustained energy restriction produces a characteristic and benign change in thyroid function tests: triiodothyronine falls, reverse triiodothyronine rises, thyroxine changes little and thyroid-stimulating hormone falls modestly or remains unchanged. This is the low-T3 pattern of adaptation to reduced energy availability, it is not hypothyroidism, and treating it as such is an error that predates this drug class by decades.
The relevant point for monitoring is that a thyroid panel drawn during rapid weight loss will frequently show a low or low-normal free T3, and that this does not indicate thyroid disease, does not require treatment, and reverses when energy balance is restored. Thyroid-stimulating hormone remains the appropriate first-line test for suspected thyroid dysfunction; adding free T3 to a panel during active weight loss reliably generates a result that requires explaining.
Separately and unrelatedly, this class carries a boxed warning in some jurisdictions derived from rodent thyroid C-cell findings. Serum calcitonin monitoring is not recommended for that purpose, and pharmacoepidemiological work examining thyroid cancer incidence in treated populations has not established the association the rodent data raised as a possibility.9 The Journal reports the boxed warning as what it is: a precaution derived from a rodent finding whose human relevance remains unestablished.
Post-bariatric micronutrient surveillance is well founded and specific. Roux-en-Y gastric bypass bypasses the duodenum and proximal jejunum, which are the principal absorption sites for iron, calcium and several B vitamins; reduced gastric acid impairs the release of food-bound B12 and the reduction of ferric iron; and the intrinsic-factor pathway is compromised by the reduction in parietal cell mass. Each of those is an identified mechanism supporting a specific test at a specific interval.
None of them applies to a receptor agonist. The gastrointestinal tract is anatomically intact, acid secretion is broadly preserved, and no absorption site is bypassed. The mechanism that does apply is reduced intake, which predicts deficiency in proportion to dietary inadequacy rather than in the bariatric pattern. Those two predictions differ: a person eating half as much of a varied diet is at different risk from a person whose duodenum has been bypassed, and the appropriate surveillance is not obviously the same.
The Journal has looked for a cohort study characterising micronutrient status in this population at twelve months or beyond and has not found one. Until one exists, monitoring schedules for this drug class are precautionary extrapolation. That is a defensible thing to do and it should be described accurately rather than presented as protocol.
Vitamin D. Concentrations frequently rise during weight loss for a reason unrelated to intake: the vitamin is fat-soluble and distributes into adipose tissue, so a smaller adipose compartment produces a higher serum concentration at unchanged total body content. A rising 25-hydroxyvitamin D during weight loss is therefore a volume-of-distribution effect as much as anything else.
Vitamin B12. The commonest confounder is co-prescription of metformin, which lowers B12 by a well-established mechanism, in a population where metformin is extremely common. Attributing a low B12 to reduced intake when the person has been on metformin for eight years is a sequencing error rather than a laboratory one.
Thiamine. The one micronutrient where the Journal thinks the precaution is well founded on mechanism: thiamine stores are small, turnover is rapid, and protracted vomiting is a recognised precipitant of deficiency. Prolonged vomiting during escalation is a plausible route to it.
Magnesium and potassium. Both fall with persistent vomiting and both are genuine findings when they do. Neither is a nutritional marker in that context; they are consequences of the losses.
One in twenty healthy people falls outside a reference interval by construction. On a comprehensive panel, the flagged result is the expected outcome.
On multiple testingA baseline panel rarely finds anything. Its value is almost entirely in what it makes possible later: within-person comparison, which for nearly every analyte on a routine panel is a more sensitive instrument than comparison against a reference interval, because within-subject biological variation is smaller than between-subject variation.
The arithmetic behind that is worth stating. For an analyte where the within-subject coefficient of variation is substantially smaller than the between-subject value — a condition satisfied by creatinine, the liver enzymes, HbA1c, the thyroid hormones and most of the electrolytes — a person’s own previous result is a better comparator than the population interval. The index of individuality formalises this, and for the analytes in question it says clearly that population intervals are relatively insensitive to change in an individual.
The practical consequence is that a person with a baseline creatinine of 62 whose value is now 78 has information that a person presenting with 78 and no baseline does not, even though both results sit inside every reference interval in use. That is the whole argument for the baseline panel, and it is a stronger argument than the one usually offered, which is that the panel might find an undiagnosed problem.
| Analytes on panel | Probability of ≥1 flag | Expected flags |
|---|---|---|
| 6 | 26% | 0.30 |
| 12 | 46% | 0.60 |
| 16 | 56% | 0.80 |
| 20 | 64% | 1.00 |
| 30 | 79% | 1.50 |
| Assumes each reference interval excludes 5% of a healthy population and that analytes are independent. Real analytes covary, so true figures are somewhat lower; the order of magnitude holds. | ||
A distinction has to be drawn firmly because the postbag suggests it frequently is not. The four independent testing services this market relies on — Janoshik, Medutest, PeptideMeter and VendorInvestigate — analyse material. They report chromatographic purity, identity by mass, peptide content where it is measured, and in the case of the verification services what could be established about a supplier. A clinical laboratory analyses a person. The two produce documents that superficially resemble each other and answer entirely unrelated questions.
A purity certificate reporting 99.1 per cent for a batch from WWB, CPC or QYB tells you nothing about anybody liver enzymes. A normal panel does not confirm that a vial contained what its label claimed, and an abnormal one does not establish that it did not. Where a person suspects a supply problem, the instrument for that is analytical testing of the material; where a person has an abnormal laboratory result, the instrument is clinical assessment. Substituting one for the other is a reliable way to spend money and learn nothing.
Compounds sold for research use only are not approved for human use in any jurisdiction, and nothing in this department should be read as guidance about using them or about monitoring their use.
Five things accompany a laboratory number in these pages. The units, because international and conventional units differ for several analytes and the same value means different things in each. The reference interval used, with a note where the interval is contested, as it is for alanine aminotransferase. The baseline, because a change of 1.8 percentage points in HbA1c from a starting value of 8.3 is a different claim from the same change from 9.5. The estimand where the figure comes from a trial. And the reference change value where we are discussing an individual delta rather than a group mean.
We also state the assay method where it matters, which is more often than one would like: HbA1c in the presence of a haemoglobin variant, thyroid function in the presence of interfering antibodies, and creatinine measured by enzymatic against Jaffe methods all behave differently, and a comparison across methods is not a comparison.
This is a heavier apparatus than most publications carry and it exists because the alternative, in our experience, is a stream of technically accurate figures that lead readers to conclusions the data does not support. Errors in this apparatus should be reported to standards@compoundjournal.com; the correction log records what came of each one.
The Laboratory Notebook reports what tests measure, how they behave, and what has been found using them. It does not recommend monitoring schedules, interpret readers’ results, or advise on treatment. A laboratory result belongs in a conversation with a clinician who has the rest of the picture, and this publication is emphatically not that conversation.
Two standing notes. Several compounds discussed in these pages are sold for research use only and are not approved for human use in any jurisdiction; the Journal reports on them as commodities and as analytical problems, not as therapies. And where we describe what the pivotal trials monitored, that is reporting on trial protocols and not a template anybody should adopt from a magazine.
Correspondence is welcome at letters@compoundjournal.com. The Journal receives a steady flow of letters containing readers’ own panel results with a request for interpretation, and we do not provide it — not from caution but because a panel without a history, an examination and a reason for ordering it cannot be interpreted by anybody, including us.
Readers who order their own panels privately, which a substantial fraction of this publication’s readership does, are in a specific position: they have the data and not the interpretive apparatus, and the apparatus is where the value is. The reference change value for the analyte in question, and the person’s own previous result, do more interpretive work than any reference interval on the printout.
The older-adult diet-and-exercise trials are the closest analogue to rapid pharmacological weight loss, and they are twenty years old.
The mechanism is well described. The variance is not.
What the regulatory dossiers actually contain on dose selection is remarkably thin, and worth knowing before treating the ladder as settled science.
A survey of the maintenance evidence, which is shorter than the survey of the withdrawal evidence.
A design note rather than a result: what the comparator was, and what that permits you to conclude.
A design note rather than a result: what the comparator was, and what that permits you to conclude.