What Is Actually Inside a "Personalised" GLP-1 Side-Effect Risk Number
Personalised risk tools give you one number. It is really two ingredients multiplied together — an evidenced base rate and a set of profile multipliers — and almost nobody shows you the second. Here are all 75 of ours, including the fact that none of them carries a citation.
Type your age and sex into a "personalised" GLP-1 side-effect tool and it returns a single percentage. Ours is two ingredients multiplied together: a base rate drawn from our evidence corpus, and a set of profile multipliers that push it up or down for you specifically. We have written at length about the first ingredient and never published the second.
So here it is — all 75 coefficients our predictor applies, together with the uncomfortable part: not one of them carries a citation. We make no claim about how anyone else builds their tool; we are publishing ours because a reader cannot check what they cannot see. This is the fifth piece in our evidence series, after which effects have enough evidence to calibrate a statistical model, the corpus-derived pooled clinical rates, what patients themselves report, and gallstones end to end.
The worked example
Ask our predictor for nausea on semaglutide at a mid-range dose, for a 40-year-old with no GI history, no diabetes, and past the first month. Both figures below are live API responses captured on 27 August 2026:
The five-point difference is one multiplier: an odds ratio of 1.25 applied to the female profile. The API says so in the response itself — unadjustedPercentage: 32, then modifiersApplied: [{ id: "sex:female", oddsRatio: 1.25, provenance: "seed-2026-04-12" }].
Now look at how differently evidenced the two ingredients are.
The 32% is backed by *26 stated rates from 18 distinct sources*, drawn from the 69 clinical and regulatory records our corpus holds for nausea at that dose tier, deduplicated to one entry per source. You can read every one of those sources, with URLs, at the nausea endpoint.
The ×1.25 is backed by nothing we can show you. It was hand-coded when the model was seeded on 12 April 2026. There is no per-modifier citation recorded anywhere in our code, our documentation, or our stored model configuration — we went looking on 27 August 2026 and found none. It is an expert-coded prior: a judgement, honestly held, but a judgement.
That asymmetry is the entire point of this article. A number carrying "26 stated rates from 18 distinct sources" beside it looks fully evidenced. Five of its 37 points are not.
All 75 coefficients
Every effect has five multipliers. An odds ratio above 1.00 raises your estimate; below 1.00 lowers it; exactly 1.00 does nothing. These are the live values on 27 August 2026, readable any time from modifiers on the per-effect endpoint.
| Effect | Female | Age 65+ | GI history | Diabetes | First month |
|---|---|---|---|---|---|
| Nausea | 1.25 | 0.90 | 1.40 | 0.85 | 2.50 |
| Vomiting | 1.30 | 1.10 | 1.50 | 0.90 | 2.50 |
| Diarrhea | 1.10 | 1.20 | 1.50 | 1.00 | 2.00 |
| Constipation | 1.30 | 1.40 | 1.30 | 1.00 | 1.50 |
| Abdominal pain | 1.15 | 1.20 | 1.80 | 1.00 | 1.80 |
| Acid reflux | 1.10 | 1.30 | 2.00 | 1.00 | 1.50 |
| Reduced appetite | 1.10 | 1.10 | 1.00 | 0.90 | 1.80 |
| Headache | 1.20 | 1.00 | 1.00 | 1.10 | 2.00 |
| Fatigue | 1.15 | 1.30 | 1.00 | 1.10 | 1.80 |
| Dizziness | 1.15 | 1.50 | 1.00 | 1.20 | 1.80 |
| Injection-site reaction | 1.10 | 1.00 | 1.00 | 1.00 | 1.50 |
| Emotional blunting | 1.10 | 1.00 | 1.00 | 1.00 | 1.20 |
| Gallstones | 1.60 | 1.40 | 1.30 | 1.20 | 0.50 |
| Hair loss | 1.40 | 1.20 | 1.00 | 1.00 | 0.30 |
| Pancreatitis | 1.00 | 1.50 | 3.00 | 1.30 | 1.20 |
Things worth noticing, stated as what the table *is* rather than what it proves:
The cap, and why it does not currently bind
Multipliers stack. To stop a five-flag profile from compounding into an absurd number, the engine caps the cumulative log-odds shift at 2.5 — roughly a 12× ceiling on the combined odds ratio — and clamps any displayed estimate to the 1–95% range.
Computing the total shift for every effect with all five flags firing, the largest achievable value across all 15 effects is 1.95 (pancreatitis, for a woman over 65 with GI history and diabetes in her first month). That is comfortably under 2.5, so the cap never binds on today's coefficients. We publish the pre-adjustment rate and the post-adjustment rate as two explicit endpoints — "32% → 37%" — rather than only the multipliers, precisely so the sentence stays true if a future coefficient update ever makes the cap bite.
Two live examples of stacking, both captured on 27 August 2026 for a woman of 70 with GI history, diabetes, and in her first month:
Why we are publishing our own weakest link
Because a bug forced the question, and we would rather answer it in public.
Until 27 August 2026 a scoping defect in our engine was checking for sex-tagged data across the wrong set of records, and the effect was that the female multiplier was being suppressed on nearly every clinical estimate — male and female predictions came out identical on 14 of 15 effects. Fixing that defect was correct. It also, in the same deploy, put an uncited ×1.25 live on every female clinical estimate, sitting immediately beside a basis line that advertised only the 26 rates and 18 sources behind the *unadjusted* half of the number.
So the same day we shipped the disclosure that now runs on the predictor API, both languages of our methodology page, the API help endpoint, and llms.txt: the pre-adjustment rate, every modifier that fired, its odds ratio, and its provenance. This article is the public version of that disclosure.
What it would take to earn these numbers
The honest path is to derive sex, age, and comorbidity effects from data rather than judgement. We checked on 27 August 2026 whether our own corpus can support that yet. It cannot:
You cannot derive a sex-stratified clinical odds ratio from a clinical evidence base in which no record states a sex. Until that changes, our options are to keep the priors and label them honestly, source them individually and drop whatever cannot be sourced, or stop applying them — which would make male and female clinical predictions identical again. That choice is open, and it is not one an automated pipeline should make quietly.
This is a third bar, not the other two
Our evidence series uses some words in narrow senses, so to be explicit about which threshold this article is and is not about:
Check it yourself
Every figure in this article was read from the live production API or database on 27 August 2026 and is stated as of that date; the corpus grows daily, so the endpoint is the current truth and this page is a snapshot.
This is educational content about how a statistical tool is built, not medical advice, and none of the numbers here should be used to decide whether to start, continue, or stop a medication — that is a conversation with a clinician who knows your history.
Query the whole dataset — all 15 effects, both tracks, every source with its URL — at magistra.health/en/data-api.
Bekijk uw eigen cijfers
Onze gratis voorspeller schat uw bijwerkingsrisico en gewichtsverloop, met bij elk cijfer het aantal vermelde percentages en afzonderlijke bronnen. Geen account nodig.
Open de voorspeller