Skip to content
Try Settle risk-free. Love it, or get your money back.
Support That Takes You Seriously
Always Free Delivery

How to Read a Supplement Study: A Practical Guide for Women Who Want Real Answers

You have probably seen it a hundred times. A brand posts something like "clinically proven" or "studies show" and suddenly a supplement feels legitimate, trustworthy, even necessary. But here is the uncomfortable truth: most of us have no idea how to check whether those claims actually hold up, and the people selling supplements are counting on that.

This guide is not going to tell you what to take or what to avoid. Instead, it is going to do something more useful. It is going to teach you how to think about supplement research yourself, so you never have to take anyone's word for it again.

You will learn how different types of studies work, why sample size matters more than you might think, and why the words "study participants" might be hiding something important. You will also learn how to spot the patterns that turn modest research findings into sweeping marketing claims.

By the end, you will have a practical checklist you can use on any supplement, any claim, any time. No science degree required. Just a willingness to ask better questions.

You Deserve Better Than 'Studies Show'

You're standing in a shop, or scrolling at midnight, and you flip a supplement to the back. There it is: "Clinically studied." Maybe even a percentage, or a reference to research. And you feel that pull, that genuine desire to make a good decision, followed almost immediately by the deflating realisation that you have no idea whether what you're reading actually means anything.

That frustration is completely reasonable, and it is not yours to carry alone.

Research shows that 9 in 10 adults need additional help interpreting health information, and that is before clinical research enters the picture. Understanding a study requires a different layer of skills again: knowing what a trial design means, whether the sample size holds up, and whether the people studied were anything like you. Nobody teaches this. It is not on school curricula. It is not explained on labels. The gap is a system failure, not a personal one.

If you are navigating perimenopause or menopause, you are already doing this on hard mode. You are intelligent, motivated, and trying to make decisions in an area where the research has historically underserved women, and where the marketing is relentless. The problem is not your ability to understand. The problem is that the tools have never been handed to you.

This guide hands them over.

By the time you reach the end, you will have a practical framework for evaluating any supplement study you encounter, plus a six-question checklist you can return to whenever a new claim crosses your path. It connects to the broader question of how to choose a supplement worth taking, and it sits at the heart of what we believe at Stillen: that you deserve more than just coping, and that starts with real, honest information.

This is not a lecture. Think of it as a conversation between equals, because your autonomy is the whole point.

Why Supplement Research Is Not the Same as Drug Research

Here is the thing that rarely gets explained: a pharmaceutical drug cannot reach a pharmacy shelf without first proving, in large-scale, long-running, independently scrutinised trials, that it actually works. Supplements operate under an entirely different set of rules. In both the UK and the US, supplements do not need to demonstrate efficacy before they go on sale. Safety reporting matters; proof of effect does not. That gap is where a lot of confusion begins, and it is worth understanding before you read a single study.

The structural differences show up in the research itself. Pharmaceutical trials typically involve thousands of participants, run for years, and are designed to meet a high regulatory bar. Supplement studies often involve far fewer people, sometimes fewer than 50, run for weeks rather than months, and face no equivalent design requirements. That does not make them worthless. It does mean you are working with a narrower and less certain type of evidence, and it is reasonable to weigh it accordingly.

There is also a gap that matters specifically to you. Historically, a large proportion of clinical research, including supplement research, has been conducted predominantly on male participants, or on mixed groups without ever reporting outcomes separately by sex. If a study does not tell you what happened in women, particularly women in hormonal transition, the headline finding may simply not apply to your body. This is not a minor detail. It is central to whether the research is relevant to you at all. You can read more about what this means in a UK context in Before You Start Supplementing: The UK Context Worth Knowing.

None of this is a reason to dismiss supplement research outright. It is context. And context is exactly what helps you ask the right questions rather than defaulting to either blind trust or blanket scepticism.

The Evidence Pyramid: Not All Studies Are Created Equal

The Evidence Pyramid: Not All Studies Are Created Equal

So now you understand why supplement research has limitations. The next step is understanding how to rank whatever evidence does exist.

Think of research designs as sitting on a ladder, with the most reliable methods at the top and the least reliable at the bottom. This is called the evidence hierarchy, and knowing where a study sits tells you a great deal about how much weight to give its findings.

Here is how it breaks down, from the bottom up:

  • Anecdotes and case reports: One person's experience, compelling but not generalisable

  • Expert opinion and editorials: Informed commentary, not original data

  • Observational studies (cross-sectional and case-control): These look at patterns across groups, but cannot prove that one thing caused another

  • Cohort studies: These follow groups of people over time, which is stronger, but participants are not randomly assigned to anything

  • Randomised controlled trials (RCTs): Participants are randomly assigned to either receive the supplement or a placebo. Random assignment matters because it removes the most common source of distortion: our tendency, conscious or not, to see what we expect to see

  • Systematic reviews and meta-analyses: These sit at the top, combining results from multiple studies to produce a more reliable overall estimate

A meta-analysis is only as good as the studies feeding into it, so it still requires critical reading. But it carries far more weight than any single trial in isolation.

Here is why this matters in practice. When a supplement brand states "a study shows magnesium improves sleep," that could mean almost anything. It might be a small observational study of 30 people conducted over four weeks. Or it might be a meta-analysis pooling results from multiple RCTs. These are not equivalent, and the label will rarely tell you which one it is.

Knowing the hierarchy means you can ask the right question immediately: what kind of study is this, actually?

Sample Size and Statistical Significance: What the Numbers Actually Mean

Knowing what type of study you're looking at is one thing. Knowing whether it involved enough people to actually mean something is another question entirely.

Think of it this way: if you asked five friends whether they liked a new restaurant, their answers would be interesting but not exactly reliable. Ask five hundred people, and you start to see a real pattern. Research works the same way. The more participants a study includes, the less likely a result is down to chance, and the more confidently you can apply the finding to people in general.

As a rough guide: a study of around 20 people tells you something worth noticing. A study of 200 people tells you something more reliable. A study of 2,000 people tells you something you can genuinely start to act on. Many supplement studies sit firmly at the lower end of that range, sometimes fewer than 50 participants, which is worth keeping in mind when a brand presents findings as settled science.

The p-value, explained without the maths

You may have seen the phrase "statistically significant" on a supplement website and assumed it meant the effect was meaningful. It does not, quite.

Statistical significance is expressed as a "p-value": the probability that a result happened purely by chance rather than because of the intervention. A p-value below 0.05 is the standard threshold, meaning there is less than a 5% chance the result was random. That sounds reassuring, but it only tells you the result probably exists. It says nothing about whether it is large enough to matter in your life.

Effect size is the number that actually matters

This is where effect size comes in. A supplement might produce a statistically significant improvement in sleep quality scores, say 0.3 points on a 10-point scale. Technically real. Practically negligible.

You do not need to calculate anything. Just ask one question: how big was the actual difference, and would I notice it? A brand citing honest evidence should be willing to answer that. If they lead with "significant results" but never tell you what changed in real terms, that is worth noting.

Who Was Actually in the Study? Why This Is the Most Important Question You Can Ask

Even a large, well-designed study with impressive numbers means very little if the people in it are nothing like you.

Every study is conducted on a specific group of people, and its findings most reliably apply to people similar to those who participated. Age, sex, hormonal status, health background, and even geography all shape how the body responds to an intervention. When those factors differ significantly from your own situation, the results become much less relevant to you personally.

This matters more for women than most people realise. The FDA Office of Women's Health has documented persistent underrepresentation of women in clinical trials. Historically, women were routinely excluded from early research, and that data gap has never been fully closed. Many supplement studies today still enrol only male participants, or include mixed groups without reporting outcomes separately by sex. A headline finding drawn from a male-only study may simply not translate to female physiology.

Hormonal context adds another layer entirely. Consider a magnesium and sleep study conducted in healthy young men, or even in postmenopausal women. Neither group is likely to reflect the experience of a woman in her mid-forties whose sleep is being disrupted by fluctuating oestrogen. Perimenopause involves a specific and shifting hormonal environment that fundamentally changes how the body responds. Understanding why women in their 40s and 50s may be particularly affected by these gaps in the research helps explain why population match matters so much.

When you open a study, go straight to the "Participants" or "Methods" section and look for three things: the average age and age range, the sex breakdown, and whether participants had the same symptoms or conditions you are trying to address. If that information is missing or the population is a poor match, note it.

Finally, finding that a study does not apply to you is not a failure of understanding. It is precisely what good critical reading looks like. It does not mean the supplement is ineffective; the bottom line is simply that you need better-matched evidence before drawing conclusions about your own situation.

What Was Actually Measured? Outcome Measures and Why They Matter

So you've confirmed the study used the right participants. Good. Now ask the next question: what did it actually measure?

Every study targets a specific outcome, and the outcome chosen matters enormously. It might be a blood test result, a score on a questionnaire, or something participants reported feeling. These are not interchangeable, and the gap between them is where a lot of marketing sleight-of-hand happens.

Surrogate outcomes vs. meaningful outcomes

A surrogate outcome is a biological marker that suggests something useful is happening, rather than measuring it directly. A magnesium study might measure serum magnesium levels (the amount of magnesium in the bloodstream) and show those levels rose after supplementation. That is interesting, and it tells you absorption is occurring. What it does not tell you is whether participants slept better, woke less at 3am, or stopped feeling tired but wired by mid-afternoon. The blood marker and the lived experience are related questions, but they are not the same question.

This distinction matters because supplement marketing often bridges that gap without acknowledging it exists. "Shown to support magnesium levels" becomes "supports restful sleep" in the headline. Technically the study showed the former. The leap to the latter is the brand's, not the evidence's.

Why validated questionnaires are worth looking for

When a study does measure subjective experience, the quality of that measurement varies considerably. Well-designed studies use validated tools; standardised questionnaires that have been tested across populations and refined to minimise bias. The Pittsburgh Sleep Quality Index, for example, is widely used in sleep research precisely because its reliability is established and its scores can be meaningfully compared across studies.

A custom questionnaire written for a single study has none of those guarantees. Its findings are harder to compare and more vulnerable to the researchers seeing what they hoped to see.

The question to keep asking

Before trusting a supplement claim, ask plainly: did this study measure what I actually want to know? If the answer is a blood marker rather than a symptom, or a custom survey rather than a validated tool, the claim deserves more scepticism than the headline suggests.

Who Paid for the Study? Understanding Funding and Conflicts of Interest

Once you know what a study measured, there is one more question worth asking: who wanted it to find what it found?

A conflict of interest, in plain terms, is when the organisation funding a study has a financial stake in the outcome. This is not a conspiracy theory; it is a documented pattern in nutrition and supplement research. A Cochrane systematic review found that industry-sponsored trials were 27% more likely to report results favouring the sponsor's product, and 34% more likely to draw favourable conclusions. The money does not corrupt every study, but it does create a consistent tilt.

Where to find this information

Scroll to the bottom of any published study and look for a section labelled "Funding," "Acknowledgements," or "Declarations." That is where researchers are required to disclose who paid for the work. If no funding is disclosed, that is itself worth noting. As a general rule, studies funded by independent bodies, such as universities, government agencies, or non-profit research institutes, carry more weight than those funded by the company selling the product being tested.

Industry funding is not an automatic disqualifier

Plenty of rigorous, genuinely useful research is funded by industry. The question is not purely who paid; it is whether the study design, analysis, and reporting hold up regardless of who paid. A well-designed, independently audited industry-funded trial can be more trustworthy than a poorly designed academic one.

Transparency is itself a signal

A supplement brand that links to specific studies, names the study design and sample size, and acknowledges limitations is behaving very differently to one that says "clinically proven" with no reference. The willingness to be scrutinised is meaningful.

The goal here is informed scepticism, not blanket distrust. Most researchers are genuinely trying to do careful work. The supplement market's looser regulatory environment simply means more of the judgement falls to you, which is exactly why this framework exists.

How Supplement Claims Get Stretched: Common Patterns to Recognise

Even when you know who funded a study, there is still another layer to work through: how the findings get translated into marketing language. This is where things get stretched, sometimes subtly enough that it is easy to miss.

Cherry-picking is the most common pattern. A brand finds one study with positive results and presents it as though the science is settled, without mentioning the four other studies that found no effect. No single claim is technically false. You are simply not seeing the full picture. If you want to know whether a body of evidence is genuinely consistent, searching for what the evidence actually says across multiple studies is worth the extra step.

Conflating correlation with causation often appears in observational research. "Women who take magnesium supplements report better sleep" sounds compelling, but it does not tell you that magnesium caused the improvement. Women who take supplements may also exercise more, eat more varied diets, or prioritise sleep hygiene. The association is real; the cause is unproven.

Exaggerating effect sizes is where "statistically significant" becomes quietly misleading. A result can be statistically real and practically negligible. A significant improvement on a sleep questionnaire might represent a difference too small to feel in daily life. Marketing copy almost never makes this distinction.

Applying group averages to individuals is easy to overlook even in well-designed research. A rigorous RCT of 500 people tells you what happened on average. Your response will be shaped by your hormonal status, stress load, diet, and individual biology. Average outcomes are a useful starting point, not a personal guarantee.

Using lab research to support human claims is perhaps the subtlest pattern. A study showing that magnesium interacts with GABA receptors in a laboratory setting is genuinely interesting science. It does not confirm that swallowing a supplement will calm your nervous system. The distance between a cell-level mechanism and a lived human outcome is significant, and brands do not always acknowledge it.

Your Practical Checklist: Six Questions to Ask Before Trusting a Supplement Study

Now that you can recognise the patterns, here is the framework to apply whenever you encounter a supplement claim. Keep these six questions somewhere handy.

1. What type of study is this? Find out where it sits on the evidence pyramid. A single case report or observational study is interesting but carries much less weight than a randomised controlled trial (RCT) or a meta-analysis that pools results across multiple studies. When a brand says "research shows," the type of research matters enormously.

2. How many people were in it? Under 50 participants means treat findings with real caution; over 200, particularly in an RCT, is more meaningful. Check the methods section to see whether the researchers justified their sample size. If they did not, that is worth noting.

3. Were women like me in this study? Check the participants section for sex, age range, and hormonal status. A study conducted in healthy young men, or exclusively in postmenopausal women, may not translate to someone in perimenopause whose sleep disruption is hormonally driven. If you cannot find this information, that absence tells you something too.

4. What was actually measured, and does it match what I care about? A study measuring serum magnesium levels in the blood is asking a different question than one measuring whether participants woke less frequently or felt more rested. Ask whether the outcome the researchers measured is the outcome the brand is now claiming. If the claim is broader than what was studied, it has been stretched.

5. Who funded it? Look for a funding or acknowledgements section. Independent funding from a university or government body adds credibility. Industry funding does not automatically disqualify a study, but it warrants closer scrutiny of the design and how results were reported.

6. How big was the actual difference? "Statistically significant" does not mean meaningfully significant. Look for the actual numbers: how much did the sleep score improve, and by how much? Ask honestly whether you would notice that change in your daily life. If the study or the brand does not answer that question clearly, be sceptical.

Putting It Into Practice: What Good Evidence Actually Looks Like

Putting It Into Practice: What Good Evidence Actually Looks Like

Now that you have the six questions, here is what you are actually looking for when you apply them.

Good evidence for a supplement does not mean one perfect trial. It means a reasonable body of converging research: multiple independent studies, findings that hold across different populations, at least some RCT evidence in the mix, and honest reporting of what the research did not show as well as what it did. When several imperfect studies point in the same direction, that matters. When only one study does, that matters too.

The perfect trial rarely exists, and this is especially true in women's health. Nutritional supplement research works within real constraints: smaller samples, shorter durations, and decades of underrepresentation of women in clinical research. Expecting pharmaceutical-grade certainty before making any decision will leave you permanently stuck. That is not a reasonable standard for this category of evidence.

What is reasonable is this: the evidence is promising but not yet definitive is a genuinely useful position to hold. It does not mean you cannot act. It means you act with appropriate expectations, not certainty you were never entitled to. That is honest. That is enough.

When Stillen formulated Settle, the starting point was the available evidence for magnesium's role in nervous system support and sleep regulation. That evidence is not flawless, but it is meaningful, and Stillen names its sources rather than hiding behind phrases like "clinically researched" with nothing to back them up. That transparency is a deliberate choice, because nothing in UK supplement regulation actually requires it. When a brand volunteers specifics, that signals something.

You now have the framework. Use it on every claim you encounter, including ours. That is the point.

Reading the Evidence Is an Act of Self-Respect

None of this requires a science degree. It requires the belief that you deserve accurate information before you spend your money or put something in your body, and that belief is worth acting on.

You now have a practical framework for doing exactly that. Whenever you encounter a new supplement claim, come back to the six questions: What type of study is it? How many people were in it? Were women like you included? What was actually measured? Who funded it? And how big was the real-world difference? Six questions. That is all it takes to move from passive reader to informed decision-maker.

The supplement market will keep generating noise. New studies, bold headlines, and confident marketing copy are not going away. But you now have signal. You know how to look past "clinically proven" to ask what that actually means. That changes how you shop, how you read labels, and how much trust you extend to any brand, including this one.

That last part matters. At Stillen, we believe that educating women before we ever ask them to buy something is not a marketing strategy; it is a baseline commitment. Women navigating perimenopause and hormonal changes have spent long enough being handed oversimplified answers to complicated questions. You deserve information that respects your intelligence, names its sources, and tells you honestly what the evidence does and does not show.

That is what honest support looks like.

Conclusion

Reading supplement research is a skill, and you have just built it. You now understand why study type matters, why sample size shapes reliability, why funding sources deserve scrutiny, and why the women in a study may or may not reflect your reality. These are not small things. They are the difference between spending money wisely and being misled by polished marketing.

You do not need to analyse every study perfectly. You need to ask the right six questions consistently, and let the answers guide your decisions.

Start today. The next time a supplement headline catches your attention, open the checklist before you open your wallet. Notice what the study does and does not actually show.

That habit, applied repeatedly, is how you stop being a target for weak evidence and start making genuinely informed choices about your health.

Frequently Asked Questions

Why don't supplements have to prove they work before being sold, unlike pharmaceutical drugs?

Unlike pharmaceutical drugs, supplements operate under a different regulatory framework in both the UK and the US. Supplements do not need to demonstrate efficacy (that they actually work) before reaching store shelves. While safety reporting is required, proof of effect is not. This regulatory gap means supplement companies can sell products without the large-scale, long-running, independently scrutinised trials that pharmaceuticals must pass. This is why understanding supplement research yourself becomes so important.

What does 'statistically significant' mean, and why doesn't it guarantee a supplement will work for me?

Statistical significance (expressed as a p-value below 0.05) means there's less than a 5% chance the study result happened by random chance. However, it only tells you the result probably exists—not whether it's large enough to matter in your actual life. A supplement might produce a statistically significant improvement in sleep quality that's practically negligible (like 0.3 points on a 10-point scale). The number that actually matters is the effect size: the real-world difference you would notice. Always ask: how big was the actual improvement, and would I notice it?

Why is it important whether women were included in supplement studies?

Women have historically been underrepresented in clinical research, including supplement studies. Many studies include only male participants or mix groups without reporting outcomes separately by sex. This is particularly critical for women navigating perimenopause or menopause, as hormonal context fundamentally changes how the body responds to interventions. A study showing magnesium improves sleep in healthy young men or postmenopausal women may not apply to someone in perimenopause with hormonally-driven sleep disruption. Population match determines whether research is relevant to your body.

How can I tell if a supplement brand is being transparent, versus just using marketing phrases like 'clinically proven'?

Transparency signals integrity. Brands that link to specific studies, name the study design and sample size, acknowledge limitations, and explain what the research actually shows are behaving very differently from those that use vague phrases like 'clinically proven' with no references. Look for brands willing to be scrutinised. When a company volunteers specifics about their evidence rather than hiding behind marketing language, that willingness to be questioned is itself meaningful and worth rewarding.

If I find that a study doesn't apply to me, does that mean the supplement doesn't work?

No. Finding that a study doesn't apply to you is not a failure—it's precisely what good critical reading looks like. It simply means that particular evidence isn't relevant to your situation, and you need better-matched research before drawing conclusions about whether it will work for you. A supplement may be effective for the population it was studied in while still being unknown for your circumstances. The bottom line is that you deserve evidence from people like you before making a decision about your own health.

Your cart

Your cart is currently empty.

Not sure where to start?
Try these collections: