Lesson 8.2Lesson 8.2 · The Evidence & the Method
Measuring the Effect
A neuroarchitecture claim usually arrives wrapped in a number - stress fell by a third, recovery was faster, scores rose - and the number is meant to end the argument, but the designer who can read it honestly asks four quieter questions instead: is the effect real or could it be chance, how big is it actually, does the study show cause or only a correlation that a hidden third factor might explain, and who exactly was measured, because a percentage with no sense of size, cause, chance or population is not evidence, it is decoration
'Stress dropped by 30 percent.' It sounds decisive. Four quiet questions decide whether it means anything.
Numbers are persuasive, and neuroarchitecture is full of them: a green wall cut stress by a third, a daylit classroom raised scores by so many points, patients with a view left hospital a day sooner. A number feels like the end of an argument - it looks precise, objective, scientific. But a number on its own is not evidence; it is a headline, and reading it honestly means asking what stands behind it. This is not advanced statistics. It is a small, humane numeracy that any designer can carry, and it changes an impressive figure from something to repeat into something to weigh.
Four questions do most of the work. Is the effect real, or could it be chance - the question of statistical significance? How big is it actually - the question of effect size, which is different and often more important? Does the study show that the space caused the outcome, or only that the two moved together - the question of correlation versus causation, where a hidden third factor often lurks? And who exactly was measured - the question of the study population, which decides whether the finding travels to your people and your place at all. Learn to ask these four, and a percentage stops being decoration and becomes something you can genuinely reason with - or honestly set aside.
Behind every number, 4 questions: 1) significance = real or chance? 2) effect size = how BIG (unpack the %)? 3) correlation != causation (reverse? confounder C -> both?) 4) who was measured (population)? Pass all four = lean on it. Fail several = hypothesis or marketing.
Significance is not size
The most common and most consequential confusion in reading research is treating statistical significance as if it meant importance. They are different questions with different answers. Statistical significance is a test of chance: it asks whether a result is unlikely to have arisen by random fluke if there were really no effect at all. When a study reports that a finding is 'significant' (often via a p-value below 0.05), it is saying, roughly, 'this is probably not just noise.' That is genuinely useful - it is a guard against mistaking randomness for a pattern - but notice what it does not say. It says nothing about how large the effect is, how much it matters in practice, or whether you should change a design because of it.
Here is the trap. With a large enough sample, almost any difference, however trivial, becomes statistically significant, because the test gets more sensitive as numbers grow. A study of ten thousand office workers might find that a certain layout 'significantly' reduces reported stress - and the reduction might be so tiny that no one inside the building would ever feel it. 'Significant' has become a technical word meaning 'probably real', but the everyday ear hears 'big and important', and marketing exploits that gap relentlessly. A headline that says an effect is significant has told you it is probably not chance; it has told you nothing about whether it is worth a single design decision.
The reverse trap matters too. A real and useful effect can fail to reach significance in a small study simply because there were not enough people to rule out chance - so 'not significant' does not always mean 'no effect', it can mean 'not enough data to be sure yet'. This is why a lone small study, significant or not, is fragile, and why replication and larger samples matter so much. The honest reader treats significance as a first gate - is this probably real? - and then immediately asks the question significance cannot answer: how big is it? That second question, effect size, is where the practical meaning lives, and it is the one most often left out of the headline. Ask both. A result that is real but minuscule is a curiosity, not a mandate; only a result that is both probably real and meaningfully large earns a place in a design decision.
Significance = is it REAL (not chance)? Effect size = how BIG? Different questions. With a huge sample even a trivial difference is 'significant'. 'Significant' means probably real, NOT big or important. Always ask both.
How big is the effect, really
Once you know an effect is probably real, the decisive question is how big. Researchers answer it with effect size - a measure of the magnitude of a difference or relationship, independent of sample size - and by convention effect sizes are loosely described as small, medium or large. You do not need the formulas; you need the habit of asking for the size and refusing to be satisfied by 'significant' alone. An effect large enough to change how a space feels or performs is worth designing for. An effect real but tiny is a footnote, whatever the press release says.
The most abused vehicle for hiding size is the bare percentage. 'Reduced stress by 30 percent' sounds dramatic, but it is meaningless until you ask: thirty percent of what baseline, measured how, and across what range? A thirty percent drop on a scale where everyone was barely stressed to begin with may be trivial; the same figure can be inflated by a relative framing that hides a small absolute change. Percentages also travel without their context - a change measured in a single lab session gets quoted as if it were a permanent life effect. Whenever you meet a percentage, mentally translate it back into something concrete: how many actual people, how much actual change, would you notice it if you were standing in the room?
Size also has to be judged against cost and reality. A design move that produces a genuine but small benefit may still be worth it if it is cheap and harmless; the same small benefit is not worth it if it drives an expensive, permanent decision or crowds out something that matters more. This is the difference between statistical significance and practical significance - between 'the effect is probably real' and 'the effect is big enough, in this context, to act on'. The literature can help you with the first; only judgement, applied to your brief and your budget, settles the second. So the discipline is simple to state and rare in practice: never let 'significant' stand in for 'big', always ask for the effect size, always unpack a percentage into something you can picture, and always weigh the size against what acting on it would cost. A number that survives those questions is worth something. Most impressive numbers do not survive them.
Correlation is not causation
The oldest warning in statistics is the one neuroarchitecture most needs, because so much of its evidence comes from field studies where researchers observe rather than control. Correlation does not imply causation. When two things move together - offices with plants have happier staff, daylit classrooms have higher scores, hospital rooms with views have faster recoveries - it is tempting to conclude that the first caused the second. But a correlation is compatible with several very different stories, and only some of them justify a design decision.
There are three explanations to weigh whenever you meet a correlation. The first is the one you hope for: A genuinely causes B - the greenery really does lift mood. The second is reverse causation: B causes A, or the arrow runs the other way than assumed - perhaps happier, better-run firms are the ones that can afford green offices, so success buys the plants rather than the plants buying success. The third, and the most treacherous, is a confounder - a hidden third factor, C, that causes both A and B and creates a correlation between them with no direct link at all. Wealthier organisations tend to have both nicer buildings and better-treated, happier staff; the building and the happiness correlate, but a common cause, resources, is driving both. The daylit classroom may also be the newer, better-funded school with better teachers. The room with a view may be the private room given to less critical patients. Untangling these is exactly what a real building makes hard, because the good features cluster together.
This is why the study design matters so much. A controlled experiment, which randomly assigns people or conditions, is the strongest tool for isolating cause, because randomisation tends to spread confounders evenly. Most neuroarchitecture evidence, though, is observational - it watches the world rather than intervening - and observational findings can strongly suggest causation but rarely prove it alone. That does not make them useless; it makes them claims to hold provisionally and to strengthen through replication, experiments and converging methods. The practical rule for a designer: when you meet 'buildings with X have more Y', do not silently upgrade it to 'X causes Y'. Ask whether the arrow might run backwards, and above all what third factor might be quietly producing both. Often the honest answer is that the causal story is plausible but unproven - which is a fine reason to treat it as a hypothesis worth testing, and a bad reason to sell it as a fact.
Who was measured - and reading the whole claim
The last question is the one most often skipped: who was in the study? Every finding is produced on some specific population, and it travels only as far as that population reasonably extends. A great deal of neuroarchitecture research is run on convenient samples - university students, Western office workers, volunteers who signed up - who are younger, healthier, wealthier and more culturally narrow than the people most buildings serve. An effect found on undergraduates in a North American lab may or may not hold for elderly patients in an Indian hospital, children in a crowded school, or a multigenerational family in a dense city. The finding is not wrong; it is simply local until shown otherwise. Ask who was measured, how many, and how similar they are to the people you are designing for, and downgrade your confidence honestly when the gap is large.
Put the four questions together and you have a compact, humane way to read any neuroarchitecture claim without a statistics degree. Is it probably real, or could it be chance (significance)? How big is it actually, once you unpack the percentage (effect size)? Does the study show cause, or only that two things moved together with a possible third factor lurking (correlation versus causation)? And who was measured, and do they resemble your people (population)? A claim that passes all four - a reasonably large, replicated effect, from a design that supports causation, on a population like yours - is strong evidence you can lean on. A claim that fails several - a tiny effect, from a small observational study on undergraduates, quoted as a dramatic percentage - is a hypothesis at best and marketing at worst.
This numeracy is not about becoming a scientist; it is about self-defence and honesty. It protects you from repeating overstated claims to clients, from spending a budget on a trivial effect, and from the neuro-washing that dresses thin numbers in scientific authority. It also keeps you appropriately humble: even the strongest study speaks to averages, not to the particular person in your particular room, and none of it displaces the codes or the clinical judgement of qualified professionals. Read the number for size, cause, chance and population - and it will tell you the truth, which is usually more modest, and more usable, than the headline.
Significance is not size
Is it real vs does it matter
Statistical significance says an effect is probably not chance; it says nothing about magnitude. With big samples trivial effects turn significant. Always ask effect size next. Modules 8.1, 9.2.
Unpack the percentage
How big, really
A bare percentage hides its baseline, its absolute change and its context. Translate it into concrete people and change you could notice, and weigh practical against statistical significance. Modules 8.3, 9.2.
Correlation is not causation
Cause vs coincidence
Two things moving together may reflect reverse causation or a hidden third factor (confounder). Only controlled designs isolate cause; observational findings are hypotheses to strengthen. Modules 9.2, 8.1.
Ask who was measured
How far a finding travels
Every finding is produced on a specific, often Western or student, population. Downgrade confidence when they differ from your people; leave clinical and health determinations to qualified professionals and the codes. Modules 9.2, 8.4.
Workshop — unpack a number
This workshop turns the four questions into a reflex. Take one neuroarchitecture statistic - ideally one quoted with a percentage or a 'significant' label - and interrogate it until you can say honestly how much weight it can bear.
One quoted statistic, its source, and a notebook. No statistical software - this trains the four questions, not calculation; any clinical or health interpretation stays with qualified professionals and peer-reviewed evidence.
Goal: read one statistic for significance, size, cause and population Inputs: one quoted neuroarchitecture number + its source + a notebook Time: ~40 minutes
- 1Write the statistic exactly as quoted (for example, 'a view of nature reduced reported stress by 30 percent' or 'daylight significantly improved test scores').
- 2Significance and size: does the source say the effect was statistically significant, and does it give an effect size or only a percentage? Unpack the percentage - of what baseline, how measured, would you notice a change that size?
- 3Cause: was the study a controlled experiment or observational? List one plausible reverse-causation story and one plausible confounder (a hidden third factor) that could explain the same result.
- 4Population: who was measured, how many, and how similar are they to the people you design for? Note honestly where the answer is unknown.
- 5Write a one-paragraph verdict: given the four answers, is this strong evidence to lean on, a hypothesis to test locally, or marketing to set aside - and what would raise your confidence?
You’ll walk away with
A one-page 'number unpacked': the statistic quoted, its significance and effect size assessed, one reverse-causation and one confounder story named, the population identified (or marked unknown), and an honest verdict on the weight it can bear. Reuse it as a checklist whenever a striking figure appears.
Three altitudes on the same idea
Read the band that fits you — or all three.
The numbers that reach you - from product reps, sustainability consultants, research summaries and your own post-occupancy data - will drive real money and permanent form, so the four questions are professional self-defence, not academic nicety. Before a statistic justifies an atrium, a facade system or a wellness feature, ask: is it probably real or could it be chance (significance); how big is it once the percentage is unpacked into something you would notice (effect size); does the study show the space caused the outcome or only that they correlated, with a possible confounder like budget or building age driving both (correlation versus causation); and who was measured, and do they resemble your occupants (population). A large, replicated, causally-supported effect on a similar population is worth building for. A tiny observational effect on undergraduates quoted as a dramatic percentage is not. Weigh statistical against practical significance - a real but small effect may not be worth a costly permanent decision. And leave every clinical, health-outcome and accessibility determination to qualified professionals, peer-reviewed evidence and the codes; your numeracy serves judgement, not certainty.
Interiors are sold with the most seductive numbers - this paint colour cuts stress by so much, this lighting lifts productivity by so many points - so the sensory designer needs the numeracy to unpack a percentage before it becomes a promise to a client. When a supplier or article quotes a figure, translate it: percent of what baseline, measured how, on how many people, and would anyone actually feel a change that size? Ask whether the study showed cause or just correlation - the calmer room in the study may also have been the quieter, greener, better-furnished one, so a confounder may be doing the work your product is being credited with. Ask who was measured, because a lab result on students may not hold for the family, the patient or the elder you are designing for. Lean on the broad, well-evidenced sensory principles (daylight, nature, quiet, warm materials) and treat dramatic specific statistics as hypotheses, not facts. And keep any health or clinical claim with the specialists and the evidence - honest numeracy protects your credibility and your clients.
A little honest numeracy is one of the highest-leverage skills you can carry into design practice, because it lets you read the studies everyone else only quotes - and it is exactly what marks out a rigorous graduate from one who repeats trendy statistics. Learn the four questions and use them on every claim you meet. Statistical significance asks whether an effect is probably real (not chance); effect size asks how big it actually is - and the two are different, so 'significant' never means 'important' on its own. Correlation versus causation asks whether the space caused the outcome or whether a hidden third factor (often money, age or building quality) produced both. Population asks who was measured, since most studies run on students or Western samples that may not generalise to India. Practise unpacking percentages into concrete, picture-able changes. You do not need advanced statistics - you need the habit of asking these four questions, which protects you from overclaiming and neuro-washing, and always leaves clinical and health determinations to qualified professionals and peer-reviewed evidence.
“The study found a statistically significant effect, and it reported a big percentage improvement, so the result is proven and important - the design clearly works, and I can tell my client this feature will deliver that improvement.”
Do it yourself
No tools needed — reason it through.
- 1Explain the difference between statistical significance and effect size, and why 'significant' does not mean 'important'.
- 2A study of 10,000 people finds a 'significant' 1 percent change. Why might a designer still ignore it?
- 3Offices with plants have happier staff. Give one reverse-causation and one confounder explanation besides 'plants cause happiness'.
- 4Why does the study population matter when deciding whether a finding applies to your project in India?
- 5Turn 'reduced stress by 30 percent' into the questions you would ask before believing it.
The one line to carry out
Peer-reviewed journals & authoritative standards
- 01Effect size — Wikipedia — Effect size, 2026.
- 02Correlation does not imply causation — Wikipedia — Correlation does not imply causation, 2026.
- 03Reproducibility — Wikipedia — Reproducibility, 2026.
- 04Peer review — Wikipedia — Peer review, 2026.
- 05Replication crisis — Wikipedia — Replication crisis, 2026.
Now you can judge how trustworthy and how large a finding is - but a finding is not a design. The next lesson takes the honest leap from evidence to a specific decision: why the research rarely gives a formula, and how disciplined judgement turns principles into a real building.
The author
Amogh N P
Architect, interior designer, and creative polymath. Studio Matrx began in his notebooks — his vision of design made honest, useful, and open to everyone. Its Academy is written and taught in his memory, and free, forever.
More about Amogh →