Studio Matrx Monthly · Volume 1 · Issue 4 · September 2026
Amogh N P
 In loving memory of Amogh N P — Architect · Designer · Visionary 
Measuring the EffectLesson 8.2
Neuroarchitecture & Design for the Brain/Module 8 · The Evidence & the Method

Lesson 8.2 · The Evidence & the Method

Measuring the Effect

A neuroarchitecture claim usually arrives wrapped in a number - stress fell by a third, recovery was faster, scores rose - and the number is meant to end the argument, but the designer who can read it honestly asks four quieter questions instead: is the effect real or could it be chance, how big is it actually, does the study show cause or only a correlation that a hidden third factor might explain, and who exactly was measured, because a percentage with no sense of size, cause, chance or population is not evidence, it is decoration

12 min Interactive lessonFree · open lessonByAmogh N P· Architect & interior designer
The hook

'Stress dropped by 30 percent.' It sounds decisive. Four quiet questions decide whether it means anything.

Numbers are persuasive, and neuroarchitecture is full of them: a green wall cut stress by a third, a daylit classroom raised scores by so many points, patients with a view left hospital a day sooner. A number feels like the end of an argument - it looks precise, objective, scientific. But a number on its own is not evidence; it is a headline, and reading it honestly means asking what stands behind it. This is not advanced statistics. It is a small, humane numeracy that any designer can carry, and it changes an impressive figure from something to repeat into something to weigh.

Four questions do most of the work. Is the effect real, or could it be chance - the question of statistical significance? How big is it actually - the question of effect size, which is different and often more important? Does the study show that the space caused the outcome, or only that the two moved together - the question of correlation versus causation, where a hidden third factor often lurks? And who exactly was measured - the question of the study population, which decides whether the finding travels to your people and your place at all. Learn to ask these four, and a percentage stops being decoration and becomes something you can genuinely reason with - or honestly set aside.

Behind every number, 4 questions: 1) significance = real or chance? 2) effect size = how BIG (unpack the %)? 3) correlation != causation (reverse? confounder C -> both?) 4) who was measured (population)? Pass all four = lean on it. Fail several = hypothesis or marketing.

Significance is not size

The most common and most consequential confusion in reading research is treating statistical significance as if it meant importance. They are different questions with different answers. Statistical significance is a test of chance: it asks whether a result is unlikely to have arisen by random fluke if there were really no effect at all. When a study reports that a finding is 'significant' (often via a p-value below 0.05), it is saying, roughly, 'this is probably not just noise.' That is genuinely useful - it is a guard against mistaking randomness for a pattern - but notice what it does not say. It says nothing about how large the effect is, how much it matters in practice, or whether you should change a design because of it.

Here is the trap. With a large enough sample, almost any difference, however trivial, becomes statistically significant, because the test gets more sensitive as numbers grow. A study of ten thousand office workers might find that a certain layout 'significantly' reduces reported stress - and the reduction might be so tiny that no one inside the building would ever feel it. 'Significant' has become a technical word meaning 'probably real', but the everyday ear hears 'big and important', and marketing exploits that gap relentlessly. A headline that says an effect is significant has told you it is probably not chance; it has told you nothing about whether it is worth a single design decision.

The reverse trap matters too. A real and useful effect can fail to reach significance in a small study simply because there were not enough people to rule out chance - so 'not significant' does not always mean 'no effect', it can mean 'not enough data to be sure yet'. This is why a lone small study, significant or not, is fragile, and why replication and larger samples matter so much. The honest reader treats significance as a first gate - is this probably real? - and then immediately asks the question significance cannot answer: how big is it? That second question, effect size, is where the practical meaning lives, and it is the one most often left out of the headline. Ask both. A result that is real but minuscule is a curiosity, not a mandate; only a result that is both probably real and meaningfully large earns a place in a design decision.

Two different questions, often confused Significance = is it real? Effect size = how big? small effect large effect real but tiny - may not matter in practice real and big - worth designing for not significant - no reliable effect shown big but not significant - could be noise; needs more data significant not sig. A tiny effect can be statistically significant and still not worth a rupee.
Zoom
Significance and size are different questions: significance asks whether an effect is probably real, size asks whether it is big enough to matter - and a tiny effect can be statistically significant while being worthless in practice.

Significance = is it REAL (not chance)? Effect size = how BIG? Different questions. With a huge sample even a trivial difference is 'significant'. 'Significant' means probably real, NOT big or important. Always ask both.

How big is the effect, really

Once you know an effect is probably real, the decisive question is how big. Researchers answer it with effect size - a measure of the magnitude of a difference or relationship, independent of sample size - and by convention effect sizes are loosely described as small, medium or large. You do not need the formulas; you need the habit of asking for the size and refusing to be satisfied by 'significant' alone. An effect large enough to change how a space feels or performs is worth designing for. An effect real but tiny is a footnote, whatever the press release says.

The most abused vehicle for hiding size is the bare percentage. 'Reduced stress by 30 percent' sounds dramatic, but it is meaningless until you ask: thirty percent of what baseline, measured how, and across what range? A thirty percent drop on a scale where everyone was barely stressed to begin with may be trivial; the same figure can be inflated by a relative framing that hides a small absolute change. Percentages also travel without their context - a change measured in a single lab session gets quoted as if it were a permanent life effect. Whenever you meet a percentage, mentally translate it back into something concrete: how many actual people, how much actual change, would you notice it if you were standing in the room?

Size also has to be judged against cost and reality. A design move that produces a genuine but small benefit may still be worth it if it is cheap and harmless; the same small benefit is not worth it if it drives an expensive, permanent decision or crowds out something that matters more. This is the difference between statistical significance and practical significance - between 'the effect is probably real' and 'the effect is big enough, in this context, to act on'. The literature can help you with the first; only judgement, applied to your brief and your budget, settles the second. So the discipline is simple to state and rare in practice: never let 'significant' stand in for 'big', always ask for the effect size, always unpack a percentage into something you can picture, and always weigh the size against what acting on it would cost. A number that survives those questions is worth something. Most impressive numbers do not survive them.

How big is the effect, really? small medium large barely noticeable changes the experience A percentage change means nothing until you ask: of what baseline, for whom, and is the difference large enough to feel? 'Statistically significant' says it is probably real; only size says it matters.
Zoom
Effect size runs from barely-noticeable to experience-changing; a bare percentage means nothing until you ask of what baseline, for whom, and whether the difference is large enough to feel.
The classic trap

Correlation is not causation

The oldest warning in statistics is the one neuroarchitecture most needs, because so much of its evidence comes from field studies where researchers observe rather than control. Correlation does not imply causation. When two things move together - offices with plants have happier staff, daylit classrooms have higher scores, hospital rooms with views have faster recoveries - it is tempting to conclude that the first caused the second. But a correlation is compatible with several very different stories, and only some of them justify a design decision.

There are three explanations to weigh whenever you meet a correlation. The first is the one you hope for: A genuinely causes B - the greenery really does lift mood. The second is reverse causation: B causes A, or the arrow runs the other way than assumed - perhaps happier, better-run firms are the ones that can afford green offices, so success buys the plants rather than the plants buying success. The third, and the most treacherous, is a confounder - a hidden third factor, C, that causes both A and B and creates a correlation between them with no direct link at all. Wealthier organisations tend to have both nicer buildings and better-treated, happier staff; the building and the happiness correlate, but a common cause, resources, is driving both. The daylit classroom may also be the newer, better-funded school with better teachers. The room with a view may be the private room given to less critical patients. Untangling these is exactly what a real building makes hard, because the good features cluster together.

This is why the study design matters so much. A controlled experiment, which randomly assigns people or conditions, is the strongest tool for isolating cause, because randomisation tends to spread confounders evenly. Most neuroarchitecture evidence, though, is observational - it watches the world rather than intervening - and observational findings can strongly suggest causation but rarely prove it alone. That does not make them useless; it makes them claims to hold provisionally and to strengthen through replication, experiments and converging methods. The practical rule for a designer: when you meet 'buildings with X have more Y', do not silently upgrade it to 'X causes Y'. Ask whether the arrow might run backwards, and above all what third factor might be quietly producing both. Often the honest answer is that the causal story is plausible but unproven - which is a fine reason to treat it as a hypothesis worth testing, and a bad reason to sell it as a fact.

They move together. Why? A (green offices) B (happier staff) they correlate But watch for three explanations: 1. A -> B the green really lifts mood (what we hope) 2. B -> A happy firms can afford green offices (reverse) C (rich firm) 3. C -> A and C -> B : a hidden third factor drives both
Zoom
When two things move together, three stories compete: the feature caused the outcome, the arrow runs in reverse, or a hidden third factor caused both - and only the last two are ruled out by careful study design.

Who was measured - and reading the whole claim

The last question is the one most often skipped: who was in the study? Every finding is produced on some specific population, and it travels only as far as that population reasonably extends. A great deal of neuroarchitecture research is run on convenient samples - university students, Western office workers, volunteers who signed up - who are younger, healthier, wealthier and more culturally narrow than the people most buildings serve. An effect found on undergraduates in a North American lab may or may not hold for elderly patients in an Indian hospital, children in a crowded school, or a multigenerational family in a dense city. The finding is not wrong; it is simply local until shown otherwise. Ask who was measured, how many, and how similar they are to the people you are designing for, and downgrade your confidence honestly when the gap is large.

Put the four questions together and you have a compact, humane way to read any neuroarchitecture claim without a statistics degree. Is it probably real, or could it be chance (significance)? How big is it actually, once you unpack the percentage (effect size)? Does the study show cause, or only that two things moved together with a possible third factor lurking (correlation versus causation)? And who was measured, and do they resemble your people (population)? A claim that passes all four - a reasonably large, replicated effect, from a design that supports causation, on a population like yours - is strong evidence you can lean on. A claim that fails several - a tiny effect, from a small observational study on undergraduates, quoted as a dramatic percentage - is a hypothesis at best and marketing at worst.

This numeracy is not about becoming a scientist; it is about self-defence and honesty. It protects you from repeating overstated claims to clients, from spending a budget on a trivial effect, and from the neuro-washing that dresses thin numbers in scientific authority. It also keeps you appropriately humble: even the strongest study speaks to averages, not to the particular person in your particular room, and none of it displaces the codes or the clinical judgement of qualified professionals. Read the number for size, cause, chance and population - and it will tell you the truth, which is usually more modest, and more usable, than the headline.

Two different questions, often confused Significance = is it real? Effect size = how big? small effect large effect real but tiny - may not matter in practice real and big - worth designing for not significant - no reliable effect shown big but not significant - could be noise; needs more data significant not sig. A tiny effect can be statistically significant and still not worth a rupee.
Zoom
Significance and size are different questions: significance asks whether an effect is probably real, size asks whether it is big enough to matter - and a tiny effect can be statistically significant while being worthless in practice.
Verify-this: the four questions behind every number

Significance is not size

Is it real vs does it matter

Statistical significance says an effect is probably not chance; it says nothing about magnitude. With big samples trivial effects turn significant. Always ask effect size next. Modules 8.1, 9.2.

Unpack the percentage

How big, really

A bare percentage hides its baseline, its absolute change and its context. Translate it into concrete people and change you could notice, and weigh practical against statistical significance. Modules 8.3, 9.2.

Correlation is not causation

Cause vs coincidence

Two things moving together may reflect reverse causation or a hidden third factor (confounder). Only controlled designs isolate cause; observational findings are hypotheses to strengthen. Modules 9.2, 8.1.

Ask who was measured

How far a finding travels

Every finding is produced on a specific, often Western or student, population. Downgrade confidence when they differ from your people; leave clinical and health determinations to qualified professionals and the codes. Modules 9.2, 8.4.

Hands-on workshop

Workshop — unpack a number

This workshop turns the four questions into a reflex. Take one neuroarchitecture statistic - ideally one quoted with a percentage or a 'significant' label - and interrogate it until you can say honestly how much weight it can bear.

One quoted statistic, its source, and a notebook. No statistical software - this trains the four questions, not calculation; any clinical or health interpretation stays with qualified professionals and peer-reviewed evidence.

Given & goal
Goal: read one statistic for significance, size, cause and population
Inputs: one quoted neuroarchitecture number + its source + a notebook
Time: ~40 minutes
  1. 1Write the statistic exactly as quoted (for example, 'a view of nature reduced reported stress by 30 percent' or 'daylight significantly improved test scores').
  2. 2Significance and size: does the source say the effect was statistically significant, and does it give an effect size or only a percentage? Unpack the percentage - of what baseline, how measured, would you notice a change that size?
  3. 3Cause: was the study a controlled experiment or observational? List one plausible reverse-causation story and one plausible confounder (a hidden third factor) that could explain the same result.
  4. 4Population: who was measured, how many, and how similar are they to the people you design for? Note honestly where the answer is unknown.
  5. 5Write a one-paragraph verdict: given the four answers, is this strong evidence to lean on, a hypothesis to test locally, or marketing to set aside - and what would raise your confidence?

You’ll walk away with
A one-page 'number unpacked': the statistic quoted, its significance and effect size assessed, one reverse-causation and one confounder story named, the population identified (or marked unknown), and an honest verdict on the weight it can bear. Reuse it as a checklist whenever a striking figure appears.

The worked example

Three altitudes on the same idea

Read the band that fits you — or all three.

For the architectDesigning buildings that support the brain, mind and wellbeing - on evidence, humbly

The numbers that reach you - from product reps, sustainability consultants, research summaries and your own post-occupancy data - will drive real money and permanent form, so the four questions are professional self-defence, not academic nicety. Before a statistic justifies an atrium, a facade system or a wellness feature, ask: is it probably real or could it be chance (significance); how big is it once the percentage is unpacked into something you would notice (effect size); does the study show the space caused the outcome or only that they correlated, with a possible confounder like budget or building age driving both (correlation versus causation); and who was measured, and do they resemble your occupants (population). A large, replicated, causally-supported effect on a similar population is worth building for. A tiny observational effect on undergraduates quoted as a dramatic percentage is not. Weigh statistical against practical significance - a real but small effect may not be worth a costly permanent decision. And leave every clinical, health-outcome and accessibility determination to qualified professionals, peer-reviewed evidence and the codes; your numeracy serves judgement, not certainty.

For the interior designerThe sensory, restorative, mood-shaping interior - the environment closest to the body

Interiors are sold with the most seductive numbers - this paint colour cuts stress by so much, this lighting lifts productivity by so many points - so the sensory designer needs the numeracy to unpack a percentage before it becomes a promise to a client. When a supplier or article quotes a figure, translate it: percent of what baseline, measured how, on how many people, and would anyone actually feel a change that size? Ask whether the study showed cause or just correlation - the calmer room in the study may also have been the quieter, greener, better-furnished one, so a confounder may be doing the work your product is being credited with. Ask who was measured, because a lab result on students may not hold for the family, the patient or the elder you are designing for. Lean on the broad, well-evidenced sensory principles (daylight, nature, quiet, warm materials) and treat dramatic specific statistics as hypotheses, not facts. And keep any health or clinical claim with the specialists and the evidence - honest numeracy protects your credibility and your clients.

For the studentHow space shapes the mind - the science, the humility, and what it means for design

A little honest numeracy is one of the highest-leverage skills you can carry into design practice, because it lets you read the studies everyone else only quotes - and it is exactly what marks out a rigorous graduate from one who repeats trendy statistics. Learn the four questions and use them on every claim you meet. Statistical significance asks whether an effect is probably real (not chance); effect size asks how big it actually is - and the two are different, so 'significant' never means 'important' on its own. Correlation versus causation asks whether the space caused the outcome or whether a hidden third factor (often money, age or building quality) produced both. Population asks who was measured, since most studies run on students or Western samples that may not generalise to India. Practise unpacking percentages into concrete, picture-able changes. You do not need advanced statistics - you need the habit of asking these four questions, which protects you from overclaiming and neuro-washing, and always leaves clinical and health determinations to qualified professionals and peer-reviewed evidence.

Misconception check

The study found a statistically significant effect, and it reported a big percentage improvement, so the result is proven and important - the design clearly works, and I can tell my client this feature will deliver that improvement.

Almost every clause here hides a mistake, and untangling them is the whole point of research numeracy. First, 'statistically significant' does not mean 'important' - it means the effect is probably not pure chance. With a large enough sample even a trivial difference becomes significant, so significance is a first gate (is it real?), not a verdict on whether it matters. The question that decides importance is effect size - how big is it? - and that is exactly what the word 'significant' does not tell you. Second, a 'big percentage' is not the same as a big effect. A percentage is meaningless until you ask: of what baseline, measured how, over what range, on how many people - a dramatic-sounding relative figure can hide a tiny absolute change, and a change measured in one lab session is not a permanent life effect. Third, 'the design works' assumes causation, but most neuroarchitecture evidence is observational - it shows correlation, and a correlation is equally compatible with reverse causation (successful firms can afford nice buildings) and with a confounder (a hidden third factor, like budget, age or staff quality, causing both the feature and the outcome). Only a controlled design that isolates the variable supports a real causal claim. Fourth, 'this feature will deliver that improvement' for your client ignores population: the study measured some specific group, often students or Western office workers, and the finding travels only as far as that group reasonably extends - averages do not predict a particular person in a particular room. The honest reading asks four questions: is it probably real (significance), how big is it actually (effect size), does it show cause or just correlation (with a possible third factor), and who was measured (population). A claim that passes all four is strong. One that fails several is a hypothesis at best - and telling a client a feature 'will deliver' a specific result is a promise the evidence cannot back, quite apart from the fact that clinical and health claims belong to qualified professionals, not to a designer with a statistic.
Try it

Do it yourself

No tools needed — reason it through.

  1. 1Explain the difference between statistical significance and effect size, and why 'significant' does not mean 'important'.
  2. 2A study of 10,000 people finds a 'significant' 1 percent change. Why might a designer still ignore it?
  3. 3Offices with plants have happier staff. Give one reverse-causation and one confounder explanation besides 'plants cause happiness'.
  4. 4Why does the study population matter when deciding whether a finding applies to your project in India?
  5. 5Turn 'reduced stress by 30 percent' into the questions you would ask before believing it.
Take this with you

The one line to carry out

A neuroarchitecture number means nothing until you ask four quiet questions of it: is the effect probably real or could it be chance (statistical significance); how big is it actually, once you unpack the percentage into something you could notice (effect size, which is a different and often more important question); does the study show that the space caused the outcome or only that they correlated, with reverse causation or a hidden third factor as live alternatives (correlation is not causation); and who exactly was measured, since a finding travels only as far as its population reasonably extends - so lean on large, replicated, causally-supported effects on people like yours, treat everything else as a hypothesis to test, and leave clinical determinations to qualified professionals.
Take it further
References & further reading

Peer-reviewed journals & authoritative standards

  1. 01Effect sizeWikipedia — Effect size, 2026.
  2. 02Correlation does not imply causationWikipedia — Correlation does not imply causation, 2026.
  3. 03ReproducibilityWikipedia — Reproducibility, 2026.
  4. 04Peer reviewWikipedia — Peer review, 2026.
  5. 05Replication crisisWikipedia — Replication crisis, 2026.
Related lessons
Recap
Neuroarchitecture claims usually arrive as numbers, and a number is a headline, not evidence, until you read what stands behind it. Four questions do the work. First, statistical significance versus effect size: significance only tells you an effect is probably not chance, and with a large enough sample even a trivial difference becomes significant, so 'significant' never means 'important' - you must separately ask how big the effect is. Effect size, described loosely as small, medium or large, is where practical meaning lives, and a bare percentage hides its size until you ask of what baseline, measured how, and whether you would notice a change that large; weigh statistical significance against practical significance and against what acting on it would cost. Third, correlation is not causation: because most neuroarchitecture evidence is observational, two things moving together may reflect the hoped-for cause, or reverse causation (successful firms can afford nice buildings), or a confounding third factor (budget, building age or staff quality driving both) - only controlled, randomised designs isolate cause well, so observational findings are hypotheses to strengthen, not proofs. Fourth, the study population: every finding is produced on some specific, often student or Western, sample, and it travels only as far as that group reasonably extends, so confidence must drop honestly when the people measured differ from the people you design for - a real concern in India, where the research base is thin. Put the four together and any claim can be weighed without a statistics degree: a large, replicated, causally-supported effect on a similar population is strong evidence to lean on; a tiny observational effect on undergraduates quoted as a dramatic percentage is a hypothesis at best. This numeracy is self-defence against overclaiming and neuro-washing, it keeps you humble because even the best study speaks to averages not to the person in your room, and it never displaces the codes or the clinical judgement of qualified professionals.
Carry forward →

Now you can judge how trustworthy and how large a finding is - but a finding is not a design. The next lesson takes the honest leap from evidence to a specific decision: why the research rarely gives a formula, and how disciplined judgement turns principles into a real building.

A

The author

Amogh N P

Architect, interior designer, and creative polymath. Studio Matrx began in his notebooks — his vision of design made honest, useful, and open to everyone. Its Academy is written and taught in his memory, and free, forever.

More about Amogh →