Studio Matrx Monthly · Volume 1 · Issue 3 · August 2026
Amogh N P
 In loving memory of Amogh N P — Architect · Designer · Visionary 
Reading a Simulation CriticallyLesson 0.4
BPS for Architecture, Planning & Urban Design/Module 0 · Foundations of Building Performance Simulation

Lesson 0.4 · Foundations of Building Performance Simulation

Reading a Simulation Critically

Garbage in, garbage out - the assumptions that dominate results, and how to spot a misleading one

13 min Interactive lessonFree · open lessonByAmogh N P· Architect & interior designer
The hook

A simulation will always give you a confident-looking number. Your job is to know when to believe it.

The dangerous thing about a simulation is that it never looks uncertain. It prints a precise figure - 87.3 kWh/m2 per year - in a clean report, and the precision is seductive. But that number rests on a tower of assumptions you chose, any one of which could be off by enough to swing the result.

Reading a simulation critically is the single most valuable skill in this course. It is what separates a trustworthy analyst from someone who launders their own guesses through software and calls the output fact. This lesson teaches you to interrogate a result: what drives it, how far to trust it, and how to catch a number that is quietly wrong.

Which input dominates? Is it plausible? What's the baseline? Read it comparatively. Interrogate every number.

Garbage in, garbage out: the assumptions that dominate

A simulation does not predict reality; it computes the consequences of your inputs, exactly and obediently. If the inputs are wrong, the output is wrong - confidently and to three decimal places. So critical reading starts by asking which inputs actually move the result, because they are not equal.

A handful of assumptions dominate most whole-building energy and comfort results. The weather file sets the whole climatic frame - run the same model on a different city's EPW and the answer changes wildly. Occupancy - how many people, how densely, doing what - drives internal gains, fresh-air needs and, in offices, a large slice of the cooling load. Schedules - when the space is used, when lights and equipment are on, when the AC runs - reshape the entire annual profile. Set-points - the thermostat target - are brutal: a cooling set-point of 24 C versus 26 C can shift cooling energy by a large margin in a hot climate. And operation - how the real building is actually run, doors propped, blinds down, systems overridden - is the biggest wildcard of all.

By contrast, some inputs matter less than beginners expect. Fabric details - a glazing U-value tweaked in the third decimal, a wall construction refined slightly - usually move the answer only a little on their own. The lesson: spend your care where the sensitivity is - weather, people, schedules, set-points, operation - not polishing inputs that barely register.

WHAT MOVES THE ANSWER MOSTWeather fileOccupancy & useSchedulesSet-pointsInfiltrationGlazing U-valueSWING IN PREDICTED RESULT ----->operation dominates; fabric alone rarely does
Zoom
Not all inputs are equal. Operation-side assumptions - the weather file, occupancy, schedules and set-points - swing a whole-building result far more than fabric details like a glazing U-value. Critical reading means spending your care where the sensitivity actually is.

Occupancy and set-points swing the answer. A third-decimal U-value tweak rarely does. Spend care where sensitivity is.

The performance gap - and why absolute numbers drift

Even a careful model diverges from the real building. This is the well-documented performance gap: measured energy use is routinely higher - sometimes much higher - than the design simulation predicted. The causes are mostly the assumptions above coming apart in reality. Occupants use the space more intensively or unpredictably than the schedule assumed. Systems are commissioned imperfectly or overridden. The building is operated with windows open and set-points nudged. Construction quality introduces thermal bridges and infiltration the model idealised away.

The gap is not a sign the simulation is broken; it is a sign that an absolute prediction of a single building's energy use is inherently uncertain. Understanding this reframes how you use any result. The headline number is an estimate with error bars you should mentally attach, not a meter reading of the future. Calibration - tuning a model against measured data once a building is running - narrows the gap and is essential for retrofit and operational work, but during design you should treat the absolute figure as indicative. Respecting the performance gap is not pessimism; it is the difference between a modeller who oversells and one who can be trusted.

This has a direct effect on how you should communicate. If you hand a client a single number - '87 kWh/m2/yr' - you have made an implicit promise the building may not keep, and the performance gap will eventually embarrass you. Present the same finding as a range, or better as a comparison against a baseline, and you have said something both honest and durable. Good practice quotes results with their uncertainty acknowledged, not stripped of it for a cleaner slide. The gap is a permanent feature of the field, not a flaw to hide - and naming it openly is what marks out a professional from an enthusiast.

THE PERFORMANCE GAPSIMULATEDMEASURED95150the gapkWh/m2.yrA vs B survives the gapOPTION AOPTION BB is ~35% lowereither wayAbsolute numbers drift; the difference between two options under the same assumptions holds.
Zoom
Left: the performance gap - measured energy routinely exceeds the design simulation, so absolute figures are estimates with error bars. Right: why comparative reading is the fix - the difference between two options run on the same assumptions survives the gap, even when each absolute number drifts.

Sanity-check every output before you believe it

Before trusting any result, check whether it is even plausible. Experienced analysts run fast reality checks, and you should build the habit early. Order of magnitude: does the energy use intensity (EUI) land where a building of this type and climate should? An air-conditioned office landing at 5 kWh/m2 per year is not efficient, it is broken input; one at 900 is equally suspect. Compare against known benchmarks - typical ranges for the building type, or code baselines like ECBC and ASHRAE 90.1.

Energy balance: the pieces should sum sensibly - cooling should dominate in a hot climate, heating in a cold one; if a Chennai office shows a big heating load, something is wrong. Direction of change: when you improve the design - add shade, better glazing, more insulation - the result should move the expected way. If a deeper overhang increases cooling energy, you have a bug, not an insight. Peaks and profiles: does the daily and seasonal shape make physical sense - cooling peaking on hot afternoons, lighting following occupancy? A result that fails any of these is telling you about your model, not your building.

These checks take minutes and catch the majority of gross errors - a floor area entered in the wrong units, a schedule left on around the clock, a construction facing the wrong way, a zone with no ventilation. Such mistakes are not rare; they are routine, and they produce outputs that look perfectly formatted and are completely wrong. The discipline is to distrust your own first result on principle and prove it plausible before you let anyone act on it. Never report a number you have not sanity-checked; the software will happily compute nonsense and format it beautifully.

EUI in a sane range? Cooling dominates in a hot climate? Improvements move the right way? If not, it is a bug.

Why comparative reading beats absolute

Here is the resolution to all this uncertainty, and the most important habit in the course: read simulations comparatively, not absolutely. When you run option A and option B through the same model with the same assumptions, the errors those assumptions carry are shared - and when you take the difference between the two results, the shared error largely cancels. So even if the absolute EUI of each is uncertain by twenty percent, the statement 'B uses 30% less than A' is far more robust than either number alone.

This is why a rough model, used to compare, is genuinely useful while a rough model quoted as an absolute is dangerous. It is also why the honest way to present findings is relative: 'the shaded option cuts overheating hours by 40% versus the base case', not 'the building will have 312 overheating hours'. The first survives the performance gap; the second pretends the gap does not exist.

The practical rule: always simulate against a baseline. A result with nothing to compare it to is hard to trust and easy to misread. A result expressed as a difference from a clear reference - the current design, a code-minimum case, the previous option - is both more reliable and more decision-useful. Comparison is not a limitation of simulation; it is how you extract its most trustworthy signal.

THE PERFORMANCE GAPSIMULATEDMEASURED95150the gapkWh/m2.yrA vs B survives the gapOPTION AOPTION BB is ~35% lowereither wayAbsolute numbers drift; the difference between two options under the same assumptions holds.
Zoom
Left: the performance gap - measured energy routinely exceeds the design simulation, so absolute figures are estimates with error bars. Right: why comparative reading is the fix - the difference between two options run on the same assumptions survives the gap, even when each absolute number drifts.

Same assumptions, take the difference, shared error cancels. 'B is 30% lower' beats 'B is 87.3'.

How to spot a misleading result

Some results are not just uncertain - they are actively misleading, and learning to smell them is a professional skill. Watch for the too-good-to-be-true: a design that shows a huge saving usually hides an optimistic assumption - a set-point relaxed, an occupancy dropped, a schedule shortened - doing the work rather than the design. Always ask what input is driving this and whether it is honest.

Watch for cherry-picked baselines: a saving is only as meaningful as what it is measured against. '40% better than a deliberately terrible base case' can be worse than the code minimum. Insist on a fair, disclosed baseline. Watch for false precision - a report quoting energy to four significant figures while the occupancy was a pure guess; the precision is decoration over a shaky input. Watch for hidden defaults: a tool's out-of-the-box schedules, constructions or climate may not be your building at all, and unexamined defaults are a classic source of confidently wrong results.

The defence is always the same short interrogation: What are the key assumptions? Which one dominates this result? What is the baseline, and is it fair? Does the number pass a sanity check? Is this being read comparatively? Run that on every simulation - your own and others' - and you will rarely be fooled. That habit of disciplined suspicion, more than any software skill, is what makes a simulationist worth trusting.

Too good to be true? Ask which optimistic input did the work. Interrogate every result, including your own.

Concepts for reading results honestly

The performance gap

Difference between simulated and measured performance

Real buildings usually use more than models predict; treat absolute figures as estimates with error bars, not meter readings.

Sensitivity analysis

Finding which inputs most change the result

Tells you where to spend care - occupancy, schedules and set-points usually dominate; fabric details often don't.

Calibration (ASHRAE Guideline 14)

Tuning a model to match measured data

Narrows the performance gap for operational and retrofit work; during design, results stay indicative.

EUI benchmark

Energy use intensity (kWh/m2.yr) for a building type

A fast sanity check - a result far outside the sane range for the type and climate signals bad input, not a great design.

Hands-on workshop

Workshop - interrogate a simulation result

No software needed - this is a thinking drill you will run on every real result for the rest of your career. You take a plausible-looking output and stress-test it with a fixed set of questions until you know how far to trust it.

None - a result and a notebook. The discipline applies to every tool you will ever use; later modules give you real outputs from Ladybug, EnergyPlus and Radiance to practise on.

Given & goal
Goal: build the habit of interrogating any simulation result
Inputs: a result (real, published, or the worked example below), a notebook
Time: ~30 minutes
  1. 1Take a result - use this example: 'A new office design is simulated at 78 kWh/m2/yr, a 45% saving over the base case.' Write it at the top of your sheet.
  2. 2List the five dominant assumptions it must rest on - weather file, occupancy, schedules, set-points, operation - and next to each write 'known' or 'guessed' as best you can judge.
  3. 3Sanity-check the number: is 78 kWh/m2/yr plausible for an air-conditioned office in this climate? Compare to a benchmark range or a code baseline. Flag it if it looks too low or too high.
  4. 4Interrogate the 45% saving: what is the baseline, and is it fair? What single assumption, if optimistic, could be doing most of that saving instead of the design?
  5. 5Rewrite the claim honestly and comparatively - e.g. 'Under identical assumptions, the shaded option uses about 25% less cooling energy than the current design; the absolute figure is indicative pending real occupancy data.' Note what you would need to check to trust it fully.

You’ll walk away with
An interrogated result: the original claim, its dominant assumptions tagged known/guessed, a sanity-check verdict, the baseline questioned, and an honest comparative restatement. This five-step interrogation is the core professional habit of the whole field.

The worked example

Three altitudes on the same idea

Read the band that fits you — or all three.

For the architectPerformance-driven design decisions

Treat every simulation result - yours or a consultant's - as a claim to interrogate, not a fact to accept. Before you let a number steer a decision, ask what dominates it, what the baseline is, and whether it passes a sanity check. Present your own findings comparatively ('this option saves 25% versus the base') rather than as absolutes, and you will be both more honest and more persuasive. The architect who reads results critically catches the flaws before they reach site.

For the interior designerComfort, daylight & healthy interiors

Your daylight and comfort results ride on assumptions that are easy to get wrong - occupancy, furniture, surface reflectances, how the blinds are really used. A dazzling sDA figure means little if the model assumed bare white walls and no one ever draws a blind. Sanity-check against how the space will actually be lived in, and report light and comfort comparatively - 'this layout is brighter and less glary than that one' - which survives the messiness of real occupation far better than a single headline percentage.

For the studentSkills, portfolio & green-building jobs

Critical reading is the skill that gets you hired and keeps you credible. Anyone can push buttons and produce a number; the value is in interrogating it - naming the dominant assumption, insisting on a fair baseline, catching a result that is too good to be true. In studio and in interviews, showing you know why a result might be wrong impresses far more than showing a pretty output. Practise the five-question interrogation until it is automatic.

Misconception check

A more detailed model is automatically a more accurate one.

Not on its own. Adding detail - finer geometry, more construction layers, more zones - improves accuracy only if the added inputs are actually known and correct. If the dominant assumptions (weather, occupancy, schedules, set-points, operation) are still guesses, a lavishly detailed model is just a guess with more decimal places, and its polish can make wrong answers look more authoritative. Real accuracy comes from getting the high-sensitivity inputs right and, where it matters, calibrating against measured data - not from piling detail onto uncertain foundations. A simple model with honest, well-chosen inputs, read comparatively, routinely beats an elaborate one built on optimistic assumptions. Detail is worth adding only where sensitivity analysis says it changes the answer.
Try it

Do it yourself

No software - reason it through.

  1. 1Name the five input assumptions that most dominate a whole-building energy result.
  2. 2What is the performance gap, and does it mean the simulation was done wrong?
  3. 3List three fast sanity checks you would run on any energy result.
  4. 4Explain why 'B uses 30% less than A' is more trustworthy than 'B uses 78 kWh/m2/yr'.
  5. 5Give two warning signs that a result is misleading rather than merely uncertain.
Take this with you

The one line to carry out

A simulation computes your assumptions exactly, so read every result critically: know which inputs dominate it, sanity-check that it is even plausible, respect the performance gap, and trust the comparison between options far more than any absolute number. Disciplined suspicion, not software skill, is what makes a simulationist worth believing.
Take it further
References & further reading

Peer-reviewed journals & authoritative standards

  1. 01Hensen, J. L. M. & Lamberts, R. (eds) — Building Performance Simulation for Design and Operation (2nd ed.)Routledge, 2019.
  2. 02Sensitivity analysisWikipedia, 2026.
  3. 03Uncertainty quantificationWikipedia, 2026.
  4. 04ASHRAE Guideline 14 — Measurement of Energy, Demand, and Water SavingsASHRAE, 2026.
  5. 05Energy modelingWikipedia, 2026.
Related lessons
Recap
Simulations are garbage-in, garbage-out: a few inputs - weather, occupancy, schedules, set-points, operation - dominate the result, while fabric details often matter little. Real buildings diverge from models (the performance gap), so absolute figures are estimates with error bars. Always sanity-check outputs against benchmarks and physical sense, read comparatively against a fair baseline, and interrogate any result - especially a flattering one - for the optimistic assumption doing the work.
Carry forward →

That completes the Foundations: you know what simulation is, why it is worth doing, what tools do it, and how to read the results honestly. From here the course goes hands-on - and it starts where every simulation starts, with the climate. Next module: analysing a climate and the EPW weather files that feed every run.

A

The author

Amogh N P

Architect, interior designer, and creative polymath. Studio Matrx began in his notebooks — his vision of design made honest, useful, and open to everyone. Its Academy is written and taught in his memory, and free, forever.

More about Amogh →