The Mathematical Modelling Process
A mathematical model uses math (an average, a rate, an equation, or a graph) to describe a real situation so you can answer a question about it. Models help people make decisions: how much food to order for a school dance, when to leave for the bus, or how long it will take to save for something. On this page you’ll learn the steps of the modelling cycle and practise them on real-life questions.
Key ideas
Section titled “Key ideas”What a model is (and isn’t)
Section titled “What a model is (and isn’t)”A model is a simplified version of reality. It keeps the parts that matter for your question and ignores the rest. For example, “my phone loses about percentage points of battery per minute of video” is a model. It ignores screen brightness, the age of the battery, and whether apps are running in the background.
That’s fine. A model doesn’t have to be perfect. It has to be good enough to answer the question and honest about what it leaves out.
Models are used everywhere to make decisions:
- City planners use models of traffic and population to decide where to build roads and schools.
- Weather forecasters use models to predict tomorrow’s temperature.
- A family can use a simple model of its water use to decide whether a low-flow shower head is worth buying.
The modelling cycle
Section titled “The modelling cycle”Building a model is a process, and you often go around it more than once.
- Ask a question of interest. Make it specific enough to answer with numbers. “Do we use a lot of water?” is vague. “How many litres of water does our household use for showers in a year?” is answerable.
- Identify the information you need. What quantities go into the answer? For the shower question: how much water the shower uses per minute, how long showers last, and how many showers there are.
- Make a plan. Decide where the data will come from: first-hand (you measure, time, count, or survey) or second-hand (a trusted source such as Statistics Canada or a product label). Write down your assumptions, and decide what varies and what stays the same.
- Collect the data. Carry out your plan and record the results in a table.
- Display and analyse the data to build a model. Choose a graph that fits the data, then find a model (see the table below).
- Answer the question and judge the model. Use the model to answer the original question. Then ask: how well does it fit? What are its limitations? What predictions can it make?
Then revise if needed: collect more data, fix an assumption, or ask a sharper question.
Assumptions, and what varies
Section titled “Assumptions, and what varies”An assumption is something you decide to treat as true so the problem becomes manageable. Good models state their assumptions out loud, because if an assumption is wrong, the answer might be too.
It also helps to sort the quantities in your situation:
- What varies: quantities that change from one measurement to the next (the length of each shower).
- What stays the same: quantities you treat as constant (the flow rate of the shower head).
Choosing a model
Section titled “Choosing a model”The kind of data you collect suggests the kind of model to build.
| Your data | A good display | A model to try |
|---|---|---|
| One quantity measured many times (shower lengths, walk times) | dot plot, histogram, or box plot | an average: mean or median |
| An amount per unit (litres per minute, dollars per week) | a table | a rate: total = rate amount |
| Two quantities that change together | scatter plot | a line of best fit |
For one-variable data, you can also use quartiles and box plots to see how spread out the values are, and different graphs to show them. The Grade 9 curriculum writes a line as ; this site usually writes . They mean the same thing: the slope is the rate of change and the -intercept is the initial value.
Judging a model
Section titled “Judging a model”When you report your answer, include three things:
- Fit. Do the data points sit close to the line, or cluster tightly around the average? A model that fits well gives more trustworthy answers.
- Limitations. Which assumptions might be wrong? Was the sample small, or collected in an unusual week?
- Predictions. What can the model predict, and how far can you trust it? Predictions inside the range of your data (interpolation) are safer than predictions far outside it (extrapolation).
Data in society
Section titled “Data in society”Today, huge amounts of data (big data) are collected automatically, often without people noticing. Phones record locations. Fitness apps record heart rates and sleep. Transit cards, like Ontario’s PRESTO card, record trips. Streaming services record everything you watch.
This data can help: traffic apps reroute drivers around a crash, and health agencies can plan where clinics are needed. But it raises real questions:
- Privacy and storage. Who can see the data? How long is it kept? Could it be sold, or stolen in a data breach?
- Use. Is it used only for what people agreed to? Could it be used to treat people unfairly?
- Representation. Data can be shown in ways that mislead: a graph whose vertical axis doesn’t start at zero, a time range picked to hide a trend, a biased sample, or a claim that one thing causes another just because they’re correlated.
When you see data in the news or an app, ask: who collected it, how, and what might they want me to think?
Worked examples
Section titled “Worked examples”Example 1: Saving for a bike
Section titled “Example 1: Saving for a bike”Maya wants a bike that costs $480 before tax. She has $60 saved. Over the last six weeks she earned $40, $55, $35, $50, $45, and $45 babysitting, and she plans to save all of it. About how many weeks will it take her to afford the bike?
Solution. Work through the cycle.
Question: How many weeks until Maya can buy the bike?
Information needed: the price with tax, what she has now, and how much she saves per week.
Assumptions: her weekly earnings stay about the same as in the past six weeks, she spends none of it, and the price doesn’t change. What varies: her weekly earnings. What stays the same: the price.
Model. Her earnings vary, so use an average. The mean weekly earnings are
so the model is “Maya saves about $45 per week”. The price with HST is
She still needs dollars. At $45 per week:
Answer. After weeks she won’t quite have enough, so it takes about weeks.
Check: , which is more than $542.40. After weeks she would have dollars, which is not enough.
Example 2: Water for showers
Section titled “Example 2: Water for showers”A family of four wants to know: how many litres of water do we use for showers in a year? They plan to use the answer to decide whether to buy a low-flow shower head.
Solution.
Information needed: the shower’s flow rate (litres per minute) and the total shower time per year.
Plan and data. They hold a L bucket under the shower and time how long it takes to fill: seconds. Then everyone writes down the length of every shower for one week. The log shows showers with a total of minutes.
Assumptions: the flow rate stays the same, the week they recorded is a typical week, and they shower like this all weeks of the year.
Model. The flow rate is a rate model:
Water per week:
Water per year:
Answer. The family uses roughly L of water for showers each year. That’s about m³, since m³ L.
Using the model to decide. A low-flow shower head uses about L/min. With the same shower times:
That would save about L per year.
Limitations. One week is a small sample. Exams, holidays, or a summer sports season could change shower times. A better plan would record two or three weeks at different times of year.
Example 3: When will my battery hit 20%?
Section titled “Example 3: When will my battery hit 20%?”Leo fully charges his phone and streams video, recording the battery level every minutes.
| Time, (min) | ||||||
|---|---|---|---|---|---|---|
| Battery, (%) |
How long can he stream before the battery reaches ?
Solution. Two quantities change together, so make a scatter plot. The points lie almost exactly on a straight line, so a linear model makes sense. Technology (a spreadsheet, graphing calculator, or Desmos) gives the line of best fit:
The slope means the battery drops about percentage points per minute. The initial value, , is close to the true starting level of .
Set and solve:
Answer. About minutes, or roughly hours and minutes.
Judging the model. The fit is excellent: the points are very close to the line. But minutes is far beyond the last measurement at minutes, so this is extrapolation. Many phones also drain at a different rate when the battery is low or warm. Leo should treat “about hours” as an estimate, and test it by streaming longer.
Example 4: Revising a model
Section titled “Example 4: Revising a model”Five weeks after making her model in Example 1, Maya has $260 saved, not the $285 her model predicted. Revise the model and update the answer.
Solution. The model predicted dollars. She has less, so her real saving rate was lower (maybe she spent some money, or had fewer jobs).
Revised rate, using what actually happened:
She still needs dollars:
After more weeks she would have dollars, just short. So she needs more weeks, for weeks in total instead of .
This is the cycle in action: real data showed the first model was a bit too hopeful, so Maya changed the rate and made a new prediction.
Common mistakes
Section titled “Common mistakes”Asking a question that can’t be measured. “Is our phone use bad?” has no numerical answer. Rewrite it as something you can collect data on, like “How many hours per day does each person in our home use a phone?”
Hiding the assumptions. Every model makes assumptions. If you don’t write them down, nobody (including you) can tell why the answer might be off. List them in step 3 and look back at them in step 6.
Trusting a prediction far outside the data. A line that fits to minutes may not hold at minutes. Say clearly when you’re extrapolating, and treat the answer as a rough estimate.
Rounding the wrong way in context. If weeks of saving are needed, weeks is not enough. When you need to reach a goal, round up to the next whole week.
Using too little data. One week of showers or one test run of a battery might be unusual. More data, collected at different times, gives a more reliable model.
Treating the model as the truth. A model is a tool, not a fact. If new data disagrees with it (like Maya’s savings), revise the model instead of ignoring the data.
Practice
Section titled “Practice”1. (Warm-up) A student asks: “How much does my family spend on groceries in a month?” List the information needed and one way to collect it.
Solution
Information needed: the cost of each grocery trip and how many trips happen in a month.
One way to collect it: keep every grocery receipt for a month (or check the family’s bank or card records, with permission) and add the totals. This is first-hand data.
2. (Warm-up) You want to know how long it takes to fill a backyard pool with a garden hose. Name one thing that varies, one thing you would treat as staying the same, and one assumption.
Solution
Sample answer. Varies: the depth of water in the pool as it fills. Stays the same: the flow rate of the hose. Assumption: the water pressure doesn’t drop while the pool fills (for example, nobody else in the house is running water).
3. (Warm-up) Which kind of model (an average, a rate, or a line of best fit) fits each question best?
- (a) How tall will my bean plant be on day , using its height each day for two weeks?
- (b) What is a typical number of text messages I send per day?
- (c) How much will L of gas cost if gas is dollars per litre?
Solution
(a) A line of best fit: two quantities (day and height) change together.
(b) An average (mean or median) of the daily counts.
(c) A rate: cost , or $54.25.
4. (Core) A garden hose fills a L bucket in seconds. About how many hours will it take to fill a pool that holds L? State an assumption.
Solution
Rate: L/s, which is L/min.
About hours. Assumption: the hose’s flow rate stays the same the whole time.
5. (Core) Ana timed her walk to school (in minutes) on seven days: , , , , , , .
- (a) Find the mean and the median.
- (b) She walks to school and back on about school days a year. Use the mean to estimate her total walking time for the year, in hours.
- (c) Why is the mean (not the median) the better choice for a total?
Solution
(a) Sum: , so the mean is min. In order: , so the median is min.
(b) Two trips a day:
(c) The mean is the total divided by the number of trips, so mean number of trips gives back the total. The median ignores how long the slow days (like min) were, so it would underestimate the total.
6. (Core) A class models the height of a sunflower with , where is the height in centimetres and is the number of days since it sprouted. They measured it for days.
- (a) Predict the height on day .
- (b) The model predicts a height of cm on day . Is that prediction trustworthy? Explain.
Solution
(a) , so about cm.
(b) Check: . But day is far beyond the days of data (extrapolation). Plants don’t grow at a steady rate forever: they slow down, stop, or die at the end of the season. The prediction is not trustworthy.
7. (Core) A free fitness app records your location, heart rate, and sleep. Describe one benefit, one risk, and one question you should ask before agreeing to share your data.
Solution
Sample answer. Benefit: it can show patterns in your sleep and exercise that help you make healthier choices. Risk: your location history could reveal where you live and go to school, and it could be shared with advertisers or exposed in a data breach. Question to ask: “Who can see my data, how long is it stored, and can I delete it?”
8. (Challenge) A store’s ad shows a bar graph of monthly sales: $960 in March and $990 in April. The vertical axis starts at $950, so the April bar looks four times as tall as the March bar.
- (a) Explain why the bars look four times different.
- (b) What was the actual percent increase?
Solution
(a) Starting at $950, the March bar shows and the April bar shows . Since , April’s bar is four times as tall, even though sales barely changed.
(b)
Sales rose by only about . A fair graph would start the axis at $0.
9. (Challenge) A family plans a km drive (about the distance from Toronto to Ottawa). Their car uses L of gas per km, and gas costs $1.55 per litre.
- (a) Build a model and estimate the cost of gas for the round trip.
- (b) List two assumptions, and say how the family could make the model better.
Solution
(a) Gas for one way:
Cost one way: , about $52.31. Round trip: , about $104.63.
(b) Sample assumptions: the car really uses L per km on this trip (highway driving, air conditioning, and a full car all change it), and the gas price stays at $1.55. To improve the model, they could check the car’s actual fuel use on a recent highway trip and look up current gas prices along the route.