IB Biology HL Practical Skills Paper 1B & IA ~13 min read

Using Tech to Process Data

A spreadsheet will happily calculate a mean to nine decimal places from three readings, and draw you a beautiful chart of something meaningless. The technology removes the arithmetic, not the thinking — and the marks are for the thinking.

📚 What you need to know

Spreadsheets

A spreadsheet does three separable things, and it is worth naming them separately because questions tend to target one at a time.

Three jobs, one piece of software1. RECORD conc trial 1 trial 2 0.2 12 14 0.4 21 23 0.6 29 28 2. CALCULATE mean of the repeats standard deviation percentage change gradient of a line correlation between two sets 3. PLOT Keep the raw data and the processed data in separate columns If a mean is wrong you can trace it back; if you typed over the raw data, you cannot
Recording, calculating and plotting are three different claims to make about your data. Questions asking “how did technology help?” usually want you to name which of the three, and why it mattered here.

Recording

Calculating

Spreadsheets perform calculations, statistical analyses and mathematical operations on whole datasets at once. Where a calculator handles one number at a time, a spreadsheet applies the same operation down a column of two thousand.

The ones you will actually use: mean of repeats, standard deviation to describe spread, percentage change, the gradient of a line to get a rate, and correlation between two variables.

Plotting

Built-in functions generate graphs and charts automatically, making it possible to visualise trends, patterns and correlations. This is particularly valuable for data with a large range — population data being the classic example, where numbers may run from tens to millions and a computer can rescale the axis instantly.

The default chart a spreadsheet offers you is almost never the right one. It will cheerfully join categorical bars with a line, start a y-axis at a value that exaggerates a tiny difference, or add a trend line through data that has no trend. Choosing the chart is a biology decision, not a software one.
Which chart the data is asking forBAR CHART LINE GRAPH SCATTER GRAPH categories, not numbers e.g. abundance in habitat A, B, C you chose the x values e.g. the temperatures you set both variables measured e.g. leaf width against light levelBars are separated because the categories are separate; a line implies values in between exist
The last line is the reason bar charts have gaps. Joining “habitat A” to “habitat B” with a line would suggest there is something halfway between them, and there is not.

Producing models

Computers use collected data to produce models that inform ongoing predictions. A model built from real measurements is fitted to the data, then used to estimate values you did not measure.

Two words are worth separating here:

The enzyme practical makes this vivid. Rate rises steadily with temperature from 10 °C to 40 °C, so a fitted straight line will confidently predict a very high rate at 70 °C. Biology says the enzyme will have denatured long before that. The model does not know about tertiary structure.

Image analysis

Images can be analysed using computer programmes. The example given in the course is images of joints in motion: video of a moving limb is broken into frames, points on the joint are marked, and software calculates the angle at the joint in each frame.

Measuring a joint angle frame by frame 155 degrees shoulder 66 degrees elbowframe 1: arm extended frame 24: arm flexedSoftware marks the joints, then calculates the angle in every frame automatically
Doing this by hand with a protractor on printed stills is possible but slow and imprecise. Software gives a value for every frame, so you get angle against time — and from that, the speed at which the joint moves.
What technology doesWhat it cannot do
Calculates a mean from 2000 values instantlyTell you whether those values were valid to collect
Draws a line of best fit through any dataKnow whether a linear relationship is biologically plausible
Extends a model beyond the measured rangeKnow that the enzyme denatures at 60 °C
Reports a correlation coefficientEstablish that one variable causes the other
Removes arithmetic slipsRemove a systematic error in the raw readings
The line examiners reward. Correlation is not causation, and software cannot tell the difference. A spreadsheet will report a strong positive correlation between ice cream sales and drowning incidents. The biology — a third variable, hot weather — has to come from you.

Worked examples

WE 1

Choosing and justifying a graph

A student measures the mean number of limpets per quadrat at five fixed distances up a rocky shore. State which type of graph should be used to display the results, and justify your choice. (3 marks)

Step 1: identify the variables Distance up the shore is continuous, and it was chosen by the student, so it is the independent variable on the x-axis. Mean number of limpets is the dependent variable. Step 2: state the graph A line graph, with distance on the x-axis and mean number per quadrat on the y-axis. Step 3: justify it Because distance is continuous, intermediate values exist, so joining the points is meaningful. A bar chart would wrongly imply the five distances are separate categories. Line graph, because the independent variable is continuous error bars showing the spread at each distance would strengthen this considerably.
WE 2

Reading a spreadsheet’s statistics

A spreadsheet reports the mean height of plants in two conditions. Condition A: mean 24.6 cm, standard deviation 3.2 cm. Condition B: mean 27.1 cm, standard deviation 2.9 cm. A student concludes that condition B causes taller growth. Evaluate this conclusion. (4 marks)

Step 1: what the means show The mean in B is 2.5 cm higher, so there is a difference in the data as collected. Step 2: what the spread shows Using one standard deviation, A spans roughly 21.4 to 27.8 cm and B spans 24.2 to 30.0 cm. These ranges overlap substantially. Step 3: the conclusion this allows The overlap means the difference between the means may be due to random variation rather than the condition. A statistical test, such as a t-test, would be needed before claiming a real difference. Step 4: the word “causes” Even a significant difference would show association. Causation requires that all other variables were controlled. The conclusion is not supported — the spreads overlap too much a spreadsheet will report a difference between any two means. Whether it means anything is your judgement, not the software’s.
WE 3

Percentage change, and the limits of a model

A population is recorded as 1200 individuals in 2015 and 1860 in 2020. Calculate the percentage increase. A model fitted to this data predicts 4400 individuals by 2035. Suggest why this prediction should be treated with caution. (4 marks)

Step 1: find the increase 1860 − 1200 = 660 individuals Step 2: express it as a percentage of the original (660 ÷ 1200) × 100 = 55 % Step 3: why the prediction is uncertain It is an extrapolation well beyond the measured range, from only two data points, and it assumes growth continues at the same rate. Step 4: the biology Real populations are limited by a carrying capacity — food, space, disease and predation increase as numbers rise — so growth is likely to slow and the model will overestimate. A 55 % increase, but the 2035 figure is an extrapolation always divide by the original value for percentage change. Dividing by 1860 gives 35 % and loses the mark.

💡 Exam tips

⚠ Common mistakes

Pulling the skill set together

Two pages, one arc. Technology lets you collect data you could not otherwise gather — faster, for longer, at intervals no person could keep up with — and then handle volumes of it that would take weeks by hand.

What it does not do is decide whether the data was worth collecting, whether the sensor was calibrated, whether the graph is honest, or whether the pattern means what it appears to mean. Those judgements stay with you, and they are where the marks live.

In your internal assessment, the strongest sentences about technology are specific ones. Not “a data logger was used to improve accuracy”, but “readings were logged every 2 s because a preliminary trial showed the reaction was over within 30 s, and manual timing would have produced only three data points.”

Want this explained one-to-one?

Book a free session with an experienced IB Biology tutor and get your trickiest topics made simple.

Book a Free Session →