IB Psychology SL Topic 5 — Methods of Research Paper 1 & 2 Core skill ~11 min read

Experiments and Their Designs

An experiment is the only research method that can tell you one thing caused another. Everything else can only tell you that two things go together. That single sentence is worth a lot of marks, and the four types of experiment below are just four different amounts of control over the same idea.

📘 What you need to know

What makes a study an experiment

Take a simple claim: “noise stops people remembering things.” To test it you need two groups doing the same memory task, one in silence and one with noise playing. Noise is the thing you change, so noise is the IV. The number of words recalled is the thing you measure, so that is the DV.

Now the important part. If the noisy group also did the task later in the day, on harder words, in a hotter room, then a lower score proves nothing. Any of those could have caused it. So the researcher freezes everything except the IV. Same words, same room, same instructions, same time of day. Now if the scores differ, there is only one thing left that could have done it.

What an experiment is actually doing change one thing, measure what happens, hold the rest still extraneous variables: noise, heat, mood, time of day held constant IV what you change DV what you measure leads to a change in as long as nothing else moved control is what buys you the right to say the IV caused it every method below is a different amount of that control
The dashed green line is the whole job of experimental design. Anything that gets past it becomes an alternative explanation for your results.
Say it like this in an exam The IV is manipulated, the DV is measured,
and extraneous variables are controlled.
If you can name the IV and the DV of a study in one sentence each, you can answer most “identify the research method” questions. Practise saying them out loud: “the IV was whether the model was aggressive or not, the DV was the number of imitative acts.”

Lab experiments

A lab experiment is run in a setting the researcher controls. That usually means a room at a university, but “lab” really means “a place where I decide what happens”, not “a place with white coats”.

Why it earns the “high internal validity” label: because the only difference between the two groups was the IV, the difference in the DV must have come from the IV. That is a cause-and-effect claim, and no other method gets to make it this confidently.

The cost is realism. Sitting in a bare room learning a list of unrelated words is not something anyone does on a normal Tuesday, so the behaviour you see may not be the behaviour you would get outside. That is low ecological validity. Participants also know they are being studied, which invites demand characteristics.

Field experiments

A field experiment keeps the manipulation but moves it into a real place: a train, a school, a shop. The researcher still decides what the IV is, so it is still an experiment.

For example, a researcher could arrange for someone to drop a folder of papers in a busy corridor, and change one thing only: whether that person is wearing a suit or a tracksuit. The IV is the clothing, the DV is how many people stop to help. The corridor is real, the passers-by are real, and nobody is trying to be a good participant, because nobody knows a study is happening.

Watch the boundary: a field experiment is not the same as a naturalistic observation. If there is no manipulated IV, it is an observation, not an experiment.

The weakness is the corridor itself. Weather, crowds, someone’s mood, a fire alarm — all of it can leak in. Those extraneous variables make the results messier and make the study harder to repeat, so reliability drops.

Natural experiments

Sometimes the IV is something no ethical researcher would ever create: a war, a flood, a hospital closure, a change in the law. In a natural experiment the world supplies the IV and the researcher simply measures what follows.

This is how psychology studies things it could never set up on purpose, which is why these studies are often described as having high ethical validity as well as high ecological validity. But because nobody controlled who ended up in each group, other differences between the groups may be doing the work, and cause and effect becomes shaky again.

Quasi-experiments

A quasi-experiment looks like a lab experiment from the outside, but the IV is a characteristic the participant already has: age, gender, being bilingual, having a diagnosis. Nobody assigned it, and nobody can.

True experiment

Participants are randomly put into the noisy or the silent condition. The groups are alike at the start, so the IV is the only difference.

Quasi-experiment

Participants are already 18 or already 70. The groups differ in dozens of ways besides age, and you cannot separate them.

That is the whole limitation, and it is called participant variables. If older participants recall fewer words, it might be age — or it might be schooling, hearing, confidence with the task, or how much sleep they had. The procedure can still be standardised and repeated, so reliability holds up. Internal validity is what suffers.

The trade-off you must be able to explain

Examiners love this because it is one idea that explains every method at once. As you move from lab to field to natural, you give away control and buy realism. There is no method that wins on both.

Control versus real life the same trade-off runs through every method in this topic more control, higher internal validity LAB EXPERIMENT FIELD EXPERIMENT NATURAL EXPERIMENT you set everything real place, your IV IV already happened more realism, higher ecological validity every step to the right buys realism and sells certainty quasi-experiments sit off this line: the IV is the person, not the place
Use this scale to build evaluation points. Whichever method you are given, say what it gained and what it gave up.
MethodWho supplies the IV?SettingStrongest atWeakest at
Lab experimentThe researcherControlled roomCause and effect, replicationRealism, demand characteristics
Field experimentThe researcherReal-world placeNatural behaviourExtraneous variables, replication
Natural experimentEvents, not the researcherWherever it happenedStudying the unstudiable ethicallyCause and effect, sample bias
Quasi-experimentA trait the person already hasOften controlledStandardised, repeatableParticipant variables

How to answer these in the exam

🧩 A four-move method for any “identify and evaluate” question

  1. Find the IV. What differed between the groups or conditions?
  2. Ask who made it differ. Researcher = lab or field. Nobody = natural. A trait of the person = quasi.
  3. Find the setting. Controlled room = lab. Everyday place = field.
  4. Evaluate using the trade-off. Name one strength and one weakness that follow from the amount of control, and link each to internal or ecological validity.

Worked examples

WORKED EXAMPLE

Identify the research method and justify your answer

Researchers measured the reading scores of children who had spent two years in a school that introduced a daily silent-reading hour, and compared them with children at a nearby school that did not. The schools made the decision themselves before the study began. Identify the research method used and give one reason for your answer. [3]

Step 1: Find the IV Whether the school ran a daily silent-reading hour. Step 2: Who supplied it? The schools did, before any researcher was involved. Step 3: Match it to a method No manipulation and no random allocation, so this is not a lab or field experiment. A natural experiment justify it in one line: the IV occurred naturally and the researcher only measured the DV afterwards
WORKED EXAMPLE

Evaluate one strength and one limitation

A researcher asks participants to come to a university room, randomly allocates them to either a quiet or a noisy condition, and measures how many words from a 20-word list they recall. Outline one strength and one limitation of this method. [4]

Name the method first Lab experiment: controlled setting, manipulated IV, random allocation. Strength Everything except noise was held constant, so a difference in recall can be put down to the noise. High internal validity. Limitation Learning a random word list in a bare room is nothing like everyday remembering, so the findings may not transfer. Low ecological validity. One point each, both explained a named strength with no explanation usually gets half the marks — always say why

💡 Exam tip

⚠️ Common mix-up

Up next: Observations: Watching Behaviour Directly — what to do when you cannot manipulate anything and can only watch.

Want this explained one-to-one?

Book a free session with an experienced IB Psychology tutor and get your trickiest topics made simple.

Book a Free Session →