IB Psychology HL Topic 5 — Research Methods Paper 3 & IA Core skill ~10 min read

Questionnaires and Survey Design

A questionnaire looks like the easiest method in psychology. Write some questions, hand them out, count the answers. The catch is that people are not very good at reporting themselves honestly — and a badly worded question can create a result that was never really there.

📚 What you need to know

Closed and open questions do different jobs

The choice is not about which is better. It is about what kind of answer your research question needs. If you want to know how many, use closed questions. If you want to know why, use open ones. Most decent questionnaires use both.

The question type decides the data type Choose the question to fit the answer you actually need. CLOSED QUESTION Are you happy? Yes / No Rate this from 1 to 5 QUANTITATIVE DATA counts, scores, percentages Quick to compare. Tells you what. OPEN QUESTION What would you change? Describe a time you felt calm QUALITATIVE DATA words, themes, meaning Rich detail. Tells you why. Closed for the pattern, open for the explanation. A questionnaire made only of closed questions can show a trend but never account for it.
This is the reason questionnaires are so often criticised for being shallow: researchers use closed questions because they are quick to analyse, then have nothing to say about why.
A common question asks you to write one closed and one open question on the same topic. Do exactly that — same topic, different formats. Writing two closed questions in different words gets you nothing.

The honesty problem

Self-report only works if people tell you the truth. Often they do not, and usually they are not even lying on purpose. They are shaping the answer into the person they would like to be. Psychologists call this social desirability bias, and it hits questions about health, effort, prejudice, money and anything with a socially “right” answer.

Why self-report data drifts Most people are not lying. They are answering as their better self. WHAT THEY ACTUALLY DID skipped the gym all month WHAT THEY TICKED exercises four times a week SOCIAL DESIRABILITY BIAS under-report the embarrassing, over-report the impressive The data can be perfectly reliable and still not be true. Anonymity, neutral wording and filler questions all reduce the gap.
Reliability and validity come apart here. Ask the same person twice and you get the same flattering answer both times — consistent, and still wrong.

How to write questions that do not break

Four things wreck a questionnaire more than anything else. Learn them as a checklist and you will spot them instantly in an exam question.

ProblemWhat it looks likeWhy it damages the data
Leading questionDo you agree that homework is pointless?Tells the participant which answer is expected, so it measures obedience, not opinion
Double-barrelledDo you find school stressful and tiring?One answer, two questions. You cannot tell which half they agreed with
Jargon or vague wordsDo you experience frequent dysregulation?Different people read it differently, so responses are not comparable
Question order biasEmotive questions placed firstEarly questions frame the mood and shift every later answer
The order fix is simple: start neutral and general, then move to the specific and sensitive. Warm people up before you ask them anything they might defend themselves against.

Checking a questionnaire is reliable

Reliability means consistency, and there are two ways to test for it. Test-retest gives the same questionnaire to the same people after a gap — six months is typical — and checks whether the scores match. That is external reliability. Split-half takes one set of responses, splits the questionnaire into two halves, and checks whether both halves give a similar score. That is internal reliability: the items are all measuring the same thing.

🧩 Building a questionnaire for the IA

  1. Start from the construct. Write down exactly what you are trying to measure before you write a single question.
  2. Mostly closed items, so the data can actually be analysed with a statistical test.
  3. Add two or three open items at the end to explain any pattern you find.
  4. Read each item out loud. If it contains “and”, “obviously”, or a word you had to look up, rewrite it.
  5. Pilot it on a few people. Ask what they thought each question meant. Their answers will surprise you.
  6. Guarantee anonymity in writing. It is the cheapest available defence against social desirability bias.

Worked examples

WORKED EXAMPLE

Identify the flaws and rewrite the item

A researcher studying student wellbeing includes this item: “Don’t you agree that social media is harmful and makes teenagers anxious?” Identify two problems and rewrite the question.

Step 1: First problem Leading question — “Don’t you agree” signals the expected answer, so agreement is inflated. Step 2: Second problem Double-barrelled — it asks about harm and about anxiety at once, so a single answer cannot be interpreted. Step 3: Name the cost Both reduce the validity of the responses: the item no longer measures the participant’s real view. Step 4: Rewrite as one neutral idea “How often does using social media make you feel anxious?” with options: never, rarely, sometimes, often, always. Leading + double-barrelled; split into one neutral scaled item a rewrite must fix both faults, or you only get half the marks
WORKED EXAMPLE

Which reliability check, and why?

A team develops a 20-item stress scale. They want to be sure the items all measure stress and are not drifting onto other topics. They also want to know that a person’s score would not swing wildly from one month to the next. Explain which reliability check answers each concern.

Step 1: Concern one is about the items themselves Use the split-half method: split the 20 items into two sets of 10 and compare the two scores. Step 2: Name what that shows A close match means high internal reliability — the scale is consistent within itself. Step 3: Concern two is about stability over time Use test-retest: give the same scale to the same people after a set gap. Step 4: Name what that shows Similar scores mean high external reliability. Split-half for internal, test-retest for external internal = inside the test, external = across time. Same words, different job

💡 Exam tip

⚠ Common mix-up

Up next: Interviews and What They Reveal — the same self-report problem, but this time with a real person sitting opposite you.

Want this explained one-to-one?

Book a free session with an experienced IB Psychology tutor and get your trickiest topics made simple.

Book a Free Session →