IB Psychology SLTopic 1 — Mental Health DisordersPaper 1 & 2Learning approach~10 min read
Operant Conditioning as an Explanation of Depression
This explanation ignores genes, chemistry and thoughts entirely. It asks one blunt question instead: what is the person’s environment currently rewarding? And its answer explains something the other theories struggle with — why depression is so good at keeping itself going.
📚 What you need to know
Operant conditioning (OC) is learning through consequences (Skinner, 1953).
Reinforcement makes a behaviour more likely; punishment makes it less likely.
Positive reinforcement = behaviour repeated to gain something pleasant.
Negative reinforcement = behaviour repeated to avoid something unpleasant.
Skinner used the Skinner box, with a lever that released food and an electroplated floor that delivered shocks.
Applied to MDD: less positive reinforcement reduces motivation, and negative reinforcement keeps withdrawal going.
Lewinsohn et al. (2006) is the study to quote.
The four consequences
Students lose marks here more than anywhere else in the learning approach, because “negative reinforcement” sounds like punishment. It is not. Reinforcement always increases behaviour. The words positive and negative just mean “something added” or “something taken away”.
Type
What happens
Effect
Example
Positive reinforcement
Something pleasant is added.
Behaviour increases
Doing homework and getting praise for it.
Negative reinforcement
Something unpleasant is removed.
Behaviour increases
Doing homework to avoid a detention.
Positive punishment
Something unpleasant is added.
Behaviour decreases
Being told off.
Negative punishment
Something pleasant is removed.
Behaviour decreases
Not being allowed to go to a party.
Skinner’s research (1953)
Skinner placed one rat at a time in a box containing a lever that released food and an electroplated floor that could deliver a mild shock. Three things happened, and each one maps onto the depression explanation.
Positive reinforcement: a hungry rat pressed the lever by accident, a food pellet dropped, and the rat quickly learned to press it repeatedly.
Negative reinforcement: a rat receiving a shock accidentally pressed the lever, the shock stopped, and it learned to press the lever immediately when placed in the box.
Avoidance learning: a light signalled that a shock was coming, and the rat learned to press the lever as soon as the light came on — acting before anything unpleasant happened.
Avoidance learning is the one that matters most here. The rat is now pressing a lever to prevent something that has not happened yet. Swap the rat for a person avoiding a social event they think will go badly, and you have the mechanism that keeps depression running.
How this explains depression
Two processes work together, and the second is the sneaky one.
🧩 The two mechanisms
Loss of positive reinforcement. Social praise and attention drop off — a friendship ends, you leave a team, you move somewhere new. With fewer rewards coming in, motivation falls and the person withdraws.
Negative reinforcement of withdrawal. Avoiding social situations reduces anxiety in the moment. That relief is a reward. So the avoidance gets stronger, which reinforces the isolation, which removes even more chances for positive reinforcement.
Read the left-hand box carefully. Avoidance is rewarding in the short term, which is exactly why it wins over the long-term cost. The relief arrives immediately; the isolation arrives slowly.
🔬 Lewinsohn et al. (2006)
Measuring positive reinforcement in everyday life
AIM
To compare how much positive reinforcement patients with MDD received compared with non-depressed participants.
PARTICIPANTS
30 patients with MDD, a group with a different (non-depression) disorder, and a control group with no MDD. Because the groups already existed, this was a quasi-experiment.
PROCEDURE
Participants completed questionnaires daily for 30 days. They rated themselves on a daily mood checklist, and completed a pleasant activities schedule listing 320 activities such as sport, meditation, reading and socialising. Each activity was rated for enjoyment and for repetition on a scale of 0 to 3. Positive reinforcement was operationalised as activities that were both enjoyed and repeated.
RESULTS
Positive correlations were found between pleasant activities and good mood: the more pleasant activities someone took part in, the higher their mood rating.
CONCLUSION
Positive reinforcement from enjoyable activities may help maintain a good mood and protect against MDD.
Praise the operationalisation. “Positive reinforcement” is an abstract idea. Lewinsohn turned it into something countable: activities that were both enjoyed and repeated. That definition is genuinely clever, because repetition is the actual evidence that reinforcement happened. Saying this in an exam shows you understand what operationalising a variable means.
But which way does it run?
The finding is a correlation, so the same problem appears as with Beck. Do pleasant activities lift mood, or does better mood make people more likely to do pleasant things? Almost certainly both, in a loop — which is exactly what the cycle diagram above is showing.
There is a practical reason this matters. If activities cause mood, then deliberately scheduling activities should help even before someone feels like doing them. That is precisely what behavioural activation in CBT does, and it does work — which is decent indirect evidence for the direction Lewinsohn suggests.
Evaluating the explanation
Strengths
Standardised procedures give reliability, and daily real-life activity data gives good ecological validity.
Reinforcement is already applied successfully in token economies, for example with schizophrenia, supporting its clinical use.
It explains maintenance and relapse better than most models.
It leads directly to a treatment that can be started immediately.
Limitations
OC cannot explain why some people keep doing harmful things — self-blame or substance misuse bring no positive reinforcement at all.
It is based on Skinner’s animal research, so it is environmentally reductionist.
Humans are far more complex than animals and operate at a higher cognitive level.
The evidence is correlational, so causality is not established.
Linking to the concepts
Causality: OC uses the mechanisms of scientific inquiry — standardised procedures, objectivity, observable behaviour only — and strives for a cause–effect answer. The trouble is that human behaviour is far more nuanced than a rat pressing a lever for food.
Perspective: environmental reductionism frames a complex disorder purely as reward-seeking and punishment-avoiding. Given the evidence for a biological basis and for faulty schemas, MDD cannot be viewed as the result of one environmental influence alone.
EXAM ANSWER
Explain one behavioural explanation of one disorder. [9 marks]
State the principle first
Operant conditioning: behaviour is shaped by its consequences (Skinner, 1953).
Apply it — both mechanisms, not one
Reduced positive reinforcement lowers motivation and causes withdrawal. Withdrawal is then kept going by negative reinforcement, because avoiding removes anxiety.
Evidence
Lewinsohn et al. found positive correlations between pleasant activities and mood across 30 days of daily self-report.
One clean limitation
Correlational data means the direction is unclear, and OC cannot explain behaviours that bring no reward at all.
Link: strongest as an account of why depression persistssay what the theory is best AT — that is a judgement, not a hedge
💡 Exam tips
Define negative reinforcement correctly, in your own words, early on. Get it wrong and the whole answer wobbles.
Always give both mechanisms. Students who only mention lost rewards miss half the marks.
Name the design: quasi-experiment, daily self-report over 30 days, 320 listed activities.
Contrast this with Beck’s model — one says thoughts maintain depression, the other says consequences do. That contrast is a ready-made evaluation paragraph.
Link forward to behavioural activation in CBT to show the theory has practical value.
This page also serves as your example for the learning approach in Paper 1.
⚠ Common mix-ups
Calling negative reinforcement a punishment. Reinforcement always increases behaviour. Always.
Saying depression is caused by punishment. The explanation is mostly about rewards stopping, not bad things starting.
Calling Lewinsohn a true experiment. Groups already existed, so it is a quasi-experiment.
Describing the Skinner box without linking it to humans. The link is where the marks are.
Mixing up positive punishment with negative reinforcement. Ask yourself: is the behaviour going up or down?
Forgetting that a correlation cannot show direction. Say it once, clearly, and move on.
Up next: Why Social Support Protects Against Depression — if losing rewarding contact with people helps depression take hold, then having people around should do the opposite. Last page of the section.
Want this explained one-to-one?
Book a free session with an experienced IB Psychology tutor and get your trickiest topics made simple.