Classical conditioning is about what happens before a behaviour. Operant conditioning is about what happens after. Do something, see what it gets you, and adjust. Simple — until you notice how many people repeat behaviour that clearly costs them.
📚 What you need to know
Operant conditioning (OC) is learning via consequences. With classical conditioning it forms the base of behaviourism.
Reinforcement makes a behaviour more likely. Punishment makes it less likely.
Positive means something is added. Negative means something is taken away. Neither word means good or bad.
Positive reinforcement: doing something to get a pleasant consequence.
Negative reinforcement: doing something to avoid or escape an unpleasant one.
The consequence itself is called the reinforcer.
The study you need is Skinner (1953) and the Skinner box.
Four boxes, two questions
Students lose easy marks here because they read ‘negative’ as ‘bad’. It does not mean that. Ask two questions instead: is the behaviour going up or down, and is something being added or taken away? The answers give you the term.
A reinforcer is the consequence, not the behaviour. Praise from a teacher is a positive reinforcer; a detention is a positive punishment.
Sometimes a behaviour is repeated not because it brings something good but because it stops something annoying. Choosing the salad so your friend stops commenting on your diet is negative reinforcement, not healthy eating.
Key study: Skinner (1953)
Skinner argued that learning is active. Organisms operate on their environment and get shaped by what their actions produce. He identified three types of consequence: neutral operants, which change nothing; reinforcers, which make a behaviour more likely; and punishers, which make it less likely.
Condition in the Skinner box
What happened
Positive reinforcement
A hungry rat exploring the box accidentally pressed a lever and received a food pellet. Adding food after the press increased lever-pressing, and rats soon pressed the lever immediately on being placed in the box.
Negative reinforcement (escape)
A mild electric current ran through the grid floor. Pressing the lever turned it off. Removing the unpleasant stimulus reinforced the behaviour, and rats learned to press quickly to escape.
Negative reinforcement (avoidance)
A light signalled that a shock was coming — a discriminative stimulus. Rats learned to press the lever as soon as the light appeared, avoiding the shock entirely.
What it showed
Both adding a pleasant stimulus and removing an unpleasant one strengthen behaviour. Reinforcement is about the effect on the behaviour, not about whether the consequence feels nice.
Evaluating operant conditioning
What is strong about it
Good application to how phobias are maintained. Someone with social phobia avoids gatherings; avoidance is negative reinforcement; each avoidance brings relief and security, so the phobia gets stronger every time.
Skinner used standardised procedures in controlled conditions, giving good reliability. The theory satisfies falsifiability, so it can be tested scientifically.
It underpins a huge amount of practical work in classrooms, therapy and animal training.
What is weaker about it
It cannot explain why people repeat behaviours that are damaging or unpleasant. Someone who self-harms may do so for the relief it brings, but that would not be recognised as a positive reinforcer. People carry on smoking while disliking the taste and smell.
It is overly simplistic — environmental reductionism. Humans operate at a far higher cognitive level than rats and can take control of their behaviour through mechanisms such as self-efficacy.
Link to concepts
Responsibility: by today’s standards this research raises real ethical problems. Keeping animals in conditions where they are repeatedly harmed by electric shocks looks unnecessarily cruel. Researchers are now expected to follow the 3 Rs: reduce the number of animals used, replace them with alternatives where possible, and refine procedures to minimise suffering.
Perspective: operant conditioning cannot explain people who are offered a route out of a harmful situation and do not take it. Some victims of domestic abuse do not leave even when escape is possible. Learned helplessness explains this better — after repeated exposure, a person comes to believe they have no power to change anything, so the chance to escape goes unused.
EXAM PRACTICE
Discuss operant conditioning as an explanation of learning. [22]
Step 1: define it precisely
Learning through consequences. Reinforcement increases behaviour; punishment decreases it. Positive means added, negative means removed.
Step 2: the study
Skinner (1953), the Skinner box, with rats.
Step 3: the conditionsFood pellet for lever-pressing increased the behaviour. Turning off a shock did the same. With a warning light, rats pressed to avoid the shock altogether.
Step 4: application
Explains how phobias are maintained: avoidance brings relief, so avoidance is reinforced.
Step 5: evaluate
Reliable and falsifiable, but reductionist, cannot explain self-damaging behaviour, ethically questionable, and learned helplessness explains some cases better.
Definition + study + conditions + application + evaluationThe self-harm and smoking examples are the sharpest limitation. They are behaviours OC says should not exist.
💡 Exam tip
Spell out that negative reinforcement is not punishment. Examiners see this error constantly.
The escape and avoidance variations are two separate findings. Use both.
Phobia maintenance is the best application. Pair it with systematic desensitisation from classical conditioning.
The 3 Rs are quick, precise ethics marks.
Learned helplessness is a strong alternative explanation worth a short paragraph.
⚠ Common mix-up
Treating negative reinforcement as punishment. It increases behaviour, so it is reinforcement.
Confusing the reinforcer with the behaviour. The reinforcer is the consequence that follows.
Mixing up classical and operant. Classical is stimulus before; operant is consequence after.
Saying the rat was taught to press the lever. The first press was accidental; the consequence did the rest.
Assuming animal findings transfer straight to humans. That assumption is itself a limitation.
Up next: The Dual Process Model of Thinking — back to cognition, and the two very different gears the mind switches between.
Want this explained one-to-one?
Book a free session with an experienced IB Psychology tutor and get your trickiest topics made simple.