Classical vs operant conditioning: how to tell them apart every time

The definitions are easy and the scenarios are not. One question sorts almost every exam item: did the consequence come before the behaviour, or after it?

7 min read · Updated


Almost nobody loses marks defining these. People lose marks on the scenario items, where a paragraph describes someone learning something and four options wait to see whether you can classify it. The good news is that one question resolves most of them.

Classical conditioning: learning what predicts what

Classical conditioning pairs a neutral stimulus with something that already produces an automatic response, until the neutral stimulus produces that response on its own. The learner does nothing deliberate. The response is a reflex.

The four terms come up constantly and are worth being exact about, because exams ask you to label the parts rather than describe the process.

TermWhat it meansIn Pavlov's study
Unconditioned stimulus (UCS)Produces the response with no learning at allFood
Unconditioned response (UCR)The automatic reaction to the UCSSalivating at food
Conditioned stimulus (CS)Was neutral, then got paired with the UCSThe bell
Conditioned response (CR)The learned reaction to the CS aloneSalivating at the bell

A reliable way to keep the labels straight: unconditioned means unlearned. If it works without any training, it is unconditioned. If it only works because of pairing, it is conditioned.

The processes that follow

  • Extinction: the CS stops being paired with the UCS, so the CR fades. The bell rings and no food arrives, so eventually the dog stops salivating.
  • Spontaneous recovery: after a rest, the CR comes back weakly without any new pairing. Examiners like this one because it looks like a contradiction of extinction.
  • Generalisation: stimuli similar to the CS trigger the response too. A different bell also produces salivation.
  • Discrimination: the opposite. The response occurs only to the specific CS and not to similar ones.

Operant conditioning: learning what works

Operant conditioning changes voluntary behaviour through its consequences. The behaviour comes first, something follows, and how often the behaviour happens changes as a result.

The four-cell grid below is where most marks on this topic live and die. Read the two words separately: the first says whether something is added or removed, the second says whether the behaviour goes up or down.

Adds somethingRemoves something
Behaviour increasesPositive reinforcement: praise for good workNegative reinforcement: the seatbelt alarm stops when you buckle up
Behaviour decreasesPositive punishment: a fine for speedingNegative punishment: phone confiscated for a week

Schedules of reinforcement

Questions about which schedule resists extinction best have a consistent answer: variable ratio. Reinforcement after an unpredictable number of responses produces high, steady rates that persist long after the reward stops, which is why slot machines use it.

ScheduleReward arrivesExample
Fixed ratioAfter a set number of responsesPaid per ten items packed
Variable ratioAfter an unpredictable numberSlot machine
Fixed intervalFirst response after a set timeWeekly payslip
Variable intervalFirst response after an unpredictable timeChecking for a reply to a message

Where the two get confused

The hard items describe a situation containing both. A child is bitten by a dog and is now frightened of dogs: that is classical, since the fear is a reflex paired with the animal. The same child then avoids the park where the dog lives, and the fear drops when they avoid it: that is operant, because avoiding is voluntary and the relief that follows reinforces it.

That combination is exactly how phobias are usually explained, and questions on it expect you to name both processes rather than pick one.

Test yourself

A driver puts on their seatbelt to stop the car making its warning noise. Over time they buckle up immediately on getting in. This is:

  • Positive reinforcement
  • Negative reinforcementCorrect
  • Positive punishment
  • Classical conditioning

The behaviour, buckling up, increases, so it is reinforcement rather than punishment. What follows is the removal of something unpleasant, the noise, so it is negative. Buckling up is voluntary and the consequence comes afterwards, which rules out classical conditioning.

The four-cell grid is quick to memorise and slow to apply under time pressure. The only thing that makes it automatic is answering enough scenario questions that you stop translating and start recognising.

Practise this

Reading it is not the same as recognising it

Learning & Conditioning questions in ExamRoad explain every option, not just the correct one, so you find out why the answer you nearly picked was wrong. All 447 of them, in one purchase.

One payment, no subscription.

Common questions

What is the simplest difference between classical and operant conditioning?
Timing. In classical conditioning the important event comes before the response and the behaviour is an involuntary reflex. In operant conditioning the important event comes after the behaviour, and the behaviour is voluntary.
Is negative reinforcement the same as punishment?
No, and this is the single most common error on the topic. Negative reinforcement removes something unpleasant in order to increase a behaviour. Punishment decreases a behaviour. Reinforcement always increases, whatever the label attached to it.
Which type is Pavlov's dog?
Classical. The dog did not choose to salivate and nothing followed the salivation. A bell was paired with food until the bell alone produced the reflex.
Which type is Skinner's box?
Operant. The rat pressed a lever, which is voluntary, and a consequence followed: food appeared, or a shock stopped. The consequence changed how often the lever was pressed.

Keep revising

Classical vs Operant Conditioning: The Difference · ExamRoad