Classical and operant conditioning are two major ways psychologists explain learning through experience. This cheat sheet helps students compare reflexive associations with behavior shaped by consequences. It is useful for identifying examples, predicting behavior changes, and avoiding confusion between similar terms.
Students in psychology, biology, and social science courses often use these ideas to explain everyday learning.
Key Facts
- Classical conditioning forms an association between two stimuli, such as when a neutral stimulus becomes linked to an unconditioned stimulus.
- In classical conditioning, UCS leads to UCR, then NS plus UCS leads to UCR, and finally CS leads to CR.
- Operant conditioning changes voluntary behavior by using consequences that follow the behavior.
- Positive reinforcement means adding a desirable stimulus after a behavior to increase that behavior.
- Negative reinforcement means removing an unpleasant stimulus after a behavior to increase that behavior.
- Positive punishment means adding an unpleasant consequence after a behavior to decrease that behavior.
- Negative punishment means removing a desirable stimulus after a behavior to decrease that behavior.
- A variable ratio schedule usually produces high and steady responding because reinforcement comes after an unpredictable number of responses.
Vocabulary
- Classical conditioning
- A type of learning in which a neutral stimulus becomes associated with a stimulus that naturally produces a reflexive response.
- Operant conditioning
- A type of learning in which voluntary behavior becomes more or less likely because of its consequences.
- Reinforcement
- A consequence that increases the likelihood that a behavior will happen again.
- Punishment
- A consequence that decreases the likelihood that a behavior will happen again.
- Extinction
- The weakening or disappearance of a learned response when the association or consequence is no longer present.
- Reinforcement schedule
- A rule that describes how often and under what conditions a behavior is reinforced.
Common Mistakes to Avoid
- Calling every learned behavior classical conditioning is wrong because classical conditioning involves automatic reflexes, while operant conditioning involves voluntary actions.
- Confusing negative reinforcement with punishment is wrong because negative reinforcement increases behavior by removing something unpleasant, while punishment decreases behavior.
- Labeling positive as good and negative as bad is wrong because positive means something is added and negative means something is removed.
- Forgetting the timing of consequences is wrong because operant conditioning depends on what happens after the behavior, not before it.
- Mixing up fixed ratio and fixed interval schedules is wrong because ratio schedules depend on number of responses, while interval schedules depend on time.
Practice Questions
- 1 A student gets a point every 5 times they answer a review question correctly. What reinforcement schedule is being used?
- 2 A seatbelt alarm stops when a driver buckles the seatbelt. Is this positive reinforcement, negative reinforcement, positive punishment, or negative punishment?
- 3 A dog hears a bell before receiving food. After many pairings, the dog salivates when it hears the bell alone. Identify the UCS, UCR, CS, and CR.
- 4 Explain why studying harder after earning praise from a teacher is an example of operant conditioning rather than classical conditioning.
Understanding Classical vs Operant Conditioning
A useful first step is to identify what the learner is doing at the moment learning occurs. Classical conditioning often involves an automatic reaction that happens before deliberate choice. A smell can bring nausea, a school bell can create excitement, and a certain song can trigger sadness because it was present during an emotional event.
The brain learns that one event predicts another. This prediction helps organisms prepare for food, danger, pain, or comfort. The learned response may be helpful, but it can become troublesome when a harmless cue produces fear or anxiety.
Operant conditioning depends on what happens after a chosen action. Consequences change the chances that the action will occur again in a similar situation. The words positive and negative describe whether something is added or removed.
They do not describe whether the result is kind, fair, or effective. A student who buckles a seat belt to stop an annoying alarm is affected by negative reinforcement. A student who loses phone time after breaking a rule experiences negative punishment.
Confusing negative reinforcement with punishment is one of the most common errors in psychology classes. Reinforcement always increases a behavior. Punishment always decreases a behavior, at least in the intended situation.
Timing matters greatly. A consequence has the strongest effect when it follows the behavior quickly and clearly. If praise arrives long after a student has finished studying, the student may not connect the praise to the study behavior.
People also learn from patterns, not from one event alone. Continuous reinforcement, where every correct response gets a reward, can build a new behavior quickly. Once the behavior is established, less predictable rewards often make it persist longer.
This helps explain why checking social media, playing some games, and gambling can become hard habits to stop. An occasional reward can keep a person responding even after many unrewarded attempts.
Extinction does not mean a memory is erased. It means the learned response becomes weaker when the expected outcome no longer occurs. A dog may gradually stop running to the kitchen if the sound of a food container no longer predicts food.
The response can return after a break, a process called spontaneous recovery. It can also reappear in a new setting or after a similar experience. When studying examples, pay attention to the exact behavior, the event before it, and the consequence after it.
Decide whether the response is reflexive or voluntary. Then ask whether the result makes that behavior more likely or less likely in the future. This method is more reliable than choosing an answer based on words such as reward, negative, or punishment.