Sign in to save

Bookmark this page so you can find it later.

Sign in to save

Bookmark this page so you can find it later.

Operant conditioning is a type of learning in which behavior changes because of its consequences. It matters because rewards and penalties shape many everyday habits, from studying for a quiz to practicing an instrument. In the classroom, on a sports team, or at home, people often repeat actions that lead to satisfying outcomes.

This makes operant conditioning a powerful tool for understanding motivation and behavior change.

The main idea is that a behavior is followed by a consequence, and that consequence affects how likely the behavior is to happen again. Reinforcement increases a behavior, while punishment decreases a behavior. Positive means something is added, and negative means something is removed, so these words do not simply mean good or bad.

A reward loop works best when the consequence is clear, consistent, and closely connected in time to the behavior.

Understanding Psychology: Operant Conditioning

Psychologist B. F. Skinner studied how consequences can build patterns over time.

In a Skinner box, an animal could press a lever to receive food or to stop an unpleasant sound. The important measurement was not simply whether the animal learned once. Skinner tracked how often the behavior occurred and how long it lasted.

This helps psychologists separate a brief reaction from a stable habit. In daily life, similar patterns appear when a student checks a phone, refreshes a game, or keeps practicing a skill. Each action has a possible result, and the brain gradually learns which actions seem worth repeating.

The timing and pattern of consequences make a major difference. A consequence that follows immediately is usually easier to connect with the action. If praise comes days after a student helped in class, the link may be weak.

Consequences can occur every time a behavior happens or only sometimes. A vending machine that works after every correct payment uses a predictable pattern. A game that gives an unexpected prize after some actions uses an unpredictable pattern.

Unpredictable rewards can produce very persistent behavior because people keep trying for the next possible reward. This is one reason notifications, online games, and some gambling systems can be hard to stop.

Behavior can be taught in small stages through shaping. Instead of waiting for a difficult final behavior, a teacher, coach, or parent can respond to steps that move in the right direction. A child learning to organize a backpack might first get recognition for putting one book away.

Later, recognition can depend on packing all needed materials without reminders. The standard rises gradually. This works because large tasks can feel too difficult at first.

Clear feedback helps a person know exactly what action led to the result. Vague comments such as good job are less useful than feedback that identifies the specific behavior.

It is important to distinguish negative reinforcement from punishment because the terms are often confused. Negative reinforcement makes a behavior more likely because something unpleasant stops or is avoided. Buckling a seat belt to stop a warning sound is one example.

Punishment aims to reduce a behavior, but it does not always teach a better replacement. Harsh punishment may create fear, anger, or hiding instead of real learning. When a previously rewarded behavior no longer receives its usual result, it may fade through extinction.

At first, the behavior can briefly increase, which is called an extinction burst. Students should pay attention to what behavior is being measured, what consequence follows it, how quickly it follows, and whether another factor such as stress or peer pressure may be affecting the result.

Key Facts

  • Operant conditioning: behavior + consequence = change in future behavior.
  • Reinforcement increases the probability of a behavior.
  • Punishment decreases the probability of a behavior.
  • Positive reinforcement adds a pleasant consequence, such as praise after homework is completed.
  • Negative reinforcement removes an unpleasant condition, such as stopping a reminder alarm after studying begins.
  • Response rate = number of responses / time, such as 30 practice problems / 15 min = 2 problems per min.

Vocabulary

Operant conditioning
Operant conditioning is learning in which the consequences of a behavior change how often that behavior occurs in the future.
Reinforcement
Reinforcement is any consequence that makes a behavior more likely to happen again.
Punishment
Punishment is any consequence that makes a behavior less likely to happen again.
Positive reinforcement
Positive reinforcement occurs when something desirable is added after a behavior, increasing that behavior.
Shaping
Shaping is the process of reinforcing small steps that gradually lead to a desired behavior.

Common Mistakes to Avoid

  • Confusing positive with good is wrong because positive means something is added, not that the consequence feels pleasant.
  • Calling every reward reinforcement is wrong because a reward only counts as reinforcement if it actually increases the behavior later.
  • Waiting too long to give feedback is ineffective because consequences are usually strongest when they follow the behavior closely in time.
  • Ignoring the behavior being measured is a mistake because operant conditioning focuses on observable actions, not vague traits like laziness or attitude.

Practice Questions

  1. 1 A student completes homework on 8 out of 10 school nights after receiving praise each time homework is turned in. What percent of school nights did the student complete homework?
  2. 2 During piano practice, a student earns 1 point for every 5 minutes of focused practice. If the student practices for 35 minutes, how many points are earned?
  3. 3 A teacher stops giving reminder prompts once students begin taking out their notebooks on time. Explain whether this is positive reinforcement, negative reinforcement, positive punishment, or negative punishment, and justify your answer.