Introduction to Operant Conditioning
Welcome to one of the most practical areas of psychology! Have you ever wondered why you keep checking your phone for likes, or why you work harder when you know there’s a reward at the end? This is all down to Operant Conditioning.
While Classical Conditioning (which you’ll study in another chapter) is about learning through association, Operant Conditioning is about learning through consequences. Essentially, if you do something and the result is good, you’ll probably do it again. If the result is bad, you’ll likely stop. It was famously developed by B.F. Skinner, who believed that almost all behavior is shaped by the environment rather than internal thoughts.
Quick Tip: Think of Operant Conditioning as "Trial and Error" learning. We "operate" on our environment, and the environment gives us feedback!
1. Reinforcement and Punishment
To master this topic, you must understand the difference between Reinforcement and Punishment. Many students get these mixed up, so here is the golden rule:
- Reinforcement: Always makes a behavior more likely to happen again (it strengthens the behavior).
- Punishment: Always makes a behavior less likely to happen again (it weakens the behavior).
The Positive and Negative Distinction
In psychology, "Positive" and "Negative" don't mean "Good" and "Bad." Instead, think of them like math symbols:
- Positive (\(+\)): Adding something new to the situation.
- Negative (\(-\)): Taking something away from the situation.
Positive Reinforcement
This involves adding a pleasant stimulus after a behavior to encourage it.
Example: Receiving a chocolate bar (adding something good) for finishing your psychology homework. You are now more likely to do your homework next time!
Negative Reinforcement
This involves removing an unpleasant stimulus to encourage a behavior.
Example: The annoying "beep-beep" sound in a car stops (removing something bad) once you put on your seatbelt. You are now more likely to wear your seatbelt to avoid the noise.
Punishment
This is designed to stop a behavior.
1. Positive Punishment: Adding something unpleasant (e.g., getting a detention for being late).
2. Negative Punishment: Taking away something pleasant (e.g., having your phone confiscated for talking in class).
Don't worry if this seems tricky at first! Just remember: Reinforcement = Repeat, Punishment = Put a stop to it.
2. Primary and Secondary Reinforcement
Not all rewards are created equal. Psychologists divide reinforcers into two categories:
Primary Reinforcers
These are things that we find naturally rewarding because they satisfy basic biological needs. We don't have to learn to like them; we are born wanting them.
Examples: Food, water, sleep, and warmth.
Secondary Reinforcers
These are things that have no value on their own, but they become rewarding because we associate them with primary reinforcers.
Example: Money. You can’t eat a \$20 bill, and it won't keep you warm, but you can use it to buy food and clothes (Primary Reinforcers). Tokens and high grades are also secondary reinforcers.
3. Schedules of Reinforcement
Skinner discovered that when and how often you reward a behavior changes how quickly a subject learns. There are two main types:
- Continuous Reinforcement: The behavior is reinforced every single time it happens. This is great for learning a new trick quickly, but the behavior stops quickly if the rewards stop.
- Partial (Intermittent) Reinforcement: The behavior is reinforced only some of the time. This makes the behavior much harder to "extinguish" (stop).
The Four Types of Partial Schedules:
1. Fixed Ratio: Reinforcement after a set number of responses (e.g., a "buy 5 coffees, get 1 free" card).
2. Variable Ratio: Reinforcement after an unpredictable number of responses (e.g., a slot machine). This is highly addictive!
3. Fixed Interval: Reinforcement after a set amount of time (e.g., getting a paycheck every Friday).
4. Variable Interval: Reinforcement after an unpredictable amount of time (e.g., checking your phone for a text message).
4. Skinner (1948): Superstition in the Pigeon
This is a prescribed study for your Unit 2 exam. You need to know what Skinner did and what he found.
The Aim
To see if pigeons would develop "superstitious" behaviors if they were given food at regular intervals, regardless of what they were doing.
The Procedure
Skinner placed hungry pigeons in a cage (the "Skinner Box"). He set a food hopper to deliver food automatically every \(15\) seconds, no matter what the pigeon was doing. The food was not a reward for any specific action.
The Findings
In \(6\) out of \(8\) pigeons, the birds developed strange, repetitive behaviors.
- One pigeon spun in circles anti-clockwise.
- Another thrust its head into a corner of the cage.
- Others developed "pendulum" movements with their heads.
The Conclusion
The pigeons were acting as if their specific movement had caused the food to appear. Because the food arrived shortly after they performed a random movement, that movement was accidentally reinforced. This explains how humans might develop superstitions (like wearing "lucky socks" for an exam because they happened to wear them once when they did well).
Key Takeaways for Revision
Quick Review Box:
- Positive Reinforcement: Add good to increase behavior.
- Negative Reinforcement: Remove bad to increase behavior.
- Punishment: Decreases behavior.
- Primary Reinforcer: Biological (Food).
- Secondary Reinforcer: Learned (Money).
- Skinner (1948): Showed that accidental reinforcement leads to "superstitious" behavior in pigeons.
Common Exam Mistakes to Avoid
- Mistaking Negative Reinforcement for Punishment: Remember, Negative Reinforcement is a good thing—it helps you escape something unpleasant! Punishment is always intended to be bad.
- Confusing Ratio and Interval: Ratio is about the number of times you do something. Interval is about the time that has passed.
- Failing to link to the scenario: If the exam asks about a student getting gold stars, make sure you identify the gold star as a secondary reinforcer and positive reinforcement.