Operant Conditioning
| English | Chinese | Pinyin |
|---|---|---|
| Law of Effect | 效果律 | xiào guǒ lǜ |
| Reinforcement | 强化 | qiáng huà |
| punishment | 惩罚 | chéng fá |
| Positive | 正 | zhèng |
| negative | 负 | fù |
| Shaping | 塑造 | sù zào |
| primary reinforcer | 原级强化物 | yuán jí qiáng huà wù |
| secondary reinforcer | 次级强化物 | cì jí qiáng huà wù |
| successive approximations | 逐步接近 | zhú bù jiē jìn |
| Continuous reinforcement | 连续强化 | lián xù qiáng huà |
| ratio | 比率 | bǐ lǜ |
| interval | 时距 | shí jù |
| Variable ratio | 可变比率 | kě biàn bǐ lǜ |
| Superstitious behaviour | 迷信行为 | mí xìn xíng wéi |
| Learned helplessness | 习得性无助 | xí dé xìng wú zhù |
Behaviour that pays is behaviour that repeats
- Classical conditioning links two stimuli. Operant conditioning links a behaviour to its consequence.
- The Law of Effect 效果律 states that behaviour followed by a good outcome is more likely to recur.
- The learner is now doing something, not just receiving something.
Four cells, two questions
- Reinforcement 强化 increases a behaviour; punishment 惩罚 decreases it.
- Positive 正 means something is added; negative 负 means something is removed.
- So negative reinforcement still increases behaviour — by removing something unpleasant.
Which cell of the grid?
Sort each consequence by whether it adds or removes, and increases or decreases.
Taking a painkiller to remove a headache, and doing it again next time, is an example of...
Something unpleasant was removed and the behaviour increased — negative reinforcement.
Primary, secondary, and shaping
- A primary reinforcer 原级强化物 satisfies a biological need; a secondary reinforcer 次级强化物 is learned, like money.
- Shaping 塑造 builds a complex behaviour by reinforcing successive approximations 逐步接近.
- Shaping is how a behaviour that never occurs spontaneously can still be reinforced.
Reinforcing successive approximations to build a complex behaviour is called ____.
Shaping is how a behaviour that never occurs spontaneously can be reinforced into existence.
Select all that are true of a secondary reinforcer.
A primary reinforcer satisfies a biological need directly; a secondary one is learned.
Negative reinforcement is not punishment. Negative means something is removed; reinforcement means the behaviour increases. Taking an aspirin to remove a headache is negative reinforcement — and you are more likely to do it again.
Which schedule produces the highest and most extinction-resistant response rate?
Variable ratio — an unpredictable number of responses per reward.
Schedules of reinforcement
- Continuous reinforcement 连续强化 rewards every response — fast learning, fast extinction.
- Partial schedules are fixed or variable, by ratio 比率 (number of responses) or interval 时距 (time).
- Variable ratio 可变比率 produces the highest, most extinction-resistant rate of responding.
Superstitious behaviour occurs when a consequence reinforces a behaviour that did not cause it.
The learner tracks what followed, not what actually caused the outcome.
When consequences mislead
- Superstitious behaviour 迷信行为 occurs when a consequence reinforces an unrelated behaviour.
- Learned helplessness 习得性无助 occurs when repeated inescapable outcomes teach that nothing one does matters.
- Both show that the learner tracks what followed, not what actually caused it.
A slot machine pays out after an unpredictable number of pulls — a variable-ratio schedule. That is exactly the schedule that produces the highest response rate and the greatest resistance to extinction, which is why the design is not accidental.
Operant conditioning links behaviour to consequence via the Law of Effect. Reinforcement increases and punishment decreases; positive adds and negative removes. Shaping builds new behaviour by successive approximations, and variable-ratio schedules produce the most persistent responding.