Unit 3 - Operant Conditioning Approach
Operant Conditioning and Its Principles
Introduction and Definition
Operant Conditioning, also known as instrumental learning, is a fundamental concept in
behavioral psychology developed primarily by B. F. Skinner (1938). It refers to the process by
which behaviors are influenced by their consequences — meaning that actions followed by
satisfying outcomes are strengthened, while those followed by unfavorable consequences are
weakened.
Operant conditioning is a learning process through which voluntary behaviors are modified
by their consequences.
First outlined by Edward Thorndike’s Law of Effect (1911) and expanded by B. F. Skinner
(1938), it is one of the most influential principles in behavioral psychology.
Definition (Feldman, 2009):
Operant conditioning is learning in which a voluntary response is strengthened or
weakened depending on its favorable or unfavorable consequences.
Definition (Miltenberger, 2008):
Operant behavior is any behavior that acts on the environment to produce
consequences and is controlled by those consequences.
Historical Foundations
The roots of operant conditioning can be traced back to Edward L. Thorndike’s (1911) Law of
Effect, which proposed that behaviors producing satisfying effects are more likely to recur.
Skinner expanded on Thorndike’s work by introducing a more systematic and experimental
analysis of how behavior is shaped and maintained by environmental contingencies.
According to Miltenberger (2008), operant behavior is any voluntary behavior that operates
on the environment to produce consequences that influence its future occurrence. The
environment, in turn, provides either reinforcing or punishing feedback that determines whether
the behavior is strengthened, maintained, or diminished.
Thus, operant conditioning emphasizes voluntary and goal-directed actions, unlike classical
(respondent) conditioning, which involves reflexive or automatic responses.
Theorist Contribution
Edward L. Formulated the Law of Effect – behaviors followed by satisfying
Thorndike outcomes are more likely to be repeated.
B. F. Skinner Distinguished operant conditioning (behavior shaped by consequences)
from respondent conditioning (behavior elicited by stimuli).
John B. Watson Advocated behaviorism: behavior is determined by environmental
stimuli, not internal states.
Miltenberger Applied these principles systematically to human learning, therapy, and
(2008) social functioning through behavior modification.
Core Concepts of Operant Conditioning
Concept Definition Example
Reinforcement Consequence that increases the Giving praise after homework
likelihood of behavior recurring. completion.
Punishment Consequence that decreases Scolding after rule-breaking.
likelihood of behavior recurring.
Positive Addition of a stimulus. Giving a reward (positive
reinforcement) or a reprimand (positive
punishment).
Negative Removal of a stimulus. Taking away pain (negative
reinforcement) or privileges (negative
punishment).
Extinction Withholding reinforcement leads Ignoring attention-seeking tantrums.
to decrease in behavior.
The ABC Model (Three-Term Contingency)
(Miltenberger, 2008)
Antecedent → Behavior → Consequence
● Antecedent (A): Environmental cue before behavior (e.g., teacher’s instruction).
● Behavior (B): Observable response (e.g., student raising hand).
● Consequence (C): Outcome following behavior (e.g., teacher praises).
This sequence determines whether behavior will increase, decrease, or maintain.
Operant Conditioning in Human Behaviour (Feldman, 2009)
Behavior modification applies operant principles to real-life improvement.
Applications:
● Education: Reinforcement for participation or accuracy.
● Clinical Psychology: Behavior therapy for phobias, addictions, autism.
● Parenting: Token economies, time-outs, praise systems.
● Health: Exercise and diet compliance.
● Workplace: Incentive programs, performance-based pay.
Example:
A couple reduced arguments using a reinforcement contract — rewarding each
other for chores completed and imposing response cost for failures.
2. The Core Concept: The Operant Conditioning Process
The operant conditioning process can be summarized as follows:
1. A behavior (response) occurs.
2. A consequence follows that behavior.
3. Depending on the nature of the consequence, the likelihood of the behavior occurring
again is increased, maintained, or decreased.
Skinner described this relationship as the three-term contingency:
Antecedent → Behavior → Consequence
or more precisely,
Discriminative Stimulus (SD) → Response (R) → Reinforcer (SR)
This model illustrates that behavior is not random — it is governed by predictable environmental
relationships.
3. Principles of Operant Conditioning
Operant conditioning is based on a few key principles that describe how consequences affect
behavior. These principles form the basis of behavior modification techniques in psychology.
A. Reinforcement
Reinforcement is the process of increasing the likelihood of a behavior by following it with a
consequence that strengthens that behavior. Reinforcers can be positive or negative, depending
on whether a stimulus is added or removed after the behavior.
1. Positive Reinforcement
● Occurs when the presentation of a pleasant or rewarding stimulus follows a behavior,
thereby strengthening it.
● Example: A student studies hard and receives praise or good grades; the praise acts as a
positive reinforcer, encouraging the student to study again.
Types of Positive Reinforcers:
● Social Reinforcers (praise, attention, approval)
● Tangible Reinforcers (money, food, toys)
● Activity Reinforcers (access to preferred activities)
● Token Reinforcers (points, tokens, stars used in token economies)
2. Negative Reinforcement
● Occurs when a behavior is strengthened because it removes or avoids an aversive
stimulus.
● Example: Taking painkillers removes a headache; wearing a raincoat avoids getting wet.
Both increase the likelihood of the behavior recurring.
Forms of Negative Reinforcement:
● Escape behavior: Behavior terminates an aversive event already present (e.g., leaving a
noisy room).
● Avoidance behavior: Behavior prevents an aversive event from occurring (e.g.,
submitting work early to avoid reprimand).
🔹 In both cases, the removal or prevention of discomfort increases the probability
of the behavior recurring.
B. Punishment
Punishment is the process of decreasing the likelihood of a behavior by following it with an
unpleasant consequence. Like reinforcement, it has two types.
1. Positive Punishment
● A behavior is followed by the presentation of an aversive stimulus, reducing its future
occurrence.
Example: A child touches a hot stove and feels pain; the painful stimulus discourages the
behavior from happening again.
2. Negative Punishment
● A behavior is followed by the removal of a pleasant stimulus, reducing its occurrence.
Example: A teenager misses curfew, and the parents take away phone privileges; loss of
privilege discourages future lateness.
NOTES: Miltenberger (2008) emphasizes that while punishment can suppress behavior
temporarily, it often produces undesirable side effects (anger, fear, avoidance) and may
not teach alternative adaptive behaviors. Thus, reinforcement-based methods are more
ethically and effectively used in behavior modification.
C. Extinction
Extinction occurs when a previously reinforced behavior no longer receives reinforcement,
leading to a gradual decrease in that behavior over time.
● Example: If a child’s tantrum no longer gains parental attention, the tantrum behavior
eventually decreases.
● During extinction, a temporary increase in the behavior called an extinction burst may
occur, followed by a gradual decline.
● Sometimes, the behavior may reappear briefly after extinction — this is called
spontaneous recovery.
Extinction is a critical behavioral principle used to eliminate maladaptive or problem
behaviors.
D. Stimulus Control and Discriminative Stimuli
Operant behaviors often occur under specific environmental cues or conditions. A
discriminative stimulus (SD) is a signal that indicates reinforcement is available following a
particular behavior.
● Example: A traffic light turning green signals that pressing the accelerator (behavior) will
be reinforced by moving forward safely.
● The absence of reinforcement (e.g., a red light) discourages the same behavior.
This process of learning when a behavior will be reinforced is called stimulus discrimination,
and the ability of a behavior to occur in similar situations is called generalization.
This three-term relationship — SD → R → SR — forms the core structure of
operant learning.
E. Schedules of Reinforcement
The schedule describes how often or when reinforcement is delivered for a particular behavior.
Different schedules produce different patterns of responding.
Type of Schedule Definition Behavioral Pattern / Example
Continuous Reinforcement follows Rapid learning but fast
Reinforcement every instance of extinction (e.g., giving a treat
(CRF) behavior every time a dog sits)
Fixed Ratio (FR) Reinforcement after a set High rate of responding with
number of responses pauses (e.g., factory
piecework)
Variable Ratio Reinforcement after an High steady rate, resistant to
(VR) unpredictable number extinction (e.g., slot
of responses machines)
Fixed Interval Reinforcement after a “Scalloped” pattern—response
(FI) fixed time interval increases as interval ends
(e.g., weekly paycheck)
Variable Interval Reinforcement after Steady moderate response rate
(VI) varying time intervals (e.g., checking email)
F. Shaping
Shaping is a technique of reinforcing successive approximations toward a target behavior.
It is particularly useful when the desired behavior is complex or not yet in the individual’s
repertoire.
Example: Teaching a child with autism to speak — the therapist reinforces initial attempts like
vocal sounds, then syllables, then full words.
Shaping involves:
1. Defining the target behavior.
2. Identifying successive steps leading to it.
3. Reinforcing each step until the final behavior is reached.
G. Establishing Operations (Motivating Operations)
An establishing operation (EO) temporarily alters the value of a reinforcer and the frequency of
the behavior it affects.
● Example: Deprivation (hunger) increases the effectiveness of food as a reinforcer.
● Conversely, abolishing operations (AOs) decrease the value of a reinforcer (e.g., being
full reduces the reinforcing effect of food).
4. Comparison: Operant vs. Respondent Conditioning
Aspect Operant Conditioning Respondent Conditioning
Type of Voluntary and goal-directed Involuntary or reflexive
Behavior
Control Behavior controlled by Behavior controlled by stimulus
consequences pairing
Process Behavior → Consequence Stimulus → Stimulus
Key Figures Thorndike, Skinner Pavlov, Watson
Example Pressing a lever → food Bell → salivation
5. Applications of Operant Conditioning
Operant conditioning principles are widely applied in various fields to modify human and
animal behavior.
● Education: Reinforcing academic performance, classroom management, token
economies.
● Clinical Psychology: Behavioral therapy for anxiety, depression, addiction, and autism
spectrum disorders.
● Parenting: Reinforcement of positive behavior and use of time-out or response cost for
undesirable acts.
● Workplace / Organizational Psychology: Incentive systems, productivity
reinforcement.
● Health and Fitness: Reinforcing exercise, medication adherence, or healthy eating.
● Animal Training: Shaping complex behaviors through systematic reinforcement.
Thorndike and Skinner — The Basics of
Operant Conditioning
1. Edward L. Thorndike and the Law of Effect (1932)
The Puzzle Box Experiment
● Thorndike placed a hungry cat inside a cage (puzzle box) with food placed outside,
just beyond its reach.
● The cat tried random actions such as clawing or pushing against the sides of the cage.
● By chance, the cat stepped on a small paddle which released the latch — the door
opened, and the cat got to the food.
Learning Through Trial and Error
● When the cat was placed in the box again:
○ It escaped faster each time.
○ After a few trials, it learned to press the paddle deliberately to escape and get
food.
Law of Effect
Thorndike concluded that:
“Responses that lead to satisfying consequences are more likely to be
repeated.”
● Behaviors followed by pleasant outcomes (rewards) are strengthened.
● Behaviors followed by unpleasant outcomes (no reward or punishment) are
weakened.
● Learning happens gradually through trial and error.
● The organism doesn’t have to “understand” — it automatically connects stimulus and
response through experience.
This shows Learning occurs through the formation of a direct connection between a situation
(stimulus) and a response, guided by its consequences.
2. B. F. Skinner and Operant Conditioning (1938)
Building on Thorndike’s Work
● Skinner extended Thorndike’s ideas and introduced the term “operant conditioning.”
● He created a controlled experimental device called the Skinner Box to study behavior
scientifically.
The Skinner Box Experiment
● The box contained:
○ A lever (or bar) that an animal (rat or pigeon) could press.
○ A food dispenser to release pellets as rewards.
○ Sometimes a light or buzzer to act as a signal (stimulus).
How Learning Occurred
1. A hungry rat was placed in the box.
2. It moved around randomly exploring the environment.
3. By chance, the rat pressed the lever → a food pellet was released.
4. The rat did not understand this connection at first, but after several trials, it learned that:
Pressing the lever = Getting food.
5. The rat began pressing the lever repeatedly to get more food.
Key Observations
● The behavior (pressing the lever) increased because it was followed by a reward (food).
● The reward acts as a reinforcer, making the behavior more likely to occur again.
● This demonstrates operant conditioning — learning based on the consequences of
voluntary actions.
3. Comparing Thorndike and Skinner
Aspect Thorndike (1932) B. F. Skinner (1938)
Focus Connection between stimulus Relationship between behavior and
and response (S-R learning) its consequence (R-C learning)
Experiment Cat in puzzle box Rat in Skinner box
Learning Type Trial-and-error learning Operant conditioning
Key Concept Law of Effect: behaviors with Reinforcement strengthens behavior;
satisfying results are repeated punishment weakens it
Conscious Not required; automatic Not required; consequence controls
Awareness association behavior
Role of Implied in satisfaction Explicitly identified and measured
Reinforcement
4. Key Concepts from the Experiments
Term Meaning Example
Operant Behavior Voluntary action that operates on Rat presses lever
environment
Reinforcement Consequence that strengthens Food pellet given after
behavior lever press
Trial and Error Repeated attempts until correct Cat learns to step on paddle
Learning response is found
Law of Effect Behavior followed by reward is Cat learns to press paddle
repeated faster
Consequence Outcome following a behavior Food, praise, or
punishment
Conditioning Process of learning through experience Rat learns lever-press =
food
5. Summary — What They Showed About Learning
● Learning depends on consequences:
When a behavior produces a satisfying result, it becomes stronger.
● Behavior is goal-directed:
Organisms learn actions that help them achieve a reward.
● Awareness is not required:
Learning can occur automatically through repetition and experience.
● Foundation for Behaviorism:
These studies laid the groundwork for modern behavioral psychology and practical
applications like behavior modification, reinforcement training, and education.
● Thorndike discovered that animals learn by trial and error, strengthening actions that
bring success.
● Skinner proved that this happens because rewards (reinforcements) make behaviors
more likely to repeat. Skinner called the process that leads the rat to continue pressing
the key “reinforce ment.”
● Both showed that learning is not about thinking — it’s about consequences.
Reinforcement and Its Types
Introduction and Definition
Reinforcement is one of the most fundamental principles of behavior in psychology and serves as
a cornerstone for operant conditioning and applied behavior analysis (ABA).
According to Miltenberger (2008), reinforcement is the process in which a behavior is
strengthened by the immediate consequence that reliably follows its occurrence.
Reinforcement is the process by which a stimulus increases the probability that a preceding
behavior will be repeated.
When a behavior is reinforced, it becomes more likely to occur again in the future under similar
conditions. Reinforcement increases either the frequency, duration, intensity, or speed of a
behavior, depending on how it is applied.
This concept was first systematically studied by Edward Thorndike (1911) through his “Law of
Effect,” and later expanded by B. F. Skinner (1938), who identified reinforcement as the primary
mechanism driving operant behavior.
History
● Thorndike’s Experiment: A hungry cat placed in a puzzle box learned to escape by
pressing a lever. With repeated trials, the cat pressed the lever faster—showing that the
behavior (lever-pressing) was reinforced by the consequence (food).
● Skinner’s Contribution: Using his “Skinner Box,” he showed that a rat pressing a lever
to receive food increased that behavior’s frequency. Reinforcement, therefore, was a key
principle explaining how voluntary behaviors are learned and maintained.
2. Characteristics of Reinforcement
Feature Description
Behavior–Consequence Reinforcement strengthens the link between a specific response
Relationship and its outcome.
Immediate Effect The reinforcer must follow the behavior immediately for
learning to occur.
Functional Definition Reinforcement is defined by its effect on behavior — if a
behavior increases, reinforcement has occurred.
Behavioral Strengthening Behavior may increase in frequency, duration, intensity, or
speed.
Operant Basis Reinforcement applies to voluntary behaviors that operate on
the environment.
3. The Two Main Types of Reinforcement
Miltenberger (2008) distinguishes positive and negative reinforcement, both of which strengthen
behavior but differ in the nature of the consequence.
A. Positive Reinforcement It occurs when, a behavior occur or is followed by the addition of
a desirable stimulus or an increase in its intensity. As a result, the behavior becomes more likely
to occur again in the future. Positive reinforcement “adds” something pleasant or rewarding after
a behavior. For Examples
● A student studies hard and receives praise from the teacher → studying increases.
● A child finishes homework and gets extra playtime.
● An employee works overtime and receives a bonus.
The added stimulus (praise, playtime, or money) acts as a positive reinforcer — it strengthens the
behavior that produced it.
B. Negative Reinforcement occurs when a behavior occurs and followed by the removal or
reduction of an aversive stimulus. The behavior is thereby strengthened or made more likely in
the future. Negative reinforcement “removes” something unpleasant to strengthen behavior. For
example Examples
● Taking painkillers removes a headache → medication-taking behavior increases.
● A driver fastens their seatbelt to stop the car alarm.
● A student studies early to avoid scolding for late submission.
Escape vs. Avoidance
Type Description Example
Escape Behavior Terminates an ongoing aversive Leaving a noisy room to escape the
condition. sound.
Avoidance Prevents an aversive condition Wearing shoes before walking on
Behavior before it starts. hot asphalt.
Clarification
● Both positive and negative reinforcement increase behavior frequency.
● Negative reinforcement is often confused with punishment — but unlike punishment, it
strengthens behavior.
Functional Analysis of Reinforcement
Reinforcement is identified not by what we intend to do but by the effect it has on behavior.
For example, praise functions as a reinforcer only if it increases the behavior. For some
individuals (e.g., children with autism), attention may not serve as a reinforcer; thus,
reinforcement is always functionally determined.
4. Types of Reinforcers
A reinforcer is any stimulus that increases the probability that a preceding behavior will
occur again. Hence, food is a reinforcer because it increases the probability that the
behavior of pressing (formally referred to as the response of pressing) will take place.
Types:
A. Primary reinforcer satisfies some biological need and works naturally, regardless of a
person’s prior experience. Food for a hungry person, warmth for a cold person, and relief
for a person in pain all would be classified as primary reinforcers.
● It is a stimulus added to the environment that brings about an increase in a
preceding response. If food, water, money, or praise is provided after a response it
is more likely that that response will occur again in the future.
● Eg: The paychecks that workers get at the end of the week, for example, increase
the likelihood that they will return to their jobs the following week.
B. Secondary reinforcer, in contrast, is a stimulus that becomes reinforcing because of its
association with a primary reinforcer. For instance, we know that money is valuable
because we have learned that it allows us to obtain other desirable objects, including
primary reinforcers such as food and shelter. Money thus becomes a secondary reinforcer.
○ It refers to an unpleasant stimulus whose removal leads to an increase in the
probability that a preceding response will be repeated in the future.
○ It teaches the individual that taking an action removes a negative condition that
exists in the environment. Like positive reinforcers, negative reinforcers increase
the likelihood that preceding behaviors will be repeated.
○ Eg: Reducing the volume of ipod when it hurts, applying an ointment on an injury
Miltenberger divides reinforcers based on their origin and learning history:
A. Unconditioned Reinforcers (Primary Reinforcers)
It is a stimuli that are naturally reinforcing because they have biological or survival value; they
do not require prior learning. For eg:
● Food, water, warmth, sleep, sexual stimulation.
● Escape from pain, extreme heat, or loud noise (negative reinforcers).
Humans and animals are born with the capacity for these reinforcers because they are essential
for survival and reproduction.
They are biologically prewired and function automatically as reinforcers from the first
experience.
B. Conditioned Reinforcers (Secondary Reinforcers)
It is a stimuli that acquire reinforcing properties through their association (pairing) with other
reinforcers — either unconditioned or already conditioned reinforcers. For example
● Social Reinforcers: Praise, attention, approval.
● Tangible Reinforcers: Money, tokens, grades
● Activity Reinforcers: Access to preferred activities (Premack Principle).
Process: A neutral stimulus becomes reinforcing through classical pairing with another reinforcer.
Example: Money becomes a conditioned reinforcer because it is repeatedly associated with food,
comfort, or other valuable goods.
Maintenance: Conditioned reinforcers continue to function only if they are periodically paired
with other reinforcers.
If money no longer buys goods, it loses its reinforcing value.
C. Generalized Conditioned Reinforcers
Conditioned reinforcers that are associated with a variety of other reinforcers, making them more
resistant to satiation and more universally motivating.
Examples
● Money: Can buy multiple reinforcers.
● Tokens: Exchanged for backup reinforcers.
● Praise: Often paired with many rewarding events across life.
Advantages
● Less likely to lose value quickly.
● Highly effective across settings and individuals.
D. Social vs. Automatic Reinforcement
Type Description Example
Social Reinforcement Delivered by another person (e.g., Teacher praises a student for
praise, attention). good work.
Automatic Behavior itself produces the Listening to music for pleasure,
Reinforcement reinforcing outcome. scratching an itch.
5. Factors Influencing the Effectiveness of Reinforcement
Factor Description Example
Immediacy Reinforcement should immediately follow Giving a treat right after a dog
the behavior. Delays weaken the sits.
behavior–consequence link.
Contingency Reinforcement must occur only if the Praise only when a child cleans
behavior occurs. up.
Establishing Operations Conditions that temporarily alter reinforcer Hunger increases food’s
(Motivating Operations) value. reinforcing power.
Magnitude/Amount Larger or more intense reinforcers produce ₹100 reward increases
stronger effects. motivation more than ₹10.
Deprivation vs. Satiation Deprivation makes a reinforcer more A thirsty person values water
effective; satiation reduces effectiveness. more.
Individual Differences What serves as a reinforcer varies among Attention may reinforce one
individuals. child but not another.
Schedules of Reinforcement
In operant conditioning, schedules of reinforcement determine when and how often a behavior
is reinforced. These schedules are crucial because they influence how quickly behavior is
learned, how often it is performed, and how resistant it is to extinction.
It is the pattern and frequency of reinforcement that can either accelerate learning or strengthen
persistence of behavior even when reinforcement stops.
In simple terms, the schedule refers to the timing or frequency of rewards following a desired
behavior.
Phase Schedule Used Purpose
Acquisition (Learning) Continuous (CRF) Build new behavior; immediate
reinforcement promotes learning.
Maintenance (Sustaining Intermittent (FR, VR, Maintain behavior long-term; reduces
Behavior) FI, VI) dependence on constant reinforcement.
2. Continuous vs. Partial (Intermittent) Reinforcement
Continuous Reinforcement (CRF)
● Every occurrence of the target behavior is followed by reinforcement.
● This schedule is most effective during initial learning (acquisition) because the
relationship between behavior and consequence becomes clear very quickly.
● Example: A child receives a sticker every time they share toys or complete homework.
Partial or Intermittent Reinforcement
● Reinforcement is given only some of the time, not after every response.
● It leads to slower learning but greater resistance to extinction — that is, the behavior
continues for longer even after reinforcement stops.
This can be explained using the vending machine vs. slot machine analogy:
● With a vending machine (continuous reinforcement), people stop trying quickly if it fails
once or twice.
● With a slot machine (intermittent reinforcement), people keep playing despite many
losses because they expect an occasional reward.
→ This shows that behaviors reinforced occasionally can be stronger and
longer-lasting than those reinforced continuously.
● The chosen behavior depends on: Frequency of reinforcement, Magnitude (amount),
Immediacy, Response effort.
3. The Four Main Schedules of Reinforcement
Miltenberger and Feldman classify intermittent schedules into two broad types — ratio
schedules (based on number of responses) and interval schedules (based on time).
A. Ratio Schedules
Continuous reinforcement schedule: Reinforcing of a behavior every time it occurs. Ratio
schedules depend on the number of behavioral responses an organism makes before receiving
reinforcement.
1. Fixed Ratio (FR) Schedule
● A schedule by which reinforcement is given only after a specific number of responses are
made.
● Reinforcement occurs after a set number of responses.
● Example: FR-10 means the behavior is reinforced after every 10 responses.
● Behavioral outcome: Produces a high rate of responding with a short pause after each
reinforcement.
● Real-life example: A factory worker is paid after producing a specific number of units.
Effect: The person or animal tends to work as quickly as possible to reach the reinforcement
point.
2. Variable Ratio (VR) Schedule
● A schedule by which reinforcement occurs after a varying number of responses rather
than after a fixed number.
● Reinforcement occurs after an unpredictable number of responses, but around an
average value.
● Example: A slot machine pays off on average once every 20 plays, but the exact number
of attempts varies.
● Behavioral outcome: Produces a high and steady rate of responding and is very
resistant to extinction.
● Example: Telemarketers or salespeople continue making calls, knowing that only some
calls will succeed.
Effect: This schedule creates persistent and strong behavior because reinforcement is
unpredictable.
B. Interval Schedules
Partial (or intermittent) reinforce ment schedule: Reinforcing of a behavior some but not all of
the time. Interval schedules depend on the passage of time rather than number of responses. The
first correct behavior after a certain time interval is reinforced.
3. Fixed Interval (FI) Schedule
● A schedule that provides reinforcement for a response only if a fixed time period has
elapsed, making overall rates of response relatively low
● Reinforcement is available only after a fixed amount of time has passed.
● Example: FI-30 min means the first correct response after 30 minutes will be reinforced.
● Real-life example: A weekly paycheck or monthly salary.
● Behavioral outcome: Produces a “scalloped” response pattern — slow responses right
after reinforcement, and faster responses as the next reinforcement time approaches.
Example: Students tend to study less just after an exam but study intensely as the next test nears.
4. Variable Interval (VI) Schedule
● A schedule by which the time between reinforcements varies around some average rather
than being fixed.
● Reinforcement is given after varying time intervals, averaging around a certain period.
● Example: Checking emails — messages arrive unpredictably, so people check regularly.
● Real-life example: Teachers giving surprise quizzes at random times.
● Behavioral outcome: Produces a steady, moderate rate of response with fewer pauses.
Effect: Because the person never knows when reinforcement will occur, behavior is maintained
consistently over time.
Key Observations and Effects
Schedule Type Basis Response Rate Resistance to Example
Extinction
Continuous Every behavior Fast learning Weak Candy vending
(CRF) machine
Fixed Ratio Set number of High, pauses Moderate Piecework pay
(FR) responses after reward
Variable Ratio Unpredictable Very high, Very strong Gambling, sales
(VR) responses steady
Fixed Interval Fixed time Moderate, Moderate Weekly
(FI) scalloped paycheck
Variable Varying time Steady, High Surprise quizzes
Interval (VI) moderate
5. Why Intermittent Reinforcement Is Powerful
● Continuous reinforcement teaches behavior quickly but makes it fade quickly once
reinforcement stops.
● Intermittent reinforcement, though slower to teach, produces enduring learning
because the organism learns that rewards may come unpredictably, so it continues the
behavior for longer periods even without immediate payoff.
This principle explains why behaviors like gambling, fishing, or sales calling persist —
reinforcement is uncertain but occasionally rewarding.
6. Reinforcement of Behavioral Dimensions
Reinforcement doesn’t just strengthen the occurrence of behavior — it can improve specific
qualities of performance.
Dimension Meaning Example
Duration Maintaining behavior for longer periods Working longer on homework
Intensity Changing the strength or force of response Speaking louder or softer
Latency Reducing the time before response Responding quickly when called
Accuracy Increasing precision or correctness Solving problems without mistakes
Example: A parent praises a child for finishing homework quickly — this reinforces
shorter latency between instruction and action.
7. Concurrent Schedules of Reinforcement
In daily life, individuals are often exposed to multiple reinforcement options at once — known
as concurrent schedules.
People tend to choose the behavior that provides:Greater or more valuable reinforcement, More
frequent or immediate payoff, and Requires less effort.
Example:
A person may choose to watch TV (immediate relaxation) instead of studying (delayed
reinforcement) because the short-term reward is stronger.
This idea helps explain behavioral choice — why we often pick quick rewards over long-term
benefits.
8. Summary
● Reinforcement schedules determine how frequently rewards follow desired behavior.
● Continuous reinforcement is best for initial learning but produces quick extinction.
● Partial reinforcement produces slower learning but stronger, longer-lasting behavior.
● Ratio schedules (FR and VR) depend on the number of responses and produce higher
response rates.
● Interval schedules (FI and VI) depend on time and produce steady but slower
responding.
● Variable schedules (VR, VI) create the most resistant and persistent behaviors.
● Reinforcement can also modify specific qualities like speed, accuracy, and intensity.
● In real life, concurrent reinforcement explains why people prefer actions that give
quicker, easier rewards.
In short: Behavior that is reinforced unpredictably tends to be stronger and
more enduring — a key lesson behind why habits, persistence, and even addictions
form the way they do.
Miltenberger’s Theory of Behavior
Modification — Study Notes
Raymond G. Miltenberger’s Behavior Modification Theory explains how behavior can be
measured, analyzed, and systematically changed using principles of learning.
It is an applied extension of behaviorism, rooted in Skinner’s operant conditioning and
informed by modern ethical and cognitive developments.
Behavior modification uses empirical assessment, reinforcement, extinction, and stimulus
control to promote adaptive behaviors and reduce maladaptive ones.
“Behavior modification is the systematic application of learning principles to assess
and improve individuals’ overt and covert behaviors.” — Miltenberger (2008)
2. Core Assumptions
Principle Explanation
Empiricism Behavior is studied scientifically through observation and data.
Determinism Behavior is lawful and influenced by environmental conditions.
Functionalism Focus on the function (purpose) of behavior, not its form.
Behavioral Focus Only observable and measurable actions are analyzed.
Environmental Behavior changes through manipulation of antecedents and
Control consequences.
3. Components of Behavior
Miltenberger defines behavior as anything a person says or does, observable and measurable,
and occurring in specific contexts.
Behavior is understood through the ABC Model (Three-Term Contingency):
Component Meaning Example
Antecedent (A) Environmental cue preceding behavior Teacher asks a
question
Behavior (B) Observable response Student raises hand
Consequence Event following behavior affecting future Teacher praises
(C) occurrence student
4. Key Principles of Behavior Modification
A. Reinforcement
● Strengthens behavior by adding (positive) or removing (negative) stimuli.
● Positive: Praise, money, approval.
● Negative: Escape from unpleasant stimulus (e.g., silencing a loud noise).
B. Punishment
● Weakens behavior by adding an aversive event (positive punishment) or removing a
pleasant one (negative punishment).
● Example: Time-out, reprimand.
C. Extinction
● Withholding reinforcement until behavior decreases.
● Often causes a temporary extinction burst before decline.
● Example: Ignoring tantrums stops attention-seeking.
D. Stimulus Control
● Behavior occurs under control of specific cues (discriminative stimuli, SDs).
● Example: A traffic light signals when “going” is reinforced.
E. Shaping
● Reinforcing successive approximations toward a final goal behavior.
● Used to teach complex or novel behaviors (e.g., speech in autistic children).
F. Schedules of Reinforcement
● Determine when reinforcement is delivered.
● Continuous: Every response (rapid learning, weak maintenance).
● Intermittent: Ratio and interval schedules (stronger maintenance).
G. Motivating Operations
● Temporary states altering reinforcer effectiveness (e.g., hunger increases food’s value).
H. Escape & Avoidance Conditioning
● Escape: Behavior ends an aversive event (e.g., leaving noisy room).
● Avoidance: Behavior prevents it (e.g., umbrella before rain).
● Both rely on negative reinforcement.
I. Differential Reinforcement
Type Description Example
DRA Reinforce an alternative, desirable Praise for polite requests instead of
behavior whining
DRO Reinforce absence of target behavior Reward 10 min without shouting
DRI Reinforce incompatible behavior Reward sitting still instead of running
J. Self-Management
● Individuals observe, record, and modify their own behavior through self-reinforcement,
goal setting, and stimulus control.
5. The Behavior Modification Process
a seven-step approach:
Step Description
1. Identify Target Behavior Define in observable, measurable terms.
2. Baseline Assessment Record frequency/intensity before intervention.
3. Functional Assessment Identify antecedents and consequences maintaining the
behavior.
4. Goal Setting Establish specific and measurable change goals.
5. Intervention Apply reinforcement, extinction, or punishment
procedures.
6. Evaluation Measure behavior change and compare with baseline.
7. Maintenance & Ensure behavior persists across settings/time.
Generalization
6. Applications of Miltenberger’s Theory
Field Examples of Application
Clinical Psychology Behavior therapy for anxiety, addictions, autism, depression.
Education Classroom management, token economies, reinforcement
programs.
Health & Rehabilitation Exercise compliance, quitting smoking, healthy eating.
Parenting & Childcare Reinforcing desirable behavior, time-outs, behavior charts.
Organizational Behavior Productivity incentives, safety reinforcement, attendance
Management (OBM) programs.
Special Education Applied Behavior Analysis (ABA) interventions for
developmental disabilities.
7. Ethical and Professional Considerations
Miltenberger stresses ethical, humane practice in behavior change.
Principle Guideline
Informed Consent Clients must understand and agree to procedures.
Least Restrictive Alternative Prefer positive reinforcement over punishment.
Beneficence Aim for client welfare and skill development.
Confidentiality Protect individual data and dignity.
Competence Only trained practitioners should apply procedures.
Data-Based Decisions Use continuous recording and evaluation for accountability.
“Ethical behavior modification teaches rather than controls, and reinforces growth
rather than compliance.” — Miltenberger (2011)
8. Integration with Broader Psychology
Miltenberger’s approach connects with other major theories:
● Skinner (Operant Conditioning): Foundation for reinforcement and punishment.
● Pavlov (Classical Conditioning): Basis for respondent behaviors and stimulus control.
● Bandura (Social Learning): Adds modeling and imitation.
● Cognitive Behaviorism: Recognizes internal processes like self-talk and awareness as
modifiable behaviors.
9. Evaluation
Strengths Limitations
Empirical and measurable. May underemphasize emotions or inner
experience.
Highly effective across contexts. Requires consistency and time.
Ethically grounded; favors positive change. Risk of mechanical application if poorly
trained.
Integrates classical and modern behavioral Focused on observable change, not existential
science. meaning.
10. Conclusion
Miltenberger’s theory represents a modern, humane evolution of behaviorism — grounded in
experimental learning principles but oriented toward ethical, goal-directed personal and social
improvement.
It views human behavior as learned, context-dependent, and modifiable, emphasizing
reinforcement, assessment, and data-driven methods to build adaptive habits and reduce
maladaptive ones.
In essence, Miltenberger demonstrates that through understanding environmental contingencies,
we can design change — scientifically, ethically, and effectively.
Punishment and Its Types
Punishment is a fundamental behavioral principle that describes the process through which a
behavior is weakened or made less likely to occur in the future because it is followed by a
specific consequence.
Punishment: A stimulus that decreases the probability that a previous behavior will occur again
Miltenberger (2008) defines punishment as:
“The occurrence of a behavior is followed by an immediate consequence, and as a
result, the behavior is less likely to occur again in the future.”
This principle contrasts with reinforcement, which strengthens a behavior. Punishment, by
definition, always refers to the decrease in future behavior frequency, regardless of the moral
or emotional meaning commonly associated with the term.
Everyday vs. Behavioral Definition
Common (Everyday) Meaning Behavioral (Scientific) Meaning
Punishment as retribution or moral correction A behavioral process where a consequence
(e.g., fines, spanking, imprisonment). reduces the likelihood of a behavior.
Implies intention, pain, or “deserved” Defined functionally — only if the behavior
suffering. decreases as a result of the consequence.
Often emotional or ethical. Empirical, objective, and measurable.
Example:
If a child is scolded but continues misbehaving, the scolding is not a punisher because the
behavior did not decrease — it may even function as reinforcement (attention-seeking).
2. The Two Major Types of Punishment
It can be classified into Positive Punishment and Negative Punishment, depending on whether
a stimulus is added or removed after the behavior.
A. Positive Punishment (Punishment by Application)
● “positive” means adding something, and “negative” means removing some thing.)
Positive punishment weakens a response through the application of an unpleas ant
stimulus. For instance, spanking a child for misbehaving or spending ten years in jail for
committing a crime is positive punishment.
● A behavior is followed by the presentation of an aversive stimulus, leading to a
decrease in the future probability of that behavior.
Behavior → Add Aversive Stimulus → Behavior Decreases
Examples
● A child touches a hot stove → feels pain → stops touching stoves.
● A student talks in class → teacher reprimands → talking decreases
● A rat presses a lever → receives a shock → pressing decreases.
Applications
● Contingent Exercise: Requiring a person to perform a physical activity after problem
behavior (e.g., a child must stand up and sit down 10 times after hitting someone).
● Overcorrection: Repeating or over-performing a correct form of behavior (e.g., cleaning
a whole classroom after scribbling on one desk).
● Aversive Stimulation: Applying mild stimuli such as water mist, facial screening, or
noise alarms contingent on behavior (Rapp & Miltenberger, 1998; Ellingson et al., 2000).
Premack-Based Punishment
Based on the Premack Principle, engaging in a low-probability activity after a problem
behavior can decrease that behavior.
Example: “If you fight, you’ll have to do extra cleaning.”
B. Negative Punishment (Punishment by Removal)
● In contrast, negative punishment consists of the removal of something pleasant. For
instance, when a teenager is told she is “grounded” and will no longer be able to use the
family car because of her poor grades, or when an employee is informed that he has been
demoted with a cut in pay because of a poor job evaluation, negative punishment is being
administered.
● A behavior is followed by the removal or loss of a reinforcing stimulus, resulting in a
decrease in the behavior’s future likelihood.
Behavior → Remove Reinforcing Stimulus → Behavior Decreases
Main Forms
1. Time-Out from Positive Reinforcement
○ The individual loses access to reinforcement temporarily.
○ Example: A child is sent to a “time-out” corner for misbehavior and loses teacher
attention and peer interaction.
2. Response Cost
○ Removal of a specific reinforcer following misbehavior.
○ Example: A child loses tokens or allowance for breaking rules.
Clarification
Negative punishment differs from extinction:
● In extinction, the reinforcer maintaining the behavior is withheld.
● In negative punishment, another reinforcer the person already possesses is removed.
Comparison: Positive vs. Negative Punishment
Aspect Positive Punishment Negative Punishment
Consequence Presentation of an aversive Removal of a positive/reinforcing stimulus
stimulus
Effect Behavior decreases Behavior decreases
Example Scolding, pain, reprimand Time-out, fine, loss of privileges
Other Terms Punishment by application Punishment by withdrawal or response cost
Ethical Used rarely, last resort Commonly preferred in modern behavior
Preference modification
3. Types of Punishers
A. Unconditioned Punishers (Primary Punishers)
Stimuli that are naturally aversive — no prior learning is required for them to function as
punishers.
Examples
● Pain, extreme heat or cold, loud noises, strong odors, bright light.
● Electric shock, physical restraint, or other sensory discomforts.
Explanation
These stimuli have biological importance. Evolutionarily, behaviors producing pain or danger
are naturally suppressed — they help ensure survival.
B. Conditioned Punishers (Secondary Punishers)
Stimuli that acquire punishing value through association with unconditioned punishers or other
conditioned punishers.
Examples
● The word “no” or scolding tone.
● Threats, angry facial expressions, disapproving looks.
● Fines, parking tickets (associated with loss of money).
Generalized Conditioned Punishers
Stimuli paired with multiple punishers or reinforcer losses (e.g., “no,” social disapproval) and
thus effective across settings.
Example: A teacher’s firm “stop that” may reduce behavior because it’s associated
with past loss of privileges or attention.
4. Factors Influencing the Effectiveness of Punishment
Miltenberger (2008) identifies several variables that determine how effective punishment will be:
Factor Explanation / Example
Immediacy Punishment must immediately follow behavior. Delays weaken
learning.
Contingency Punisher should occur only when the behavior occurs. Consistency
is key.
Establishing Some antecedents increase or decrease the punisher’s value (e.g.,
Operations warning increases fear of punishment).
Individual Differences Punishers vary per person due to experience (e.g., scolding may
affect one student more than another).
Magnitude/Intensity More intense punishers are generally more effective, but risk side
effects.
Schedule of Continuous punishment (after every behavior) is more effective
Punishment initially than intermittent punishment.
5. Problems and Side Effects of Punishment
While punishment can reduce problem behavior, Miltenberger warns about several undesirable
side effects:
1. Emotional Reactions – Fear, anxiety, or aggression may be elicited.
2. Escape and Avoidance – The punished individual may avoid the punisher or the setting
(e.g., skipping school).
3. Negative Reinforcement for the Punisher – The person administering punishment may
find relief when the behavior stops, reinforcing their use of punishment and leading to
overuse.
4. Modeling Aggression – Observing punishment (especially physical) can increase
aggressive behavior in others (Bandura, 1969).
5. Ethical and Social Concerns – Painful punishers are viewed as inhumane and socially
unacceptable.
For these reasons, positive punishment is considered a last resort in behavior
modification and always combined with reinforcement of alternative behaviors
(DRA/DRO).
6. Ethical Considerations in Using Punishment
● Informed consent: Client/ Guardian must fully understand the procedure and agree
voluntarily
● Alternative treatments: Non aversive, Reinforcement - based must be tried first
● Severity of problem: Punishment justified only for severe, self-harming, or dangerous
behaviours
● Recipient safety: The procedure must never cause in jury or distress
● Training and Supervision: Implementers must be trained and monitored
● Written guidelines and peer review: procedures must be documented and reviewed by
professionals peers
● Accountability: accurate record-keeping and regular review of data to prevent misuse
7. Summary Table: Punishment in Behavior Modification
Aspect Positive Punishment Negative Punishment
Definition Behavior followed by presentation Behavior followed by removal of
of aversive stimulus → behavior reinforcing stimulus → behavior
decreases. decreases.
Example Spanking, scolding, facial screening, Time-out, fines, loss of privileges.
loud noise.
Ethical Use Rare; last resort; for severe cases. Common; less aversive and more
socially acceptable.
Underlying Adds unpleasant stimulus. Removes pleasant stimulus.
Principle
Effect on Decreases likelihood of future Decreases likelihood of future
Behavior behavior. behavior.
8. Conclusion
Punishment is a powerful but ethically sensitive tool in behavior modification. It operates by
weakening behaviors through aversive or loss-based consequences. However, due to its potential
for misuse and side effects, professionals emphasize the reinforcement of alternative adaptive
behaviors and the use of non-aversive methods first.
Miltenberger (2008) emphasizes that ethical application, informed consent, and concurrent
reinforcement are essential whenever punishment is used.
Why Reinforcement is Better?
Punishment can sometimes be the quickest way to stop dangerous or harmful behavior,
especially when immediate correction is needed. For instance, if a child runs into a busy street, a
parent may use punishment right away to prevent serious injury. Similarly, in rare and severe
cases—such as self-injurious behavior in children with autism—controlled forms of
punishment, like a brief electric shock, may be used only when all other treatments fail and to
keep the individual safe until positive reinforcement methods can begin (Salvy, Mulick, &
Butter, 2004; Ducharme et al., 2007).
However, despite its occasional necessity, punishment has major drawbacks that make it
unreliable and often unethical as a routine method of behavior modification.
1. Limitations and Disadvantages
Timing and consistency are critical. Punishment is effective only if it follows the behavior
immediately and every time. Delays or inconsistency weaken learning. If a person can escape
the punishment, no behavior change occurs—like a teenager borrowing a friend’s car after losing
the use of the family car.
Modeling aggression is another danger. Physical punishment may teach that aggression is
acceptable. A parent who hits or yells teaches the child to respond aggressively in similar
situations. Over time, this normalizes violence as a way to solve problems.
Emotional effects are also common. Harsh punishment can lead to fear, resentment, and low
self-esteem, especially if the person doesn’t understand the reason behind it. When delivered in
anger, it can easily become excessive, leading to emotional harm and avoidance of the punisher.
Another problem is that punishment does not teach alternative behavior. It tells a person what
not to do but offers no guidance on what to do instead. A child punished for daydreaming in
class might stop looking out the window but start staring at the floor instead. Unless the
punishment is followed by clear instruction and reinforcement of desirable behavior, it
accomplishes little.
Finally, frequent punishment may cause the person to fear or avoid the punisher rather than
changing the targeted behavior. This can damage trust and communication, making long-term
improvement difficult.
2. Why Reinforcement Is Better
Psychologists emphasize that reinforcement is more effective and humane for shaping
behavior. Reinforcing desired actions teaches what is expected and motivates repetition through
positive outcomes. Unlike punishment, reinforcement builds confidence, cooperation, and
lasting change.
Research consistently shows that rewarding appropriate behavior works better than
punishing bad behavior (Pogarsky & Piquero, 2003; Sidman, 2006). Punishment should therefore
be used only when absolutely necessary, applied calmly and consistently, and always followed
by guidance and reinforcement of correct behavior.
Applications: Extinction, Decreasing Behavior,
Escape and Avoidance Conditioning
1. Extinction
In learning, extinction refers to the process by which a previously learned response decreases
and eventually disappears when it is no longer followed by reinforcement (operant) or when
the conditioned stimulus is no longer paired with the unconditioned stimulus (classical).
Extinction: A basic phenomenon of learning that occurs when a previously
reinforced or conditioned response weakens because reinforcement or pairing is
withdrawn.
In Classical Conditioning (Feldman)
● A conditioned response (CR) decreases when the conditioned stimulus (CS) is
repeatedly presented without the unconditioned stimulus (US).
● Example: A dog trained to salivate at a bell (CS) stops salivating when the bell rings but
food (US) is no longer provided.
● This shows that the learned association fades over time.
● Spontaneous Recovery: After a rest period, the extinguished response may reappear
temporarily, indicating that extinction suppresses but doesn’t erase the original learning.
In Operant Conditioning (Miltenberger, 2008)
Extinction occurs when a behavior that was previously reinforced no longer produces the
reinforcing consequence, and as a result, the behavior decreases in frequency.
Example
● A child throws tantrums to get attention.
● When parents stop responding (no attention given), the tantrums gradually decrease and
stop.
Key Features
Feature Description
Mechanism Reinforcer maintaining behavior is removed.
Effect Behavior frequency decreases over time.
Temporary Increase (Extinction Behavior may initially increase in frequency/intensity
Burst) before declining.
Spontaneous Recovery Behavior may reappear briefly after extinction.
Research Examples
● Williams (1959): Extinction used to eliminate a child’s night tantrums by instructing
parents to ignore them.
● Pinkston et al. (1973): Teachers reduced aggression by withdrawing attention (a
reinforcer).
● Lovaas & Simmons (1969): Extinction reduced self-injurious behavior in a child with
developmental disabilities.
Sensory Extinction (Miltenberger)
Used when behavior is maintained by automatic (sensory) reinforcement — not social
consequences.
● Example: A child repeatedly spins a plate because the sound is reinforcing.
● Intervention: Remove or change the sensory feedback (e.g., use soft surfaces).
● Result: Behavior decreases as reinforcement is eliminated.
Practical Applications
● Managing tantrums or attention-seeking behavior in children.
● Reducing self-stimulatory behaviors in autism.
● Eliminating addictive habits by removing reinforcement cues.
● Reducing maladaptive work behaviors (e.g., ignoring disruptive conduct).
Summary
Extinction is a non-punitive method of behavior reduction that relies on withdrawal of
reinforcement. Though initially behavior may spike (extinction burst), consistent application
leads to long-term decrease.
2. Decreasing Behavior (Behavior Reduction Strategies)
Behavior modification seeks not only to increase adaptive behaviors but also to decrease
maladaptive or problem behaviors using systematic procedures (Miltenberger, 2008; Feldman,
2009).
Major Behavior Reduction Procedures
Method Description Example
Extinction Withholding reinforcement for a Ignoring a child’s
problem behavior. tantrum.
Differential Reinforcing alternative or other Praising quiet sitting
Reinforcement (DRO, behaviors while withholding while ignoring shouting.
DRA, DRI) reinforcement for undesired behavior.
Time-Out (Negative Removing access to reinforcement for “Time-out” chair after
Punishment) a brief period following problem aggression.
behavior.
Response Cost Removing a specific reinforcer Losing tokens for
following a problem behavior. misbehavior.
Overcorrection Requiring corrective or exaggerated Cleaning up entire area
behavior following a problem action. after spilling milk.
Reinforcement of Reinforce behaviors physically Reinforcing keeping
Incompatible Behavior incompatible with the problem hands folded instead of
(DRI) behavior. hitting.
Key Principles (Miltenberger)
1. Functional Assessment: Identify reinforcers maintaining problem behavior.
2. Consistency: Apply extinction or punishment systematically.
3. Reinforce Alternatives: Always pair reduction techniques with reinforcement for
positive behaviors.
4. Ethical Practice: Prefer reinforcement-based methods; use punishment only when
necessary and humane.
Applications (Feldman, 2009)
Behavior modification programs combine extinction with reinforcement:
● Example: A couple reduced household conflict by using verbal praise for cooperation
and a $1 response cost for missed chores.
● Used in clinical, educational, and health settings — e.g., reducing aggression, smoking,
or overeating.
3. Escape and Avoidance Conditioning
Both escape and avoidance conditioning are forms of negative reinforcement, where behavior
increases because it removes or prevents an aversive stimulus.
A. Escape Conditioning
Concept
● Behavior terminates an ongoing aversive stimulus.
● The person “escapes” from the unpleasant situation.
Example
● A rat jumps to the other side of a box to escape an electric shock.
● A person leaves a noisy theater to escape the loud noise.
Mechanism
Sequence Explanation
Aversive stimulus present → Behavior occurs → Behavior strengthens (negative
Aversive stimulus removed reinforcement).
Everyday Examples
● Turning down loud music to stop discomfort.
● Taking medication to stop pain.
● Closing a window to escape cold wind
B. Avoidance Conditioning
Concept
● Behavior prevents the aversive stimulus before it occurs.
● A warning cue (conditioned stimulus) signals that an aversive event is coming.
Example (Miltenberger’s Rat Study)
● A tone (warning signal) precedes an electric shock.
● The rat learns to jump to the other side when it hears the tone, thus avoiding the shock
altogether.
Mechanism
Sequence Explanation
Warning signal → Behavior → Prevents Behavior strengthens through negative
aversive stimulus reinforcement.
Everyday Examples
● Putting on shoes before walking on hot asphalt.
● Submitting homework early to avoid teacher scolding.
● Taking an umbrella after seeing cloudy skies to avoid getting wet.
Escape vs. Avoidance
Aspect Escape Conditioning Avoidance Conditioning
Timing of After the aversive event begins Before the aversive event occurs
Behavior
Goal Terminate existing unpleasant Prevent occurrence of unpleasant
stimulus stimulus
Type of Direct negative reinforcement Classical + Operant conditioning
Learning (warning signal)
Example Leaving a room to stop loud Wearing earplugs before entering noisy
noise room
Applications
● Therapy: Treating anxiety and phobias through exposure and prevention (breaking
avoidance patterns).
● Clinical training: Teaching adaptive avoidance (e.g., safety skills).
● Education: Encouraging homework submission before deadlines (avoidance of
scolding).
● Rehabilitation & Health: Escape conditioning used to manage chronic pain or
discomfort.
Ethical Considerations
Miltenberger warns that escape and avoidance responses can inadvertently reinforce
maladaptive behavior — for instance, avoiding social events due to anxiety maintains the
anxiety.
Hence, clinicians use exposure therapy to extinguish such avoidance patterns.
4. Integrated Applications in Behavior Modification
Application Area Extinction / Decreasing Behavior Escape & Avoidance
Conditioning
Clinical Extinction for phobias, habits, Avoidance reduction in
Psychology addictions. anxiety therapy.
Parenting & Ignoring misbehavior (extinction); Time-outs or escape
Education reinforcement of compliance. extinction for tantrums.
Rehabilitation Reducing self-injury, improving Escape training to tolerate
compliance with therapy. medical procedures.
Self-Management Removing cues and reinforcers for bad Avoidance of triggers (e.g.,
habits. smoking cues).
Health Behavior Extinguishing maladaptive eating or Avoiding exposure to stressors
sedentary habits. or relapse cues.
5. Summary
● Extinction: Weakens behavior by withholding reinforcement; may cause temporary
bursts before decline.
● Decreasing Behavior: Combines extinction, differential reinforcement, time-out, and
response cost ethically.
● Escape Conditioning: Behavior terminates ongoing aversive stimuli.
● Avoidance Conditioning: Behavior prevents anticipated aversive stimuli.
Operant Conditioning and Generalizing
Behavioural Change
Generalizing Behavioural Change (Miltenberger, 2008, Ch. 19)
Behavioral change is not complete until the new behavior generalizes — that is, it occurs
beyond the training situation in all relevant contexts.
Definition:
Generalization is the occurrence of learned behavior in the presence of stimuli
similar to those present during training or in new environments.
Example
● A child taught to greet peers during therapy must also greet peers at school or home for
generalization to be achieved.
Importance of Generalization
● Ensures lasting, functional behavior change.
● Prevents relapse into problem behaviors.
● Promotes maintenance across time, people, and places.
● The ultimate goal of behavior modification is behavior that continues naturally
without external control.
Strategies to Promote Generalization
(Adapted from Miltenberger, 2008; Stokes & Baer, 1977)
Strategy Explanation Example
1. Reinforce Reinforce behavior when it Praise a child for polite speech at
Generalization occurs in new settings. home and in school.
2. Use Similar Stimuli Make training conditions Practice assertiveness with real
resemble real-life contexts. coworkers during role-play.
3. Train with Teach responses under varied Use different teachers,
Multiple Exemplars conditions. classrooms, or settings during
training.
4. Incorporate Use reinforcers available in real Encourage self-praise and natural
Natural Reinforcers environments. social feedback.
5. Teach Enable individuals to monitor Self-charting progress or using
Self-Management and reinforce their own behavior. reminders.
6. Program Common Include familiar cues in training Teaching safety rules in actual
Stimuli that also exist in the real setting. playground environments.
7. Eliminate Remove any outside Stop peers from teasing students
Punishment Barriers consequences that discourage who interact with disabled
desired behavior. classmates.
Maintenance of Behavioural Change
To maintain generalization:
1. Continue reinforcement intermittently in natural settings.
2. Train caregivers, teachers, or peers to reinforce behavior.
3. Monitor and measure behavior across time and contexts.
4. Encourage self-monitoring and self-reinforcement.
Example:
A student praised for submitting work early during training continues the behavior
after reinforcement fades because it is naturally rewarded (less stress, teacher trust).
Integration: Operant Learning and Generalization
Operant Principle Role in Generalization
Reinforcement Maintains behavior across new settings.
Discrimination Training Teaches when and where to perform behavior.
Stimulus Generalization Behavior transfers to similar cues.
Behavioral Skills Training Promotes transfer through real-life rehearsal.
Natural Contingencies Sustain learned behavior after intervention ends.
Summary
Aspect Operant Conditioning Generalization of Behavioural
Change
Focus How consequences shape How learned behavior spreads to
voluntary behavior new situations
Key Theorists Thorndike, Skinner, Miltenberger Stokes & Baer, Miltenberger
Main Reinforcement, punishment, Reinforcement in multiple contexts
Mechanisms extinction
Goal Modify and strengthen desired Ensure behavior persists across
behavior settings
Example Learning to study through praise Continuing to study at home and
school