Behavioral Mechanics Slides Behavioral Mechanics
Defining Operant Conditioning & Reinforcement
Lesson 1 The ABCs of Behavior
The Hot/Cold Challenge
The Task
A volunteer must complete a specific, secret task (e.g., placing a pencil in a specific drawer) using only verbal feedback.
The Feedback
"Hot" = Closer to goal (Reinforcement)
"Cold" = Further from goal
Silence = No change
Observe!
How does the volunteer's behavior change as they receive feedback? How do they "learn" the secret task?
The Mechanics of Choice
Operant Conditioning
Learning where behavior is controlled by consequences . An organism "operates" on the environment to produce an effect.
B.F. Skinner's insight:
"The consequences of behavior determine the probability that the behavior will occur again."
Classical Conditioning
Automatic/Reflexive responses to a stimulus (Pavlov's Dogs).
Operant Conditioning
Voluntary behaviors followed by consequences (Skinner's Rats).
The ABC Data Loop
A
Antecedent
What happened before the behavior? (The cue or trigger)
e.g., Phone pings
B
Behavior
The action performed. Must be observable and measurable.
e.g., Checking notification
C
Consequence
What happened after ? (The result)
e.g., Friend's funny text
"Behavior is a function of its consequence."
Positive Reinforcement
Adding a stimulus increases the likelihood of the behavior.
Addition (+)
Strength (↑)
Example A
Work hard (B) → Get a bonus (C). Work hard again? Yes.
Example B
Tell a joke (B) → Friends laugh (C). Tell more jokes? Yes.
The Golden Rule
It is only reinforcement if the behavior actually increases in the future.
If you give a kid a sticker for cleaning but they stop cleaning next week, that sticker was NOT a reinforcer.
Crucial Clarification
"Positive" does NOT mean "Good" or "Happy".
In behavioral science, Positive means to add (like a plus sign).
If a student shouts out and the teacher yells at them, and the student shouts out more often because they wanted attention...
That is Positive Reinforcement.
Behavior Lab Worksheet Behavioral Lab: ABC Analysis
Topic: Operant Conditioning & Positive Reinforcement
Name:
Date:
The ABC Framework
In behavioral science, we don't just look at an action; we look at the environment surrounding it. Use this worksheet to dissect behaviors into three components: Antecedent (the trigger), Behavior (the observable action), and Consequence (the immediate result).
Part 1: Scenario Dissection
Scenario 1: While studying, Jamie's phone vibrates. Jamie checks the notification and finds a funny meme from a friend. Jamie continues to check their phone every time it vibrates for the rest of the night.
Antecedent (A)
Behavior (B)
Consequence (C)
Scenario 2: A toddler sees a candy bar in the checkout line and starts screaming. The exhausted parent gives the toddler the candy bar. The next time they go to the store, the toddler screams as soon as they reach the checkout line.
Antecedent (A)
Behavior (B)
Consequence (C)
Part 2: Deep Dive Analysis
1. In Scenario 2, whose behavior was positively reinforced? (Hint: It might be more than one person!) Explain your reasoning using the definition of reinforcement.
2. Define a behavior of your own that you perform frequently. What is the most likely consequence that maintains this behavior? Is it positive reinforcement (adding something)?
My Behavior
The Consequence
Yes
No
3. Reflection: How does the ABC model help us understand behaviors that might seem "irrational" or "annoying" to others?
Unit: Behavioral Mechanics
Behavior Facilitation Guide Facilitator Protocol
Lesson 1: The Reinforcement Loop
TEACHER ONLY
Lesson Overview
This lesson shifts students from viewing reinforcement as "getting a reward" to seeing it as a biological mechanism for survival and learning. The goal is to master the ABC (Antecedent-Behavior-Consequence) model.
Duration
45-60 Minutes
Key Concept
Positive Reinforcement
Learning Targets
Distinguish between Classical and Operant Conditioning.
Identify A, B, and C in behavior loops.
Apply "positive reinforcement" correctly (adding/increasing).
01
The Hook: The Hot/Cold Game (10 mins)
Facilitation Notes:
Ask for one volunteer to leave the room.
While they are gone, pick a multi-step task for them (e.g., "Walk to the back, pick up a blue marker, and place it on the center of the teacher's desk").
Instruct the class: We will only use reinforcement. When they get closer, say "Hot." When they deviate, stay silent or say "Cold."
Bring student back. DO NOT explain the goal. Start the process.
The Reveal
"How did you know what to do without being told? You were 'operated' on by the consequences of your movement."
02
Direct Instruction (15 mins)
Use the Behavioral Mechanics Slides to cover:
Operant vs. Classical: Highlight that operant involves voluntary choice based on outcome.
The ABC Model: Emphasize that the Antecedent doesn't cause the behavior; it's a "setting event" that signals a consequence is available.
The Misconception: Spend time on "Positive ≠ Good." Use the "Teacher yelling at a disruptive student" example. If yelling increases behavior, yelling is positive reinforcement.
03
The Behavior Lab (20 mins)
Hand out the Behavior Lab Worksheet . Circulate and check for the "Toddler/Candy" scenario. Students often miss that BOTH parties are being reinforced:
Toddler: Reinforced for screaming (Gets candy).
Parent: Reinforced for giving candy (Screaming stops—this is actually Negative Reinforcement , which you can use as a teaser for the next lesson!).
Exit Ticket / Closure
Persistence Pulse Slides The Persistence Pulse
Schedules of Reinforcement
Lesson 2 Ratio vs. Interval
Why do we keep going?
"Why do we check our phones even when we don't have a notification?"
"Why does a person pull the lever of a slot machine for 5 hours straight?"
The answer is in the Timing.
The Reinforcement Matrix
FIXED RATIO (FR)
Reinforcement after a set number of responses.
e.g., "Buy 10 coffees, get 1 free"
VARIABLE RATIO (VR)
Reinforcement after an average, unpredictable number of responses.
e.g., Slot machines, Loot boxes
FIXED INTERVAL (FI)
Reinforcement for the first response after a set amount of time.
e.g., Paychecks every Friday
VARIABLE INTERVAL (VI)
Reinforcement for the first response after unpredictable time chunks.
e.g., Checking email / Social media feed
The Power of Unpredictability
Which schedule creates the highest response rate and the most resistance to "extinction" (giving up)?
Variable Ratio
This is the "Strongest" schedule. Since you never know which click will win, you keep clicking forever.
FR
VR
Time Cumulative Responses
Fixed Interval Scalloping
The "Cramming" Pattern
In a Fixed Interval schedule (like a test every 2 weeks), behavior drops to near-zero right after a reward and then spikes right before the next reward is due.
Predictable Lulls & Bursts
Real World Example
Legislative voting spikes right before elections. Student studying spikes right before the midterm.
The VI Advantage
Variable Interval creates slow, steady responding because you have to "check in" constantly to see if the window is open.
Schedule Simulator Lab Schedule Simulators
Topic: Patterns of Reinforcement & Extinction
Collaborator:
Lab Station:
The Mission
How does it feel to be on different schedules? In this lab, you will perform a simple repetitive task (e.g., tallying marks on a page) while your partner provides reinforcement (a high-five or a "Good job!") according to a specific schedule.
1
Fixed Ratio (FR-5)
Protocol:
Reinforcer given every 5th tally mark precisely.
Student Response Data Area
Describe the "feel" / motivation level:
2
Variable Ratio (VR-5)
Protocol:
Reinforcer given on an average of 5 marks (Randomly at 2, 8, 4, 6).
Student Response Data Area
Describe the "feel" / motivation level:
3
Fixed Interval (FI-15s)
Protocol:
Reinforcer given for the first tally after 15 seconds have passed.
Student Response Data Area
Describe the "feel" / motivation level:
4
Variable Interval (VI-15s)
Protocol:
Reinforcer given for the first tally after random time (5s, 25s, 10s, 20s).
Student Response Data Area
Describe the "feel" / motivation level:
The Persistence Check
1. Resistance to Extinction: If the partner stopped reinforcing altogether, which schedule would lead the student to keep working the LONGEST? Why?
2. Digital Design: How does a social media "feed" (scrolling) mimic one of these schedules? Identify the schedule and explain the mechanics of the "reward."
3. The Vending Machine Paradox: If you put money in a vending machine (Fixed Ratio 1) and nothing comes out, you stop immediately. If you put money in a slot machine (Variable Ratio) and nothing comes out, you keep going. Explain this difference using the concept of "unpredictability."
Lab Protocol #002: PERSISTENCE PULSE
Evolution of Action Slides The Evolution of Action
Shaping & Chaining Complex Behavior
Lesson 3 Successive Approximations
How do you reinforce a behavior that never happens?
Reinforcement requires the behavior to occur first (the "B" in ABC).
But what if you want to teach a dog to ride a skateboard? Or a toddler to tie their shoes? Or a pilot to land a 747?
The Paradox:
You can't reinforce what isn't there.
The Solution:
Break it down. Reinforce the "almosts."
Shaping
Formal Definition:
Reinforcing Successive Approximations of a desired behavior.
1
Find a starting behavior (something they ALREADY do).
2
Only reinforce moves that look a little more like the goal.
3
Stop reinforcing the old "approximations."
Example: Dog Skateboard
Look at board
Walk toward board
Sniff board
Touch board with one paw
Push off with back paw
Chaining
The "Assembly Line"
Linking together a specific sequence of behaviors to complete a complex task. Each step in the chain serves as the consequence for the previous step and the cue for the next.
Forward Chaining
Teach step 1 first. Then step 1+2. Then step 1+2+3.
Backward Chaining
Teacher does all steps except the LAST one. Student completes last step and gets the REWARD.
STEP 1
STEP 2
STEP 3
Know the Difference
Shaping
Teaching a new quality or dimension of behavior.
Like sculpting clay.
Goal: A single final action.
Chaining
Teaching a sequence of existing behaviors.
Like assembling a machine.
Goal: A multi-step complex routine.
Shaping Planner Worksheet The Architect of Action
Topic: Shaping Complex Behavior Through Approximations
Behavioral Designer:
Date:
Phase 1: Defining the Terminal Goal
Choose a complex behavior that an organism (human or animal) does not currently perform. This is your Terminal Goal .
Target Behavior Description:
e.g., A dog putting its own toys away in a bin...
The Golden Rule
Shaping is about patience. If you move too fast, the learner gets frustrated. If you move too slow, they get bored or stay stuck.
Phase 2: The Shaping Staircase
Break your goal down into 5 "Successive Approximations." Each step must be slightly more difficult than the last.
5
Terminal Behavior
4
3
2
1
Starting Point (Current Behavior)
Phase 3: Troubleshooting the Chain
1. Stagnation: What if your learner gets stuck on Step 3 and won't move to Step 4? What is your behavioral strategy to "nudge" them forward?
Think about the magnitude or timing of reinforcement...
2. Extinction Burst: If you stop reinforcing Step 2 to move to Step 3, the learner might get frustrated and try Step 2 even harder/louder. How will you handle this "burst" of old behavior?
Explain why staying consistent is key...
3. Reflection: How is shaping different from simply "explaining" how to do something? Why is shaping more effective for non-verbal or highly complex motor tasks?
Behavioral Architecture Protocol // Evolution of Action
Currency of Choice Slides The Currency of Choice
Selecting Effective Reinforcers
Lesson 4 Primary, Secondary, and Premack
One person's treasure...
A reinforcer is only effective if the individual values it. There is no such thing as a "universal reward."
Candy
Reinforcing for a child, maybe not for a diabetic adult.
Praise
Reinforcing for an extravert, punishing for an introvert.
The Preference Rule
To find an effective reinforcer, you must conduct a Preference Assessment .
Observe free-time choices
Ask (Surveys/Interviews)
Forced choice (A vs B)
Types of Reinforcers
PRIMARY
Unlearned & Biological
Food, water, sleep, warmth, social touch. We are born needing these.
SECONDARY
Learned & Conditioned
Money, grades, tokens, likes, praise. These are valuable because they represent or buy primary reinforcers.
GRANDMA'S RULE
The Premack Principle
A "high-probability" behavior can reinforce a "low-probability" behavior.
Low Prob (Work)
Eat Broccoli
High Prob (Play)
Go Outside
Think About It:
You aren't adding a "thing." You are using an activity as the reward.
"If you finish your homework (low prob), you can play video games (high prob)."
The State of the Learner
Satiation
"Too much of a good thing."
If you just ate a 5-course meal, pizza is no longer a reinforcer. The reward loses its power.
Deprivation
"Hungry for the reward."
The longer you go without a reinforcer, the more powerful it becomes. (e.g., Checking your phone after a 4-hour exam).
Effective behavioral designers vary reinforcers to avoid satiation.
Reward Audit Worksheet Reward Audit
Preference Assessment Inventory
Subject Name:
Researcher:
Part 1: The Reinforcement Menu
Rank the following categories from 1 (Most Motivating) to 5 (Least Motivating) for YOURSELF.
Rank
Edibles
(Snacks, treats)
Rank
Activities
(Gaming, free time)
Rank
Social
(Praise, group work)
Rank
Tokens
(Money, points)
Rank
Tangibles
(Stickers, merch)
Part 2: Forced Choice Trials
Interview a partner. For each pair below, they MUST choose one, even if they like both. Check the box of their preference.
A: $10 Cash
B: No homework for a week
A: Public praise at an assembly
B: Private "Great job" note
A: A free gourmet pizza
B: Control of the aux/music for an hour
Part 3: Activity Reinforcement
The Premack Principle: High-Probability (High-P) behaviors reinforce Low-Probability (Low-P) behaviors.
Subject's "High-P" Activity
(What do they choose to do during free time?)
Subject's "Low-P" Activity
(What do they often avoid or need prompting for?)
Write a "Grandma's Rule" statement for this subject:
"If you ____________________________, then you can ____________________________."
Part 4: Managing Satiation
Your subject has been using "Social Media Time" as a reward for every hour of studying. Lately, they are studying less and checking the social media feed even when they haven't worked. They are "satiated."
How would you adjust their behavioral plan to regain the power of the reinforcer?
Preference Protocol // Unit: Behavioral Mechanics
Behavior Blueprint Project Guide The Behavior Blueprint
Final Capstone Project
Level 12 / Behavioral Engineering
The Challenge
You are a Behavioral Consultant hired to design a comprehensive modification plan for one of the following scenarios. You must use only positive reinforcement strategies. Your goal is to create a sustainable, ethical, and effective system for long-term behavioral change.
Step 1: Select a Scenario
A: The Screen-Free Family
A family wants to increase outdoor play and reading time for their children while decreasing passive screen time.
B: The Productivity Pod
A small startup company wants to increase collaborative problem-solving and meeting punctuality among its remote team.
C: The Shelter Success
An animal shelter wants to teach highly-stressed dogs "calm waiting" behaviors at the kennel doors to increase adoption rates.
Step 2: The Components
Your final report must address each of the following sections in detail:
I. The Target (B)
Define the terminal behavior in observable, measurable terms.
Drafting space for target behavior:
II. Reinforcement Menu
Identify 3 diverse reinforcers. Justify why they are likely effective based on the learner's demographic.
Reinforcer 1
Reinforcer 2
Reinforcer 3
III. The Schedule
Determine the starting schedule and the transition schedule for long-term maintenance.
Why start with a Fixed Ratio 1 (Continuous) schedule?
What is the "Maintenance" schedule (e.g., Variable Ratio)?
IV. Shaping Protocol
Outline at least 3 successive approximations leading to the goal.
Ethical Consideration
Ensure your plan respects the agency of the learner. How will you ensure the learner is a willing participant in this modification?
Behavior Blueprint Rubric Blueprint Rubric
Behavioral Analysis & Plan Evaluation
/ 100 PTS
Criterion Advanced (4) Proficient (3) Developing (2-1) Behavioral Definition Target behavior is defined in strictly observable and measurable terms. No internal states (e.g., "feeling happy") used. Target behavior is clear but includes some minor subjective language. Behavior is vague or defined by internal feelings rather than actions. Reinforcement Selection Plan includes diverse reinforcers (Primary/Secondary/Premack) with a strong justification based on demographics. Plan includes effective reinforcers, but lacks diversity or strong justification. Reinforcers are arbitrary or likely to lead to rapid satiation. Schedules of Reinforcement Strategic shift from continuous (acquisition) to variable (maintenance) schedules is clearly explained. Both acquisition and maintenance schedules are identified but the transition logic is weak. Schedule selection is inappropriate for the target behavior. Shaping Protocol At least 4 logical successive approximations are defined. The "distance" between steps is consistent. 3 logical approximations are defined. Steps may be slightly too large or too small. Steps are missing or do not logically lead to the terminal behavior. Ethical Integrity Plan avoids punishment and focuses on learner agency. Potential pitfalls (e.g., coercion) are addressed. Plan is ethical and positive-only but lacks depth on learner agency. Plan relies on negative reinforcement or punishment disguised as positive.
Feedback & Critical Analysis
Assessment Tool // Unit: Behavioral Mechanics