Reward Loop Slides HOOKED ON THE LOOP
Predictable vs. Random Rewards
The Great Debate
The Vending Machine
You put money in, you get a snack out. Every. Single. Time.
The Slot Machine
You put money in... maybe nothing, maybe a jackpot. It's a mystery.
Which one is harder to walk away from? Why?
The Psychology Checklist
Continuous Reinforcement
A reward is given for every correct behavior. (The Vending Machine)
Fixed Schedule
A reward is given after a set number of times or time period. (Buy 10 coffees, get 1 free)
Variable Schedule
A reward is given after an unpredictable number of times. (Slot Machines, Loot Boxes)
The Power of Variable Rewards
Resistance to Extinction: When the rewards stop, you keep trying for much longer.
High Response Rates: You perform the behavior faster and more often.
The Cycle of "Just One More"
Simulation Phase
We are going to test these schedules ourselves. Grab your dice and your data sheets.
Group A
Fixed Ratio 3
Group B
Variable Ratio
Reward Simulation Worksheet The Reward Simulation
Case Study: Behavioral Persistence
Agent Name:
Date:
Mission Objective
You will act as a "user" performing a simple behavior (tapping your desk). Your partner will act as the "app" and deliver rewards (a tally mark). We want to see which reinforcement schedule makes you keep "using" the app even when the rewards stop.
Phase 1: Training
Collect 10 total rewards according to your assigned schedule.
Phase 2: Extinction
The rewards stop. Keep tapping until you feel like quitting. Count every tap.
Data Log
Assigned Schedule Phase 1: Taps to Reward Phase 2: Total Taps (No Reward) Fixed Ratio (3)
Variable Ratio
Check the one assigned to you.
|
Record the number of taps it took to get each of the 10 rewards:
1
2
3
4
5
6
7
8
9
10
|
|
Post-Sim Analysis
1. Compare your Phase 2 (Extinction) score with a partner who had the DIFFERENT schedule. Who tapped more? Why?
2. How did it feel when the rewards stopped? Was there a difference in "frustration" between the Fixed and Variable groups?
3. Predict: Why do social media apps use the Variable Ratio schedule instead of a Fixed one?
Reinforcement Simulation Teacher Guide TEACHER GUIDE
Predictable vs. Random Rewards
Lesson 1 of 5
Reinforcement Schedules
Learning Objectives
Define continuous, fixed-ratio, and variable-ratio schedules.
Observe "resistance to extinction" in variable-ratio schedules.
Connect behavioral psychology to digital design (e.g., loot boxes).
Materials Needed
Reward Loop Slides, Simulation Worksheets, 6-sided dice (1 per pair), pens/pencils.
Pacing Guide
Hook / Slides 15 min
Simulation Setup 5 min
Simulation Phase 1 10 min
Simulation Phase 2 10 min
Debrief 10 min
Facilitating the Simulation
1
Assigning Groups
Divide the class into pairs. One student is the User (tapper) and one is the App (rewarder). Half the pairs are Fixed-Ratio 3 , the other half are Variable-Ratio .
2
Phase 1: Training (10 Rewards)
Fixed Ratio: The App gives a reward (tally mark) every 3rd tap.
Variable Ratio: The App rolls a die before each session. If they roll a 1, the user gets a reward on the 1st tap. If they roll a 6, they get it on the 6th tap, etc. The App must keep this hidden from the User.
3
Phase 2: Extinction (No Rewards)
The App secretly stops giving rewards entirely. The User continues to tap until they decide to stop (give them a maximum of 2 minutes). The App counts how many times the User taps before giving up.
Debrief Discussion Points
Why do variable schedules resist extinction?
Because the user is used to "droughts." They don't know if the reward is truly gone or if the next tap is finally the winner. Fixed users notice immediately when the pattern breaks.
How does this apply to real apps?
Infinite scroll (Variable), Daily Login Streaks (Fixed), Leveling Up (Fixed). The variable nature of "content" makes social media hard to put down.
Gamification Lab Slides Lesson 02
THE GAMIFICATION LAB
Why points feel like "real" progress
The Definition
"The application of game-design elements and game principles in non-game contexts."
The Goal
To increase user engagement and loyalty through positive reinforcement.
Badges
Leaderboards
Points
Streaks
Secondary Reinforcers
A "ding" or a digital badge has no value in the real world. You can't eat it or spend it.
"They become powerful because we associate them with achievement, status, or progress."
🍔
Primary
Biological (Food, Water)
🏆
Secondary
Conditioned (Money, Points, Likes)
Duolingo vs. Snapchat
How do these apps use streaks to keep you coming back every single day without fail?
Analyze
Evaluate
Redesign
Gamification Lab Handout THE GAMIFICATION LAB
App Analysis & Reinforcement Patterns
Researcher:
Date:
Objective
Gamification isn't just about fun; it's about behavior modification . Below are two case studies. Identify the "reinforcers" (the rewards) and the "schedule" (how they are given) for each.
DUO
Case Study 1: The Language Owl
"You've learned 10 new words! You earned 20 XP and a 5-day streak! Don't let your streak die—your owl is waiting."
Analysis Questions
List 3 secondary reinforcers in this app:
What is the "punishment" for missing a day?
Reinforcement Log
What is the behavior being reinforced?
Does this feel like a "game" or "school"? Why?
CHIRP
Case Study 2: The Social Sphere
"Post shared! You now have a Level 3 'Influencer' badge. Your leaderboard rank increased from 50th to 42nd. Tap to see who liked your post!"
Analysis Questions
Identify the social reinforcer:
Why is a leaderboard more effective than just "points"?
Behavioral Loop
Action -> Reward -> Investment
How does the user "invest" their time to get the next reward?
Synthesizing the Lab
FINAL VERDICT: If these apps stopped giving rewards (XP, streaks, likes) tomorrow, which one would users quit first? Why?
Social Media Slot Machine Slides Lesson 03
THE SLOT MACHINE
IN YOUR POCKET
Social Media & Variable Ratios
Pull to Refresh
Why do we pull down on a screen to see new content?
It’s a physical gesture that mimics pulling the lever on a slot machine.
The short delay (the loading spinner) creates a moment of anticipation. Will you win? (Will the new post be good?)
The Infinite Scroll
Variable Rewards
Some posts are boring. Some are funny. Some are exciting. You never know when the "jackpot" post will appear.
Frictionless Experience
No pages to click = no "stopping points" for your brain to decide to leave.
The "Zombie" State
When you realize you've been scrolling for 30 minutes and don't remember a single thing you saw.
The "Red Dot" Effect
1
Intermittent Reinforcement: You don't get a notification every minute, but when you do, it triggers a Dopamine Spike .
Anticipation: Is it a like? A DM? A tag? The mystery is more rewarding than the actual notification.
FOMO: Variable schedules create a fear that if you don't check now, you'll miss out on the "reward."
Your Mission
Open your Digital Loop Tracker . For the next part of class, we are going to audit our own physical reactions to the "Ding."
Observe
Track
Break the Loop
My Digital Loop Tracker My Digital Loop Tracker
Module 03 // Self-Audit
User Identification Name: __________________________________
Session Date Date: ____________
1 The "Ding" Response
Think back to the last time your phone buzzed or "dinged." Answer the following honestly:
Physical Reaction:
(e.g., increased heart rate, reaching for pocket, phantom vibration)
Mental Urge:
(e.g., curiosity, anxiety if you can't check it, anticipation)
2 Auditing the Scroll
When you use an app with "infinite scroll" (TikTok, Reels, Twitter/X), complete the behavioral log below:
Target Reinforcer My Experience / Observation The "Boredom" Gap How many "boring" posts do you scroll past to get to one you actually like?
| |
| The "Zombie" State
How long does it take before you lose track of time?
| |
| The Social Jackpot
Which "reward" feels better: a post from a friend or a "viral" video from a stranger?
| |
3 Loop-Breaking Strategy
"If a variable ratio schedule is designed to make me persist, how can I introduce a 'stopping point' manually?"
Propose one specific change you can make to your phone settings or habits to break the reinforcement loop:
Why will this work psychologically?
Dark Patterns Slides Lesson 04
DARK PATTERNS
The Ethics of Behavioral Design
The "Dark Pattern"
A user interface that has been carefully crafted to trick users into doing things they didn't mean to do.
The Ethical Line
Engagement (helpful/fun) vs. Extraction (manipulative/harmful).
Deceptive Layouts
Forced Action
Sneaking Things in Carts
Hacking the Brain
Loss Aversion
"Maintain your 365-day streak!"
The fear of losing a conditioned reinforcer (the badge) is stronger than the reward of the badge itself.
Social Pressure
"Read Receipts" and "Typing..." indicators force you to stay in the app and respond immediately to avoid social "punishment."
Is it manipulation or design?
✅
Ethical
Helps user achieve their own goals (e.g., fitness, learning).
VS
❌
Unethical
Prioritizes company metrics (ad revenue) over user well-being.
THE BIG DEBATE
Should tech companies be legally banned from using "variable reinforcement" on teenagers?
Team: Protection
Team: Freedom
Dark Patterns Field Guide DARK PATTERNS FIELD GUIDE
Identifying Manipulative UX Design
Sneak into Basket
The site adds an extra item to your shopping cart through an opt-out radio button or checkbox on a prior page.
Roach Motel
The design makes it very easy for you to get into a certain situation (like a subscription) but very hard to get out of it.
Confirmshaming
The act of guilting the user into opting into something. The decline button might say: "No thanks, I prefer to pay full price."
Bait and Switch
You set out to do one thing, but a different, undesirable thing happens instead (like clicking 'X' to close an ad and it opens the store).
The Behavioral Trap
Many of these patterns rely on Positive Reinforcement or Negative Reinforcement (removing an annoying popup only if you sign up).
Social Proof
"12 other people are looking at this hotel right now!" (Forces immediate action through social anxiety).
Infinite Scroll
A variable ratio schedule that removes "stopping points" to prevent the brain from disengaging.
Activity: Spot the Pattern
Think of an app you use daily. Describe one "Dark Pattern" or "Manipulative Design" choice you've noticed within it:
App Name:
Description of the Pattern:
How does it use "Reinforcement" to keep you there?
Humane Design Slides Lesson 05 // Finale
HUMANE DESIGN
Building for Balance, Not Burnout
The Humane Shift
Humane design moves away from maximizing "time on screen" and toward maximizing user agency.
Intentional Reinforcement
Clear Stopping Points
Transparent Mechanics
The Challenge
Can we use psychology to reward stopping?
New Reinforcement Models
Positive Friction
Instead of infinite scroll, add a "Load More" button or a summary page every 10 posts.
This breaks the automatic behavior and forces a conscious choice.
Offline Rewards
Earn "Focus Points" or badges for keeping the phone locked for 2 hours.
Turning "not using the app" into the desired behavior.
Design Inspiration
Screen Time
Visualizing data to help users self-regulate.
Sleep Mode
Automatic muting of reinforcers during rest.
Batching
Delivering notifications only once per day.
APP REBUILD CHALLENGE
Pick your most addictive app. Redesign one core feature to be Humane while still being Useful .
Final Deliverable
A wireframe mockup and a behavioral justification.
App Rebuild Challenge Project APP REBUILD CHALLENGE
Final Project: Humane Design Mockup
Designer:
Phase 1: Problem Identification
App to Redesign:
Target "Toxic" Mechanic:
(e.g., infinite scroll, streak pressure, variable notifications)
Psychological Analysis
How does the current design use reinforcement to keep users engaged past the point of well-being?
Phase 2: The Humane Solution
New Reinforcement Strategy:
How will you reward the user for healthy behavior or create a "stopping point"?
"Good design doesn't just ask for attention; it respects it."
Interface Mockup
Visual Layout
DRAW MOCKUP HERE
Feature Annotation 1
Describe the first major change to the UI:
Feature Annotation 2
Describe the second major change to the UI:
The "Agency" Test
"How does this redesign give the power back to the user, rather than the algorithm?"