Variable rewards in the hook model skill

The reward phase is what keeps users coming back.

by wondelai·MIT license·★ 2,235 Stars on the repo·GitHub ↗

Use now

Files of Variable rewards in the hook model

wondelai/main1 file
rewards.md
Show the full text240 lines

Variable Rewards in the Hook Model

The reward phase is what keeps users coming back. Predictable rewards lose power—variability is what maintains engagement.

The Science of Variable Rewards

Key insight: Dopamine is released in anticipation of reward, not upon receiving it. The brain craves the hunt, not the prize.

The slot machine effect: Variable ratio reinforcement (uncertain outcomes) creates the strongest behavioral response.

Autonomy matters: Users must feel in control. Forced or manipulative rewards backfire.


Variable Reinforcement Schedules

B.F. Skinner identified four primary schedules that predict how behavior responds to rewards:

Fixed Ratio

Reward after a set number of actions (e.g., every 5th action).

Behavior pattern: Steady work, pause after reward Example: Loyalty cards (buy 10, get 1 free) Product use: Completion bonuses, milestone rewards

Variable Ratio

Reward after random number of actions.

Behavior pattern: High, steady response rate—most powerful for sustained behavior Example: Slot machines, social media likes Product use: Feed refresh, message notifications, loot boxes

Fixed Interval

Reward after a set time period.

Behavior pattern: Slow start, activity increases near reward time Example: Daily login bonuses, weekly digest emails Product use: Scheduled content drops, time-gated rewards

Variable Interval

Reward at random times regardless of actions.

Behavior pattern: Steady, moderate response rate Example: Email inbox (messages arrive unpredictably) Product use: Push notifications, "new content" indicators

Schedule Selection Guide
Goal Best Schedule
Sustained engagement Variable ratio
Routine/habit building Fixed interval
Completion motivation Fixed ratio
Checking behavior Variable interval
Combining Schedules

Most successful products combine multiple schedules:

Instagram example:

  • Variable ratio: Likes (unknown timing/quantity)
  • Variable interval: Feed updates
  • Fixed ratio: Post completion (edit, filter, share)
  • Fixed interval: Stories expiration (24 hours)

Three Types of Variable Rewards

1. Rewards of the Tribe

Social rewards—validation, acceptance, and status from others.

Examples:

Product Tribe Reward Variability
Instagram Likes, comments, followers Unknown who will engage
LinkedIn Endorsements, profile views Unknown when/how many
Stack Overflow Upvotes, reputation points Unknown if answer is appreciated
Discord Reactions, mentions, role upgrades Unknown social interactions

Design patterns:

  • Show who engaged (not just numbers)
  • Notify on social validation
  • Display social proof (X people liked this)
  • Create status hierarchies (badges, levels)

Risks:

  • Can create anxiety and comparison
  • Social rejection is painful
  • Metrics can become obsessive
2. Rewards of the Hunt

The search for resources—information, money, or material possessions.

Examples:

Product Hunt Reward Variability
Twitter feed Interesting content Unknown what you'll find
Pinterest Visual inspiration Infinite scroll discovery
Amazon Deals, new products Unknown sales/discoveries
Email Important messages Unknown when/what arrives

Design patterns:

  • Infinite scroll (always more to discover)
  • Variable content feeds
  • Deals and limited offers
  • Personalized recommendations with surprises
  • Pull-to-refresh mechanism

Risks:

  • Can enable compulsive checking
  • FOMO exploitation
  • Endless consumption without satisfaction
3. Rewards of the Self

Intrinsic motivation—mastery, completion, competence, and consistency.

Examples:

Product Self Reward Variability
Duolingo Streak maintenance, XP Unknown difficulty of challenges
Video games Level progression, achievements Varying challenges and rewards
Codecademy Skill completion Unknown what you'll learn
Inbox Zero Completion satisfaction Unknown how hard it will be

Design patterns:

  • Progress bars and levels
  • Streaks and consistency rewards
  • Achievement unlocks
  • Skill progression visualization
  • "You completed X" celebrations

Risks:

  • Hollow achievements feel manipulative
  • Streaks create anxiety about breaking
  • Completion focus may not equal actual value

Designing Variable Rewards

Principles

1. Connect reward to core value The reward should reinforce why the user came. Random dopamine hits without value lead to regret.

  • Good: Learning app rewards progress with skill validation
  • Bad: Productivity app rewards with confetti animations

2. Maintain autonomy User must feel in control. Forced engagement or "dark patterns" create resentment.

  • Good: "You might also like..." (optional)
  • Bad: "You must complete X to continue"

3. Vary the type, not just the frequency Mix all three reward types when possible.

  • Tribe: "3 people liked your work"
  • Hunt: "Here's a resource we found for you"
  • Self: "You've completed 5 lessons this week"

4. Balance anticipation and delivery Too little reward = frustration and abandonment Too much reward = habituation and boredom

5. Tie reward to effort Earned rewards are more satisfying than free ones.

Variable Reward Checklist

For each reward in your product:

  • Is it genuinely valuable to the user?
  • Is there variability (not the same every time)?
  • Does it connect to an internal trigger?
  • Does user have autonomy/choice?
  • Does it load the next trigger (lead to investment)?
  • Which type is it? (Tribe/Hunt/Self)

Reward Timing

Immediate vs. Delayed
Timing When to Use Example
Immediate Core action completion Like animation, sound effect
Short delay Building anticipation "Your post is being processed..."
Scheduled Creating routine Daily rewards, weekly reports
Accumulated Encouraging return Points that build up over time
The "Loaded" Next Trigger

Best rewards "load" the next external trigger:

Action Reward Loaded Trigger
Post content Likes/comments Notification of engagement
Complete profile Matches shown Match notification
Add friends Friend activity Friend update notification
Create playlist "Made for you" playlist New music notification

Common Reward Mistakes

Mistake Why It Fails Solution
Predictable rewards Lose power quickly Add variability
Rewards unconnected to value Feel hollow/manipulative Align with user goals
Too much gamification Feels like work Use sparingly, meaningfully
Removing autonomy Creates resentment Always give choice
Same reward type only Incomplete motivation Mix Tribe/Hunt/Self
Delayed core reward User leaves before feeling value Front-load initial reward
Rewards without investment Low retention Connect reward to investment

Measuring Reward Effectiveness

Metric What It Shows Target
Time in reward phase Engagement depth Increasing over time
Return rate after reward Reward → Next cycle Higher = working
Sharing/screenshots Reward worth showing Organic social proof
Session length Overall engagement Consistent or growing
"Completed" actions Self-reward effectiveness High completion rate
1# Variable Rewards in the Hook Model
2 
3The reward phase is what keeps users coming back. Predictable rewards lose power—variability is what maintains engagement.
4 
5## The Science of Variable Rewards
6 
7**Key insight:** Dopamine is released in anticipation of reward, not upon receiving it. The brain craves the hunt, not the prize.
8 
9**The slot machine effect:** Variable ratio reinforcement (uncertain outcomes) creates the strongest behavioral response.
10 
11**Autonomy matters:** Users must feel in control. Forced or manipulative rewards backfire.
12 
13---
14 
15## Variable Reinforcement Schedules
16 
17B.F. Skinner identified four primary schedules that predict how behavior responds to rewards:
18 
19### Fixed Ratio
20 
21Reward after a set number of actions (e.g., every 5th action).
22 
23**Behavior pattern:** Steady work, pause after reward
24**Example:** Loyalty cards (buy 10, get 1 free)
25**Product use:** Completion bonuses, milestone rewards
26 
27### Variable Ratio
28 
29Reward after random number of actions.
30 
31**Behavior pattern:** High, steady response rate—most powerful for sustained behavior
32**Example:** Slot machines, social media likes
33**Product use:** Feed refresh, message notifications, loot boxes
34 
35### Fixed Interval
36 
37Reward after a set time period.
38 
39**Behavior pattern:** Slow start, activity increases near reward time
40**Example:** Daily login bonuses, weekly digest emails
41**Product use:** Scheduled content drops, time-gated rewards
42 
43### Variable Interval
44 
45Reward at random times regardless of actions.
46 
47**Behavior pattern:** Steady, moderate response rate
48**Example:** Email inbox (messages arrive unpredictably)
49**Product use:** Push notifications, "new content" indicators
50 
51### Schedule Selection Guide
52 
53| Goal | Best Schedule |
54|------|---------------|
55| Sustained engagement | Variable ratio |
56| Routine/habit building | Fixed interval |
57| Completion motivation | Fixed ratio |
58| Checking behavior | Variable interval |
59 
60### Combining Schedules
61 
62Most successful products combine multiple schedules:
63 
64**Instagram example:**
65- Variable ratio: Likes (unknown timing/quantity)
66- Variable interval: Feed updates
67- Fixed ratio: Post completion (edit, filter, share)
68- Fixed interval: Stories expiration (24 hours)
69 
70---
71 
72## Three Types of Variable Rewards
73 
74### 1. Rewards of the Tribe
75 
76Social rewards—validation, acceptance, and status from others.
77 
78**Examples:**
79 
80| Product | Tribe Reward | Variability |
81|---------|--------------|-------------|
82| Instagram | Likes, comments, followers | Unknown who will engage |
83| LinkedIn | Endorsements, profile views | Unknown when/how many |
84| Stack Overflow | Upvotes, reputation points | Unknown if answer is appreciated |
85| Discord | Reactions, mentions, role upgrades | Unknown social interactions |
86 
87**Design patterns:**
88- Show who engaged (not just numbers)
89- Notify on social validation
90- Display social proof (X people liked this)
91- Create status hierarchies (badges, levels)
92 
93**Risks:**
94- Can create anxiety and comparison
95- Social rejection is painful
96- Metrics can become obsessive
97 
98### 2. Rewards of the Hunt
99 
100The search for resources—information, money, or material possessions.
101 
102**Examples:**
103 
104| Product | Hunt Reward | Variability |
105|---------|-------------|-------------|
106| Twitter feed | Interesting content | Unknown what you'll find |
107| Pinterest | Visual inspiration | Infinite scroll discovery |
108| Amazon | Deals, new products | Unknown sales/discoveries |
109| Email | Important messages | Unknown when/what arrives |
110 
111**Design patterns:**
112- Infinite scroll (always more to discover)
113- Variable content feeds
114- Deals and limited offers
115- Personalized recommendations with surprises
116- Pull-to-refresh mechanism
117 
118**Risks:**
119- Can enable compulsive checking
120- FOMO exploitation
121- Endless consumption without satisfaction
122 
123### 3. Rewards of the Self
124 
125Intrinsic motivation—mastery, completion, competence, and consistency.
126 
127**Examples:**
128 
129| Product | Self Reward | Variability |
130|---------|-------------|-------------|
131| Duolingo | Streak maintenance, XP | Unknown difficulty of challenges |
132| Video games | Level progression, achievements | Varying challenges and rewards |
133| Codecademy | Skill completion | Unknown what you'll learn |
134| Inbox Zero | Completion satisfaction | Unknown how hard it will be |
135 
136**Design patterns:**
137- Progress bars and levels
138- Streaks and consistency rewards
139- Achievement unlocks
140- Skill progression visualization
141- "You completed X" celebrations
142 
143**Risks:**
144- Hollow achievements feel manipulative
145- Streaks create anxiety about breaking
146- Completion focus may not equal actual value
147 
148---
149 
150## Designing Variable Rewards
151 
152### Principles
153 
154**1. Connect reward to core value**
155The reward should reinforce why the user came. Random dopamine hits without value lead to regret.
156 
157- Good: Learning app rewards progress with skill validation
158- Bad: Productivity app rewards with confetti animations
159 
160**2. Maintain autonomy**
161User must feel in control. Forced engagement or "dark patterns" create resentment.
162 
163- Good: "You might also like..." (optional)
164- Bad: "You must complete X to continue"
165 
166**3. Vary the type, not just the frequency**
167Mix all three reward types when possible.
168 
169- Tribe: "3 people liked your work"
170- Hunt: "Here's a resource we found for you"
171- Self: "You've completed 5 lessons this week"
172 
173**4. Balance anticipation and delivery**
174Too little reward = frustration and abandonment
175Too much reward = habituation and boredom
176 
177**5. Tie reward to effort**
178Earned rewards are more satisfying than free ones.
179 
180### Variable Reward Checklist
181 
182For each reward in your product:
183 
184- [ ] Is it genuinely valuable to the user?
185- [ ] Is there variability (not the same every time)?
186- [ ] Does it connect to an internal trigger?
187- [ ] Does user have autonomy/choice?
188- [ ] Does it load the next trigger (lead to investment)?
189- [ ] Which type is it? (Tribe/Hunt/Self)
190 
191---
192 
193## Reward Timing
194 
195### Immediate vs. Delayed
196 
197| Timing | When to Use | Example |
198|--------|-------------|---------|
199| Immediate | Core action completion | Like animation, sound effect |
200| Short delay | Building anticipation | "Your post is being processed..." |
201| Scheduled | Creating routine | Daily rewards, weekly reports |
202| Accumulated | Encouraging return | Points that build up over time |
203 
204### The "Loaded" Next Trigger
205 
206Best rewards "load" the next external trigger:
207 
208| Action | Reward | Loaded Trigger |
209|--------|--------|----------------|
210| Post content | Likes/comments | Notification of engagement |
211| Complete profile | Matches shown | Match notification |
212| Add friends | Friend activity | Friend update notification |
213| Create playlist | "Made for you" playlist | New music notification |
214 
215---
216 
217## Common Reward Mistakes
218 
219| Mistake | Why It Fails | Solution |
220|---------|--------------|----------|
221| Predictable rewards | Lose power quickly | Add variability |
222| Rewards unconnected to value | Feel hollow/manipulative | Align with user goals |
223| Too much gamification | Feels like work | Use sparingly, meaningfully |
224| Removing autonomy | Creates resentment | Always give choice |
225| Same reward type only | Incomplete motivation | Mix Tribe/Hunt/Self |
226| Delayed core reward | User leaves before feeling value | Front-load initial reward |
227| Rewards without investment | Low retention | Connect reward to investment |
228 
229---
230 
231## Measuring Reward Effectiveness
232 
233| Metric | What It Shows | Target |
234|--------|---------------|--------|
235| Time in reward phase | Engagement depth | Increasing over time |
236| Return rate after reward | Reward → Next cycle | Higher = working |
237| Sharing/screenshots | Reward worth showing | Organic social proof |
238| Session length | Overall engagement | Consistent or growing |
239| "Completed" actions | Self-reward effectiveness | High completion rate |
240 

Discussion