Shape complex behaviors through successive approximations
Reinforce progressively closer approximations to the target behavior rather than waiting for the full behavior to appear.
Why it works
Shaping works because complex behaviors cannot be reinforced until they occur, and they rarely occur spontaneously. By reinforcing steps toward the target — each one a closer approximation — the person (or animal) is guided through a behavior that would never emerge through waiting. This is the principle behind all graded task assignment in therapy and progressive overload in training: you reward the direction, not the destination.
How to do it
- Define the terminal behavior precisely (what would "fully there" look like?).
- Work backward to identify the step just below the current level of performance.
- Reinforce that step consistently until it is fluent, then raise the criterion to the next approximation.
- Move the criterion up only when the current step is stable — never when it is still being acquired.
Evidence
Shaping is a foundational operant technique with extensive experimental support in animal learning, applied behavior analysis, and behavioral therapy (graded task assignment). (clinical)
Setting the criterion steps requires judgment; moving too fast stalls progress, moving too slowly creates fixation at a sub-target level.
Sources
- Cooper, Heron & Heward, "Applied Behavior Analysis" — standard reference for shaping procedures
Common mistake
Setting the initial criterion too high so reinforcement never occurs, which extinguishes the behavior before it can be shaped.
Practice this with IX Coach
7 days free, then $40/month (~$1.30/day).
More practices for Operant Conditioning and Schedules of Reinforcement
- Design a genuine positive reinforcer for the target behavior
Identify something that genuinely increases your likelihood of repeating the behavior — not what should work, but what actually does.
- Use variable-ratio reinforcement to make habits persistent
Once a behavior is established, shift to an unpredictable reward schedule to make it resistant to extinction.
- Extinguish unwanted behaviors by removing their reinforcement
Stop a behavior by consistently withholding the reinforcer that maintains it — not by punishing it.
- Use differential reinforcement to increase desired behavior while reducing unwanted behavior
Reinforce the behavior you want while withholding reinforcement from the one you don’t — at the same time.
- Make consequences immediate to bridge the reward delay problem
The closer in time a consequence follows a behavior, the stronger its effect on that behavior.
- Modify antecedents to trigger behavior before it depends on motivation
Change the cues that precede a behavior to make it more or less likely to occur.
Related concepts
- Atomic Habits, Made Practical
The four laws, the real mechanisms, and where the science is strong
- Classical Conditioning and Habit Triggers
How neutral cues become powerful habit triggers — and how to use that fact
- The Power of Habit, Made Practical
The habit loop, craving, keystone habits, and an honest read on the science