Reinforcement · Adding something

Positive Reinforcement

The workhorse of operant conditioning: add something after a behavior, and the behavior grows. Here is exactly how it works, what counts as a reinforcer, and why so many attempts at it fail.

Updated 12 min read

Definition

Positive reinforcement is the process in which a behavior is followed by the addition of a stimulus, and as a result the behavior becomes more frequent, intense, or likely in the future. The added stimulus is called a positive reinforcer.

"Positive" means something is added (think plus sign), not that the outcome is pleasant — although positive reinforcers usually are. Whether a stimulus is a reinforcer is determined only by its effect on behavior.[1]

In brief

  • Positive reinforcement adds a stimulus after a behavior, and the behavior becomes more frequent, intense, or likely in the future.
  • A reinforcer is defined by its effect on behavior, not by whether the giver thinks it is nice.
  • Reinforcement must be immediate and contingent; reinforce every occurrence while learning, then thin to an intermittent schedule so the behavior persists.

How positive reinforcement works

The sequence is always the same: an antecedent sets the occasion, a behavior occurs, and immediately afterward a stimulus appears that was not there before. If the behavior then happens more often under similar conditions, positive reinforcement has taken place.

Notice that in each case the consequence is delivered because of the behavior (contingency) and right after it (contiguity). Remove either and the effect weakens sharply. A bonus paid in December for effort in March reinforces very little of that effort; the pellet that arrives ten seconds after the press teaches the rat almost nothing about pressing.[2]

The functional definition, again

A reinforcer is not a "reward." A reward is something the giver thinks is nice. A reinforcer is something that demonstrably increases the behavior it follows. Praise that a teenager finds embarrassing is not a reinforcer for them. Attention — even scolding — often is a reinforcer for a child who gets little of it. You find out what reinforces a behavior by watching what happens to the behavior.

Check any scenario in two questions

Not sure whether a consequence is this quadrant or a neighbor? Answer the two questions and the checker names it.

Examples of positive reinforcement

SettingBehaviorStimulus addedResult
HomeToddler uses the pottySticker and enthusiastic praiseUses the potty more often
HomeChild clears the table without being asked"That was really helpful — thank you."Clears the table more often
ClassroomStudent raises hand instead of calling outTeacher calls on themHand-raising increases
ClassroomClass transitions quietlyMarble added to the class jar (token)Quiet transitions increase
WorkplaceEmployee submits a report earlyPublic recognition in the team meetingEarly submissions increase
WorkplaceSalesperson closes a dealCommissionClosing behavior increases
Dog trainingDog sits on cueClick, then a treatSitting on cue increases
Dog trainingDog returns when calledPlay with a favorite toyRecall improves
TechnologyUser opens the appNew content, likes, badgesApp-opening increases
HealthPatient attends a treatment sessionVoucher (contingency management)Attendance increases
Self-managementYou put on running shoes at 6:30 a.m.Marking the habit done; a satisfying checkShoes go on more mornings
NatureBee visits a flowerNectarVisits to that flower type increase

Not every example of "adding something nice" is positive reinforcement. If a parent gives a child candy to stop a tantrum in the store, the candy positively reinforces the tantrum (and the end of the tantrum negatively reinforces the parent's candy-giving). Both people learned something; neither learned what they intended.

Types of positive reinforcers

Primary (unconditioned) reinforcers

Stimuli that reinforce without any learning history because they relate to biological needs: food, water, warmth, sexual contact, relief from pain, and — for social species — physical contact. Their effectiveness depends on deprivation: food is a powerful reinforcer for a hungry rat and a weak one for a full one.

Secondary (conditioned) reinforcers

Neutral stimuli that acquire reinforcing power by being paired with existing reinforcers. Money is the classic example; so are grades, praise, a clicker's sound, and a green checkmark. Conditioned reinforcers are what make delayed real-world consequences workable: the click bridges the gap between the dog's sit and the treat that follows a second later.[3]

Generalized reinforcers

Conditioned reinforcers paired with many different reinforcers, so they work regardless of the organism's current state. Money, tokens, and social approval are generalized reinforcers; that is why token economies work in classrooms and psychiatric wards.[4]

Social, tangible, activity, and sensory reinforcers

Practitioners often sort reinforcers by category: social (attention, praise, a smile), tangible (a sticker, a toy, a bonus), activity (getting to play, choosing the music), and sensory (a pleasant sound or texture). The Premack principle is the rule for activity reinforcers: a higher-probability behavior can reinforce a lower-probability one. "You can play video games after you finish your homework" is Premack in action.[5] The Premack principle in depth ›

What makes positive reinforcement effective

  1. Immediacy. Deliver the reinforcer within seconds. If you can't, deliver a conditioned reinforcer (a click, a word, a checkmark) immediately and the real one later.
  2. Contingency. The reinforcer follows the behavior and only the behavior. Non-contingent goodies are nice, but they don't teach.
  3. Magnitude and quality. Bigger and better reinforcers work better up to a point: rats press faster for larger and for sweeter food rewards, with diminishing returns as amount grows.[8] But the effect is relative. Crespi found that rats switched from a large reward to a small one ran slower than rats that had only ever had the small one — a negative contrast effect — and rats shifted upward briefly ran faster than rats always given the large amount.[9] Cutting a reinforcer is felt as a loss, not as a smaller gain. And small, immediate reinforcers usually beat large, delayed ones.
  4. Motivating operations. Deprivation makes a reinforcer stronger; satiation makes it weaker. Training a dog after dinner with kibble is a losing game.
  5. Schedule. Reinforce every occurrence while the behavior is being learned, then shift to an intermittent schedule so it persists. Schedules of reinforcement ›
  6. Individualization. Reinforcers are personal. What works is discovered by preference assessment and by watching the data, not assumed.

Positive reinforcement vs. negative reinforcement vs. bribery

Three things get confused constantly:

AspectPositive reinforcementNegative reinforcementBribery
What happensStimulus is added after the behaviorStimulus is removed after the behaviorReward is offered before the behavior, often to stop misbehavior
EffectBehavior increasesBehavior increasesUsually reinforces the misbehavior that prompted the bribe
ExamplePraise after homework is doneNagging stops when homework is done"If you stop screaming, I'll buy you the toy"

The difference between reinforcement and bribery is timing and target. Reinforcement follows the desired behavior. Bribery precedes it and is typically triggered by an undesired behavior, so the undesired behavior is what gets strengthened. More on negative reinforcement ›

Common mistakes

Does positive reinforcement undermine intrinsic motivation?

This is the most serious research-based objection, and the honest answer is "sometimes, under specific conditions." A meta-analysis of 128 studies found that expected, tangible rewards for doing an already-interesting activity reduced free-choice engagement with it afterward; verbal praise and unexpected rewards did not, and informational feedback generally increased intrinsic motivation.[6] A competing meta-analysis found the negative effect small and limited to a narrow set of conditions.[7] The practical guidance that follows from both:

How to use positive reinforcement: a six-step protocol

  1. Define the behavior so a camera could see it. "Raises her hand before speaking," not "is respectful." If you cannot count it, you cannot reinforce it.
  2. Find a reinforcer that works for this individual. Watch what the person chooses when free, ask, or test. A reinforcer is proven by its effect, not by your intentions; a sticker that a teenager finds embarrassing is not one.
  3. Deliver it immediately and every time — at first. Within seconds, and after every occurrence, until the behavior is reliable. If the real reinforcer must wait, mark the behavior instantly with a conditioned reinforcer: a word, a click, a check.
  4. Make it contingent and only contingent. The reinforcer follows the behavior and nothing else. Free access to the same reinforcer at other times drains its power.
  5. Thin the schedule. Once the behavior is steady, reinforce it only sometimes, unpredictably. Intermittent reinforcement is what makes the behavior survive days when no one is watching. How to thin a schedule ›
  6. Measure, and fade to natural reinforcers. Count the behavior before and after. Then hand the job to consequences the world already provides — the finished essay, the dog's walk, the satisfaction of the check mark — so the behavior no longer depends on you.

Praise, done properly

Praise is the cheapest positive reinforcer available, and most of it is wasted. Jere Brophy's analysis of teacher praise found that it changed behavior only when it was contingent (delivered for the behavior, not as a reflex), specific (named what was done well), and credible (sincere, and not inflated).[10] A later review added a fourth condition: praise for effort and strategy supports motivation, while praise for ability or for merely finishing can undermine it.[11] "You kept the paragraph to one idea — that made it clear" reinforces; "great job" does not. Praise, token economies, and the Good Behavior Game in the classroom ›

A worked example

A second-grade teacher wants a student, Maya, to start her worksheet without a reminder. Behavior: pencil on paper within one minute of the worksheet landing on her desk. Reinforcer: the teacher notices that Maya lights up when asked to hand out materials, so "hand out the next set" becomes the consequence. Delivery: the moment the pencil moves, a quiet "you started on your own — you're handing out the readers at ten." Every time, for a week. Thinning: in week two, the job goes to two starts out of three, then to unpredictable ones. Result: starts rose from one in five worksheets to almost all of them, and by the end of the month the teacher's occasional nod was enough. Every step is on this page; none of it required a sticker chart.

Quick check

Positive reinforcement in practice

Positive reinforcement is the default procedure of applied behavior analysis, the core of reward-based dog training, the basis of classroom systems like PBIS and token economies, the engine of contingency management in addiction treatment, and the mechanism behind most successful habit-formation methods. The through-line: identify the behavior, find a reinforcer that actually works for that individual, deliver it immediately and contingently, and thin the schedule over time.

Key takeaways

Explain it to a friend. Explain why a reinforcer is not the same thing as a reward, using an example from your own week and without using the word "reward" itself.

Frequently asked questions

What is positive reinforcement in simple terms?

Adding something after a behavior so the behavior happens more often. A dog sits, gets a treat, and sits more often. The treat is the positive reinforcer.

What is an example of positive reinforcement?

A child puts their toys away and a parent says, "Thank you for cleaning up — that was a big help." If the child cleans up more often afterward, the praise was a positive reinforcer. Other examples: a paycheck for hours worked, a like on a post, a treat for a dog's trick.

Is positive reinforcement the same as a reward?

Not exactly. A reward is something intended to be pleasant. A positive reinforcer is defined by its effect: it must actually increase the behavior it follows. Many rewards fail to reinforce, and some unpleasant things (like scolding, which is attention) do reinforce.

What is the difference between positive and negative reinforcement?

Both increase behavior. Positive reinforcement adds a stimulus (a treat appears). Negative reinforcement removes one (an alarm stops). "Positive" and "negative" mean added and removed, not good and bad.

Who developed positive reinforcement?

The principle traces to Edward Thorndike's law of effect (1898). B. F. Skinner developed the concept of reinforcement, the term "positive reinforcement," and the experimental science around it, beginning with The Behavior of Organisms (1938).

Is positive reinforcement effective for adults?

Yes. Organizational behavior management uses it to improve safety and productivity, contingency management uses it in addiction treatment with strong evidence, and it is the mechanism behind effective self-management and habit apps. Adults simply have more complex reinforcers (money, status, autonomy, feedback) than treats.

References

  1. Skinner, B. F. (1953). Science and Human Behavior. Macmillan.
  2. Lattal, K. A. (2010). Delayed reinforcement of operant behavior. Journal of the Experimental Analysis of Behavior, 93(1), 129–139.
  3. Williams, B. A. (1994). Conditioned reinforcement: Experimental and theoretical issues. The Behavior Analyst, 17(2), 261–285.
  4. Ayllon, T., & Azrin, N. H. (1968). The Token Economy: A Motivational System for Therapy and Rehabilitation. Appleton-Century-Crofts.
  5. Premack, D. (1959). Toward empirical behavior laws: I. Positive reinforcement. Psychological Review, 66(4), 219–233.
  6. Deci, E. L., Koestner, R., & Ryan, R. M. (1999). A meta-analytic review of experiments examining the effects of extrinsic rewards on intrinsic motivation. Psychological Bulletin, 125(6), 627–668.
  7. Cameron, J., & Pierce, W. D. (1994). Reinforcement, reward, and intrinsic motivation: A meta-analysis. Review of Educational Research, 64(3), 363–423.
  8. Hutt, P. J. (1954). Rate of bar pressing as a function of quality and quantity of food reward. Journal of Comparative and Physiological Psychology, 47(3), 235–239.
  9. Crespi, L. P. (1942). Quantitative variation of incentive and performance in the white rat. American Journal of Psychology, 55(4), 467–517.
  10. Brophy, J. (1981). Teacher praise: A functional analysis. Review of Educational Research, 51(1), 5–32.
  11. Henderlong, J., & Lepper, M. R. (2002). The effects of praise on children's intrinsic motivation: A review and synthesis. Psychological Bulletin, 128(5), 774–795.