# Applications of Operant Conditioning: Therapy, Education, Parenting, Work, Health, Technology, and More

> Operant conditioning in the classroom, parenting, the workplace, therapy, animal training, technology, and AI — and how strong the evidence is for each.

- Source: https://operantconditioning.com/applications/
- Author: Ryan Martinson (https://operantconditioning.com/about/)
- Publisher: Operant Conditioning Inc.
- Published: 2026-09-07 · Updated: 2026-09-10
- License: https://operantconditioning.com/terms/#copyright (quote with attribution and a link; free to reproduce for non-commercial teaching)

---

*Applications*

The same three-term contingency runs a therapy session, a classroom, a family dinner, a factory floor, a dolphin show, and a slot machine. Here is where operant conditioning is used, how, and how strong the evidence really is in each field.

> **Definition**
>
> The **applications of operant conditioning** are the deliberate arrangements of antecedents and consequences to change behavior that matters — in clinics, schools, homes, workplaces, zoos, software, and in the design of one's own life.
>
> The discipline built to do this systematically is **applied behavior analysis**, founded as a field in 1968, but the principles are used far beyond it, in some places rigorously and in others loosely.[1] The table below rates each field's evidence as candidly as the literature allows.

**In brief**

- Applications of operant conditioning are deliberate arrangements of antecedents and consequences to change behavior in clinics, schools, homes, workplaces, zoos, and software.
- Applied behavior analysis, founded in 1968, is the discipline built to do this systematically; its core tool is the functional analysis, and treatment follows function.
- The evidence varies by field: strong for classroom programs, parent training, contingency management, and behavioral activation; weaker for early intensive autism intervention and gamification.

| Field | Core techniques | Evidence strength | Learn more |
| --- | --- | --- | --- |
| **Applied behavior analysis** | Functional analysis, reinforcement, differential reinforcement, shaping, prompting and fading | Strong for targeted behavior change in single-case research; low-certainty and contested for early intensive intervention in autism | ABA › |
| **Education** | Token economies, Good Behavior Game, PBIS, precision teaching, direct instruction, praise ratios | Strong: randomized trials with follow-ups into adulthood; large meta-analyses for direct instruction | Classroom › |
| **Parenting** | Specific praise, planned ignoring, time-out, response cost, consistency | Strong for behavioral parent-training programs; strong evidence *against* corporal punishment | Parenting › |
| **Self-management** | Antecedent control, tiny behaviors, self-monitoring, immediate self-administered consequences | Moderate: self-monitoring and implementation intentions well supported; "self-reinforcement" mechanism debated | Self-management › · [Habits guide ›](https://operantconditioning.com/habits/) |
| **Health and clinical** | Contingency management, behavioral activation, exposure with response prevention, habit reversal | Strong: contingency management and behavioral activation are among the best-supported psychosocial treatments in their areas | Health › |
| **Animal training** | Marker signals, shaping, reinforcement-based husbandry training | Strong for reinforcement-based methods; aversive methods linked to welfare costs | Animals › · [Dog training ›](https://operantconditioning.com/dog-training/) |
| **Workplace** | Pinpointing, performance feedback, behavior-based safety, incentives | Moderate to strong within-organization studies; feedback reliably works; gamification mixed | Workplace › |
| **Technology and product design** | Variable-ratio feeds, notifications, streaks, loot boxes | Strong that the mechanisms drive engagement; contested whether population-level harms are large | Technology › |
| **Behavioral economics and AI** | Matching law, delay discounting, reinforcement learning | Strong: quantitative laws replicated across species; reinforcement learning is foundational to modern AI | Economics & AI › |

> **How to read the evidence column**
>
> "Strong" means randomized trials or many replicated single-case experiments with durable effects. "Moderate" means consistent results from weaker designs, or strong results for some components and not others. Where a field's reputation outruns its data, the table says so.

## Applied behavior analysis (ABA)

**Applied behavior analysis** is the science in which procedures derived from the principles of behavior are applied to improve socially significant behavior, and experimentation is used to show that the procedures were responsible for the change.[2] Baer, Wolf, and Risley's founding paper set out seven dimensions that still define the field: work should be applied, behavioral, analytic, technological, conceptually systematic, effective, and general.[1]

Its core tool is the **functional analysis**: experimentally testing what consequence maintains a problem behavior before trying to change it, a method established by Iwata and colleagues' 1982 study of self-injury.[3] Treatment then follows function. If a child hits to get attention, the analyst teaches a replacement — tapping a shoulder and saying "excuse me" — and reinforces it while the hitting no longer produces attention. That is **differential reinforcement of alternative behavior** (DRA), the most-used procedure in the field; its cousins reinforce the *absence* of the behavior for a set interval (DRO) or a behavior physically *incompatible* with it (DRI).[4] [How ABC data and functional analysis work ›](https://operantconditioning.com/abc-model/)

The evidence deserves a careful statement. For targeted behavior change — reducing self-injury, teaching communication, building daily-living skills — hundreds of single-case experiments show large, reliable effects. The claim that early intensive behavioral intervention (EIBI) transforms outcomes in autism rests on a smaller base. Lovaas reported in 1987 that 47% of children receiving roughly 40 hours a week of one-to-one therapy reached normal-range IQ and unsupported first-grade placement, versus 2% of a comparison group.[5] Later reviews were more modest: a Cochrane review found only low-certainty evidence, from a handful of mostly non-randomized studies, that EIBI improves adaptive behavior and IQ, and a 2020 meta-analysis found support for behavioral approaches weakened substantially when limited to randomized trials with outcomes not reported by caregivers.[6][7]

ABA also faces a principled critique from autistic adults and the neurodiversity movement: that its historical goal of making autistic children "indistinguishable" from peers, its early use of aversives (Lovaas's 1960s studies used contingent electric shock), and its emphasis on compliance were harmful whatever the outcome measures said.[8] Contemporary practice has moved toward assent, client-chosen goals, and reinforcement-only procedures. Practitioners are credentialed by the Behavior Analyst Certification Board (the BCBA credential dates from 1998), and the profession is licensed in most U.S. states.

## Operant conditioning in the classroom

Schools were the first institutions to use operant technology at scale, and they still produce some of its best evidence. The token economy — points earned for target behaviors and exchanged later for backup reinforcers — moved from a psychiatric ward into classrooms almost as soon as Ayllon and Azrin described it, and it works because tokens are generalized conditioned reinforcers that can be delivered the instant a behavior occurs.[9] The Good Behavior Game, a team-based contingency introduced in a fourth-grade classroom in 1969, is among the most thoroughly studied classroom procedures of all: in randomized trials that followed first-graders into their twenties, the boys who had played it showed lower rates of substance-use and antisocial-personality disorders than controls.[10][11] School-wide positive behavioral supports scale the same logic across a building, and Skinner's own contribution — programmed instruction, with its small steps, active responding, and immediate feedback — survives in adaptive software and Direct Instruction.[12][13][14] Praise, the cheapest reinforcer in the room, works only when it is contingent, specific, and credible.[56] [The classroom in depth: praise, token economies, the Good Behavior Game, PBIS, and what the evidence says ›](https://operantconditioning.com/classroom/)

## Operant conditioning in parenting

A parent's most powerful reinforcer is attention, and the most common parenting error is spending it on misbehavior. "Catching them being good" reverses the allocation: notice the behavior you want and name it specifically and immediately. Gerald Patterson's observations of families showed how the opposite pattern escalates — the child's tantrum is negatively reinforced when the parent gives in, and the parent's giving in is negatively reinforced when the tantrum stops, a *coercive family process* that trains both sides.[15] Parent management training, the best-supported treatment for childhood conduct problems, teaches the reverse contingencies: specific praise, planned ignoring, effective commands, brief time-out from reinforcement, and point systems.[16][17] On the punishment side, the largest meta-analysis of spanking — 75 studies and more than 160,000 children — found it associated with worse outcomes and no benefit.[18] [Parenting in depth: tantrums, time-out, sticker charts, bedtime, and the evidence ›](https://operantconditioning.com/parenting/)

## Self-management

The three-term contingency turned inward. Skinner argued that a person controls their own behavior with the same tools used on anyone else's — changing the stimulus, restraining themselves physically, arranging deprivation and satiation, and delivering their own consequences.[19] The best-supported components are antecedent control (implementation intentions have a medium-to-large meta-analytic effect on goal attainment) and self-monitoring (recording your own behavior reliably changes it).[20][21] Whether a self-delivered reward is "really" reinforcement is debated; what is not debated is that a cue you cannot miss, a behavior small enough to emit, and a consequence that arrives immediately outperform resolve. [The complete habit-building protocol ›](https://operantconditioning.com/habits/)

## Health and clinical applications

**Contingency management** for substance use is the clearest case of operant principles applied to a hard medical problem. In 1991 Stephen Higgins and colleagues offered cocaine-dependent outpatients vouchers exchangeable for retail goods, contingent on cocaine-negative urine samples, with the value escalating for consecutive clean samples and resetting after a positive one; retention and abstinence far exceeded standard counseling.[22] Two 2006 meta-analyses confirmed moderate, reliable effects across drugs, strongest for stimulants and opioids, and a 2008 meta-analytic review found contingency management produced the largest effects of any psychosocial treatment examined.[23][24][25] Its limits are what the theory predicts: effects fade after incentives stop unless natural reinforcers have taken over, and uptake has been slowed more by regulation and squeamishness about "paying people to stay clean" than by the data.

Chronic pain was one of the first medical problems given an operant analysis. Wilbert Fordyce observed that "pain behaviors" — guarding, limping, grimacing, resting, taking medication, talking about pain — are behaviors, and that they are reinforced: by attention and sympathy, by relief from unwanted duties, and by medication given in response to complaints. His inpatient program at the University of Washington, reported in 1973, reversed the contingencies. Staff attended to activity rather than complaints, exercise quotas were raised gradually with rest as the reinforcer for meeting them rather than for pain, and medication was given on a time schedule instead of on demand; patients' activity rose and their medication use and reported pain fell.[57][58] Operant-behavioral treatment remains a component of modern pain programs; in a randomized trial for fibromyalgia, operant-behavioral and cognitive-behavioral treatments each outperformed an attention-placebo condition, with the operant program's largest gains in physical functioning.[59] Nobody in this literature claims the pain is imaginary; the claim is that what a person does about pain is learned, and can be re-learned.

**Behavioral activation** treats depression as a collapse in response-contingent positive reinforcement, an analysis Charles Ferster offered in 1973.[26] The treatment schedules activities, grades difficult tasks, and dismantles avoidance so the person contacts reinforcement again. A 1996 dismantling study found the activation component alone matched the full cognitive therapy package; a 2006 randomized trial found it comparable to antidepressant medication and superior to cognitive therapy among more severely depressed patients; and a 2016 trial found it non-inferior to CBT when delivered by junior mental-health workers at lower cost.[27][28][29]

**Exposure therapy** has an operant engine inside its classical shell. In Mowrer's two-factor account, fear is classically conditioned, but *avoidance* is negatively reinforced by the relief it brings, and avoidance is what prevents the fear from extinguishing.[30] Exposure with response prevention blocks the escape so extinction can proceed. **Habit reversal training** — awareness training, a competing response, and social support — was introduced by Azrin and Nunn in 1973 for tics and nervous habits and is now the core of the leading behavioral treatment for Tourette syndrome.[31][32] **Biofeedback** is operant conditioning of physiological responses; Neal Miller's 1969 claims of visceral learning in rats proved hard to replicate, but clinical biofeedback has a real evidence base for specific problems such as incontinence and tension headache.[33] And financial incentives for **medication adherence** consistently improve adherence while they are in place.[34]

## Animal training

Modern animal training is operant conditioning with the lecture removed. A **marker signal** — a clicker or a word — is a conditioned reinforcer that bridges the gap between the behavior and the treat; [shaping](https://operantconditioning.com/shaping/) builds complex behavior from approximations; and reinforcement-based methods have displaced the dominance-and-correction approaches of the mid-twentieth century. [Dog training with operant conditioning ›](https://operantconditioning.com/dog-training/)

The commercial lineage runs through Keller and Marian Breland, Skinner's former students, who founded Animal Behavior Enterprises in 1943 and trained thousands of animals for advertising, exhibitions, and their "IQ Zoo." Their 1961 paper "The misbehavior of organisms," reporting raccoons that "washed" coins instead of depositing them, remains the classic demonstration that reinforcement works with, not against, an animal's evolved tendencies.[35] Marine-mammal training grew from the same roots; in a famous 1969 study, dolphins reinforced only for behaviors they had never shown before began producing novel actions.[36]

The quietest revolution is in zoos and laboratories. **Husbandry training** teaches animals to present a limb for a blood draw, open their mouths for dental checks, and enter a crate voluntarily, replacing restraint and anesthesia with cooperation and reducing measurable stress.[37] Reviews of aversive dog-training methods, meanwhile, link them to stress and problem behaviors without any advantage in effectiveness.[38]

## Operant conditioning in the workplace

**Organizational behavior management** (OBM) applies the same analysis to employees: pinpoint a behavior, measure it, give feedback, and reinforce it. The field has had its own journal since 1977. The signature demonstration is **behavior-based safety**: in a 1978 study at a food-manufacturing plant, researchers defined specific safe behaviors, posted graphed feedback on how often they occurred, and watched safe performance rise sharply — then fall when feedback was withdrawn and rise again when it returned.[39] A review of performance-feedback studies from 1985 to 1998 found feedback effective in most applications, most consistently when it was graphic, frequent, and combined with goals and reinforcement; a meta-analysis of behavior-modification programs across organizations reported an average performance improvement of about 17%.[40][41]

Aubrey Daniels, the field's best-known popularizer, summarizes the timing problem with a three-letter test: consequences that are positive, immediate, and certain control behavior; consequences that are negative, delayed, or uncertain barely register.[42] The annual performance review fails on every count. It arrives months after the behavior it addresses, it is delivered once, and for most people it is aversive, so it mainly evokes escape — the polished self-assessment, the defensive meeting — rather than changing what anyone does on Tuesday. Daily feedback from a supervisor who knows what to look for costs less and does more.

Two cautions. **Gamification** — points, badges, leaderboards — reliably produces short-term engagement and unreliably produces lasting change; a review of the empirical studies found positive but context-dependent effects that often faded with novelty.[43] And reinforcing a proxy reinforces the proxy. When sales targets and incentives at Wells Fargo rewarded accounts opened, employees opened millions of accounts customers had not asked for. The contingency worked as designed; the design was the problem.

## Technology and product design

Consumer software is the largest deployment of operant conditioning in history, and much of it is aimed at the user rather than for them. A social feed is a [variable-ratio schedule](https://operantconditioning.com/schedules-of-reinforcement/): pull to refresh, and the reinforcer — a like, a message, something novel — arrives after an unpredictable number of pulls, the schedule that produces the highest response rates and the most persistence when reinforcement stops. Notifications are discriminative stimuli for opening the app. Streaks convert a habit into an avoidance contingency — the reinforcer becomes not losing the count — and borrow loss aversion to make a missed day feel like a fine.

Gambling is the original engineered variable-ratio schedule, and slot machines add a refinement the laboratory did not anticipate: the **near miss**. Two cherries and a lemon is a loss, but it looks like almost winning, and Reid argued in 1986 that near misses encourage continued play as if they were partial reinforcement.[60] Brain-imaging work later found that near misses recruit some of the same reward circuitry as wins and increase the urge to keep playing, more so in people who gamble more heavily.[61] Game designers borrowed the whole toolkit early; a 2001 article by the psychologist John Hopson laid out how to apply schedules of reinforcement to game design so that players keep playing; the industry's later name for the resulting anticipation–activity–reward cycle is the **compulsion loop**, and most games since have one.[62]

**Loot boxes** in video games are the starkest case. Researchers who examined popular games found that a large share of loot-box systems met the psychological criteria for gambling, and spending on them correlates with problem-gambling severity; Belgium's gaming regulator ruled paid loot boxes illegal gambling in 2018.[44][45] BJ Fogg's **behavior model** — behavior occurs when motivation, ability, and a prompt converge — is a design-oriented cousin of the four-term contingency: the prompt is the antecedent, ability stands in for response effort, and motivation for the motivating operation.[46] When those tools trick users into choices they would not otherwise make, the designs are called **dark patterns**, and they are increasingly the target of regulation.

Honesty requires a caveat: whether the resulting engagement causes large population-level harm is contested. Large-sample analyses of adolescent well-being find the association with digital technology use to be tiny.[47] That the mechanisms work is not in doubt; how much damage they do is. The same science runs the other way: environment design against the feeds — the phone in the kitchen, the app logged out — and tools that make the user set the antecedent and consequence for behavior they have chosen. [Using the loop for your own goals ›](https://operantconditioning.com/habits/)

## Behavioral economics and artificial intelligence

Operant research turned quantitative in 1961, when Richard Herrnstein showed that pigeons offered two keys, each paying on its own variable-interval schedule, distributed their pecks in proportion to the reinforcement each key delivered — the **[matching law](https://operantconditioning.com/matching-law/)**.[48] Matching turned choice into something that could be modeled with the tools of economics, and by the 1990s laboratory animals had been shown to obey demand curves, substitute between goods, and respond to price much as human consumers do.[49]

The most consequential offshoot is **delay discounting**. Using an adjusting procedure, James Mazur showed that the value of a delayed reinforcer falls along a hyperbola rather than the exponential curve standard economics assumed.[50] George Ainslie had already worked out the implication: hyperbolic curves cross, so a person who prefers the larger, later reward from a distance will flip to the smaller, sooner one as it approaches — a mathematical account of impulsiveness and of why commitment devices are needed.[51] Steeper discounting has since been documented in smokers, in people with substance-use disorders, and in problem gamblers, and it is now studied as a process cutting across many conditions.[52]

Public policy has borrowed from both halves of the contingency. **Nudges** — Thaler and Sunstein's term for changes in "choice architecture" that steer behavior without changing incentives, such as making retirement saving the default — are antecedent interventions: they alter the stimulus conditions under which a choice is made, not its consequences, and they are cheap precisely because no reinforcer has to be delivered.[63] Incentive programs — conditional cash transfers, deposit contracts, sin taxes — work on the consequence side. The behavioral analysis predicts what the evaluations find: nudges produce modest, durable effects when the behavior is a one-off choice, and incentives produce larger effects that fade when the incentive stops unless a natural reinforcer takes over.

The law of effect also became an algorithm. **Reinforcement learning**, the branch of machine learning in which an agent learns by acting and receiving reward signals, descends directly from Thorndike and Skinner by way of temporal-difference learning, and its standard textbook says so in its opening pages.[53] In 1997 Schultz, Dayan, and Montague showed that midbrain dopamine neurons fire in the pattern of a temporal-difference prediction error — the brain running the same computation.[54] [The neuroscience of operant conditioning ›](https://operantconditioning.com/neuroscience/) Reinforcement learning drove the systems that mastered Go, and reinforcement learning from human feedback is one of the methods used to train large language models to behave as people prefer.[55] [The history from Thorndike to reinforcement learning ›](https://operantconditioning.com/history/)

## Military training

The most frequently cited military application is also the most disputed. After the Second World War, the U.S. Army historian S. L. A. Marshall reported, on the basis of group interviews, that only about 15–20% of riflemen had fired their weapons at the enemy in combat.[64] Training changed in response: bullseye targets were replaced by man-shaped silhouettes that pop up briefly and fall when hit — immediate feedback on a realistic discriminative stimulus — and the conditioned response was practiced hundreds of times. Dave Grossman argued in *On Killing* that this was operant conditioning in all but name, and credited it with raising reported firing rates to about 55% in Korea and over 90% in Vietnam.[65] The account should be read with care: Marshall's figures have been challenged by historians who found no record of the systematic interviews he described, and the later rates are not measured the same way.[66] What is not in dispute is the design of the training itself, which is a textbook contingency, or Skinner's own wartime contribution: Project Pigeon, in which he shaped pigeons to peck at an image of a target so that their pecks could steer a glide bomb — a working system that the military declined to deploy.[67]

## Coercion: the dark side of the contingency

The same principles that build a token economy can hold a person in a harmful relationship or a toxic workplace, and behavior analysts have said so plainly. Murray Sidman's *Coercion and Its Fallout* catalogued what aversive control does to the people subjected to it: it produces escape and avoidance, countercontrol, aggression, and a generalized suppression of behavior that reaches far beyond the punished response.[68]

**Traumatic bonding** is the clearest case. Donald Dutton and Susan Painter proposed that two features of abusive relationships — a power imbalance and the *intermittency* of abuse, with cycles of cruelty and reconciliation — produce unusually strong attachment to the abuser, on the same logic by which [intermittent reinforcement](https://operantconditioning.com/schedules-of-reinforcement/) produces the most persistent behavior in the laboratory. Their follow-up study of women who had left abusive partners found that intermittency of abuse and dominance predicted continued attachment months later.[69][70] Popular accounts of psychological manipulation list the same operations — intermittent reward, negative reinforcement by the removal of hostility, punishment, and one-trial traumatic learning — and they are recognizable to anyone who has read this far.

Organizations run the same contingencies at scale. Blake Ashforth's analysis of "petty tyranny" in management described supervisors who rule through arbitrary punishment, belittling, and the withholding of consideration, and traced the consequences the laboratory predicts: helplessness, low initiative, and the disappearance of any behavior not strictly required.[71] A culture of fear is a culture of avoidance, and avoidance, as [its own page](https://operantconditioning.com/avoidance-learning/) explains, is the behavior least sensitive to whether the threat is real. The ethical guidance of behavior analysis — reinforcement first, the least restrictive procedure, consent, the client's own goals — exists because the field knows exactly how effective the alternative is.

## Key takeaways

- The same three-term contingency runs every application. Applied behavior analysis is the discipline built to apply it systematically, and its core tool is the functional analysis: find what consequence maintains a behavior, then teach and reinforce an alternative while the problem behavior no longer pays.
- Evidence strength varies by field, and a field's reputation can outrun its data. The Good Behavior Game, behavioral parent training, contingency management, and behavioral activation have randomized trials with durable effects; early intensive behavioral intervention in autism rests on low-certainty evidence and faces a principled critique from autistic adults.
- Consequences control behavior when they are positive, immediate, and certain; delayed, infrequent, and aversive consequences such as the annual review mainly evoke escape. Reinforcing a proxy reinforces the proxy.
- Consumer software is the largest deployment of operant conditioning in history: feeds run variable-ratio schedules, notifications are discriminative stimuli, streaks are avoidance contingencies, and loot boxes reproduce the structure of gambling. That the mechanisms work is not in doubt; how much population-level harm they do is.
- The principles cut both ways. Intermittent abuse produces traumatic bonding on the same logic that makes intermittent reinforcement persistent, and coercive management produces avoidance and helplessness. The field's ethics, reinforcement first, the least restrictive procedure, and consent, exist because it knows how effective the alternative is.

### Check yourself

**A parent gives in to a tantrum and the tantrum stops. A classmate says the child has been rewarded and the parent has been punished by losing the standoff. What is actually happening, in operant terms?**

Both behaviors are being strengthened, not one. The child's tantrum is negatively reinforced when the parent gives in, and the parent's giving in is negatively reinforced when the tantrum stops. This is Patterson's coercive family process: each side trains the other, which is why parent management training teaches the reverse contingencies.

**A clinic offers vouchers for cocaine-negative urine samples. A critic objects that paying people to stay clean cannot work because motivation has to come from within. What does the evidence say, and what is the program's real limitation?**

The evidence says it works: meta-analyses find moderate, reliable effects across drugs, strongest for stimulants and opioids, and a review of psychosocial treatments found contingency management had the largest effects of any approach examined. Its real limitation is what the theory predicts: effects fade after the incentives stop unless natural reinforcers have taken over.

**A bank pays incentives on the number of accounts opened, and the number of accounts opened soars. Did the contingency fail?**

No. It worked exactly as designed, and the design was the problem. Reinforcing a proxy reinforces the proxy, so employees opened millions of accounts customers had not asked for. Organizational behavior management starts by pinpointing the behavior that matters, not a stand-in for it.

**One agency wants more people to enroll in a retirement plan. Another wants people to keep exercising for a year. Which tool suits each, a nudge or an incentive, and why?**

Enrollment is a one-off choice, so a nudge fits: making saving the default changes the stimulus conditions under which the choice is made, delivers no reinforcer, and produces a modest, durable effect. Exercising for a year is ongoing behavior, so an incentive works on the consequence side and produces a larger effect, but that effect will fade when the incentive stops unless a natural reinforcer has taken over.

**Explain it to a friend.** Explain what a classroom token economy and a slot machine have in common, and what separates them, without using the word "reinforcement."

## Frequently asked questions

**What are the main applications of operant conditioning?**

Applied behavior analysis (including autism and developmental-disability services), classroom management and instruction, parenting and parent training, animal training, organizational behavior management and workplace safety, clinical treatments such as contingency management and behavioral activation, product and game design, self-management and habit formation, and — in its mathematical form — behavioral economics and reinforcement learning in AI.

**How is operant conditioning used in the classroom?**

Through token economies, the Good Behavior Game, school-wide PBIS, high praise-to-reprimand ratios, and instructional methods built on immediate feedback such as programmed instruction, precision teaching, and Direct Instruction. The Good Behavior Game has randomized trials with follow-ups showing benefits into early adulthood.

**How is operant conditioning used in parenting?**

By reinforcing wanted behavior with specific, immediate attention; using planned ignoring for attention-maintained misbehavior; using brief, calm time-out and response cost for serious misbehavior; and being consistent rather than severe. Behavioral parent-training programs such as PMT and PCIT package these skills and are among the best-supported treatments for childhood behavior problems.

**How is operant conditioning used in the workplace?**

Organizational behavior management pinpoints specific behaviors, measures them, and delivers frequent feedback and reinforcement. Behavior-based safety programs and graphic performance feedback have decades of supporting studies. Annual reviews fail as consequences because they are delayed, infrequent, and aversive.

**Is ABA therapy the same as operant conditioning?**

No. Operant conditioning is the basic learning process; applied behavior analysis is the professional discipline that applies its principles (and others) to socially important behavior, with its own methods, credentials, and ethics code. ABA is one application of operant conditioning among many.

**Is contingency management effective for addiction?**

Yes. Meta-analyses find it produces reliable reductions in drug use, with the strongest effects for stimulants and opioids, and a broad review of psychosocial treatments found it had the largest effects of any approach. Its main limitation is that gains can fade after incentives end unless natural reinforcers have taken over.

**How do apps and games use operant conditioning?**

Feeds deliver unpredictable rewards on a variable-ratio schedule, notifications act as cues to open the app, streaks turn use into loss-avoidance, and loot boxes reproduce the structure of gambling. The same principles can be used deliberately for your own goals by controlling the cues in your environment and attaching immediate consequences to behaviors you choose.

**Does operant conditioning apply to artificial intelligence?**

Yes. Reinforcement learning — the AI method in which an agent learns from reward signals — is a direct mathematical descendant of the law of effect, and dopamine neurons in the brain have been shown to compute the same prediction-error signal its algorithms use. Reinforcement learning underlies game-playing systems and is used in training large language models.

## References

1. Baer, D. M., Wolf, M. M., & Risley, T. R. (1968). Some current dimensions of applied behavior analysis. *Journal of Applied Behavior Analysis, 1*(1), 91–97.
2. Cooper, J. O., Heron, T. E., & Heward, W. L. (2020). *Applied Behavior Analysis* (3rd ed.). Pearson.
3. Iwata, B. A., Dorsey, M. F., Slifer, K. J., Bauman, K. E., & Richman, G. S. (1982/1994). Toward a functional analysis of self-injury. *Analysis and Intervention in Developmental Disabilities, 2*(1), 3–20. Reprinted in *Journal of Applied Behavior Analysis, 27*(2), 197–209.
4. Carr, E. G., & Durand, V. M. (1985). Reducing behavior problems through functional communication training. *Journal of Applied Behavior Analysis, 18*(2), 111–126.
5. Lovaas, O. I. (1987). Behavioral treatment and normal educational and intellectual functioning in young autistic children. *Journal of Consulting and Clinical Psychology, 55*(1), 3–9.
6. Reichow, B., Hume, K., Barton, E. E., & Boyd, B. A. (2018). Early intensive behavioral intervention (EIBI) for young children with autism spectrum disorders (ASD). *Cochrane Database of Systematic Reviews*, Issue 5, CD009260.
7. Sandbank, M., Bottema-Beutel, K., Crowley, S., et al. (2020). Project AIM: Autism intervention meta-analysis for studies of young children. *Psychological Bulletin, 146*(1), 1–29.
8. Lovaas, O. I., Schaeffer, B., & Simmons, J. Q. (1965). Building social behavior in autistic children by use of electric shock. *Journal of Experimental Research in Personality, 1*, 99–109.
9. Ayllon, T., & Azrin, N. H. (1968). *The Token Economy: A Motivational System for Therapy and Rehabilitation*. Appleton-Century-Crofts.
10. Barrish, H. H., Saunders, M., & Wolf, M. M. (1969). Good behavior game: Effects of individual contingencies for group consequences on disruptive behavior in a classroom. *Journal of Applied Behavior Analysis, 2*(2), 119–124.
11. Kellam, S. G., Brown, C. H., Poduska, J. M., et al. (2008). Effects of a universal classroom behavior management program in first and second grades on young adult behavioral, psychiatric, and social outcomes. *Drug and Alcohol Dependence, 95*(Suppl. 1), S5–S28.
12. Bradshaw, C. P., Mitchell, M. M., & Leaf, P. J. (2010). Examining the effects of schoolwide positive behavioral interventions and supports on student outcomes: Results from a randomized controlled effectiveness trial in elementary schools. *Journal of Positive Behavior Interventions, 12*(3), 133–148.
13. Skinner, B. F. (1954). The science of learning and the art of teaching. *Harvard Educational Review, 24*(2), 86–97. See also Skinner, B. F. (1958). Teaching machines. *Science, 128*(3330), 969–977.
14. Stockard, J., Wood, T. W., Coughlin, C., & Rasplica Khoury, C. (2018). The effectiveness of Direct Instruction curricula: A meta-analysis of a half century of research. *Review of Educational Research, 88*(4), 479–507.
15. Wolf, M. M., Risley, T. R., & Mees, H. (1964). Application of operant conditioning procedures to the behaviour problems of an autistic child. *Behaviour Research and Therapy, 1*, 305–312.
16. Azrin, N. H., & Holz, W. C. (1966). Punishment. In W. K. Honig (Ed.), *Operant Behavior: Areas of Research and Application* (pp. 380–447). Appleton-Century-Crofts.
17. Patterson, G. R. (1982). *Coercive Family Process*. Castalia.
18. Gershoff, E. T., & Grogan-Kaylor, A. (2016). Spanking and child outcomes: Old controversies and new meta-analyses. *Journal of Family Psychology, 30*(4), 453–469.
19. Skinner, B. F. (1953). *Science and Human Behavior*. Macmillan.
20. Gollwitzer, P. M., & Sheeran, P. (2006). Implementation intentions and goal achievement: A meta-analysis of effects and processes. *Advances in Experimental Social Psychology, 38*, 69–119.
21. Harkin, B., Webb, T. L., Chang, B. P. I., et al. (2016). Does monitoring goal progress promote goal attainment? A meta-analysis of the experimental evidence. *Psychological Bulletin, 142*(2), 198–229.
22. Higgins, S. T., Delaney, D. D., Budney, A. J., Bickel, W. K., Hughes, J. R., Foerg, F., & Fenwick, J. W. (1991). A behavioral approach to achieving initial cocaine abstinence. *American Journal of Psychiatry, 148*(9), 1218–1224.
23. Prendergast, M., Podus, D., Finney, J., Greenwell, L., & Roll, J. (2006). Contingency management for treatment of substance use disorders: A meta-analysis. *Addiction, 101*(11), 1546–1560.
24. Lussier, J. P., Heil, S. H., Mongeon, J. A., Badger, G. J., & Higgins, S. T. (2006). A meta-analysis of voucher-based reinforcement therapy for substance use disorders. *Addiction, 101*(2), 192–203.
25. Dutra, L., Stathopoulou, G., Basden, S. L., Leyro, T. M., Powers, M. B., & Otto, M. W. (2008). A meta-analytic review of psychosocial interventions for substance use disorders. *American Journal of Psychiatry, 165*(2), 179–187.
26. Ferster, C. B. (1973). A functional analysis of depression. *American Psychologist, 28*(10), 857–870.
27. Jacobson, N. S., Dobson, K. S., Truax, P. A., et al. (1996). A component analysis of cognitive-behavioral treatment for depression. *Journal of Consulting and Clinical Psychology, 64*(2), 295–304.
28. Dimidjian, S., Hollon, S. D., Dobson, K. S., et al. (2006). Randomized trial of behavioral activation, cognitive therapy, and antidepressant medication in the acute treatment of adults with major depression. *Journal of Consulting and Clinical Psychology, 74*(4), 658–670.
29. Richards, D. A., Ekers, D., McMillan, D., et al. (2016). Cost and outcome of behavioural activation versus cognitive behavioural therapy for depression (COBRA): A randomised, controlled, non-inferiority trial. *The Lancet, 388*(10047), 871–880.
30. Mowrer, O. H. (1960). *Learning Theory and Behavior*. Wiley.
31. Azrin, N. H., & Nunn, R. G. (1973). Habit-reversal: A method of eliminating nervous habits and tics. *Behaviour Research and Therapy, 11*(4), 619–628.
32. Piacentini, J., Woods, D. W., Scahill, L., et al. (2010). Behavior therapy for children with Tourette disorder: A randomized controlled trial. *JAMA, 303*(19), 1929–1937.
33. Miller, N. E. (1969). Learning of visceral and glandular responses. *Science, 163*(3866), 434–445.
34. DeFulio, A., & Silverman, K. (2012). The use of incentives to reinforce medication adherence. *Preventive Medicine, 55*(Suppl.), S86–S94.
35. Breland, K., & Breland, M. (1961). The misbehavior of organisms. *American Psychologist, 16*(11), 681–684.
36. Pryor, K. W., Haag, R., & O'Reilly, J. (1969). The creative porpoise: Training for novel behavior. *Journal of the Experimental Analysis of Behavior, 12*(4), 653–661.
37. Laule, G. E., Bloomsmith, M. A., & Schapiro, S. J. (2003). The use of positive reinforcement training techniques to enhance the care, management, and welfare of primates in the laboratory. *Journal of Applied Animal Welfare Science, 6*(3), 163–173.
38. Ziv, G. (2017). The effects of using aversive training methods in dogs — A review. *Journal of Veterinary Behavior, 19*, 50–60.
39. Komaki, J., Barwick, K. D., & Scott, L. R. (1978). A behavioral approach to occupational safety: Pinpointing and reinforcing safe performance in a food manufacturing plant. *Journal of Applied Psychology, 63*(4), 434–445.
40. Alvero, A. M., Bucklin, B. R., & Austin, J. (2001). An objective review of the effectiveness and essential characteristics of performance feedback in organizational settings (1985–1998). *Journal of Organizational Behavior Management, 21*(1), 3–29.
41. Stajkovic, A. D., & Luthans, F. (1997). A meta-analysis of the effects of organizational behavior modification on task performance, 1975–95. *Academy of Management Journal, 40*(5), 1122–1149.
42. Daniels, A. C. (2000). *Bringing Out the Best in People: How to Apply the Astonishing Power of Positive Reinforcement* (2nd ed.). McGraw-Hill.
43. Hamari, J., Koivisto, J., & Sarsa, H. (2014). Does gamification work? — A literature review of empirical studies on gamification. *Proceedings of the 47th Hawaii International Conference on System Sciences*, 3025–3034.
44. Drummond, A., & Sauer, J. D. (2018). Video game loot boxes are psychologically akin to gambling. *Nature Human Behaviour, 2*(8), 530–532.
45. Zendle, D., & Cairns, P. (2018). Video game loot boxes are linked to problem gambling: Results of a large-scale survey. *PLoS ONE, 13*(11), e0206767.
46. Fogg, B. J. (2009). A behavior model for persuasive design. *Proceedings of the 4th International Conference on Persuasive Technology*, Article 40.
47. Orben, A., & Przybylski, A. K. (2019). The association between adolescent well-being and digital technology use. *Nature Human Behaviour, 3*(2), 173–182.
48. Herrnstein, R. J. (1961). Relative and absolute strength of response as a function of frequency of reinforcement. *Journal of the Experimental Analysis of Behavior, 4*(3), 267–272.
49. Kagel, J. H., Battalio, R. C., & Green, L. (1995). *Economic Choice Theory: An Experimental Analysis of Animal Behavior*. Cambridge University Press.
50. Mazur, J. E. (1987). An adjusting procedure for studying delayed reinforcement. In M. L. Commons, J. E. Mazur, J. A. Nevin, & H. Rachlin (Eds.), *Quantitative Analyses of Behavior: Vol. 5. The Effect of Delay and of Intervening Events on Reinforcement Value* (pp. 55–73). Erlbaum.
51. Ainslie, G. (1975). Specious reward: A behavioral theory of impulsiveness and impulse control. *Psychological Bulletin, 82*(4), 463–496.
52. Bickel, W. K., Odum, A. L., & Madden, G. J. (1999). Impulsivity and cigarette smoking: Delay discounting in current, never, and ex-smokers. *Psychopharmacology, 146*(4), 447–454.
53. Sutton, R. S., & Barto, A. G. (2018). *Reinforcement Learning: An Introduction* (2nd ed.). MIT Press.
54. Schultz, W., Dayan, P., & Montague, P. R. (1997). A neural substrate of prediction and reward. *Science, 275*(5306), 1593–1599.
55. Christiano, P. F., Leike, J., Brown, T., Martic, M., Legg, S., & Amodei, D. (2017). Deep reinforcement learning from human preferences. *Advances in Neural Information Processing Systems, 30*.
56. Brophy, J. (1981). Teacher praise: A functional analysis. *Review of Educational Research, 51*(1), 5–32.
57. Fordyce, W. E., Fowler, R. S., Lehmann, J. F., DeLateur, B. J., Sand, P. L., & Trieschmann, R. B. (1973). Operant conditioning in the treatment of chronic pain. *Archives of Physical Medicine and Rehabilitation, 54*(9), 399–408.
58. Fordyce, W. E. (1976). *Behavioral Methods for Chronic Pain and Illness*. Mosby.
59. Thieme, K., Flor, H., & Turk, D. C. (2006). Psychological pain treatment in fibromyalgia syndrome: Efficacy of operant behavioural and cognitive behavioural treatments. *Arthritis Research & Therapy, 8*(4), R121.
60. Reid, R. L. (1986). The psychology of the near miss. *Journal of Gambling Behavior, 2*(1), 32–39.
61. Clark, L., Lawrence, A. J., Astley-Jones, F., & Gray, N. (2009). Gambling near-misses enhance motivation to gamble and recruit win-related brain circuitry. *Neuron, 61*(3), 481–490.
62. Hopson, J. (2001, April 27). Behavioral game design. *Gamasutra*.
63. Thaler, R. H., & Sunstein, C. R. (2008). *Nudge: Improving Decisions About Health, Wealth, and Happiness*. Yale University Press.
64. Marshall, S. L. A. (1947). *Men Against Fire: The Problem of Battle Command in Future War*. William Morrow.
65. Grossman, D. (1995). *On Killing: The Psychological Cost of Learning to Kill in War and Society*. Little, Brown.
66. Spiller, R. J. (1988). S.L.A. Marshall and the ratio of fire. *RUSI Journal, 133*(4), 63–71.
67. Skinner, B. F. (1960). Pigeons in a pelican. *American Psychologist, 15*(1), 28–37.
68. Sidman, M. (1989). *Coercion and Its Fallout*. Authors Cooperative.
69. Dutton, D. G., & Painter, S. L. (1981). Traumatic bonding: The development of emotional attachments in battered women and other relationships of intermittent abuse. *Victimology, 6*(1–4), 139–155.
70. Dutton, D. G., & Painter, S. (1993). Emotional attachments in abusive relationships: A test of traumatic bonding theory. *Violence and Victims, 8*(2), 105–120.
71. Ashforth, B. (1994). Petty tyranny in organizations. *Human Relations, 47*(7), 755–778.


## About the author

Ryan holds a master's degree from UCLA, where he studied animal behavior in Daniel Blumstein's lab and was part of the university's Evolutionary Medicine Program, which applies findings from evolutionary biology and animal behavior to human health. He founded Operant Conditioning Inc. and built the Operant habit app (https://operantconditioning.com/app/). Every page here is written from the primary literature and cites it. How pages are checked: https://operantconditioning.com/about/#editorial-standards

## Related

- [Build habits with operant conditioning](https://operantconditioning.com/habits/): The self-management protocol, step by step.
- [Dog training](https://operantconditioning.com/dog-training/): Markers, shaping, and the evidence on aversive methods.
- [50+ examples](https://operantconditioning.com/examples/): Every quadrant, every setting.
