Abstract
Punishment, which MeSH classifies under reinforcement, is the operation of following a response with a consequence that reduces its future probability — the functional mirror of reinforcement within operant conditioning. A folk tradition holds that punishment does not work, that it suppresses behavior while present and teaches nothing lasting. The experimental record is more exact: whether punishment produces transient or durable suppression depends on parameters — intensity, immediacy, consistency, and an alternative response — mapped in the operant laboratory across the mid-twentieth century. This article traces punishment from Thorndike's law of effect through Estes's temporary suppression, the parametric analyses of Azrin and Church, Solomon's rehabilitation, and Skinner's case against its use, to the modern developmental science of corporal punishment. Three demonstrations let the reader manipulate the contingencies, the parameters of suppression, and the recovery of a punished response.
Keywords: punishment, operant conditioning, response suppression, corporal punishment, law of effect
The word carries two meanings that the psychology of punishment must keep apart. In everyday and legal use, punishment is a deserved penalty — retribution measured against an offense. In the analysis of behavior it is something narrower and purely functional: a consequence is a punisher if, and only if, it makes the behavior it follows less likely in the future (Azrin & Holz, 1963). The definition says nothing about pain, justice, or intent; a consequence that a person finds unpleasant but that does not change behavior is not, technically, a punisher at all. This functional definition is what let punishment be studied with the same precision as reward, and it is the reason the laboratory findings so often contradict the folk intuitions that share the word.
- Punishment is defined functionally as a consequence that reduces the future probability of the response it follows, the mirror of reinforcement rather than a synonym for pain or penalty.
- Positive punishment adds an aversive stimulus; negative punishment removes a valued one. Neither should be confused with negative reinforcement, which strengthens behavior.
- Estes showed that mild punishment suppresses responding only temporarily, seeding the durable belief that punishment is ineffective.
- Parametric work by Azrin, Church, and Solomon showed that intensity, immediacy, and consistency determine whether suppression is transient or complete and lasting.
- Skinner argued against punishment on practical grounds, and modern meta-analyses link corporal punishment of children to worse developmental outcomes.
What Punishment Is
In the operant framework, behavior is selected by its consequences, and punishment is one of the four basic contingencies that arrange them. Two axes define the set: whether a stimulus is presented or removed, and whether the future rate of the behavior rises or falls. Reinforcement raises the rate; punishment lowers it. Positive punishment presents an aversive stimulus after the response — a shock, a reprimand — while negative punishment removes a valued one, as when a privilege is withdrawn or a token is lost (Azrin & Holz, 1963). Both reduce responding; they differ only in whether the operative event is an onset or an offset.
The distinction most often confused is between punishment and negative reinforcement. Both involve an aversive stimulus, but they act in opposite directions. Negative reinforcement strengthens a response by removing or preventing something aversive — the behavior that turns off a loud alarm is reinforced. Punishment weakens a response by producing an aversive consequence. The shared vocabulary of aversion hides a categorical difference in outcome, and much applied error follows from collapsing the two (Skinner, 1953).
The four operant contingencies
Two choices define every consequence: whether a stimulus is presented or removed, and whether the future behavior increases or decreases. Set both and see which of the four cells you land in. The two decrease cells are the punishers.
Because the definition is functional, whether a given consequence counts as a punisher is an empirical question about a particular organism at a particular time, not a property of the stimulus. The same reprimand can punish one child's behavior and, by supplying attention, reinforce another's. MeSH files the descriptor under reinforcement, reflecting an indexing decision to group the consequence operations together rather than any claim that punishment is a species of reinforcement; functionally the two are opposites arranged along the same dimension.
Punishment and the Law of Effect
Thorndike's original law of effect was symmetric: responses followed by a satisfying state of affairs were stamped in, and responses followed by an annoying state of affairs were stamped out. His own later experiments led him to abandon the second half. Finding that saying “wrong” to a learner did not weaken a response the way “right” strengthened it, he concluded that reward and punishment are not mirror images and that punishment acts, if at all, only indirectly (Thorndike, 1927). This asymmetry — reward strengthens directly, punishment does not simply weaken — became the founding puzzle of the field.
Estes gave the puzzle its most influential experimental form. In a series of studies he trained rats to press a lever for food, then punished the press with a brief shock. Responding dropped sharply, but when he continued to observe, the rate recovered toward its unpunished level, and the total number of responses eventually emitted was little changed — as if the punishment had rescheduled the behavior in time rather than eliminating it (Estes, 1944). The natural reading, which hardened into textbook orthodoxy, was that punishment produces only a temporary emotional suppression and no lasting weakening of the response. That conclusion was correct for the mild, brief punishment Estes used, and wrong as a general law — a distinction the next two decades of parametric work would draw sharply.
The Parameters of Effective Punishment
The laboratory answer to “does punishment work?” is that it depends on how it is arranged, and the relevant variables can be stated precisely. Azrin's parametric studies, using pigeons whose key-pecking was maintained on a schedule of food reinforcement and simultaneously punished with electric shock, showed that response suppression is a graded function of shock intensity: weak shock barely dented the rate, moderate shock depressed it, and intense shock drove it to zero and held it there (Azrin, 1960). The same program established that punishment delivered on every response suppresses more than intermittent punishment, and that suppression is deepest when the punisher follows the response immediately (Azrin, Holz, & Hake, 1963).
What makes punishment effective
Suppression is not fixed by the punisher alone. Move intensity, immediacy, and consistency and watch the surviving response rate. A delayed or intermittent punisher acts as if it were weaker, which is why mild, late, or occasional punishment barely dents behavior. (Schematic, after Azrin, 1960; Church, 1963.)
Table 1
Parameters That Determine Whether Punishment Suppresses Responding
| Parameter | Effect on response suppression | Key evidence |
|---|---|---|
| Intensity | A graded function: weak punishers barely dent the rate, while sufficiently intense ones drive it to near zero and hold it there. | Azrin (1960) |
| Immediacy | Suppression is deepest when the punisher follows the response with little delay; a delayed punisher loses much of its effect. | Azrin, Holz, & Hake (1963) |
| Consistency | Punishing every response suppresses behavior more than punishing intermittently, which leaves unpunished occasions to sustain the response. | Azrin, Holz, & Hake (1963) |
| Alternative response | Suppression is far more durable when another route to the same reinforcer is available; with no alternative, motivation and suppression reach an equilibrium. | Church (1963) |
Church's review assembled these effects into a coherent picture and added the crucial role of an alternative. Punishment suppresses a response far more effectively, and more durably, when the organism has another route to the same reinforcer; punishing a behavior while leaving no alternative pits suppression against the ongoing motivation to respond, and the two reach an equilibrium rather than elimination (Church, 1963). Solomon drew the practical moral in a review whose purpose was explicitly corrective: the widespread conviction that punishment is ineffective, he argued, was a “legend” built on studies that used punishment too mild, too delayed, or too inconsistent to reveal its real effects. Arranged with sufficient intensity, immediacy, and consistency, punishment produces suppression as orderly and durable as any reinforcement effect (Solomon, 1964).
Figure 1
The Four Operant Contingencies
Skinner's Case Against Punishment
Skinner accepted that punishment could suppress behavior, yet argued against relying on it, and his objection was practical rather than definitional. Suppression, he held, is often temporary; the punished behavior is not unlearned but held down, and reappears when the threat of punishment is removed (Skinner, 1953). More importantly, punishment generates unwanted by-products: it conditions fear and avoidance to the whole situation, including the punishing agent, and it produces escape and counter-aggression rather than the desired alternative behavior. This coupling of a classically conditioned emotional response with the instrumental suppression it drives is the core of the two-process theory of punishment, associated with Solomon, which holds that a punisher works through both Pavlovian fear conditioning and the operant suppression that fear supports (Solomon, 1964). Because punishment specifies only what not to do, it leaves the reinforcement that maintains the unwanted behavior in place and supplies no route to a better response.
Skinner's recommended alternatives — extinction of the unwanted behavior combined with positive reinforcement of an incompatible one — followed directly from this analysis, and his position shaped decades of applied practice in education and clinical work toward reinforcement-based methods. The laboratory qualification matters, though: Azrin and Church had shown that Skinner's “temporary suppression” was itself a parameter-dependent result, characteristic of mild punishment and of arrangements offering no alternative. The disagreement was less about the facts of suppression than about which arrangements are typical, ethical, and wise to use with people.
Corporal Punishment and Child Development
The question that carried punishment out of the operant laboratory and into public health is whether parents should use physical punishment on children. Here the functional definition and the empirical outcomes diverge sharply from folk practice. Gershoff's meta-analytic review of the corporal-punishment literature found that while physical punishment was associated with immediate compliance, it was also associated with a consistent set of detrimental outcomes — increased child aggression, antisocial behavior, and impaired mental health — and with a weaker moral internalization than reasoning-based discipline (Gershoff, 2002).
Suppression and recovery: does the behavior come back?
The rate is at baseline, then punished, then the punisher is withdrawn. Toggle the strength of the punisher and watch what happens after it stops — the difference between Estes’s transient dip and Solomon’s durable suppression. (Schematic, after Estes, 1944; Solomon, 1964.)
A recurring objection was that such studies confounded ordinary spanking with physical abuse. Gershoff and Grogan-Kaylor addressed it directly in a later meta-analysis restricted to spanking as customarily defined, excluding harsher physical force, and still found spanking reliably associated with the same adverse outcomes and with no evidence of benefit (Gershoff & Grogan-Kaylor, 2016). The developmental evidence thus converges with the operant lesson from a different direction: the consequence that most efficiently secures momentary compliance is not the one that best builds durable, self-regulated behavior, and its collateral effects can outweigh its suppressive value.
Worked Example
The graded relation between punishment intensity and response suppression can be made concrete with a simple model. Let a behavior occur at a baseline rate R0 in the absence of punishment, and suppose that adding a punisher of intensity I reduces the rate multiplicatively, R = R0 e−kI, where k measures the organism's sensitivity to the punisher. Take a baseline of R0 = 60 responses per minute and k = 0.02 per volt, in the range Azrin's shock-intensity functions suggest (Azrin, 1960).
At I = 0 the rate is the full 60 per minute. A mild punisher of 25 V lowers it to 60 e−0.5 = 36.4 per minute, a 39% suppression; 50 V yields 60 e−1 = 22.1 per minute, 63% suppression; 100 V yields 60 e−2 = 8.1 per minute, 87% suppression; and 150 V drives it to 60 e−3 = 3.0 per minute, 95% suppression. The response rate is halved at I = ln 2 / k = 34.7 V and falls by 90% at I = ln 10 / k = 115.1 V. The curve captures Azrin's central finding: suppression is not all-or-none but a smooth, accelerating function of intensity, so that weak punishment barely disturbs the behavior while sufficiently intense punishment approaches complete suppression.
What the one-parameter model deliberately omits is the dimension Estes measured — time. It describes the rate while the punisher is in force and contains no term for recovery once it is withdrawn, which for mild punishers is substantial and for intense ones minimal (Estes, 1944; Solomon, 1964). Reading suppression and recovery as two faces of the same contingency, rather than as rival verdicts on whether punishment “works,” is precisely the correction the parametric tradition supplied.
Current Directions
Contemporary research on punishment has shifted decisively toward the developmental and neural consequences of its most common human form. Cuartas and colleagues used functional neuroimaging to show that children who had experienced corporal punishment, even in the absence of more severe maltreatment, displayed a heightened neural response to threatening facial cues in prefrontal regions — a pattern resembling that seen after more severe adversity, and evidence that ordinary physical punishment leaves a measurable neurodevelopmental trace (Cuartas et al., 2021).
The second line is a synthesis of the longitudinal evidence to inform policy. A narrative review of prospective studies, which follow children over time and so better support causal inference than cross-sectional work, concluded that physical punishment predicts increases in behavioral problems and detects no benefit, and that the associations are consistent across countries and cultural contexts (Heilmann et al., 2021). This convergence of behavioral, developmental, and neural evidence has underwritten a global movement toward legal prohibition of corporal punishment, moving the century-old laboratory concept of the punisher into the center of child-health policy.
Key Researchers
Edward L. Thorndike (1874-1949). Teachers College, Columbia University; his law of effect first proposed that consequences select behavior, and his later revision — that reward strengthens more reliably than punishment weakens — framed the asymmetry the field has studied ever since. Wikipedia
B. F. Skinner (1904-1990). Harvard University; he defined punishment operationally within the operant framework and argued from its temporary suppression and aversive by-products that reinforcement of alternative behavior is the better tool. Wikipedia
William K. Estes (1919-2011). His 1944 monograph gave punishment its first rigorous experimental analysis, showing that mild punishment suppresses responding only transiently — the result that seeded the durable belief that punishment is ineffective. Wikipedia
Nathan H. Azrin (1930-2013). His parametric studies established that response suppression is a graded function of shock intensity, schedule, and immediacy, converting punishment from a vague deterrent into a quantitative operant relation. Wikipedia
Richard L. Solomon (1918-1995). University of Pennsylvania; his 1964 review dismantled the “legend” that punishment is ineffective, showing that intensity, immediacy, and consistency determine whether suppression is transient or durable. Wikidata
Elizabeth T. Gershoff. University of Texas at Austin; her meta-analyses moved punishment from the operant laboratory into developmental science, showing that corporal punishment of children predicts worse behavioral and mental-health outcomes. ORCID
Andrew Grogan-Kaylor. University of Michigan; with Gershoff he separated ordinary spanking from harsher physical force in a meta-analysis and still found reliable associations with detrimental child outcomes and no benefit. ORCID
Jorge Cuartas. New York University; his neuroimaging work links corporal punishment to a heightened neural response to threat in children, extending the behavioral evidence into developmental neuroscience. ORCID
Commonly Confused With
- Negative Reinforcement
- Both arrangements involve an aversive stimulus, which is why they are so often swapped, but they move behavior in opposite directions. Negative reinforcement removes an aversive stimulus and thereby strengthens the response that removed it; punishment delivers an aversive consequence and thereby weakens the response that produced it. The reliable test is not whether the event feels unpleasant but what happens to the behavior afterward: if the response becomes more likely the contingency was reinforcement, and if it becomes less likely the contingency was punishment.
Discussion
Punishment holds a peculiar place in the psychology of learning: it is the exact functional complement of reinforcement, yet for most of a century it was treated as the weaker, less orderly, and less trustworthy of the pair. The reason was largely historical accident. Thorndike's abandonment of the negative law of effect, and Estes's demonstration that mild punishment only briefly suppresses responding, established early that punishment was different in kind — a mere emotional damping rather than a genuine selection of behavior (Thorndike, 1927; Estes, 1944). The parametric tradition corrected the record without overturning the observations: punishment does produce transient suppression when it is mild, delayed, or inconsistent, and produces deep, durable suppression when it is intense, immediate, and certain, especially where an alternative response is available (Azrin, 1960; Church, 1963; Solomon, 1964).
That the concept works as an orderly operant relation does not settle whether it should be used, and here the modern evidence and Skinner's early caution point the same way. Punishment specifies only what to stop, generates aversive by-products, and — in its commonest human application, physical punishment of children — is associated with worse developmental outcomes and a measurable neural signature, with no detected benefit over reasoning-based discipline (Skinner, 1953; Gershoff & Grogan-Kaylor, 2016; Cuartas et al., 2021). The trajectory of the field — from establishing that punishment can work, to specifying when, to asking whether it ought to be used — tracks the maturation of a behavioral concept into a matter of public health and policy (Heilmann et al., 2021).
Glossary
- Aversive stimulus.
- A stimulus whose presentation tends to suppress the behavior it follows or whose removal tends to reinforce behavior.
- Corporal punishment.
- The use of physical force, such as spanking, intended to cause discomfort in order to correct or control a child's behavior.
- Immediacy.
- The delay between a response and its punisher; suppression is deepest when the punisher follows the response with little delay.
- Law of effect.
- Thorndike's principle that consequences select behavior; its negative half, that annoyers stamp out responses, he later revised.
- Negative punishment.
- Reducing the future probability of a response by removing a valued stimulus after it, as in loss of a privilege or a token.
- Negative reinforcement.
- Strengthening a response by removing or preventing an aversive stimulus; opposite in effect to punishment despite the shared aversion.
- Operant conditioning.
- Learning in which the consequences of a response alter its future probability; the framework within which punishment is defined.
- Positive punishment.
- Reducing the future probability of a response by presenting an aversive stimulus after it, such as a shock or a reprimand.
- Punisher.
- A consequence that, by definition, reduces the future probability of the response it follows; a functional, not a physical, category.
- Punishment.
- The operation of following a response with a consequence that reduces its future probability; the subject of this article.
- Reinforcement.
- The strengthening of behavior by its consequences; the functional opposite of punishment and the MeSH category under which it is filed.
- Response suppression.
- The reduction in the rate of a response produced by punishment, which may be transient or durable depending on the parameters.
- Suppression ratio.
- A measure of punishment's effect expressing the punished response rate as a proportion of the unpunished baseline.
- Two-process theory.
- The account, associated with Solomon, that punishment acts through both classically conditioned fear and the instrumental suppression of responding.
Frequently Asked Questions
What is punishment in psychology?
In the analysis of behavior, punishment is the operation of following a response with a consequence that reduces the future probability of that response; it is defined by its effect on behavior, not by whether the consequence is painful (Azrin & Holz, 1963).
What is the difference between positive and negative punishment?
Positive punishment presents an aversive stimulus after the response, such as a reprimand; negative punishment removes a valued stimulus, such as a privilege. Both reduce the behavior and differ only in whether something is added or taken away (Skinner, 1953).
How is punishment different from negative reinforcement?
They act in opposite directions. Negative reinforcement strengthens a response by removing something aversive, whereas punishment weakens a response by producing an aversive consequence (Skinner, 1953).
Does punishment actually work?
It depends on how it is arranged. Mild, delayed, or inconsistent punishment produces only temporary suppression, but intense, immediate, and consistent punishment can produce deep and durable suppression, especially when an alternative response is available (Solomon, 1964; Church, 1963).
Why did psychologists once believe punishment was ineffective?
Estes showed that mild punishment suppressed responding only temporarily, with the total number of responses eventually recovering; this result, correct for weak punishment, was overgeneralized into a belief that punishment never produces lasting change (Estes, 1944).
Why did Skinner oppose the use of punishment?
Skinner argued that punishment often suppresses behavior only temporarily, conditions fear and avoidance of the punishing agent, and specifies only what not to do, so reinforcing an incompatible behavior is the more effective strategy (Skinner, 1953).
Is spanking harmful to children?
Meta-analyses find that spanking, even when separated from harsher physical force, is reliably associated with increased aggression and other detrimental outcomes and with no detected benefit over reasoning-based discipline (Gershoff & Grogan-Kaylor, 2016).
Does corporal punishment affect the developing brain?
Neuroimaging evidence indicates that children who experienced corporal punishment, even without more severe maltreatment, show a heightened neural response to threat cues, a pattern resembling that seen after more severe adversity (Cuartas et al., 2021).
References
Azrin, N. H. (1960). Effects of punishment intensity during variable-interval reinforcement. Journal of the Experimental Analysis of Behavior, 3(2), 123-142. https://doi.org/10.1901/jeab.1960.3-123
Azrin, N. H., Holz, W. C., & Hake, D. F. (1963). Fixed-ratio punishment. Journal of the Experimental Analysis of Behavior, 6(2), 141-148. https://doi.org/10.1901/jeab.1963.6-141
Church, R. M. (1963). The varied effects of punishment on behavior. Psychological Review, 70(5), 369-402. https://doi.org/10.1037/h0046499
Cuartas, J., Weissman, D. G., Sheridan, M. A., Lengua, L., & McLaughlin, K. A. (2021). Corporal punishment and elevated neural response to threat in children. Child Development, 92(3), 821-832. https://doi.org/10.1111/cdev.13565
Estes, W. K. (1944). An experimental study of punishment. Psychological Monographs, 57(3), i-40. https://doi.org/10.1037/h0093550
Gershoff, E. T. (2002). Corporal punishment by parents and associated child behaviors and experiences: A meta-analytic and theoretical review. Psychological Bulletin, 128(4), 539-579. https://doi.org/10.1037/0033-2909.128.4.539
Gershoff, E. T., & Grogan-Kaylor, A. (2016). Spanking and child outcomes: Old controversies and new meta-analyses. Journal of Family Psychology, 30(4), 453-469. https://doi.org/10.1037/fam0000191
Heilmann, A., Mehay, A., Watt, R. G., Kelly, Y., Durrant, J. E., van Turnhout, J., & Gershoff, E. T. (2021). Physical punishment and child outcomes: A narrative review of prospective studies. The Lancet, 398(10297), 355-364. https://doi.org/10.1016/s0140-6736(21)00582-1
Skinner, B. F. (1953). Science and human behavior. Macmillan. ISBN 9780029290408.
Solomon, R. L. (1964). Punishment. American Psychologist, 19(4), 239-253. https://doi.org/10.1037/h0042493
Thorndike, E. L. (1927). The law of effect. The American Journal of Psychology, 39(1/4), 212-222. https://doi.org/10.2307/1415413