Conditioned Reinforcer and Marker Signal¶
Definition¶
A conditioned reinforcer is an initially meaningless signal—a sound (whistle, click), a word ("Good!"), a light, a gesture—that is deliberately paired with a primary reinforcer (food, petting, praise) until the organism learns to recognize it as a predictor of the real reward. The signal then becomes reinforcing in itself. When used specifically to mark the exact moment a desired behavior occurred (before delivering the actual reinforcer), the conditioned reinforcer functions as a "marker signal" that provides immediate feedback about what earned the reward. This bridging function is crucial: it can span the time between a behavior and its physical reward, making the connection instant and unmistakable.
In the Book¶
Pryor explains that dolphin trainers traditionally used a police whistle as a conditioned reinforcer because it's easily heard even underwater and leaves the trainer's hands free. She describes training a dolphin to recognize the whistle as a signal for "you just earned a fish, go get it," which allows the trainer to mark a behavior occurring in midair (a leap) and have the dolphin understand the connection before swimming to collect the reward. The formal term Keller Breland gave was "bridging stimulus" because the signal bridges the time gap.
In the 1990s, dog owners began using a plastic clicker as the signal instead of a whistle, and the term "clicker training" emerged. Pryor emphasizes that the click does more than simply indicate reinforcement: it is an "event marker" that identifies exactly which behavior earned the reward. When a dog is learning to sit, a click at the precise moment its rear touches the ground tells the dog "that exact position is what I want," far more precisely than training without a marker. The marker also shifts something psychological in the learner: "It puts control in the hands, paws, fins, whatever, of the learner. The subject no longer just repeats the behavior; the subject exhibits intention." Trainers call this moment when understanding dawns "the light bulb goes on."
Pryor notes that conditioned reinforcers can become enormously powerful—marine mammals work past the point of satiation on conditioned reinforcers (the whistle alone), and horses and dogs work for hours with few primary reinforcers. Money in human society is a generalized conditioned reinforcer because it's paired with so many different primary reinforcers (food, shelter, entertainment, safety). The key limitation: once established, conditioned reinforcers lose power if used meaninglessly ("Good job!" when the learner hasn't done anything), so they must be reserved for actual reinforcement moments.
The book also describes "conditioned aversive signals" (like a cat accidentally learning "No!" as a warning that something bad is about to happen, due to an unfortunate timing with a falling tray). A modern addition is the "no-reward marker"—a neutral word like "Wrong" that tells the learner "that won't work," which can help sophisticated learners try alternative approaches.
Why It Matters¶
The conditioned reinforcer principle reveals how to provide real-time feedback on complex behaviors where the primary reinforcer cannot be delivered instantly. It explains why immediate feedback is so much more powerful than delayed feedback: a teacher giving a "Good!" immediately after a correct answer strengthens the behavior far more than marking the answer correct on a returned test a week later. It also provides a universal communication tool—Pryor shows how a marker signal can communicate across species barriers and even across language gaps. For anyone trying to teach rapid learning (whether training animals, coaching athletes, instructing children, or managing employees), establishing a clear, reliable marker signal ("Well done," a thumbs-up, public recognition) and using it precisely in the moment of desired behavior can dramatically accelerate learning and clarify what actually counts as success.