4.4 Primary vs. Secondary Reinforcers & The Premack Principle
Key Takeaways
- Unconditioned reinforcers depend less on learned pairing; conditioned reinforcers acquire value through history.
- Reinforcer value is individual and changes with deprivation, satiation, health, competition, effort, and context.
- A generalized conditioned reinforcer has been related to multiple backup reinforcers, but a dog-training marker is not automatically generalized.
- The Premack principle uses access to a higher-probability activity contingent on a lower-probability behavior.
- Relief or distance can reinforce behavior through negative reinforcement and must be planned with safety and emotional welfare.
4.4 Reinforcers and the Premack Principle
Quick Answer: A reinforcer is defined by an increase in behavior, not by what trainers usually call rewarding. Food, play, touch, distance, sniffing, access, and social interaction can function differently across dogs and moments. Test value in context and update the plan.
Unconditioned and Conditioned Reinforcers
Unconditioned reinforcers derive value from biological and developmental processes without requiring the specific learning history that creates a conditioned reinforcer. Food, water when needed, warmth, or relief may serve, but none is effective in every moment.
Conditioned reinforcers acquire value through relations with other reinforcers. A click, praise word, token, or release cue may become reinforcing. If its backup history weakens or the context changes, its function can change. “Good dog” is not automatically reinforcement because it sounds positive to the person.
A generalized conditioned reinforcer is associated with more than one kind of backup reinforcer and is therefore less dependent on one motivating condition. Money is a common human example. A dog marker paired with food, play, and access might approach this function, but calling every versatile marker “universal” overstates the evidence. Its effect must still be demonstrated.
Value Changes
Food value can fall after eating or with nausea and rise with appetite; play depends on fatigue, environment, and relationship; touch can be appetitive in one context and aversive in another. Competing stimuli and response effort also matter.
Motivating operations alter the current effectiveness of a consequence and behavior that has produced it. Ethical planning does not create motivation by withholding necessary water, using medically unsafe food restriction, or depriving a dog of basic welfare. Coordinate dietary use with the owner and veterinarian when health is relevant, and count training food in daily intake.
Premack Principle
The Premack principle uses a behavior with higher current probability to reinforce a lower-probability behavior. A dog that wants to sniff may gain access after a few steps on a loose leash; a dog that wants to greet may first orient to the handler. The arrangement is “first this, then that,” but only if the second activity is actually valuable at that moment.
Avoid turning access into coercion. If the dog is too fearful to approach, requiring proximity before retreat can create unsafe pressure. Criteria should be attainable and the high-probability activity safe and lawful.
Functional Reinforcement and Distance
Behavior that makes a trigger move away or increases distance may be negatively reinforced. A bark that causes a visitor to retreat can strengthen; a calm orienting response followed by a planned U-turn can also gain value through relief. The trainer should not repeatedly provoke severe reactions to use distance as a reinforcer. Work within a range that preserves learning and safety.
Preference and Reinforcer Assessments
Offer safe choices, record selections and consumption or engagement, and then test whether contingent delivery increases the target behavior. Preference is not identical to reinforcing effect. Rotate consequences to reduce satiation and preserve variety, but avoid abrupt switches that confuse the contingency.
For multi-dog settings, control competition and guarding. Deliver separately when needed. Consider allergies, swallowing risk, toy conflict, handler ability, and facility rules.
Reinforcer Variety and Agency
Choice can help identify value and support participation. A dog may choose food on one trial and play or distance on another. Offer options without creating competition or overwhelming the learner. Consent-seeking matters especially for touch: approach, pause, and observe whether the dog reinitiates rather than assuming petting is rewarding. For environmental access, confirm that the destination is safe and permitted. Reinforcing with release to chase wildlife, greet an unwilling dog, or enter traffic is not acceptable merely because the dog values it.
When using relief, avoid setting up unnecessary distress to manufacture reinforcement. The environment should allow retreat before panic, not require intense exposure to earn distance.
For social reinforcers, watch whether the dog approaches again, remains loose, or moves away. The human’s affection is not evidence that touch reinforced the dog’s behavior.
Reinforcement Plans
Define what earns reinforcement, what will be delivered, where, how quickly, and what happens if the dog declines. Monitor rate, latency, errors, and welfare. If behavior does not increase, change the consequence, timing, effort, or environmental conditions rather than blaming motivation.
Reinforcement is a measured relationship. It is not praise, food, or freedom by definition, and its value cannot be inferred solely from a neurochemical story.
What makes an event a reinforcer?
Which is a Premack arrangement?
Is a clicker automatically a generalized conditioned reinforcer?