4.4 Primary vs. Secondary Reinforcers & The Premack Principle

Key Takeaways

  • Unconditioned reinforcers depend less on learned pairing; conditioned reinforcers acquire value through history.
  • Reinforcer value is individual and changes with deprivation, satiation, health, competition, effort, and context.
  • A generalized conditioned reinforcer has been related to multiple backup reinforcers, but a dog-training marker is not automatically generalized.
  • The Premack principle uses access to a higher-probability activity contingent on a lower-probability behavior.
  • Relief or distance can reinforce behavior through negative reinforcement and must be planned with safety and emotional welfare.
Last updated: September 2026

4.4 Reinforcers and the Premack Principle

Quick Answer: A reinforcer is defined by an increase in behavior, not by what trainers usually call rewarding. Food, play, touch, distance, sniffing, access, and social interaction can function differently across dogs and moments. Test value in context and update the plan.

Unconditioned and Conditioned Reinforcers

Unconditioned reinforcers derive value from biological and developmental processes without requiring the specific learning history that creates a conditioned reinforcer. Food, water when needed, warmth, or relief may serve, but none is effective in every moment.

Conditioned reinforcers acquire value through relations with other reinforcers. A click, praise word, token, or release cue may become reinforcing. If its backup history weakens or the context changes, its function can change. “Good dog” is not automatically reinforcement because it sounds positive to the person.

A generalized conditioned reinforcer is associated with more than one kind of backup reinforcer and is therefore less dependent on one motivating condition. Money is a common human example. A dog marker paired with food, play, and access might approach this function, but calling every versatile marker “universal” overstates the evidence. Its effect must still be demonstrated.

Value Changes

Food value can fall after eating or with nausea and rise with appetite; play depends on fatigue, environment, and relationship; touch can be appetitive in one context and aversive in another. Competing stimuli and response effort also matter.

Motivating operations alter the current effectiveness of a consequence and behavior that has produced it. Ethical planning does not create motivation by withholding necessary water, using medically unsafe food restriction, or depriving a dog of basic welfare. Coordinate dietary use with the owner and veterinarian when health is relevant, and count training food in daily intake.

Premack Principle

The Premack principle uses a behavior with higher current probability to reinforce a lower-probability behavior. A dog that wants to sniff may gain access after a few steps on a loose leash; a dog that wants to greet may first orient to the handler. The arrangement is “first this, then that,” but only if the second activity is actually valuable at that moment.

Avoid turning access into coercion. If the dog is too fearful to approach, requiring proximity before retreat can create unsafe pressure. Criteria should be attainable and the high-probability activity safe and lawful.

Functional Reinforcement and Distance

Behavior that makes a trigger move away or increases distance may be negatively reinforced. A bark that causes a visitor to retreat can strengthen; a calm orienting response followed by a planned U-turn can also gain value through relief. The trainer should not repeatedly provoke severe reactions to use distance as a reinforcer. Work within a range that preserves learning and safety.

Preference and Reinforcer Assessments

Offer safe choices, record selections and consumption or engagement, and then test whether contingent delivery increases the target behavior. Preference is not identical to reinforcing effect. Rotate consequences to reduce satiation and preserve variety, but avoid abrupt switches that confuse the contingency.

For multi-dog settings, control competition and guarding. Deliver separately when needed. Consider allergies, swallowing risk, toy conflict, handler ability, and facility rules.

Reinforcer Variety and Agency

Choice can help identify value and support participation. A dog may choose food on one trial and play or distance on another. Offer options without creating competition or overwhelming the learner. Consent-seeking matters especially for touch: approach, pause, and observe whether the dog reinitiates rather than assuming petting is rewarding. For environmental access, confirm that the destination is safe and permitted. Reinforcing with release to chase wildlife, greet an unwilling dog, or enter traffic is not acceptable merely because the dog values it.

When using relief, avoid setting up unnecessary distress to manufacture reinforcement. The environment should allow retreat before panic, not require intense exposure to earn distance.

For social reinforcers, watch whether the dog approaches again, remains loose, or moves away. The human’s affection is not evidence that touch reinforced the dog’s behavior.

Reinforcement Plans

Define what earns reinforcement, what will be delivered, where, how quickly, and what happens if the dog declines. Monitor rate, latency, errors, and welfare. If behavior does not increase, change the consequence, timing, effort, or environmental conditions rather than blaming motivation.

Reinforcement is a measured relationship. It is not praise, food, or freedom by definition, and its value cannot be inferred solely from a neurochemical story.

Loading diagram...
Reinforcer Test
Test Your Knowledge

What makes an event a reinforcer?

A
B
C
D
Test Your Knowledge

Which is a Premack arrangement?

A
B
C
D
Test Your Knowledge

Is a clicker automatically a generalized conditioned reinforcer?

A
B
C
D