You're Not a Positive Reinforcement Trainer
By Nando Brown
The label gets thrown around constantly — but what does the science actually say about how you're using reinforcement?

Most trainers describe themselves as positive reinforcement trainers. It sounds ethical. It sounds scientific. It sounds safe.
It is also, in most cases, technically wrong.
That is not an insult. It is a problem of language and a misunderstanding of how learning theory actually works when applied properly.
Skinner wasn't designing a training philosophy
Skinner did not design a training philosophy. He created a framework for describing learning after it happens. His work allows us to classify behaviour. It does not give you permission slips before you train.
Reinforcement and punishment are not methods. They are outcomes.
A consequence is only reinforcement if the behaviour increases after it. A consequence is only punishment if the behaviour decreases after it. You do not decide this in advance. You measure it after the fact.
That is why saying "I am going to use positive reinforcement" is scientifically inaccurate. You can only say "that was positive reinforcement" once you have seen what happened to the behaviour.
What the quadrant actually says
The quadrant itself is simple when stripped of emotion.
Positive means something was added. Negative means something was removed. Reinforcement means behaviour increased. Punishment means behaviour decreased.
That is all it describes. There is no ethical judgement built into the model.
Here is where discomfort creeps in.
If a dog bites you and, as a consequence of the bite, you kick the dog, something has been added. If the frequency of biting increases afterwards, that consequence functioned as positive reinforcement. Something was added and the behaviour increased.
Most people object to this example because they believe reinforcement must involve something pleasant. That belief breaks the quadrant.
If reinforcement only counts when something pleasant is added, then adding something aversive that increases behaviour no longer fits anywhere in the model. You would need eight categories instead of four. Skinner did not design it that way. He deliberately removed emotion because it clouds classification.
The quadrant does not care whether the consequence was kind, harsh, mild, or extreme. It only cares about what happened to the behaviour.
Intent does not matter. Outcome does.
One quadrant per behaviour
A common error is the idea that multiple quadrants act at the same time on the same behaviour. They do not.
Only one quadrant applies per behaviour being measured. When trainers think they see overlap, it is usually because they have switched which behaviour they are measuring mid-explanation. For example, a trainer might apply negative reinforcement to reduce pulling and simultaneously reinforce a loose lead. Those are two separate behaviours, two separate consequences, two separate quadrant entries. Not one event in multiple quadrants.
That is a measurement problem, not a learning theory one.
The quadrant is a diagnostic tool, not an ethical framework
All of this matters because the quadrant describes what already happened. It does not tell you how good a trainer you are, and it does not excuse poor practice.
This is where ethics come back in.
Saying that outcomes define reinforcement does not mean the outcome is all that matters. It does not mean the journey is irrelevant. It does not mean anything goes as long as behaviour changes.
What makes a trainer skilled, modern, and ethical is trajectory.
Is the behaviour becoming clearer, more stable, and more flexible across contexts? Or is it escalating, fragmenting, or leaking into new problems?
Is the dog more confident in situations it once struggled with? Faster to recover? More willing to engage? Or is the dog compliant but flatter, slower, and more hesitant?
Look at change over time. Across sessions. Within sessions. In the dog's body language, not just the checklist of behaviours.
A dog that once avoided now approaches. A dog that once froze now offers behaviour. A dog that once exploded now pauses.
That is progress.
Ethical training is not about pretending punishment does not exist. It is about not needing coercive tools to get results that last.
What you actually are
Learning theory tells you what happened. Trajectory tells you whether what you are doing is worth continuing.
So no, you are probably not a positive reinforcement trainer in the way the label is usually used.
You are a trainer who applies consequences, measures outcomes, and judges success by where the dog is heading and how it gets there.
That distinction matters.
Both the Behaviour Bible and PuppyLab include a dedicated learning theory module built around exactly this. Not quadrant-quoting for social media, but practical science you can actually use in real cases to make better decisions and track whether your work is genuinely moving dogs in the right direction.