How Reinforcement Shapes Your Dog’s Behavior (And Why It Matters For Dog Training)
Susan GarrettDogs That · 영상 · 3분
강화가 반려견의 행동을 형성하는 방식 (그리고 반려견 훈련에 중요한 이유)
How Reinforcement Shapes Your Dog’s Behavior (And Why It Matters For Dog Training)
0:00
밥 베일리는 훌륭한 훈련사와 좋은 훈련사의 차이는 '충분히 좋다'라는 말에 있다고 말합니다. 그러니 반려견과 맺고 있는 관계 어딘가에서, 만약 여러분이 훌륭한 결과를 얻지 못하고 있다면, 무의식적으로든 아니든 '충분히 좋다'라는 말이 훈련 과정 어딘가에 스며들었을 가능성이 있습니다. 왜냐하면 밥이 말하는 또 다른 훌륭한 명언이 있는데, 정말 맞는 말입니다. 행동은 결과의 함수입니다. 여러분 눈앞의 행동은 어딘가 다른 곳에서 강화되어 온 것입니다. 여러분이 보고 있는 것은 계속 강화되고 있는 중입니다. 그러니 여러분은 반려견이 하고 있는 행동이 여러분이 하는 것과 반대되는 방식으로 강화를 얻고 있지는 않은지 파악하기만 하면 됩니다. 이 팟캐스트에서 저에게 주어졌던 예를 하나 들어보겠습니다. 어떤 신사분이 반려견에게 힐(heel) 큐를 주면, 반려견이 한두 걸음 정도만 힐 위치에 있다가 냄새를 맡으러 가버린다고 했습니다. 저의 경우, 반려견에게 무언가를 요구해서 힐 위치로 오게 한다면, 제 큐는 제 옆에서 편하게 걷는 것이 될 겁니다. 그리고 제 반려견은 멋대로 돌아다니지 않을 텐데, 그 이유는 제가 요구한 것과 반대되는 행동을 함으로써 강화를 얻는 꼴이 되기 때문입니다. 그래서 냄새 맡는 것이 제 반려견에게 얼마나 강화 효과가 큰지 안다면, 저는 반려견을 제 곁에, 그게 한 걸음이든, 다섯 걸음이든, 10분이든 일단 걷게 할 겁니다. 그러고 나서 저는 이렇게 말하겠죠. 냄새 맡는 것이 매우 큰 보상이라는 걸 알기에 릴리즈(release) 단어를 써서 이렇게 말할 겁니다. 가서 냄새 맡아, 혹은 가서 뛰어놀아. 만약 다른 개가 있다면 그 개를 만나러 가라거나 사람에게 가보라고 할 수도 있겠죠. 저는 의도적으로 무엇이 제 반려견을 강화하는지 인지하고 있을 겁니다. 그리고 강화 구역(reinforcement zone)에 머무른 보상으로 그 강화물에 접근할 수 있게 해줄 겁니다. 하지만 제 반려견이 원할 때 그냥 떠나버린다면, 그건 저나 어떤 행동과도 무관하게 강화를 챙겨가는 셈이 됩니다. 매칭 법칙(matching law)은 여러분 곁에 머물게 하려면 최소한 그와 같거나 더 큰 가치의 보상이 필요하다고 말해줍니다. 그리고 외부 환경으로 나가서 냄새를 맡는 것이 실제로 더 높은 가치 강도를 가진다면, 당신 곁에 머무는 것에 대해 훨씬 더 많은 보상을 해주어야 합니다. 물론, 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게 이렇게
Bob Bailey says, the difference between a great trainer and a good trainer are the words good enough. So, somewhere along the relationship you have with the dog, if you aren't getting those great results, potentially subconsciously or otherwise the words good enough have crept into the protocol somewhere. Because another great line Bob says is, and it's true, behavior is a function of consequence. What you're seeing has been reinforced somewhere else. What you're seeing is being reinforced. And you just need to figure out is what the dog doing getting reinforcement in a way that's contrary to what you're doing. And I'll give you an example that was given to me on this podcast that when the gentleman says, gives his dog a cue to heal, the dog can stay in heel position for like one or two steps and then go off and sniff. Now for me, if I asked my dog to do something, come into heel position and my cue would be with me for an informal walk beside me. And my dog wouldn't wander off because that would be gaining reinforcement for doing something in opposition to what I asked you to do. And so, knowing how reinforcing sniffing would be for my dog, I might have them walk in with me position for whatever, one step, five steps, 10 minutes. And then I would say, give them the release word that they could, if I knew sniffing was very, very high value, I would say, go sniff or go for a run or go play. If there was another dog there, maybe go see if there was a person. I would intentionally know, I would be aware of what it is that is reinforcing my dog. And I would give them access to that reinforcement as a release for hanging out in reinforcement zone. But if my dog just leaves when they want, they're taking the reinforcement that is not contingent on me or any behavior, matching law tells us you need at least the same value or greater for hanging out beside you. And if getting out in the environment and sniffing actually has a higher intensity of value, then you need a lot more reinforcements for hanging out beside you. Of course, böylece böylece böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle böyle