Should We All Stop Using Non-Reward Markers In Dog Training? #251
Susan GarrettShaped by Dog with Susan Garrett · 팟캐스트 · 17분
반려견 훈련에서 비보상 마커를 제거하고 환경 설계를 최적화하여 명확성을 높인다. 훈련의 성공 여부는 반려견의 지능이나 고집이 아니라 훈련사의 계획과 환경 조성 능력에 달려 있다. 반려견이 실수하는 것은 선행 사건 설정이 미흡함을 의미하므로, 훈련사는 비보상 마커에 의존하기보다 반려견이 올바른 선택을 할 수밖에 없는 환경을 구축해야 한다. 야망보다 반려견을 향한 연민을 우선할 때 훈련사는 더 평온해지고, 반려견과의 관계는 더욱 돈독해진다. 선행 조건과 환경을 정밀하게 조정하면 반려견은 스스로 행동을 선택하고 성공을 경험하며, 결과적으로 비보상 마커를 사용할 필요성이 사라진다. 훈련의 모든 과정은 ABC 모델, 즉 선행 사건, 행동, 결과의 체계 속에서 이해되어야 하며, 훈련사는 실패 시 반려견을 탓하는 대신 실험 설계를 수정하여 재접근한다.
우리 모두 개 훈련에서 비보상 마커 사용을 중단해야 할까요? #251
Should We All Stop Using Non-Reward Markers In Dog Training? #251
0:00
오늘 팟캐스트에서 여러분과 나누고 싶은 이야기가 정말 많습니다. 그리고 저는 이것이 쌓여온 결과라고 생각합니다. 제가 전문 반려견 훈련사로 일해온 30년 동안 훈련 방식이 어떻게 발전해 왔는지를 지켜본 결과 말이죠. 글쎄요, 아마 30년이 넘었을 겁니다. 이제 33년이 된 것 같네요. 신체적 교정이나 언어적 위압을 사용하지 않고 반려견을 훈련하기로 선택한 사람들의 공동체로서 우리가 어디로 향하고 있는지에 대한 이야기입니다. 만약 여러분이 그런 분이라면, 반려견을 훈련할 때 무엇이 더 큰지 확인해 보세요. 반려견에 대해 가지고 있는 연민의 수준인가요, 아니면 반려견이 무엇을 할 수 있는지에 대한 여러분의 야망인가요? 안녕하세요, 저는 수잔 가렛입니다. 'Shaped by Dog'에 오신 것을 환영합니다. 지난주에 제가 관찰한 많은 것들이 있습니다. 온라인 수강생들의 과제를 읽으면서, 그리고 몇몇 강사들에게 받은 개인적인 메시지들을 통해 느낀 점들이죠. 또한 일반적인 공간에서 사람들이 반려견을 어떻게 훈련하는지 지켜보면서 느낀 것들입니다. 하루를 마치며 제가 깨달은 점은, 제가 교류하거나 소셜 미디어에서 게시물을 보는 사람들의 99.9%가 저만큼이나 자신의 반려견을 진심으로 사랑한다고 믿는다는 것입니다. 그리고 그들이 반려견을 위해 내리는 결정은 현재 그들이 우선순위로 두는 것을 바탕으로 한 최선의 결정이라는 점입니다. 하지만 저는 여러분께 그 우선순위에 대해 한번 의문을 가져보라고 제안하고 싶습니다. 예를 들어, 새 강아지를 집에 데려왔을 때를 생각해 보세요. 이 새로운 강아지와 함께하는 첫날, 강아지가 여러분을 깨물 때, 실망하거나 충격을 받거나 좌절하거나 당황하시나요? 어떤 행동을 취하시나요? 단호하게 '안 돼'라고 하시나요? 아니면 뒷덜미를 잡으시나요? 여러분은 무엇을 하시나요? 왜냐하면 여러분의 행동은 물론 여러분의 과거 경험, 멘토가 가르쳐준 것, 주변 사람들로부터 영향을 많이 받지만, 동시에 그 강아지에 대해 여러분이 가진 야망에도 근거하기 때문입니다. 저는 이 강아지가 자라서 어떤 모습이 되기를 바라시나요? 아이들과 잘 지내고, 어떤 환경에서도 믿을 수 있으며, 절대 사람에게 입을 대지 않을 강아지 말이에요. 그것이 제 성견에게 바라는 목표입니다. 그래서 배가 너무 고프거나, 새로운 환경에 지나치게 자극을 받았을지도 모를 강아지에 대해 우리가 느끼는 동정심은, 새로운 환경에 과도하게 자극받았을지도 모르고, 형제자매와 엄마로부터 새로운 집으로 옮겨온 큰 변화로 인해 너무 피곤해졌을지도 모를 강아지에게 우리가 갖는 동정심의 수준은, 우리가 가진 목표 수준보다 결코 높지 않습니다. 팟캐스트 64회에서 제가 비보상 마커(non-reward marker)의 사용에 대해 언급하며 그 이야기를 했습니다. 사람들이 강아지가 잘못된 선택을 하지 못하게 멈추게 하려고 무언가를 사용하는 방식에 대해서요. 때때로 그 비보상 마커는 '아, 안 돼, 야, 멈춰'와 같은 처벌의 의미를 그대로 담고 있기도 합니다.
There's so much I want to share with you in today's podcast. And I believe it's an accumulation of looking at how training has progressed in the 30 years that I've been a professional dog trainer. Well, probably more than 30 years. I think it's 33 now. Where we are headed as a community of people who choose to train without the use of physical corrections or verbal intimidation in our dogs. And if that is you, you can check what's greater when you're training your dog. Your level of compassion you have for your dog or your level of ambition you have for what your dog can do. Hi, I'm Susan Garrett. Welcome to Shaped by Dog. And there's so many things that I've observed over the past week, reading some of our online students' work, some private messages that I've received from some of our instructors, as well as just being out there in the general space, looking at what people or how people are training their dog. And I recognize at the end of the day, 99.9% of the people that I interact with or that I see post things on social media, I assume absolutely love their dogs the way that I love my dogs. And the decisions they make for their dogs are the best decisions that they can based on what they have as their priorities now. But I want to ask you to maybe question the priorities. Like take for example, if you get a brand new puppy home, like first day with this new puppy and the new puppy nips at you, are you like disappointed, shocked, frustrated, alarmed? What action do you take? Do you give it a firm no? Do you give it a scruff? What is it that you do? Because your actions are based, yes, a lot on your history, what your mentors have taught you, the people who surround you, but they're also based on your ambition you have for that puppy. I want this puppy to grow up to be a dog who's good with children, who I can trust in any environment that would never put their mouth on people. That's the ambition I have for my adult dog. And so, the compassion we have for this puppy who might be over hungry, who might be overstimulated by their new environment, who might be overtired from the big day of transferring from their littermates and their mom to your new home, the level of compassion we have is not higher than the level of ambition that we have. In podcast episode number 64, I spoke about that when I referred to the use of non-reward markers and how people will use something to freeze a dog from making a poor choice. And sometimes that non-reward marker would right out and out be a punisher like, ah, no, hey, stop.
2:52
이 모든 것들은 강아지를 멈추게 하려는 의도입니다. 더 이상 앞으로 나아가지 말라는 거죠. 그것들은 '웁스(oops)'나 '다시 시도해 봐(try again)' 같은 비보상 마커일 수도 있습니다. 보통 이렇게 겉보기에 더 부드러운 비보상 마커들은 노래하듯 가볍게 말합니다. 강아지가 무언가를 시도하고 있는데 여러분이 보상을 주지 않을 때, 좋은 어조로 이렇게 알려주고 싶어 하죠. '웁스. 이런, 무슨 일이야? 다시 해봐.' 지난주에 이 주제에 대해 생각하게 만드는 글을 하나 읽었습니다. 그들의 학생은 강아지가 잘못된 선택을 하면 '루저(loser)'라는 단어를 사용하더군요. 여러분은 어떠실지 모르겠지만, 그 말을 하는 것만으로도 제 마음이 아팠습니다. 최선을 다해 강아지를 훈련하는 과정에서 강아지를 '루저'라고 부르는 사람들을 비판하고 싶지는 않지만, 이 팟캐스트에서 그런 식의 접근법으로 흐려지고 싶지는 않습니다. 저는 일반적인 비보상 마커의 사용에 대해서만 이야기하고 싶습니다. 그래서 강아지를 훈련할 때 사람들은 이렇게 말합니다. '수잔, 비보상 마커는 정말 유용해요.' 왜냐하면 강아지가 똑같은 선택을 반복해서 하는 것을 막아주기 때문이죠. 그래서 사실 저는 비보상 마커를 사용하는 것이 매우 친절한 행동이라고 생각할 수 있지만, 저는 사실 훈련할 때 비보상 마커를 절대 사용하지 않으려고 노력합니다. 사람들은 흔히 자기 개가 고집이 세다거나 배우는 속도가 느리다고 말하곤 합니다. 하지만 저는 그 말을 도저히 믿을 수가 없습니다. 왜냐하면 제가 지금까지 훈련해 온 모든 개들 중에서, 고집이 세거나 배우는 속도가 느린 개는 단 한 마리도 본 적이 없기 때문입니다. 제가 관찰하는 것은 오직 제가 개를 놓아둔 환경에서 개가 보이는 행동, 그리고 제가 그 개에게 제공한 교육뿐입니다. 그래서 팟캐스트 64회에서 비보상 마커에 대해 이야기할 때 제가 이렇게 말했습니다. 훈련할 때는 개가 보상으로 여겨 열심히 하려고 하는 강화물을 준비하는 것이 중요하며, 환경이 개의 선택에 영향을 주지 않도록 잘 준비하고, 정말 훌륭한 훈련 계획을 세우는 것이 중요하다고 말이죠. 그리고 그 훌륭한 훈련 계획은 개가 선택할 수 있는 반응의 가짓수를 최소화하거나 최소화해야 합니다. 왜냐하면 우리가 하는 모든 학습은 ABC 모델, 즉 선행사건-행동-결과 모델을 통해 이루어지기 때문입니다. 그렇죠? 선행사건, 행동, 결과요. 앞서 말씀드렸던 것처럼 우리는 문제 행동을 교정하려고 노력하지만, 사실 ABC는 개가 무언가를 배우는 방식이기도 합니다. 따라서 개가 당신에게 어떤 반응을 보였을 때, 그 행동을 반복하지 못하게 막아야겠다는 생각이 든다면, 그것은 당신의 계획에 결함이 있거나, 선행사건 배열이 개가 성공하도록 설정되지 않았기 때문일 수 있습니다. 예를 들어볼게요. 만약 제가 개에게 머핀 트레이에서 공을 꺼내게 하고 싶다고 가정해 봅시다. 당신은 왜 그런 걸 시키냐고 할지도 모르죠. 하지만 저는 개가 코로 공을 굴리는 것을 배우게 하고 싶을 수도 있습니다.
That all of those things are meant to get the dog to freeze. Don't take any forward motion. They also could be a non-reward marker like oops or try again. And usually these seemingly more pleasant non-reward markers are sung, like a little sing song. Dog is trying to do something and you're not going to reward them and you would like them to know with a nice little sing song. Oops. Oh no, what happened? Try again. I read one last week that kind of got me thinking about this thread. Their student uses the word loser when their dog makes a wrong choice. Now, I don't know about you, but me just saying it, that kind of hurt my heart. And I try not to be judgmental of the person who's doing the best they can calling their dog a loser in the middle of training, but I don't want to get wrapped up in that with this podcast. I want to talk about just the general use of non-reward markers. So, if you're training your dog, people will say, Susan, a non-reward marker is so valuable because it stops my dog from making and repeating the same choice over and over again. So, I'm actually being super kind by using the non-reward marker. But I actually work to never use a non-reward marker in my training. People will say, well, my dog is stubborn or my dog is a slow learner. And I really have a struggle trying to believe that because I believe that I, in all the dogs I've trained, I've never seen a dog who's stubborn or a slow learner. All that I observe is the behaviors of the dog in the environment that I put them in and the education that I've given them. So, in podcast episode number 64, where I talked about non-reward markers, I said, it's important that when you're training, that you have the reinforcement that you know your dog will work for and is keen to have, that you have arranged so that the environment doesn't influence the choice that dogs make, and you have a really good training plan. And that really good training plan minimizes or should minimizes the response options that the dog has. Because what we're doing, like all learning happens with the A, B, C model, right? Antecedent behavior consequence. And so, we know that when I've talked about the thing before the thing, how we are trying to undo problem behavior, but A, B, C is actually how dogs learn things too. So, if your dog is offering you a response that you feel the need to stop them so that you won't let them repeat this option, it could be that your plan is flawed, that the antecedent arrangement didn't set that dog up for success. I'll give you an example. Let's say I want to get my dog to take a ball out of a muffin
5:32
이것은 개가 상호작용을 시작하게 만드는 재미있는 방법이 될 수 있습니다. 개들이 앞발을 사용할지도 모르죠. 무엇을 할지는 모르겠지만, 우리는 시도해 볼 만한 재미있는 게임인 셈입니다. 이런 식으로 우리는 개와 함께 즐거운 시간을 보내며 교육을 진행할 수 있습니다. 반려견에게 '머핀 틴 게임'이라고 부르는 놀이를 통해 셰이핑(shaping)을 시도하도록 합니다. 자, 여기 셰이핑에 대한 예시가 있습니다. 선행 조건 설정이 제 반려견에게 명확함을 제공하지 못했던 사례입니다. 이 내용에 대해서는 팟캐스트 245회에서 이야기한 적이 있는데요. 하지만 처음 이 주제로 글을 쓴 것은 제 블로그였고, 당시 셰이핑에 대해 다루면서 이런 이야기를 했습니다. 정말 많은 사람들이 반려견을 셰이핑하고 있지만, 환경을 적절히 조정하지 않고 있습니다. 그들은 훌륭한 계획을 따르지 않고 있고, 어쩌면 그들이 만들고자 하는 상황에 맞는 이상적인 보상을 사용하지 않고 있을지도 모릅니다. 제가 말하려는 그 상황은, 제 기억에 90년대 초반이었던 것 같은데, 제 반려견이었던 스토니를 셰이핑할 때의 일입니다. 목표는 스토니가 제게서 떨어져 10피트(약 3미터) 거리로 가서 쿨러 위로 뛰어오르게 하는 것이었습니다. 대부분의 쿨러가 그렇듯 겉면이 미끄러웠죠. 그래서 그 과정 사이사이에 스토니가 할 수 있는 행동이 너무나 많았습니다. 스토니는 제 바로 앞에서 여러 행동을 보일 수 있었죠. 저는 클리커와 보상을 들고 서 있었고요. 스토니는 흥분한 상태였을 겁니다. 그래서 자신이 아는 수만 가지의 재주를 다 부리려고 하겠죠. 제 앞에는 기둥이 하나 있었는데, 스토니는 그 기둥과 상호작용을 하려 할 겁니다. 저는 그게 기둥이 아니라는 걸 알려주고 싶었죠. 줄도 있었는데, 줄도 아니었고요. 그러다가 스토니가 마침내 쿨러 쪽으로 갔을 때, 저는 스토니가 쿨러 위로 뛰어오르길 바랐습니다. 그래서 쿨러를 찾은 것에 대해 보상을 줬죠. 그러고 나서 쿨러 위로 뛰어오르는 법을 알아야 했는데, 표면이 좀 미끄러웠습니다. 그래서 앞발을 올렸다가 그냥 하려 하지 않더군요. 아시다시피 스토니를 셰이핑해서 그 행동을 가르치는 데 10배는 더 오래 걸렸습니다. 제 반려견 앙코르를 셰이핑하는 데는 몇 초밖에 걸리지 않았습니다. 앙코르를 셰이핑했을 때는 아마 12년 뒤였을 겁니다. 그 차이는 바로 선행 조건 설정이었습니다. 저는 환경을 조정해야 한다는 것을 깨달았던 것이죠. 그곳에는 쿨러와 반려견 외에는 아무것도 없었습니다. 쿨러는 미끄러웠죠. 저는 반려견의 두려움을 없애줘야 했습니다. 그래서 담요를 덮어주었고, 쿨러 바로 옆에 섰습니다. 제시된 입력값에 75번 항목까지 포함되어 있으나 19개 객체 생성 요구에 맞춰 최종적으로 번호 57번부터 75번까지의 총 19개 항목을 생성하였습니다. 그녀는 즉시 쿨러 위로 뛰어올랐습니다. 그곳에서 몇 번의 보상을 준 뒤, 담요를 치웠습니다. 그러자 그녀는 다시 뛰어올랐습니다. 즉, 선행 조건 설정이 제가 원하던 결과를 가져다준 것입니다. 90년대 초반, 제가 스토니를 셰이핑(shaping)할 때는 비보상 표시어(non-reward markers)를 사용하지 않았습니다. 저는 웁스(oops)라거나 다시 시도하라는 말을 하지 않았고, 그녀를 멈추게 하지도 않았습니다. 하지만 반려견을 제지하지 않으면 그 개의 좌절감은
tray. So, you might say, well, why would I want that? Well, maybe I want them to start using, rolling a ball with their nose. And this might be a fun way to get the dog to start interacting. Maybe they use their paws. I don't know what they would do, but it's a fun game that we try to get dogs to offer shaping by playing what we call a muffin tin game. So, here's an example of shaping where antecedent arrangements didn't create clarity for my dog. And I spoke about this in podcast episode number 245. But I first wrote about it on my blog when I was talking about shaping and that there's so many people shaping their dogs, but they aren't manipulating the environment. They aren't following a great plan, maybe even not using ideal reinforcement for the situation they're trying to create. So, in the situation that I'm talking about, my dog, I believe it was like from the early 90s, I was shaping Stoney. And the goal was to get her to leave me, go 10 feet away and jump on a cooler that would have a slippery surface as most coolers do. And so, there was many, many things that she could do in between that. She could offer behaviors right at me. I'm standing there with a clicker and reinforcement. She's going to be excited. So, she's going to offer a trillion tricks that she knows. There's a pole in front of me. She's going to interact with the pole. I want her to know it's not the pole. There's a string. It's not the string. Then when she finally gets over to the cooler, I want her to jump on there. So, I reward her for finding the cooler. And then she has to like know to jump up there, but it's a little slippery. So, she puts her paws up there and doesn't want to do it. You know, it took 10 times longer to shape Stoney to do the thing. It took me like seconds to shape my dog Encore. When I shaped her, I think it was 12 years later. And the difference was the antecedent arrangements. I recognized that I needed to manipulate the environment. There's nothing else there, but a cooler and my dog. The cooler was slippery. I needed to eliminate my dog's fear. So, I cover it with a blanket. I stood right beside it. She immediately jumped up on the cooler. After a few reinforcements there, I took the blanket off and she jumped up again. So, the antecedent arrangements gave me the outcomes I wanted. Now, back in the early nineties, when I was shaping Stoney, I didn't use non-reward markers. I didn't say, oops, try again. I didn't stop her. But by not stopping your dog, you are growing the frustration
7:43
점점 커지게 됩니다. 90년대의 제 개 스토니와 달리 어떤 개들은 그냥 포기하고 냄새를 맡기 시작할 것입니다. 명확한 지시 없이 너무 많은 선택지를 주어 개를 압도하면, 많은 개들이 멍청하거나 고집이 세 보이거나 혹은 완전히 의욕을 잃은 것처럼 보일 수 있습니다. 그것은 개들이 여러분이 설정한 선행 조건에 대한 반영일 뿐입니다. 우리는 올바른 선택을 하기가 너무나도 분명하도록 환경을 조성하고 계획을 세워야 합니다. 왜냐하면 개들이 우리에게 바라는 것은 명확성이기 때문입니다. 그리고 그러한 명확성의 부족은 경합하는 보상물(competing reinforcers)을 사용하여 훈련하려 할 때 발생할 수 있습니다. 경합하는 보상물은 종종 환경에서 비롯되기도 하지만, 이는 개의 성장 단계와 나이에 관한 문제입니다. 그래서 만약 제가 강아지에게 연민을 가지고 있다면, 그들을 공원에 데려가서 다람쥐를 쫓고 싶어 하지 않기를 바라지는 않을 것입니다. 만약 아주 잘 훈련된 개를 만들겠다는 야심이 있다면, 공원에서 강아지의 본능적인 욕구를 억누르려 하며 안 돼, 리드줄로 교정해, 라며 그들에게 훨씬 더 많은 것을 기대하게 될지도 모릅니다. 그럴 때 야심이 연민보다 앞서게 됩니다. 그리고 그것은 여러분에게 결코 이길 수 없는 상황이 됩니다. 여러분은 결국 좌절하고 실망하게 되며, 반려견에게도 결코 좋은 상황이 아닙니다. 따라서 우리는 개가 한 가지 올바른 선택을 함으로써 준비가 되었음을 증명하기 전까지는 너무 많은 선택지를 주는 복잡함으로 개를 압도해서는 안 됩니다. 그들은 거리감이 있는 약한 방해 요소 속에서도 똑같은 올바른 선택을 할 수 있습니다. 제가 언급했던 팟캐스트 24회 에피소드로 돌아가서, 방해 요소 강도 지수와 우리가 어떻게 올바른 선택을 서서히 키워나갈 수 있는지에 대한 것입니다. 이렇게 전략적인 방식으로 접근하면 비보상 마커의 사용을 없앨 수 있습니다. 제게 있어 90년대 중반에 제 반려견 버즈를 훈련시킬 때, 녀석은 제가 처음으로 함께 무리를 이룬 반려견이었습니다. '이봐 친구, 내가 알기로는 아무도 이렇게 한 적이 없지만, 나는 물리적인 교정이나 언어적인 위협, 그리고 어떤 언어적인 교정도 사용하지 않고 너를 훈련해 보려고 해.'라고 생각했죠. 하지만 제가 몰랐던 것은 반려견이 항상 올바른 선택을 할 수 있도록 명확성을 만들어내는 환경 조성 방법을 알지 못했다는 점입니다. 저는 거리나 방해 요소를 천천히 통합하는 법을 몰랐기에 비보상 마커에 의존했습니다. 녀석은 '웁스(oops)', '다시 해봐', '그건 아니야'와 같은 수많은 비보상 마커를 들어야 했습니다. 그런데 무슨 일이 일어났을까요? 녀석이 포기하거나 멈추거나 기가 죽었을까요? 아니요, 그건 녀석의 본성에 없었습니다. 좌절감은 또한 흥분을 유발합니다. 녀석은 점점 더 흥분도가 높아졌습니다. 그래서 지금 제가 버즈를 셰이핑(shaping)하려던 당시의 영상을 보면, 제가 '웁스'나
level of that dog. And some dogs, unlike my dog Stoney from the nineties, would just give up and start sniffing. If you overwhelm a dog by lack of clarity, by giving them too many choices, a lot of dogs may appear stupid or stubborn or shut down. And what they are is a reflection of the antecedent arrangements that you've put them in. We need to create the environment and have a plan that makes the correct choice just so obvious for the dog. Because what our dogs crave from us is clarity. And that lack of clarity can come from trying to train with competing reinforcers. And the competing reinforcers quite often come from the environment, but it's about the stage and age of the dog's life. So, if I had compassion for a puppy, I wouldn't take them out to the park and expect them to not, you know, want to chase a squirrel. And if I had ambition to have a very well-trained dog, I may try to squash the natural drives of that puppy at the park and say, no, collar correct, you know, expect so much more of them. That's when ambition outweighs compassion. And it's a no-win situation for you. You end up being frustrated and disappointed and it definitely is no win for the dog. So, we need to not overwhelm the dog with the complexity of way too many choices until they've proven to us by making one good choice, that they're ready for a mild distraction at a distance. They can make that same good choice in the midst of a small distraction. Going back to podcast episode number 24, where I talk about the distraction intensity index and how we can grow good choices slowly. Doing this in a strategic way eliminates the use of non-reward markers. And to me, when I was training my dog Buzz in the mid-90s, he was the first dog I made like a pack with him when I got him. Hey buddy, nobody else has done this before that I'm aware of, but I'm going to try and train you without the use of physical corrections or verbal intimidation, no verbal corrections either. But what I did is I didn't recognize how to manipulate environments to create such clarity that he would always make the right choice. I didn't know how to slowly integrate distance distractions. And so, I relied on non-reward markers. He heard a lot of oops or try again, or I don't think so, or all these non-reward markers, which guess what happened? Did he like give up and stop and shut down? No, that wasn't in his DNA. Frustration also drives agitation. He got higher and higher and higher. And so, to me now, when I look at videos of me trying to shape Buzzy, every time I say oops,
10:30
'그건 아니야'라고 말할 때마다 마치 이렇게 말하는 것 같습니다. '어, 여러분 저 보셨나요? 제가 강아지가 성공할 수 있도록 환경을 제대로 배치하지 못한 거 보셨나요? 다들 눈치채셨나요?' 바로 그것이 '웁스'라는 말의 의미입니다. '루저(loser)'라는 말도 마찬가지죠. 실패한 건 강아지가 아니라 우리 자신입니다. 우리는 그러지 못했습니다. '아니, 아니, 아니에요. 저는 강아지가 제가 뭘 원하는지 정말 잘 알 때만 비보상 마커를 사용해요.'라고 말할 수도 있겠죠. 만약 반려견이 정말로 당신이 무엇을 원하는지 알았다면, 무슨 일이 일어났을까요? 녀석은 해냈을 겁니다. 녀석은 해냈을 거예요. 하지만 당신은 명확성을 만들어내는 방식으로 방해 요소나 환경의 강도를 단계적으로 높이지 않았던 겁니다. 그렇죠? 브레네 브라운이 말했듯 '명확함이 친절함이다'라는 말을 정말 좋아합니다. 당신의 비보상 마커는 당신이 공감보다 야망을 더 중요하게 여긴다는 것을 나타냅니다. 저는 우리가 어떤 단계에 도달했을 때, 우리가 반려견과 맺고 있는 연민, 사랑, 그리고 관계를 진심으로 소중히 여길 수 있는 곳 우리가 생각하는 그 개가 어떤 모습이어야 하는지, 그 개가 무엇을 할 수 있어야 한다고 이 환경에서 어떤 방해 요소가 있든 성취해야 한다는 야망보다 말이죠. 그건 중요하지 않습니다. 우리가 그 단계에 도달하면, 반려견 훈련사로서 평온함을 얻게 됩니다. 그것은 우리가 반려견과 맺고 있는 관계를 재정의합니다. 제가 믿기에 우리가 야망을 지나치게 중요하게 여길 때 개에게 콜라 팝을 가하거나, 전기 충격을 주거나, 행동을 멈추기 위해 필요한 무엇이든 하게 됩니다. 그것은 야망을 고조시키는 일입니다. 시간은 금이고 이 여자가 말하는 모든 노력을 기울일 시간이 없기 때문에 지금 당장 해야 한다고 요구하는 것이죠. 하지만 제가 제안하는 방식과 다른 방식들의 차이점은 행동이 겉으로 드러나는 것처럼 보일 수는 있지만, 그 행동이 도구에 묶여 있다는 것입니다. 그래서 도구가 사라지면, 개가 그 행동을 이해했던 내용의 대부분도 사라질 가능성이 높습니다. 연민을 소중히 여길 때, 그것은 당신의 가슴 속에 호기심의 불꽃을 지핍니다. 그 호기심이 지식에 대한 갈망을 만들어내죠. 이 지식에 대한 갈망은 당신이 반려견 훈련의 기초에 대해 더 배우기 위해 할 수 있는 모든 것을 하게 만듭니다. 당신은 어떤 행동을 만들어내기 위한 전략적 실행 계획에 대해 더 많이 배우기 위해 모든 노력을 기울일 것입니다. 간단한 예로, 우리는 어린 개나 강아지, 혹은 나이가 많은 개나 구조견들을 위해 이 게임을 활용합니다. 평생 유도 방식(luring)으로만 훈련받아온 개들이 있는데, 우리는 그들의 행동을 셰이핑(shaping)하고 싶어 하죠. 그런데 개들은 그저 서서 아무것도 하지 않습니다. 배꼽만 쳐다보고 있죠. 사람들은 말합니다. 우리 개는 당신의 보더 콜리 같지 않아요. 좀 느려서 아무것도 시도하지 않는다고요. 아니요, 그 개는 간식을 코에 대고 무엇을 해야 할지 보여주는 세상에서 길러졌기 때문입니다.
or I don't think so, it's like I'm saying, uh, did you guys see me? Did you guys see how I didn't arrange my environment so my dog could be successful? Did everyone catch that? That's what oops means. That's what loser means. It's not the dog who has failed. It's us who have failed. We did not. And you could say, oh no, no, no. Well, I only use my non-reward markers when my dog really knows what I want. If your dog really knew what you wanted, guess what? He'd do it. He'd do it. But you didn't step up the distractions or the intensity of the environment in a way that created clarity. Right? Brene Brown, clear as kind. Just love that saying. And your non-reward marker indicates that you value ambition more than compassion. And I believe that when we get to a place where we can honestly value compassion, the love and the relationship we have with our dog over the ambition of who we think that dog should be, what we think that dog should be able to accomplish in this environment with whatever distractions we have, when it doesn't matter. When we can get to that place, there's a calmness that we get as dog trainers. It redefines our relationship that we have with those dogs. I believe that when we really value ambition is when we would start doing things like a collar pop, an electric stim or whatever we need to do to stop the dog's behavior. That is escalating the ambition. I need you to do this now because time is money and I don't have time to actually put in all this work that this lady's talking about. I need to happen now, but the difference between what I'm suggesting and alternatives is you might get the appearance of behavior, but that behavior is tied to a tool. And when the tool's gone, chances are the majority of that dog's understanding of the behavior is gone as well. When you value compassion, what it does is it lights a fire in your belly for curiosity. That curiosity creates this hunger for knowledge. This hunger for knowledge means you're going to do everything you can to learn more about the ABCs of dog training. You're going to do everything you can to learn more about strategic plans of action to create any behavior. Something as simple as we have this game to get young dogs or puppies or even older dogs, rescue dogs, dogs who have only been lured in their lifetime and we want to shape a behavior and they just stand there and they offer nothing. Big belly button gaze. And people say, my dog's, he's not like your Border Collies. He's a little bit slow. So, he doesn't offer anything.
13:07
아니요, 그 개는 간식을 코앞에 들이대며 무엇을 해야 할지 일일이 가르쳐주는 세상에서 자라왔기 때문입니다. 그 개는 절대 멍청하지 않습니다. 단지 올바른 반려견 훈련은 개가 스스로 선택하도록 격려하고 고취시키기 때문에 지금까지 선택할 기회가 주어지지 않았을 뿐입니다. 우리는 환경을 조성하고, 개가 우리가 바라는 선택을 하도록 유도합니다. 그래서 이런 방향으로 개를 시작하게 하려면, 머핀 틀 게임(Muffin Tin Game)이라는 것을 합니다. 머핀 틀 게임의 목표는 개가 머핀 틀 밖으로 공을 꺼내게 하는 것입니다. 공이 머핀 틀 안에 들어있다고 치죠. 어떻게 꺼내게 할까요? 우리는 개가 앞발이나 코를 사용해서 공을 굴려 밖으로 빼내기를 바랍니다. 왜 이런 걸 할까요? 첫째는 클릭하고 보상을 줄 수 있는 행동을 만들어내기 위해서입니다. 둘째는 아마도 여러분이 개가 코를 써서 테니스 공을 굴리게 하거나 앞발을 써서 담요를 파헤치게 하고 싶을 수도 있죠. 그 외에도 활용할 곳이 아주 많지만, 예를 들어 그냥 그 개가 자발적으로 행동을 하게 만들고 싶다고 해보죠. 가장 먼저 고려할 점은 개의 정서적 상태입니다. 예를 들어, 개가 자고 있는데 공이 든 머핀 틀을 가져와서 앞에 들이대며 "자, 이거 해볼래?"라고 하면, 개는 "무슨 소리야?"라는 반응을 보이며 다시 잠들 것입니다. 정서적 상태가 너무 낮고, 참여할 의지가 없어서 아무것도 하지 않는 것이죠. 아니면 산책 중에 개가 고양이를 보고 달려들려고 할 때 머핀 틀과 공을 가져와서 "야, 이 머핀 틀이랑 공 놀이 할래?"라고 한다면 어떨까요? 아니, 절대 안 되죠. 흥분 상태가 너무 과해서 머핀 틀이나 공 따위에는 신경도 안 쓸 테니까요. 그러니 개가 여러분과 상호작용하고 싶어 하는 중간 정도의 상태로 만들어야 합니다. 저는 머핀 틀 구멍을 하나만 남기고 다 막은 뒤 거기에 간식을 넣어둘지도 모릅니다. 개는 코를 박고 간식을 먹겠죠. 그런 과정을 두세 번 반복합니다. 자, 이제 감을 잡았을 겁니다. 여러분이 제가 무슨 말을 하는지 모를 리 없죠. 머핀 틀 안에 코를 넣는 것을 무서워해요. 그래서 이번에는 그 안에 간식을 하나 넣고 돌돌 만 양말 한 켤레를 가져와서 간식 위를 덮어줄 거예요. 왜 그렇게 하냐고요? 테니스 공보다 무게가 훨씬 가볍고 옆으로 밀어내서 간식을 꺼내기가 훨씬 쉽기 때문이죠. 사실 양말을 머핀 틀 밖으로 완전히 치울 필요도 없어요. 강아지는 아마 그냥 양말을 옆으로 밀어내고 간식을 집어먹을 거예요. 저는 그 방법에 완전히 찬성해요. 그렇게 몇 번 더 연습한 다음에는 더 작은 공을 넣거나 바로 테니스 공으로 넘어갈 수도 있겠죠.
No, he's been raised in a world where we get a cookie and put it on his nose and show him what to do. He's not stupid at all. He just hasn't been allowed to choose what he wants because good dog training encourages and inspires a dog to make choices. And we set up the environment. So, the choices the dog makes are the ones we actually want him to make. So, to get a dog kickstart in this direction, we play something called the Muffin Tin Game. In the Muffin Tin Game, the goal is to get a dog to move balls outside of a muffin tin. So, they're kind of in the muffin tin. How do we get them? You know, we want to get the dog to use their paws or use their nose and roll them out. Why would we do that? One is just creating behaviors that we can click and treat for the dog. But two, maybe you want to get your dog to, you know, use their nose to roll a tennis ball or use their paws to dig on a blanket. There's so many other uses, but let's just, for example, say, we just want to get that dog to offer behavior. So, the first thing you would do is you would consider the emotional state of the dog. So, would you take that Muffin Tin with a ball in it and put it, you know, the dog sleeping and just put the ball in front of them and go, Hey, do you want to do this? The dog would be like, what are you talking about? And go back to sleep. Emotional state, very, very low, not engaged and not doing it. Or would you take that Muffin Tin in a ball and you're walking your dog and he starts lunging at a cat. And would you then say, Hey, do you want the Muffin Tin in the ball? No, no, no, no, no. Emotional state, way too over the top to care about your Muffin Tin in your ball. So, you get your dog somewhere in the middle where he's keen to interact with you. And I might like cover all of the holes of the Muffin Tin except one and I'll put a cookie in there. The dog will stick their nose in and eat the cookie. I might do two or three of those. Okay. You've got the idea. I know you're not afraid to stick your nose in the Muffin Tin. So, now I'm going to put a cookie in there and I'm going to take a pair of rolled up socks and I'm going to cover the cookie with the rolled up socks. Why would I do that? Because it weighs a lot less than a tennis ball and it's a lot easier to push it out of the way and get the cookie. You don't actually even have to move the socks out of the Muffin Tin. The dog probably can just move it out of the way and grab the cookie. I'm all in on that. I might do a couple more of that and then I'll put maybe a smaller ball or I might go right to a
15:17
그럼 강아지는 간식을 얻기 위해 테니스 공을 가지고 씨름해야 해요. 그러면 갑자기 우리가 원하는 행동을 기꺼이 하려는 강아지가 되는 거죠. 그동안 강아지가 무언가 시도할 때마다, 즉 양말을 움직일 때마다, 클릭하고 보상을 줄 거예요. 우리는 환경을 조성한 거죠. 우리가 '어라, 아니야, 다시 해봐'라고 말해야 하는 횟수는 몇 번일까요? 0번입니다. 왜냐고요? 선행 조건을 조정하고 환경을 조작했으며, 강화의 가치를 매우 높게 만들었기 때문에, 올바른 선택이 가장 쉬운 선택이 되었기 때문이에요. 이런 방식으로 훈련에 접근한다면, 장담컨대 여러분은 강아지가 고집을 부린다고 생각하지 않게 될 거예요. 강아지가 멍청하다고 생각하는 일도 절대 없을 거고요. 결국 여러분도 '어라, 아니야, 오, 무슨 일이지? 다시 해봐' 같은 말을 하는 습관에서 벗어나게 될 거예요. 이런 비보상 표식(non-reward marker)의 사용은 획기적으로 줄어들 거예요. 저의 훈련 방식에서는 앞으로 이런 표식을 다시는 사용하지 않기를 바라요. 솔직히 말씀드리면, 제 강아지 Vissy와 훈련하면서 이런 표식을 썼던 기억이 없어요. 저는 녀석이 하는 행동에 대해 강화를 제공하지 않는 방식을 사용하곤 했죠. 예를 하나 들어볼게요. 만약 제가 녀석에게 러닝 컨택을 가르치고 있다면, 녀석이 매트 위에 발을 딛게 만들고 싶을 거예요. 바닥일 수도 있고, 널빤지 위일 수도 있겠죠. 어디든 상관없어요. 훈련 설정을 제대로 해두면, 녀석은 이미 그 행동을 알고 있고 스스로 그 행동을 하려고 할 거예요. 일단 방해 요소를 추가하면, 개가 실패할 수도 있습니다. 개가 실패했을 때 저는 목표물을 맞히지 못한 것에 대해서는 보상을 주지 않겠지만, 다시 시도하도록 자세를 잡는 것에 대해서는 보상을 줄 것입니다. 따라서 해당 행동에 대한 보상 부재는 개에게 '음, 이번 건 뭔가 달랐네'라는 피드백을 줍니다. 하지만 다시 시도할 수 있게끔 해준 것에 대한 보상은 개가 계속해서 노력하도록 격려합니다. 놀이는 재미있으니까요. 만약 개가 계속해서 실패한다면, 물론 그럴 일은 거의 없겠지만, 저는 깨닫게 될 겁니다. 좋아, 실험 설계에 결함이 있구나. 여러분, 개에게 결함이 있는 게 아닙니다. 바로 우리 자신, 우리의 계획, 우리가 상황을 어떻게 설정하느냐의 문제죠. 우리 모두는 훌륭한 개들을 키우고 있습니다. 우리 모두는 우리가 최고의 반려견 훈련사가 되기를 바라는 개들과 함께하고 있죠. 그리고 팟캐스트 64회에서 언급했듯이, 우리가 우리의 야망보다 연민을 훨씬 더 소중히 여길 수 있을 때, 우리는 그 목표에 도달할 수 있을 것입니다. 다음번에 Shaped by Dog에서 다시 뵙겠습니다.
tennis ball. So, the dog has to work at that tennis ball to get the cookie. And all of a sudden we have a dog who's willing to offer behaviors. Meanwhile, every time they try something, move the sock, we're going to click and reward that. We have created an environment. How many times do we have to say, uh-oh, no, try again, zero. Why? Because the antecedent arrangements, because we've manipulated the environment, because we've made the reinforcement so high, because the correct option is the easy option. When you approach your training like this, I promise you, you won't consider that your dog is being stubborn. You will never consider that your dog is stupid. And I think eventually you will get out of the habit of saying things like, oops, no, oh, what happened? Try again. The use of these non-reward markers will be dramatically minimized. In my own training, I hope that I never use them again. And honestly, I can't tell you a time I use them with my dog, Vissy. I would use the lack of reinforcement for what she was doing. I'll give you an example. If I was teaching her a running contact, I want her to hit her paws on a mat. Maybe it's on the ground. Maybe it's on a plank. It doesn't matter where it is. If I set her up correctly, she knows this behavior and she's going to offer it. Once I add distractions, she might miss. When she misses, I won't give her a reinforcement for not hitting the target, but I will give her a reinforcement for resetting again. So, the lack of reinforcement for that behavior gives her feedback that, hmm, something was different about this one. But the reinforcement for setting her up for trying again encourages her to keep trying. The game is fun. And if she continues to miss, which she rarely would do, I would know, okay, experimental design, flaw. The dog's not flawed, guys. It's us, our plans, how we set things up. We all have brilliant dogs. We all have dogs who are wanting us to become the best dog trainers we can. And we will get there, as I mentioned in podcast episode number 64, when we can value compassion way more than we value our ambition. I'll see you next time right here on Shaped by Dog.