Why Isn't My Dog Learning What I'm Training? #125
Susan GarrettShaped by Dog with Susan Garrett · 팟캐스트 · 16분
반려견 훈련에서 보호자가 의도한 교육 내용과 강아지가 실제로 학습하는 행동 사이에는 간극이 존재한다. 이 간극은 주로 강화 과정(reinforcement process)의 각 요소 사이에서 발생하며, 보호자가 미세한 타이밍을 놓치거나 보상 전달 과정에서 부주의한 신호를 보낼 때 심화된다. 강아지는 보호자의 언어적 신호보다 신체적 움직임을 더 빠르게 포착하므로, 훈련의 기계적 기술은 행동의 정교함을 결정짓는 핵심 요소이다. 보호자는 강화 과정의 관찰, 인식, 마킹, 보상 전달, 배치, 그리고 정지 자세에서의 해제라는 6가지 단계를 의도적으로 통제해야 한다. 간식을 꺼내는 과정에서 발생하는 불필요한 움직임이나 위치 오류는 강아지에게 잘못된 행동을 보상으로 각인시킬 위험이 있다. 따라서 모든 간식은 훈련 전에 미리 준비되어야 하며, 보상은 강아지가 목표로 하는 위치에 도달했을 때 정확하게 제공되어야 한다. 보호자가 이 과정을 의도적으로 관리할 때 강아지가 학습하는 내용은 의도한 교육 목표와 일치하게 된다.
내 반려견은 왜 내가 훈련하는 내용을 배우지 못할까? #125
Why Isn't My Dog Learning What I'm Training? #125
0:00
모든 교육 과정에는 가르치는 사람이 가르치는 내용과 배우는 사람이 배우는 내용이 있습니다. 대수학 수업을 조금이라도 들어본 사람이라면 누구나 알겠지만, 이 둘은 종종 일치하지 않죠. 안녕하세요, 저는 수잔 가렛이고 'Shaped by Dog'에 오신 것을 환영합니다. 이 개념은 개를 키우고 사랑하며 예절 바른 개로 기르고 싶어 하는 우리에게 매우 중요합니다. 교실에서는 선생님이 가르치는 내용이 있지만, 숙제를 얼마나 열심히 하는지나 수업에 얼마나 집중하는지와 같은 다른 요소들도 존재하기 때문이죠. 선생님이 "이거 배웠잖아, 이미 배웠으니 당연히 알아야지"라고 말하며 허리에 손을 얹고 답답해할 만한 상황이 많습니다. 그런데 난독증이 있는 아이로 교육 과정을 겪었던 사람으로서 말하자면, 선생님이 가르쳤다고 해서 제가 배운 것은 아니라는 점입니다. 선생님이 가르쳤다는 사실이 곧 내가 배웠다는 뜻은 아니거든요. 확실히 그렇습니다. 아무튼 딴 길로 샜네요. 이제 강아지 이야기를 해보죠. 강아지에게는 "숙제했니?"라거나 "집중했니?"라고 물을 수 없으니까요. 모든 책임은 우리에게 있습니다. 여러분, 100% 우리 책임입니다. 우리가 가르쳤다고 생각하는 것과 개가 실제로 보여주는 행동 사이에 간극이 있다면, 누구를 탓할 수도 없습니다. 문제는 결국 개가 무엇을 보여주고 있느냐 하는 것이고, 그건 우리 교육 시스템에 결함이 있다는 뜻입니다. 오늘 저는 그 결함의 큰 부분을 어떻게 해결할 수 있는지 알려드리려고 합니다. 그 모든 것은 '강화 과정(reinforcement process)'이라는 것과 관련이 있습니다. 여기에는 다섯 가지, 때로는 여섯 가지 요소가 포함됩니다. 제가 무슨 말을 하는지 모르는 것처럼 들릴 수도 있겠지만, 저를 믿어보세요. 그 여섯 번째 요소가 무엇인지 나중에 설명해 드리겠지만, 강화 과정에는 분명히 다섯 가지, 때로는 여섯 가지 요소가 있습니다. 공원에서나 길을 걸을 때 제가 보는 대부분의 사람들은 이 부분을 제대로 이해하지 못하고 있습니다. 솔직히 이 팟캐스트를 듣는 것만으로도 해결할 수 있습니다. 저를 믿으세요, 해결할 수 있습니다. 먼저, 강화 과정이 어떤 것인지 알려드리겠습니다. 자, 예를 들어서 반려견을 훈련하고 싶은 것이라면 무엇이든 좋습니다. 하지만 오늘은 반려견과 함께 길을 걷는 것에 대해 이야기해보겠습니다. 여러분의 목표는 루즈 리시 워킹(느슨한 줄 산책)입니다. 저는 제 개가 제 옆에 머물기를 원합니다, 대략적으로요. 조금 앞서 있어도 괜찮습니다. 크게 신경 쓰지 않아요. 아무도 점수를 매기지 않으니까요. 하지만 길을 건너려고 멈출 때, 개가 제 옆에서 앉았으면 좋겠어요. 그건 단지 안전을 위한 거죠. 그리고 제가 이제 움직인다고 말하면, 우리는 움직이는 겁니다. 그게 전부예요. 별일 아닙니다. 그게 제 목표예요. 자, 여기에는 강화 과정이 있습니다. 첫 번째 요소는 우리가 반려견이 무엇을 하는지 관찰하는 것입니다.
In any education process, there's what the educator is teaching and there is what the learner is learning. And anybody who's ever spent any time in an algebra class knows those who are often not the same thing. Hi, I'm Susan Garrett and welcome to Shaped by Dog. This concept is incredibly important for those of us who own and love and want well-behaved dogs. Because in a classroom, there is what the teacher is teaching, but there's other elements like how much you apply yourself doing your homework or how well you pay attention in class. There's a lot of things that the teacher can put their hands on their hips and go, we took this. You should know this. We took this already. And let me just share with you that as a dyslexic child going through an education process, just because you thought you taught it doesn't mean I was learning it, sister. I can tell you that much for shizel, but I digress. Let's talk dogs because with dogs, there isn't any, did you do your homework or did you pay attention? It's all on us, guys. It is a hundred percent on us. If there is a distance between what we thought we trained and what the dog is showing us they know, Ty doesn't go to the runner here. It all is about what the dog is demonstrating because there's a flaw in the system. And today I'm going to share with you how you can fix a big part of that flaw. It all comes down to something called a reinforcement process. And there's six, well, five elements, sometimes six. It sounds like I don't know what I'm talking about. Trust me on this. I'll get to what that sixth element may be, but there's for sure five elements and sometimes six elements to the reinforcement process. And the vast majority of the people that I see with their dogs at the park, walking down the street are messing this up. And honestly, you can fix it just by listening to this podcast. Trust me, you can fix it. First, let me share with you what the reinforcement process looks like. Now let's take an example like you can take anything you want to train your dog. But today we're going to talk about walking your dog down the street. Your goal is loose leash walking. I want my dog to hang out at my side, sort of. They can be a little bit in front, a little bit ahead. I'm not going crazy. No one's scoring us on this. But if I stop to cross the street, I'd like my dog to sit when I stop. And that's just safety. And when I say, you know, we're moving, then we're moving. That's all. No big deal. That's my goal. Now, there is a reinforcement process. The first element is us observing what our dog is doing,
2:57
왜냐하면 우리는 반려견이 멋진 행동을 하는 순간을 포착하고 싶기 때문입니다. 원래 있어야 할 위치에서 걷거나 원래 앉아야 할 위치에 앉아 있는 것을 말이죠. 그래서 우리는 관찰할 것입니다. 이제 우리의 뇌는 반려견이 하는 모든 행동에 대한 데이터를 샅샅이 뒤져야 합니다. 개가 무엇을 하고 있는지, 개가 무엇을 하고 있는지요. 그리고 두 번째 단계는 인식입니다. 그래서 우리는 관찰하고 인식할 것입니다. 우리가 정말로 마음에 드는 것을요. 오, 그거 좋네. 그런 다음 표시(마킹)를 할 겁니다. 우리는 반려견에게 '이봐, 네가 뭘 하려는지 나도 알아차렸어'라고 알려주는 겁니다. 네가 하는 행동이 좋다는 것을 아는 거죠. 그래서 마킹을 합니다. 이것이 강화 과정의 세 번째 요소입니다. 그런 다음 우리는 그들이 정말 좋아하는 보상을 제공할 것입니다. 물론 마킹을 한 것도 좋지만, 그 뒤에 긍정적인 강화를 덧붙이는 것이죠. 우후! 우리는 여기서 매우 특별한 것의 기초를 다지고 있는 겁니다. 이제 반려견이 보상을 받을 수 있도록 위치를 지정할 것입니다. 하지만 우리가 알겠고 이제 전달하는 거예요. 강아지가 그걸 확실히 얻게 해야 하죠. 이제 여섯 번째이자 때때로 강화 과정의 요소인데, 이것은 릴리즈(release, 해제 신호)라고 불리는 것으로, 오직 통제된 자세일 때만 사용합니다. 앉아서 움직이지 않거나, 엎드려서 움직이지 않거나, 서서 움직이지 않는 것 같은 상황이죠. 즉, 우리가 강아지에게 움직이지 말라고 요구하는 상황이라면, 언제 끝나는지 알아야 합니다. 거기서 릴리즈가 필요한 거죠. 그러니까 6가지 요소, 때로는 5가지, 대부분의 경우엔 5가지죠. 맞죠? 관찰하고, 인식해서 딩, 딩, 딩. '저거야, 내가 원하던 거.' 마킹을 해서 강아지가 알게 하고, 전달하고, 그리고 그 보상이 최종적으로 놓일 위치를 결정하는 것, 그리고 때때로 릴리즈까지요. 좋아요. 다시 길을 걷는 상황으로 돌아가 봅시다. 자, 이런 일이 벌어집니다. 우리가 귀여운 골든두들을 데리고 산책하는데 강아지가 갑자기 다람쥐를 보고 왼쪽으로 휙 달려들어서 어깨가 빠질 정도로 잡아당기지만, 곧바로 다시 돌아옵니다. 아, 다람쥐가 사라졌네. 괜찮아. 다시 옆으로 돌아오면 우리 뇌는 이렇게 생각하죠. '오, 봐, 제자리에 있네.' 원래는 완전히 벗어나 있었지만, 지금은 다시 제자리로 돌아왔어. 잘했네. '딩, 딩.' 우리가 그걸 인지하는 거예요. 그래서 주머니에 있는 간식을 꺼내려고 손을 움직이기 시작하죠. 그러면서 마킹을 합니다. '굿'. 그러니까 어떤 상황에는 클리커를 사용할 수도 있어요. 그게 마커니까요. '착한 강아지'라고 할 수도 있고, '굿'이라고 할 수도 있고, '예스'라고 할 수도 있죠. 저는 '굿'이나 '예스'라는 단어를 좋아해요. 요즘은 '굿'이라는 단어를 더 자주 씁니다.
because we want to catch our dog doing something amazing. Walking in the position they're supposed to be in or sitting in the position they're supposed to be in. So, we're going to observe. Now, our brain has to go scatter through all the data of what the dog is doing, what the dog is doing, what the dog is doing. And the second step is recognize. So, we're going to observe and recognize the thing that we really, really like. Oh, that's good. Then we're going to mark it. We're going to acknowledge to the dog, Hey, I get, I'm picking up what you're putting down. I know what you're doing is good. So, we mark. It's a third element of the reinforcement process. And then we're going to deliver them a reinforcement, something they really like. Because sure, it's good that we marked it, but then we follow it up with some positive reinforcement. Woo-hoo! We're building the foundation of something very, very special here. Now, we're going to place the reinforcement so that the dog can get it. It's no good that we've got it and we're delivering it. We've got to make sure the dog gets that. Now, the sixth and sometimes element of this reinforcement process, and that is something called the release, which we only use if it's a control position. Like you sit and don't move, you lay down, don't move, stand and don't move. So, if it's something that we're asking the dog to don't move, then they have to know when it's over. And that's where the release comes in. So, six elements, sometimes five, most of the time it's five. Right? Observing, recognizing, ding, ding, ding. That's the one I want. Marking so that the dog knows out, I know, and delivering, and then the placement of where the final resting place is of that reinforcement and sometimes the release. Okay. Let's get back to walking down the street. So, here's what's happening. We're walking our cute little golden doodle and the dog suddenly sees a squirrel and lurches over to the left and kind of pulls her shoulder out of the socket, but immediately comes back in. Oh, the squirrel's gone. It's okay. Comes back in beside and our brain goes, Oh, look, they're right in position. They were way out of position, but now they're right back in position. That's good. That's a ding, ding. We recognize it. So, we start to move our hand to get into our pocket to get our cookie. And then as we're doing that, we mark it, good. So, you could use a clicker for some things. Like that's a marker. You could say good dog. You could say good. You could say yes. I like the word good or yes. I'm using the word good more often now.
5:24
그게 바로 강아지에게 '네가 내 옆에 아주 잘 있다는 걸 나도 알고 있어'라고 알려주는 마커예요. 이제 주머니에서 간식을 꺼내는데, 아 이런, 간식이 너무 크네요. 그래서 조각을 하나 떼어서 반은 다시 집어넣고, 입에서 꺼내서 간식을 부수고 있으면 반려견은 이렇게 생각하죠. 입에 뭐가 들었어? 그러고는 반려견이 위치를 잡고 앞으로 나와서 시작하는 거죠. 당신이 걸을 때 뒤로 걷기 시작해요. 반려견은 당신 앞에서 방방 뛰고 있죠. 그거 맛있는 간식인가요? 왜냐면 난 착한 아이니까요. 그러고 당신은 입에서 꺼내서 부수고 있어요. 이제 반려견은 당신 가슴 위로 뛰어오르고, 당신은 그래, 친구야. 그래, 잘했어. 라고 말하며 간식을 주죠. 반려견의 입이 바로 여기 있으니까요. 보상 전달과 위치 선정이 동시에 일어날 수 있는 거죠. 그게 얼마나 좋았을까요? 바로 그 점 때문에 개들은 우리가 무엇을 원하는지 배우기 어려워합니다. 이건 과장된 예시입니다. 하지만 보상 과정에서 여러분이 원하는 것은 행동을 형성하는 데 도움이 되는 간극입니다. 왜냐하면 다섯 가지 혹은 여섯 가지 요소 사이의 간극, 그 간극에서 행동이 사라지기 때문입니다. 자, 한번 확인해 보죠. 방금 그 시나리오를 다시 살펴보겠습니다. 길을 걷고 있는데, 반려견이 왼쪽으로 나가서 리드줄 끝까지 당겨진다면, 그건 사실 반려견의 신체 구조상 정말 좋지 않은 일입니다. 분명 어딘가 다칠 거예요. Shaped by Dog의 16번째 에피소드, 행동 이전의 행동에 대해 들으셨다면, 여러분은 방금 멋진 행동 연쇄를 만드신 겁니다. 왜냐하면 반려견이 원래 위치에서 너무 멀리 떨어져 있었기 때문에 당신의 뇌가 딩, 딩, 딩 하고 울린 거죠. 반려견이 즉시 당신 옆으로 돌아왔을 때 그건 잭팟이라고 생각하는 겁니다. 그리고 당신은, 반려견이 있던 곳과 지금 있는 곳의 차이가 너무 크니 이건 분명 좋은 거야. 라고 생각하죠. 그래서 보상 마킹을 하고 싶어 합니다. 하지만 입으로 말을 뱉기도 전에, 팔이 먼저 간식을 향해 움직이기 시작하죠. 개들은 언어적 신호보다 신체적 신호를 훨씬 더 쉽게 포착합니다. 그래서 그들은 '굿'이라는 단어를 듣지도 못합니다. 그저 손이 움직이는 것을 보고, 아, 간식이구나 하고 움직이기 시작하죠. 간식을 얻으려고 손을 따라가는 위치, 그렇죠? 하지만 마크하고 잘했어라고 말하는 거예요. 그래서 지금 당신이 마크한 것은 위치로 다시 달려들어 온 그 개가 아닙니다. 당신이 마크한 건 주머니를 뒤지는 당신의 움직이는 손을 보고 따라가기 시작한 그 개죠. 그런 다음 이제 보상이 전달됩니다. 쿠키를 꺼냈는데 오 이런, 너무 크네요. 그래서 부숴줘야 합니다. 그 틈이 바로 행동이 변질될 수 있는 빈틈인 거죠. 자, 이제 개가
That's the marker that tells the dog, you know what I know. I'm being good here right beside me. So, now you're getting in your pocket, you get your cookie out and then, oh man, that's a really big cookie. So, you break off a piece of it, put half of it back in, you take that out of your mouth and then you're breaking it up and your dog's going, what do you got in your mouth? And he comes into position and he comes out in front and he starts walking backwards as you're walking. He's out in front of you bouncing up and down. Is that a good cookie? Because I'm a good boy. And then you get out of your mouth and you're breaking it up. And now he's jumping up on your chest and you go, yeah, buddy. Yeah, you were good. And you give him a cookie because like his mouth is right here. The delivery and the placement can happen simultaneously. How good was that? And that's why dogs have a hard time learning what it is that we want. Now that's an exaggerated example. But what you want in a reward process is you want gaps that help build the behavior because the gaps between those five elements or six, the gaps are where behavior is lost. So, let's check this out. Let's look at that whole scenario again. You're walking down the street, your dog exit stage left, gets to the end of your leash, which by the way, is really not a good thing for your dog structurally. Something's going to hurt, I'm sure. And if you listen to episode 16 here on Shaped by Dog, the thing before the thing, you've just created a lovely behavior chain because your dog was so far out of position that your brain went ding, ding, ding. That's a jackpot when he immediately came crashing back beside you. And you're like, the difference between where he was and where he is now was so great. That has to be good. And so, you want to mark it. But before your mouth says anything, your arm starts to move for the cookie. To dogs, they pick up on physical cues so much easier than verbal ones. And so, they don't even hear the word good. They just see the hand moving and they go, oh, I'm in. And they start moving at a position to follow where your hand's going for the goods, isn't it? But you mark and say good. And so, what you did marked now isn't that dog who came crashing back into position. You marked the dog who was starting to move to follow your hand that they saw moving that was digging into your pocket. And so, now the delivery happens. You've got the cookie out and you go, oh yeah, that's too big. And you have to break it up. That is a gap that allows behavior to be morphed. So, now the dog
7:57
당신에게 앞발을 올립니다. 거기 뭐 들었어? 하는 거죠. 당신은 내 옆에 잘 붙어 있는 개에게 보상을 줘야겠다고 생각해서, 개가 앞발을 올리고 있는 상태에서 쿠키를 줍니다. 산책을 가려는데 엉뚱한 방향으로 걸으면서요. 그건 개가 느슨한 줄로 걷는 것과는 아무런 상관이 없습니다. 다시 말하지만 과장된 예시지만, 그런 일의 일부는 일 년 내내 매일 일어납니다. 예시를 하나 들어볼게요. 어떤 분이 레슨을 받으러 왔는데, 흔히들 가장 많이 불평하는 게 개가 짖기만 한다는 거예요. 그래서 저는 평소대로 그분을 관찰하고 있습니다. 저는 항상 개와 학생 모두를 관찰하고 인지하고 마크하는 이 과정을 거치며 일하거든요. 그분의 큰 불평 중 하나가 개가 너무 많이 짖는다는 것이었어요. 그래서 개를 크레이트에서 꺼냈죠. 아니나 다를까, 신호라도 받은 듯 개가 빙글 돌면서 짖기 시작했습니다. 그래서 훈련을 시작하려고 개가 짖지 않게 하려고 간식을 주기 시작합니다. 강화는 행동을 구축합니다. 자, 방금 무엇을 보상하셨나요? 그걸 적어두세요, 좋은 문신 문구가 될 거예요. 방금 그 쿠키가 무엇을 보상했나요? 내가 방금 무엇을 보상했지? 많은 경우 의도조차 하지 않죠. 그러니 개가 달려들었다가 다시 위치로 돌아왔을 때를 생각해서 의도적으로 행동하세요. 나는 그걸 마크하고 싶다. 그리고 개 훈련은 기계적인 기술입니다. 제 멘토인 밥 베일리가 항상 입버릇처럼 하던 말이죠. 반려견 훈련은 기계적인 기술입니다. 그래서 여러분은 행동을 마킹할 때 목소리를 사용해야 합니다. 그런 다음 움직이세요. 정말 의도적으로 행동하세요. 훈련 영상을 촬영해 보세요. 길을 걸을 때는 힘들다는 걸 알지만, 마킹을 하고 나서 움직이세요. 지금 주머니를 뒤적거리고 계시죠. 지금 몸에 딱 붙는 꽉 끼는 청바지를 입고 있어서 간식을 꺼내기가 힘들군요. 그래서 캥거루 파우치, 트레이닝 파우치가 없다면 아주 좋은 대안이 됩니다. 만약 트레이닝 파우치가 없다면, 제가 좋아하는 걸 하나 공유해 드릴 수 있어요. 몇 개 가지고 있거든요. 쇼 노트에 링크를 올려둘게요. 유튜브로 보고 계시다면 설명란에 있을 겁니다. 훈련 중에 간식을 절대 부수지 마세요. 말씀드리고 싶은 게 있는데, 우리 맹세 하나 하죠. 다들 손을 들어보세요. 뭐, 운전 중이라면 한 손은 핸들에 두시고요. "나는 훈련하기 전에 항상 간식을 미리 잘라두겠다." 제가 하는 간단한 방법이 있어요. 로스트 비프를 사서 인스턴트팟에 넣고, 잘게 썬 다음 작은 플라스틱 용기에 나누어 냉동실에 넣고 일주일 치씩 꺼내 씁니다. 물론 로스트 비프만 가지고 훈련하지는 않아요. 그건 고가치 보상용이죠. 하지만 다른 간식들도 똑같이 해서 훈련용 간식은 모두 필요하기 훨씬 전에 미리 잘라둡니다. 한참 전에요. 자, 주머니를 뒤적거릴 때 간식을 찾느라 더 오래 걸릴수록 그 간격은 더 커집니다.
puts their paws up on you. What do you got there? And you're going, well, I got to reward my dog for being good, being at my side. And so, then you give them a cookie with them, their paws up on you, walking in the wrong direction as you're trying to go for a walk. None of that has anything to do with the dog walking on a loose leash. Again, an exaggerated example, but parts of that happen every single day of the year. So, I'll give you an example. Somebody was in for a lesson and often a big complaint is her dog barks all the time. So, I'm observing her as I do. I'm always working through this process of observing and recognizing and marking both with dogs and with students. And one of her big complaints is her dog just barks way too much. So, she gets a dog out of the crate. Sure enough, on cue, the dog spins and starts barking. So, she starts feeding it to get it to not bark so that she can start working. Reinforcement builds behavior. So, what did you just reward? Write that, you know, that would be a great tattoo. What did that cookie just reward? What did I just reward? A lot of times you aren't even being intentional. So, think about the dog lunge, they came back into position and be intentional. I want to mark that. And dog training is a mechanical skill. My mentor, Bob Bailey used to say that over and over again, dog training is a mechanical skill. And so, what you are doing, you have to mark with your voice and then move. Be really intentional. Video your training. I know it's hard when you're walking down the street, but mark and then move. Now, you're digging into the pocket. Now you've got these tighty tight little jeans on and that's hard to get those cookies out. And so, kangaroo pouches, that's a great place if you don't have a training pouch. But if you don't have a training pouch, I can share with you one that I like. I've got a couple. I'll put some links in the show notes. If you're watching this on YouTube, it'll be in the description. Your cookies are never broken up while you're training. Let me say, let's take an oath. Everyone put your hands up. Well, enough you're driving, keep one on the wheel. My cookies will always be broke up before I train. Something simple I do. I'll get a roast, throw it in the Instapot, cut it up, put it into little Tupperware containers, throw them in the freezer, take one out for a week. Now, of course, I don't just train with roast beef. Those are my high value rewards, but I'll do the same thing with other things so that all of those training treats are cut up long before I need them. Long before I need
10:41
무슨 일이 벌어지는지 기억하시나요? 보상 과정의 요소들 사이의 간격이 행동을 변형되게 만듭니다. 우리는 행동이 변형되는 것을 원치 않아요. 깔끔한 행동을 원하죠. 그래서 마킹을 잘하고, 베이트 파우치나 캥거루 파우치 등 간식이 있는 곳으로 손을 뻗어 이제 전달을 하는 겁니다. 만약 제가 간식을 꺼내려고 손을 뻗을 때 반려견이 제 앞으로 온다면, 그건 우리 훈련 프로그램의 핵심 요소인 '반려견의 선택'이라는 개념이 조금 흐려졌다는 신호입니다. 그래서 저는 그냥 간식을 가져갈 겁니다. 그리고 저는 그냥 손에 간식을 쥐고 앞으로 내밀어 봅니다. 정말로 먹고 싶니? 그러면 개는 그제야 '아 맞다, 여기 있으면 안 되지, 그렇지?' 하고 제 옆자리로 다시 돌아옵니다. 그래서 저는 간식을 다시 주머니에 넣고, 개의 머리를 토닥인 뒤 다시 연습을 시도합니다. 다시 간식을 꺼내면서 '자세 유지할 수 있니? 그래, 잘했어.' 하고요. 그러니까 전달은 여러분이 손을 뻗어 개에게 간식을 가져다주는 행위이지만, 보상을 주는 위치인 배치야말로 가장 중요합니다. 보상, 즉 강화의 위치는 여러분이 만들고자 하는 행동을 강화하거나 무너뜨립니다. 사람들이 개에게 앉으라고 시키는 모습을 얼마나 많이 봤는지 모릅니다. 개는 앉았고 주인은 '잘했어'라고 말하며 간식을 주려는데, 개는 앞발을 땅에서 떼고, 주인은 그 상태에서 간식을 줍니다. 오늘 친구인 린다 오튼 힐이 놀러 와서, 제가 나쁜 반려견 훈련사의 역할을 해달라고 부탁했습니다. 유튜브로 보고 계신다면, 린다와 그녀의 개가 아주 잘못된 배치와 보상 방식을 보여주고 있을 겁니다. 중요한 건 보상 과정 그 자체이며, 방금 준 간식이 무엇을 보상했는지 이해하는 것입니다. 내가 무엇을 강화하고 있었던 걸까요? 만약 개에게 엎드려라고 시킨 후 간식을 주려는데, 개가 간식을 받아먹기 편하려고 팔꿈치를 바닥에서 뗀다면, 보상 과정은 이미 작동하고 있는 겁니다. 강화 과정은 항상 작동합니다. 여러분은 방금 개에게 엎드린 자세란 팔꿈치가 바닥에 닿아 있든 아니든 상관없다고 강화해 버린 셈입니다. 알겠죠? 그러니 항상 차이를 인지하세요. 관찰과 인식 사이의 간극 말이죠. 관찰하고 인식해야 합니다. 저게 바로 내가 원하는 행동이야라고요. 하지만 내가 원하는 행동 직후에 원치 않는 행동이 바로 따라온다면, 저는 마킹하지 않을 겁니다. 일단 관찰하며 몇 걸음 더 가서 다시 기회를 노릴 것입니다. 그러고 나서 저는 '좋아, 잘했어'라고 말할 겁니다. 무언가를 쫓아간 뒤 돌아오는 건 아주 훌륭해요. 제자리를 지키는 것도 정말 정말 좋고요. 자, 그래서 보상은 신속하고 의도적으로 이루어져야 합니다.
them. All right. So, when you dig into your pocket, the more you break that up, the bigger the gap. Remember what happens? The gap between the elements of the reward process are what allows behavior to morph. We don't want behavior to morph. We want behavior to be clean. So, we mark good. We reach into our bait pouch or wherever we have it, kangaroo pouch. And now it's the delivery. If when I reach to get a cookie, my dog comes in front, they're telling me the fundamental element of our dog training program, it's your choice, is now a little blurry. So, I'll just take the cookies and I'll just put them in my hands and put them out in front. Do you really want them? And the dog will then go, Oh yeah, right. I shouldn't be here. Should I? And they'll go back beside me. So, I'll put those cookies back in my pouch, pat the dog on the head and try to rehearse, bringing them out again. Will you stay in position? Yeah, that's good. So, the delivery is about you reaching and moving the cookie to the dog. But the placement, the placement is gold. The placement of the reward, of the reinforcement either builds or tears down the behavior you're trying to create. I can't tell you how many times I've seen people ask their dog to sit. And when their dog sits and they say good, and they go to give them a cookie and the dog brings their paws off the ground and they feed it. My friend Linda Orton Hill was over today. So, I asked her to be my demo of the bad dog trainer. And so, her and her dog, if you're watching this on YouTube, her and her dog are demonstrating some really very bad positioning and bad reinforcement. It's all about the reinforcement process and understanding what did that cookie just reward? What was I rewarding? And so, if you ask your dog to down and you go to give them the cookie and they lift up off their elbows to help facilitate you giving that cookie, the reward process is always in play. The reinforcement process is always in play. You just reinforce for the dog that the down position is with or without the elbows on the ground. All right. So, constantly be aware of the gaps, the gaps between the observation and the recognition. You've got to be observing and recognizing that's the behavior I want. But if the behavior I want just came immediately before a behavior I do not want, I am not going to mark that. I will observe it and see if I can get it for a few more steps. And then I will say, okay, that's good. Coming back after you chase something is great. Holding position is really, really good. All right. And so, the delivery has got to be swift and intentional
13:40
반복되었으면 하는 행동에 대해 강화물을 제공하는 것이죠. 만약 제가 '좋아'라고 말했는데 개가 앞에서 날뛴다면, 그런 식으로 느슨한 줄로 걷는 법을 가르칠 수 있을까요? 결국엔 가능하겠죠. 하지만 그건 개가 앞에 있는 것을 강화하는 꼴이 되며, 이는 느슨한 줄로 걷는 방식과 반대되는 행동입니다. 왜냐하면 개가 내 앞에 있으면 함께 걸을 수 없기 때문입니다. 걸려 넘어질 수 있어요. 위험하죠. 안전하지 않습니다. 그러지 맙시다. 따라서 보상은 옆쪽에서 이루어져야 합니다. 결국 우리가 개에게 무엇을 가르치고 싶은지가 핵심입니다. 이 강화 과정에는 다섯 가지, 때로는 여섯 가지 요소가 있는데, 그 간극을 줄이는 것이 중요합니다. 어떻게 간극을 줄일까요? 아주 뛰어난 훈련 기술을 사용하는 겁니다. 준비하고, 의도를 가지고, 개가 하길 바라는 행동을 만들어내는 것이죠. 개에게 '핫 존(hot zone)'으로 들어가거나 어디 위로 올라가라고 하는 것처럼 간단한 것일 수도 있습니다. 만약 개가 중간까지 가다가 돌아오거나, 중간쯤 가다가 돌아서는데 여러분이 '안 돼, 안 돼, 내 말은 핫 존으로 들어가라는 거였어'라고 말한다면, 이제 여러분은 그 행동에 간극을 만들어 버린 겁니다. 그 행동은 '저기 가서 침대 위로 올라가'가 아닙니다. 그 행동은 '중간까지만 가보고 주인이 정말 진심인지 확인해 봐'가 되어버리죠. 중간까지만 가는 행동을 하는 이유는, 침대로 가는 것이 간식을 잔뜩 들고 있는 거대한 인간 페즈 디스펜서로부터 멀어지는 일이기 때문입니다. 그리고 개 입장에선 간식을 다 가지고 있는 여러분에게서 멀어지고 싶지 않겠죠. 침대에는 간식이 없으니까요. 그런 경우에는 보상의 위치가 침대에서 내려오는 개에게 주어져서는 안 된다는 점을 확실히 해야 합니다. 저는 개가 침대에 머물 수 있도록 침대 위로 간식을 던져줄 겁니다. 그러면 개는 배우게 되죠. '아, 침대로 올라가라고 해서 올라가면 간식을 받는구나'. 페즈 디스펜서는 이제 지금 던지세요. 여러분, 제 말을 듣고 계신 모든 분이 페즈 디스펜서가 뭔지 꼭 아셨으면 좋겠네요. 그건, 있잖아요, 사탕 디스펜서예요. 만약 그게 비건 초콜릿 칩 쿠키 디스펜서라면 훨씬 더 멋질 텐데요. 하지만 딴소리는 여기까지 할게요. 그게 다예요. 강화 과정 말이죠. 의도적으로 하세요. 그러면 장담하건대 여러분은 여러분이 가르치고 있다고 생각하는 것과 강아지가 배우고 있는 것 사이의 간격을 좁히게 될 겁니다. 강화 과정에 정말로 의도적으로 임하게 되면 그 간격은 점점 더 좁혀질 거예요. 그리고 그동안 기억하세요. 여러분이 무엇을 가르치고 있다고 생각하는지는 중요하지 않습니다. 오직 중요한 것은 강아지가 무엇을 보여주는지, 무엇을 배우고 있는지뿐입니다. 제가 대수학 수업을 들었던 것처럼, 당신이 무엇을 가르쳤든 상관없어요. 제가 배운 것은 바로 이것이니까요. 강아지들은 여덟 살이었던 저보다 훨씬 더 잘 반영하죠. 그럼 다음에 'Shaped by Dog'에서 뵙겠습니다. Shaped by Dog.
at creating a reinforcement for the behavior you would like to see repeated. So, if I say good and my dog flips out in front, is it possible to teach loose leash walking like that? Eventually, sure. But you're also reinforcing your dog for being in front, which is counter to loose leash walking because you can't walk with your dog in front of you. You're going to trip. It's dangerous. It's a hazard. Let's not do that. So, reinforcement needs to be on the side. It's all about what is it we want our dog to learn. The five elements or six sometimes in this reinforcement process, getting those gaps and minimizing them. And how do we minimize them? With really good training mechanics. Being prepared, being intentional, and creating the behaviors that you want your dog to do. Now, it could be something as simple as asking your dog to go in their hot zone or go hop it up somewhere. If they go partway and they come back or they go partway and they spin and then you say, no, no, no, I did say go in your hot zone. You now have built in a gap in that behavior. The behavior isn't go over and hop it up in your bed. The behavior is go partway and see if I really mean it. The behavior is go partway because going in the bed is getting further away from the big human Pez dispenser that holds all the cookies. And I don't really want to be far away from you because you got all the cookies. This bed, no cookies over here. We want to make sure that the placement of the reinforcement in that case isn't with the dog coming off the bed. I would throw cookies to the dog so they land in the bed. So, the dog learns, oh, if you ask me to go on the bed and I get in the bed, I'll get cookies. The Pez dispenser can throw now. You guys, I hope everyone listening to me really knows what a Pez dispenser is. It's, you know, it's a candy dispenser. It would be even more awesome if it was vegan chocolate chip cookie dispenser. But I digress. That's it. Reinforcement process. Be intentional. I promise you that you will be tightening that distance between what you think you are training and what your dog is learning. That will become closer and closer when you become really intentional about that reinforcement process. And in the meantime, remember, it doesn't matter what you think you're teaching. The only thing that's important is what that dog is demonstrating, what they're learning. Because just like me taking that algebra class, I don't care what you taught, this is what I've learned. Dogs reflected even better than I did as an eight-year-old. I'll see you next time here on Shaped by Dog.