Transfer Of Value In Dog Training: Bubbles To Treats For German Shepherd Retrieve Case Study
Susan GarrettDogs That · 영상 · 8분
반려견 훈련에서의 가치 이전: 저먼 셰퍼드 리트리브 사례 연구를 통한 거품에서 간식으로
Transfer Of Value In Dog Training: Bubbles To Treats For German Shepherd Retrieve Case Study
0:00
안녕하세요, 저는 수잔 개럿입니다. 어느 팟캐스트에서 제가 애틀랜타에서 가르쳤던 저먼 셰퍼드에 대해 언급한 적이 있습니다. 그건 적어도 20년도 더 된 일입니다. 제가 세미나를 진행하고 있었는데, 한 여성이 바로 하루 일과가 시작될 때 나타났습니다. 그녀는 이 개가 제 반려견이라고 말했죠. 벨라라고 부르기로 하죠. 그녀가 말하길, 복종 훈련을 위해 벨라에게 덤벨 물어오기를 가르쳐야 한다고 했습니다. 제 강사는 강제 가져오기(force fetch)를 하고 싶어 했는데, 그건 개가 입을 벌릴 때까지 귀를 꼬집겠다는 뜻이었습니다. 그러면 덤벨을 입에 물리는 거죠. 그것을 부적 강화라고 합니다. 개가 원하는 행동을 하면 고통을 제거해 주는 것이죠. 하지만 이게 제가 어떻게든 그 여성분이 그 행동을 가르치도록 돕지 못했다면 벨라가 겪게 될 운명이었습니다. 제가 그 여성분에게 가장 먼저 물어본 것은, 무엇이 가치 있는 보상인가였어요. 그랬더니 엄청난 뷔페를 열어보더군요. 그녀는 작은 타파웨어 통을 잔뜩 가지고 있었는데, 하나에는 잘게 썬 로스트 비프가, 또 하나에는 잘게 썬 닭고기, 다른 하나에는 잘게 썬 삶은 달걀, 잘게 썬 치즈, 그리고 작은 동결 건조 빙어들도 있었어요. 아직도 기억나는 게, 와, 정말 운 좋은 개라고 생각했죠. 이건 모두 꽤나 엄청난 고가치의 보상 같아 보였습니다. 그래서 제가 말했죠. 좋아요. 그럼 클리커는요? 클리커를 알고 계신가요? 네, 클리커 소리가 나면 보상을 받는다는 걸 알고 있어요. 알겠습니다. 두 번째로 그녀는 개에게 물어오기를 가르치고 싶다고 했죠. 하지만 제가 이전에도 훈련의 감정에 대해 이야기했던 것을 기억하세요. 물어오기 셰이핑 과정의 어떤 부분을 시작하기 전에, 이 개가 훈련에 어떤 감정을 가지고 임할지 알아야 합니다. 그래서, 저는 훈련의 초점을 중요한 것에서 다른 곳으로 돌립니다. 여러분이 꼭 알아야 할 매우 중요한 점인데, 단순히 가르치고 싶은 것부터 시작하지 마세요. 중요한 것에서부터 관심을 돌려보세요. 그래서 어질리티 훈련 중 개에게 도그워크 끝에서 멈추도록 가르친다면, 저는 그것을 그 상황에서 빼내어 훈련에서 계단 끝에서 멈추도록 가르치고 높은 수준의 반응을 얻을 때까지 기다릴 겁니다. 세상에, 이건 정말 좋네요. 그다음에는 중요한 것으로 넘어가겠습니다. 그래서 우리는 중요한 것에서 훈련을 떼어놓을 것이고, 거기에는 타겟 스틱이 포함되었습니다. 그래서 저는 개가 자기 코를 타겟 스틱에 대기만을 바랍니다. 클릭하고 보상을 줄 겁니다. 자, 그래서 우리는 이걸 그룹으로 하고 있습니다. 그녀가 타겟을 제시했고, 개가 그것을 건드렸습니다. 그런데 만약 곰돌이 푸의 이요르를 의인화한 개가 있다면 바로 이 녀석이었을 겁니다. 그녀가 로스트비프 한 조각을 주었는데 아주 천천히 씹어 먹더군요. 그다음 그녀가 다시 타겟을 제시했습니다. 그러자 개는 그녀를 쳐다보는 둥 마는 둥 하더니 아주 느릿느릿 시선을 돌리고는 의도적으로 타겟을 건드렸습니다. 그녀가 클릭했죠. 개가 다시 그녀를 쳐다보자
Hi, I'm Susan Garrett. In a podcast, I mentioned a German Shepherd that I was teaching in Atlanta. And this was at least or more than 20 years ago. I was giving a seminar. This woman showed up right at the beginning of the day. She said, this is my dog. Let's call her Bella. She said, I need to teach Bella how to retrieve the dumbbell for obedience. And my instructor said that they would like to do a force fetch, which means they were going to ear pinch the dog until the dog opened its mouth. Then they would put the dumbbell in. And that's called negative reinforcement. They remove the pain when the dog does what they want. But this was what was going to be Bella's fate if I couldn't help this woman somehow do that behavior. The first thing I said to this lady is, what is of value? And she opened this amazing buffet. She had all these little Tupperware containers, chopped up roast beef in this one, chopped up chicken in this one, a chopped up hard boiled egg in this one, chopped up cheese, and some little freeze-dried minnows. I remember it like it was like, wow, this is one lucky dog. These seem like all pretty outrageously high value rewards. And I said, okay. And what about a clicker? Does you know a clicker? Yeah. He knows a clicker means it's going to get a reward. All right. Number two, she said, I want to teach him to retrieve. But remember I talked about the emotion of training before I go to what's important, any part of the shaping process for the retrieve, I need to know what kind of emotion this dog is going to be bringing into the training. And so, I take my training away from what's important. Super important for you to know, you don't just start training what you want to train, take it away from what's important. So, if I was teaching in agility, a dog to stop at the end of a dog walk, I would take that away from training and teach them to stop at the end of a set of stairs until I get a high level of, oh my gosh, this is so good. Then I would take it into what's important. So, we're going to take the training away from what's important and that involved a target stick. So, I want the dog to just touch his nose to the target stick. We're going to click and reward that. Okay. So, we're doing this in a group. She presented the target, the dog touched it. And if ever there was a dog personifying Eeyore from Winnie the Pooh, he did it. She gave him a piece of roast beef and he chewed it up so slowly. Then she presented it. And then the dog kind of was looking at her, looked away very slowly and deliberately touched the target. She clicked. He looked back at her,
2:37
그녀가 닭고기를 주었습니다. 그녀가 아무리 엄청난 고가치 보상을 주어도 반응은 똑같았습니다. 그래서 세션이 끝날 때, 우리는 개들을 들여보내고 제가 말했습니다. 저는 확신합니다. 이게 당신의 엄청난 고가치 보상이라면, 내일 당신의 반려견은 강사에게 귀를 꼬집힐 겁니다. 왜냐하면 개가 강화물에 전혀 관심이 없는 상태에서 가져오기(retrieve)를 형성(shape)할 수는 없기 때문입니다. 왜냐하면 셰이핑(shaping)은 가치의 전달에 관한 것이기 때문입니다. 개가 당신이 사용하는 강화물에 대해 느끼는 가치가 당신이 훈련하고 있는 행동으로 옮겨가는 것이죠. 그래서 물었습니다. 터그 놀이감을 좋아하나요? 아니요. 테니스 공은요? 잘 모르겠어요. 자, 당신의 반려견이 본 것, 한 것, 먹은 것, 혹은 냄새 맡은 것 중에 귀가 쫑긋해지고, 발끝으로 서게 되며, 꼬리가 살랑거리기 시작하는 게 있나요? 아, 있어요. 있어요. 그녀가 말하길, 아이들이 뒷마당에서 비눗방울 놀이를 할 때, 그 있잖아요, 비눗방울을 불면 방울이 나오는 거요. 제가 말했죠. 좋아요, 다이소 가서 비눗방울 좀 사 오세요. 비눗방울을 가져왔습니다. 자, 핵심은 여기 있습니다. 이제 우리의 최고의 강화물을 찾은 거죠. 분할을 완료했습니다. 마커도 준비됐고요. 이제 환경을 조작해야 합니다. 그래서 저는 비눗방울을 의자 위에 두고 말했죠. 자, 이제 여기서 행동을 형성해 볼 거예요. 이 개가 비눗방울에 어떻게 반응하는지 보고 싶었거든요. 그래서 타겟을 제시했습니다. 개가 타겟을 건드렸을 때, 제가 말했어요. 이제 강화 과정을 만들 겁니다. 그래서 그가 타겟을 건드리면 '비눗방울'이라고 말해주세요. 이 개에게는 아무 의미 없는 단어죠. 비눗방울이 뭔지 몰랐으니까요. 개가 그녀를 돌아봤습니다. 그녀는 25피트 떨어진 의자까지 달려가 비눗방울을 집어 들고는, 비눗방울을 불기 시작했어요. 이 개는 제가 본 적 없는 모습으로 돌변했습니다. 그녀의 머리 위로 뛰어넘고 악어처럼 입을 딱딱거리며 비눗방울을 낚아채려 했죠. 그녀가 비눗방울을 다시 준비하는 동안 개의 꼬리는 헬리콥터처럼 돌아갔습니다. 세상에, 비눗방울이잖아. 그래서 비눗방울을 두 번 불었습니다. 거기까지. 다시 제자리에 두고, 이쪽으로 돌아오세요. 대부분의 사람들은 이렇게 말하겠죠, 세상에, 개가 정말 좋아하는 걸 찾았어. 남은 평생 비눗방울로만 훈련해야겠어. 여기에는 두 가지, 사실 세 가지 문제가 있습니다. 첫째, 온 집안에 끈적거리는 비눗물 찌꺼기가 묻게 됩니다. 누가 그걸 원하겠어요? 둘째, 결국 개도 지치게 됩니다. 방금은 꽤나 기운 빠지는 소동이었거든요. 그리고 셋째, 항상 비눗방울을 챙겨 다닌다는 건 얼마나 불편한가요? 그래서 우리는 이 엄청난 고가치의 강화제 가치를
she gave him chicken. Same response, no matter what kind of outrageous high value reward she gave. So, at the end of the session, we put the dogs away and I said, I'm very confident. If this is your outrageous high value reward, your dog will be getting an ear pinch from your instructor tomorrow. Because there is no way we are going to shape a retrieve when he really doesn't care about the reinforcement. Because shaping is about the transfer of value. The value the dog has for the reinforcement you're using goes into the behavior you're training. So, I said, does he like tug toys? No. Does he like tennis balls? Not really. Okay. Is there anything that your dog has ever seen or done or eaten or smelled that makes his ears come up and he's on the balls of his feet and the tail just starts going? Oh yeah. Oh yeah. She said, when the kids play with the bubbles in the backyard, you know, those soap bubbles you blow and the bubbles come out. I said, all right, go get me some dollar store bubbles. We got the bubbles. Now, here's the key. We've got our best reinforcement. We've got a split. We've got our marker. And now we've got to manipulate the environment. So, I put the bubbles on a chair and I said, we are going to go and shape behavior over here. And I said, I just want to see what this dog's like with bubbles. So, we presented the target. And when he touched it, I said, we're going to create a reinforcement process. So, when he touches it, I want you to say the word bubbles, which is meaningless to this dog. He didn't know they were bubbles. He turned to her. She started running the 25 feet to the chair, picked up the bubbles and doing some streams. This dog turned into something the likes I've never seen before. He was jumping over the woman's head, snapping like an alligator, grabbing those bubbles. And in between when she was loading, his tail was going like a helicopter. Oh my gosh, there's bubbles. So, two streams of bubbles. That's it. Put it back. Come on back over here. Now, what most people would say, oh my gosh, we found something that the dog loves. We're just going to train with bubbles for the rest of his life. There's two problems. Actually, probably three problems with that. Number one, you get that icky soap scum all over your house. Who wants that? Number two, eventually the dog's getting tired. Like that was a pretty exhausting exhibition. And number three, how inconvenient is it to always be packing bubbles? So, what we have to do is take the value of the high outrageous reinforcer
5:04
우리의 다른 강화제들로 옮겨야 합니다. 이를 가치 전이라고 하죠. 과정은 이렇습니다. 그녀가 타겟을 제시했습니다. 개가 그것을 건드리자 클릭하고 로스트 비프 한 조각을 주었죠. 개는 천천히 씹어 먹었습니다. 두 번째로 타겟을 제시했습니다. '비눗방울'이라고 말하고 바닥을 가로질러 달려가 다시 한번 비눗방울을 불어주었습니다. 두 번 만에 이 개는 '그래, 그 공을 건드리면 되겠구나'라고 깨달았죠. 그래서 그녀는 시작했습니다. 공을 치고 나면 벨라는 스스로 버블을 향해 달려가려고 했어요. 하지만 아무도 버블이라고 말하지 않았죠. 마커가 없으면 강화도 없습니다. 그렇기 때문에 강아지가 스스로 강화물을 그냥 낚아채게 두지 않는 것이 정말 중요합니다. 이리 와, 벨라. 녀석이 다시 공을 쳤고 어느 정도 움직였지만, 클릭은 간식을 받는다는 의미예요. 그러고 나서 다시 클릭과 간식을 주고, 또 한 번 터치를 하고 버블을 줬어요. 우리는 달려가서 버블을 더 놀았죠. 그렇게 세션이 끝났습니다. 세션의 나머지 부분에서는 코 터치 다섯 번에 간식을 한 번 주고 버블을 한 번 주는 방식으로 시작했어요. 그다음엔 아마 코 터치 한 번에 바로 버블을 줬죠. 그러다 코 터치 10번을 한 뒤에 버블을 주기도 했어요. 그러고 나서 세 번 하고 다시 버블을 주고, 그다음엔 20번까지 갈 수도 있었겠죠. 그리고 맞춰보세요. 하루가 끝날 무렵엔 벨라가 우리가 원했다면 나무로 무엇이든 깎아왔을 것 같아요. 녀석은 클릭과 간식 받는 일에 정말 신이 났는데, 그것을 많이 받을수록 버블에 더 가까워진다는 것을 알았기 때문이죠. 곧 버블은 세션당 한 번, 혹은 일주일에 한 번 정도만 나와도 충분할 테고, 결국엔 굳이 사용하지 않아도 되는 특별한 강화물이 될 거예요. 간식이 그저 참을 만한 수준에서 아주 좋은 것으로 바뀌었기 때문이죠. 그것이 버블에 더 가까워질 기회를 의미하게 되었으니까요. 그러니까 틀을 깨는 사고방식이란 모든 강아지에게 선형적인 경로가 존재하지 않는다는 뜻입니다. 강화 과정을 살펴보세요. 강아지에게 무엇이 열광적인 보상인지 확인하는 겁니다. 그리고 어떻게 흥분을 키울 수 있을지 고민해 보세요. 버블을 25피트 떨어진 곳에 둠으로써 우리는 강아지가 달리게 하고, 생리적 반응을 바꾸며, 나와 함께 일하는 것이 즐겁다는 게임 속의 게임을 만들어낼 수 있었습니다. 나와 함께 움직이는 것이 즐거운 일이죠. 강아지가 물체를 가져오게 하거나 뒷다리를 가리키게 만드는 셰이핑 등, 무엇이든 가능했을 겁니다. 이제 필요한 것은 훌륭한 훈련 계획과 그 방법을 아는 것뿐입니다. 당신의 환경은 어떤지, 방해 요소는 무엇인지, 그리고 만들어내고자 하는 행동은 무엇인지 확인해보세요. 그리고 가장 적절한 강화물은 무엇인지 생각해보세요. 왜냐하면 지금 단계에서는 사용할 수 있는 강화물이 아주 많기 때문입니다.
and put it into our other reinforcers, AKA the transfer of value. And so, this is the process. She presented the target. The dog touched it, click, give a piece of roast beef. And he did his slow chew. Presented the target the second time. We say bubbles, run across the floor, another stream of bubbles. It only took two before this dog said, yeah, I'll touch that ball. So, she started going like hit the ball and then she was turning to sprint to the bubbles on her own. But no one said bubbles. Without a marker, there's no reinforcement. That's why it's just super important that we're not letting a dog just grab reinforcement on their own. Come on back here, Bella. She hit it again and she was kind of moved, but click means you're getting food. Then we did a click and food again, and then another touch and bubbles. Off we go running and more bubbles. So, that was the end of the session. The rest of the sessions is we started going five nose touches with food and one bubbles. And then maybe we do one nose touch and bubbles. And then we went to 10 nose touches before we got the bubbles. And then we might've done three and then one again with bubbles. And then we might go to 20. And guess what? By the end of the day, I think Bella would have whittled us something out of wood if we wanted. She was so excited about the click and getting food because the more she got to that, the closer she got to bubbles. Very soon the bubbles would only have to come out, I don't know, once a session or once every week, and eventually it would be just a special reinforcement that you wouldn't have to use because food had gone from barely tolerable to really good because they meant the chance to get closer to bubbles. So, outside the box thinking means there isn't a linear path for all dogs. You look at the reinforcement process. You look at what is outrageous to the dog. And you look at how can you grow excitement? By putting the bubbles 25 feet away, we got the dog to run, changing the physiology, building in a game within a game that working with me is exciting. Moving with me is exciting. Shaping that dog to retrieve an object or point at her back leg, anything would have been possible. Now, all that would be required is a great training plan and knowing what's your environment, what are the distractions, what is the behavior you want to create, and what's the most appropriate reinforcement because you have so many at this point.