Shaping Q&A: Non-Reward Markers, Reinforcement Strategies, And Making Training Work For Every Dog
Susan GarrettDogs That · 영상 · 4분
셰이핑 Q&A: 비보상 마커, 강화 전략, 모든 반려견에게 효과적인 훈련 만들기
Shaping Q&A: Non-Reward Markers, Reinforcement Strategies, And Making Training Work For Every Dog
0:00
자, 그럼 첫 번째 질문부터 시작해 보겠습니다. 셰이핑에 관한 여러분의 모든 궁금증을 해결해 드릴게요. 와, 정말 정말 훌륭한 질문들을 많이 보내주셨네요. 여러분의 질문들을 하나씩 살펴보고 있습니다. 자, 첫 번째 질문입니다. 셰이핑을 할 때 '비보상 마커(non-reward marker)'를 사용해야 할까요? 아니요. 여러분의 선행 조건 설정은 여러분이 원하는 반응이 명확하게 나오도록 배치되어 있어야 합니다. 그러니까, 여러분의 도움은 필요 없습니다. 모든 것은 여러분의 계획과 강아지를 준비시키는 과정, 그리고 강아지가 이전에 보상을 받았던 반응들에 달려 있습니다. 강아지가 스스로 반응을 제안(offering responses)하는 것을 편안하게 느끼도록 돕는 좋은 훈련은 무엇일까요? 글쎄요, 아주 간단하게는 위치 기반 강화 마커인 '서치(search)'와 담요를 활용하는 방법이 있습니다. 259회 에피소드에서 말씀드렸듯이, 그 훈련은 모든 강아지가 앞발 타겟팅을 하도록 동기를 부여할 것입니다. 만약 그렇지 않다면, 아마도 보상이 강아지에게 충분히 가치 있지 않기 때문일 가능성이 높습니다. '행동에는 클릭을, 위치에는 보상을'이라는 말을 들어본 적이 있습니다. 동의하시나요? 그건 정적인 행동을 셰이핑하는 것인지, 아니면 움직이는 행동을 강화하는 것인지 구분하는 데 도움이 되는 아주 훌륭한 경험 법칙입니다. 사람들은 종종 강아지가 자신으로부터 멀리 뛰어가길 바라면서도, 클릭하고 보상은 다시 자기 쪽으로 불러서 줍니다. 하지만 저는 뭔가를 던져서 강아지를 보상할 것이고, 사실 클릭도 하지 않을 겁니다. 저는 그냥 '굿(good)' 같은 언어적 마커를 사용하고 강아지에게 보상을 던져줄 것입니다. 자, '행동에는 클릭, 위치에는 보상'이라는 말이 항상 우리가 하는 '서치' 방식과 100% 일치하는 것은 아닙니다. 우리는 위치 때문에 보상하는 것이 아니라, 의도적으로 보상을 함으로써 강아지가 우리가 원하는 행동을 할 수 있도록 리셋(reset)을 만들어내는 것입니다. 셰이핑을 시도해 보지만 제 테크닉이 엉망이라 강아지도 저도 너무 답답해요. 그래서 셰이핑은 초보자를 위한 게 아니라는 말을 들었습니다. 사실인가요? 제 생각에 셰이핑은 모두를 위한 것입니다. 왜냐하면 더 많이 연습할수록 더 좋아지기 때문이죠. 자, 그럼 첫 번째 질문부터 시작해 보겠습니다. 셰이핑에 관한 여러분의 모든 궁금증을 해결해 드릴게요. 와, 정말 정말 훌륭한 질문들을 많이 보내주셨네요. 여러분의 질문들을 하나씩 살펴보고 있습니다. 자, 첫 번째 질문입니다. 셰이핑을 할 때 '비보상 마커(non-reward marker)'를 사용해야 할까요? 아니요. 여러분의 선행 조건 설정은 시작하세요. 만약 여러분과 반려견이 좌절감을 느끼고 있다면, 그것은 선행 조건 배치로 돌아가야 한다는 신호입니다. 저는 앞발 타겟팅으로 돌아가서 간단한 것부터 시작한 다음, 거기서부터 발전시켜 나갈 것입니다. 다시 말하지만, 점진적 접근법(successive approximations)은 행동 블록(behavioral blocks)을 사용하는 셰이핑보다 여러분과 반려견을 더 좌절하게 만들 가능성이 높습니다. 앞서 언급했듯이, 반려견이 좌절하면, 짖기, 낑낑거리기, 발로 긁기 같은 속임수 행동이 나타날 것입니다. 원치 않는 행동들이 그 행동 속에 자리 잡게 될 것입니다. 그리고 그것은 여러분에게 더 큰 좌절감을 줄 것입니다. 수잔, 셰이핑 세션을 시작할 때 별도의 신호를 주시나요? 아니요, 저에게는 그저 반려견 훈련 세션일 뿐이니까요. 그래서,
And here we go with the first one, your questions answered all about shaping. And wow, did we get a lot of really, really great questions. I'm getting to your questions. So, question number one, should I be using a non-reward marker when I'm shaping? No. Your antecedent arrangements are arranged in a way that the response you're looking for is the obvious response. So, no help from you. It is all on your planning, your setting up of the dog and the dog's offered previously reinforced responses. What is a good exercise to help a dog to learn to be okay with offering responses? Well, something as simple as the location-specific reinforcement marker of search and the blanket. As I spoke about in episode number 259, that exercise will get every dog motivated to offer paw targeting. And if it doesn't, chances are your reinforcement isn't high enough value to the dog. I've heard the comment, click for action and reward for position. Do you agree? You know, that is a great little rule of thumb to help differentiate between are you shaping a stationary behavior or are you reinforcing a behavior of motion? So, so often people want their dogs to run away from them and then they click and reward them back at them. But I would click and reward the dog by throwing something and I wouldn't even click, honestly. I would just use a verbal marker like good and throw their reinforcement out there to them. Now, click for action, reward for position isn't always 100% what we do by saying search. We aren't really reinforcing for position, but we're intentionally reinforcing to create a reset that allows a dog to do what we're looking for. I try to shape, but my mechanics suck and my dog and I get very frustrated. So, I've heard shaping isn't for novices. Is this true? So, shaping is for everybody in my opinion, because the more you do it, the better you get at it. If you and your dog are getting frustrated, that comes back to your antecedent arrangements. And I would go back to the paw targeting, start with something simple and then grow from that. Again, successive approximations are probably going to frustrate you and your dog more than shaping with behavioral blocks. And as I mentioned earlier, that when the dog gets frustrated, you'll get cheat behaviors like barking, whining, pawing at you. Things you don't want are going to get built into that behavior. And that's going to be even more frustrating for you. Susan, do you cue shaping sessions? No, because to me, they're just dog training sessions. And so,
2:39
아시다시피, 제가 선행 조건이나 환경을 어떻게 배치했는지가 제 반려견들에게는 우리가 새로운 것을 배우거나 과거에 해왔던 작업을 하려 한다는 아주 큰 신호가 됩니다. 반려견이 계속해서 틀리면 어떻게 하시나요? 저는 세션을 끝내고, 반려견을 핫존(hot zone)으로 점프하게 한 뒤, 그에 대한 보상을 줍니다. 그리고 나서 영상을 확인하며 어떤 부분의 선행 조건 배치가 제가 실제로 반려견이 하길 원했던 것과 어긋났는지 평가할 것입니다. 이제, 특정 트릭으로 셰이핑을 받은 반려견이 계속해서 그 트릭을 반복적으로 시도한다면, 칼라를 잡아 움직이게 함으로써 방해할 수 있습니다. 하지만 다시 말하지만, 저는 반려견이 스스로 방법을 찾아내는 것을 정말 좋아하지만, 누구도 좌절하지 않는 방식으로 하고 싶습니다. 그래서 만약 제가 선행 조건을 재배치하여 올바른 행동을 하기가 매우 분명한 환경을 만들 수 있다면, 그것이 저의 첫 번째 선택이 될 것입니다. 셰이핑은 가르칠 수 있는 모든 행동과 트릭에 효과가 있나요, 아니면 특정 상황에서만 사용하나요? 모든 것에 셰이핑을 적용하시나요? 글쎄요, 셰이핑으로 가르칠 수 없는 것은 떠오르지 않네요. 어떤 것들은 어떻게 셰이핑해야 할지 도저히 상상이 안 가긴 하지만요. 알겠습니다. 댓글을 남겨주세요. 어떤 방법이었는지 알려주세요. 이게 모든 견종에게 효과가 있을까요, 심지어 지능이 낮은 견종에게도요? 윽. 저는 개인적으로 지능이 낮은 견종은 없다고 생각합니다. 저는 어떤 기술에는 다른 견종보다 더 적합한 견종이 있다고 믿습니다. 네, 셰이핑은 모든 대상에게 효과가 있습니다. 개뿐만 아니라 앵무새, 햄스터, 쥐, 그리고 마당에 있는 까마귀나 다람쥐에게도요. 제 말은, 정말 많은 것들이 있다는 거죠. 야생 동물에게 먹이를 주라고 부추기고 싶지는 않지만, 모든 동물은 셰이핑을 통해 배웁니다. 심지어 우리 사람들도요.
you know, how I've arranged the antecedents or my environment is a pretty big cue to my dogs that we're about to learn something new or work on something that we've been working on in the past. What do you do if the dog keeps getting it wrong? I would end the session, have them jump in the hot zone, give them a reinforcement for that. And then I would go to my video and evaluate what part of the antecedent arrangements were in opposition to what I really wanted my dog to do. Now, if you have a dog that's been shaped a certain trick and they just keep offering that trick over and over and over again, you can interrupt it by maybe doing a collar grab and moving them. But again, I really like the dog to figure things out for themselves, but in a way that doesn't frustrate anybody. And so, if I can rearrange the antecedents and create an environment where the correct is super obvious, that would be my first choice. Does shaping work with all behaviors and tricks that can be taught or shaping only for specific things? Are you shaping for everything? So, I can't think of something that it can't be taught with. Some things I just can't fathom how to shape. Okay, leave me a comment. Let me know what it was. Does this work with all breeds, even unintelligent breeds? Yikes. I personally don't think there are unintelligent breeds. I believe that there are breeds that are better suited for some skills than others. And yes, shaping works for all, not just dogs, but parrots and hamsters and rats and your backyard crows and squirrel. I mean, there's so many things. I don't want to encourage you to feed wildlife, but all animals learn by shaping, even yes, us people.