Training Without Conflict Podcast Episode Fifteen: Dr. Michael Perone
Ivan BalabanovIvan Balabanov · 영상 · 1시간 40분
행동 분석의 원리는 기초 실험 연구와 반려견 훈련이라는 서로 다른 분야를 관통하며, 정적 강화와 부적 강화에 대한 이분법적 사고를 지양할 것을 제안한다. 부적 강화는 일상적인 위험 회피나 사회적 자선 활동 등 긍정적 결과를 도출하는 자연스러운 기제이며, 혐오 자극이 포함되었다는 이유만으로 배척 대상이 되어서는 안 된다. 쥐를 대상으로 한 기초 연구는 혐오 자극을 사용하더라도 동물이 건강하고 편안하게 행동을 수행할 수 있음을 보여준다. 처벌의 경우, 효과적인 억제를 위해 반드시 강한 자극이 필요하지는 않으며 아주 약한 자극으로도 충분히 목적을 달성할 수 있다. 학습의 맥락과 강화 스케줄은 행동의 결과에 지대한 영향을 미치며, 특정 보상이 오히려 내재적 동기를 저해하거나 상황에 따라 처벌제로 작용할 수 있음을 인지해야 한다. 유전적 구성은 환경의 산물로서 환경과의 끊임없는 상호작용 속에서 발현되므로, 생물학적 프로그래밍을 억제하기보다는 자연스러운 행동 양식을 수용하고 더 나은 방향으로 유도하는 것이 훈련의 핵심이다. 오류 없는 학습과 시행착오 학습은 각기 다른 목적과 장단점을 가지며, 숙련된 신체 수행 능력을 습득하는 데는 시행착오가 필수적인 역할을 한다. 타임아웃은 그 자체로 부적 처벌에 해당하며, 환경이 풍부할수록 더욱 강력한 처벌 기제로 작동한다.
Training Without Conflict 팟캐스트 에피소드 15: Dr. Michael Perone
Training Without Conflict Podcast Episode Fifteen: Dr. Michael Perone
0:04
[음악] 네, 여러분 안녕하세요, 여기는 트레이닝 위드아웃 컨플릭트(Training Without Conflict) 팟캐스트 제15회입니다. 오늘 게스트는 웨스트 버지니아 대학교 심리학과 교수이신 마이클 페로니 박사님입니다. 박사님은 30년 넘게 조작적 행동의 실험적 분석 분야에서 활발히 활동해 오셨으며, 조건 강화와 회피, 그리고 여러 연구 프로그램으로도 잘 알려져 계십니다. 실제로 그 업적이 매우 방대해서 저희가 모든 자격을 팟캐스트 아래에 기재하도록 하겠습니다. 박사님은 행동 분석 분야의 주요 저널에서 핵심 편집 위원직을 맡으셨고, 행동 심리 및 인지 과학 연합회에서 행동 분석을 대표하셨으며, 수많은 위원회에서 활동하셨습니다. 소개할 내용이 상당히 많이 누락된 것 같지만, 박사님께서 이 분야에서 워낙 오랫동안 활동하셨고 정말 훌륭한 업적을 많이 남기셨기 때문입니다. 그러니 잠시 직접 자기소개를 부탁드리겠습니다. 제가 페로니 박사님을 모신 이유는 제가 우연히 접하게 된 몇 가지 연구 때문입니다. 그 연구들이 매우 흥미로웠고, 반려견 훈련, 동물 학습, 인간 학습 심리학 등이 어떤 면에서는 서로 겹치는 부분이 있다고 생각합니다. 우리는 한 분야에서 배운 것을 다른 분야에 적용하고, 다시 또 상호 교류하며 배우기 때문입니다. 배울 점이 정말 많죠. 네, 서부 버지니아 대학교 교수님이신 마이클 페로니 박사님을 다시 한번 환영합니다. 네, 반갑습니다. 오늘 정말 기대됩니다. 저도 마찬가지입니다. 감사합니다. 교환의 그리고 어 원래 제 정말 무엇이 저를 입문하게 한 것은 한 연구인데 2003년 당시 정적 강화의 부정적 영향에 관한 것이었고 그리고 당신이 수행했던 또 다른 연구가 하나 있었죠 함께 음 동료 몇 분들과 함께한 타임아웃을 통한 부적 강화와 회피에 관한 연구였고 기본적으로 저는 우리가 몇 가지 흥미로운 대화를 나눌 수 있을 거라 생각해요 당신의 연구 결과에 대해 많이 듣고 싶고 비록 지금은 당신이 분야를 옮겨서 제가 알기로는 실험실에서 동물들과는 연구를 많이 하지 않으시지만 저는 반려견 훈련사로서 우리가 얻어갈 수 있는 교훈이 있다고 생각해요 그래서 음 본인에 대해 조금만 더 말씀해 주시겠어요 제가 좀 서둘러 소개하긴 했지만 음 글쎄요 저는 늙은이라서 만약 제 약력의 모든 세부 사항을 다 읊는다면 꽤 지루할 겁니다 음 저는 이곳 웨스트버지니아 대학교 교수이고 1984년부터 재직해 왔습니다 음 저는 제 경력 전반에 걸쳐 사람과 동물을 대상으로 기초 조작적 연구를 병행해 왔고 최근 몇 년 동안은 주로 쥐와 비둘기 같은 동물을 포함한 연구에 집중해 왔으며 제 학생들과 저는 다양한 기초 과정들을 연구합니다 정적 강화와 부적 강화 둘 다요 네, 그래서 당신이 그 부분을 실제로 지금도 여전히 어
[Music] all right hello everybody this is training without conflict podcast number 15 and today's guest is Dr Michael peronic who is a professor in the department of psychology at Western Virginia West Virginia University um has been actively involved in experimental analysis of operant behavior for over 30 years also well known for his uh programmatic research on condition reinforcement avoidance and um it's actually quite extensive and and we will list all of the um credentials uh below the podcast um my has been appointed to key editorial position for major journals in Behavior Analysis represented behavior analysis on the Federation of behavior psychology psychological and cognitive sciences and served on numerous committees I I know I'm missing quite a bit of the introduction because again you've been on the field for a very long time and very well accomplished so again I will let you introduce yourself for a moment the reason I invited um Dr perrona is because of few of the studies that I I stumbled upon and they were I found them very interesting and I think you know um dog training animal learning human learning psychology they in some ways overlap and we learn from one and we use it to the other and back and forth there is a lot of exchange and uh so that originally my like what really introduced me to is one one study that was the negative effects of positive reinforcement back in 2003 and and then there was another one that you did in collaboration with um a few other colleagues of yours on the negative reinforcement by time out and avoidance and and basically I I believe that we may have some interesting conversation of of you know I I want to hear a lot of your findings and even though at the moment you you you you've moved on and not so much work from what I understand with with animals in the laboratory but I believe that as dog trainers we we have some takeaways to you know we benefit from so um just just give me a little bit of more about yourself I know I kind of rushed through everything but um well I'm an old man so if you if you went through all my uh all the little details of my biography it would be pretty boring um well uh I'm a professor here at West Virginia University I've been here since 1984. um I uh I uh over my career I've done a mix of uh basic Opera research with humans and with animals uh in the last few years I've concentrated mainly on studies involving animals rats and pigeons and my students and I study a variety of basic processes uh involving both positive and negative reinforcement yeah so I that's great that you you're actually still currently uh doing some
3:51
동물들과 무언가를 하고 있다는 건 정말 정말 멋진 일이에요. 어디서부터 시작해야 할지 잘 모르겠지만 우리 그 부분에 대해 조금 다뤄볼 수 있을 것 같아요 정적 강화의 부정적인 부작용에 대해서요 보통 적어도 반려견 훈련 업계에서는 우리 반려견 훈련사들은 굉장히 음 저는 확신할 수 없지만 우리는 꽤 나뉘어 있어요 음, 매우 중립적인 입장의 훈련사들도 있지만 두 가지 극단적인 견해가 있는데 엄격하게 말해서 정적 강화가 정답이고 다른 어떤 방식보다 훨씬 더 우월하다는 쪽과 그리고 제가 가진 그러니까 당신도 알다시피 우리가 정적 강화에 대해 이야기할 때 물론 그것에는 어떤 단점들도 있고 때로는 다른 접근 방식이 필요할 때도 있다는 거죠 음, 하지만 당신의 일반적인 입장은 무엇인가요 당신은 어떤 한 가지 접근법이나 강화 방식이 다른 것보다 우월하다고 믿으시나요? 물론 항상 존재하듯이 거슬러 올라가서 예를 들어 1975년에 마이클이 연구했던 것처럼 그는 강화에 부적 강화와 정적 강화가 모두 포함된다는 개념을 도입했고 우리는 그런 이분법적인 생각을 버리고 나누어야 한다는 거죠 음, 이 부분에 대해 조금 말씀해 주시겠어요? 이것은 아주 반려견 훈련사 청중들에게 매우 흥미로울 만한 주제입니다. 네, 그럼요. 음, 당신이 몇 년 전에 제가 쓴 논문을 참고하셨었죠. '정적 강화의 부정적 효과'라는 제목의 논문인데, 그 논문은 사실 국제 행동 분석 학회(The Association for Behavior Analysis International)에서 제가 했던 회장 연설이었습니다. 음, 그 논문의 목적은 정적 강화가 나쁘고 부적 강화가 좋다고 주장하려는 것이 아니었습니다. 오히려 그 흔한 가정, 즉 정적 강화는 항상 좋고 부적 강화는 항상 나쁘다는 그 가정, 그 가정이 절대적으로 잘못되었다는 것을 지적하기 위함이었습니다. 제가 그 논문에서 보여주고자 했던 것은 첫째, 음, 때로는 정적 강화와 부적 강화를 구별하기가 어렵다는 점입니다. 그리고 그것은 일부분 방금 언급하신 논문에서 잭 마이클(Jack Michael)이 제기한 주장 때문이기도 합니다. 음, 그리고 또 다른 이유는 부적 강화가 자연 환경에서 불가피한 요소이며, 그것이 아주 많은 긍정적인 효과를 지니고 있다는 점입니다. 그러니 그저 마치 부적 강화가 금기시되어야 할 것인 양, 마치 무슨 수를 써서라도 피해야 할 것처럼 치부해 버리는 것은 말이 되지 않습니다. 아시다시피, 모든 종류의 훈련을 할 수 있죠. 부적응적이고 자기 파괴적인 행동들 정적 강화로 말이죠, 맞아요 초보 반려견 훈련사들이 나쁜 행동을 하도록 개를 훈련하는 모습을 분명 보신 적 있을 겁니다. 사실 이런 말이 있죠. 모든 나쁜 행동들은 대부분
stuff with with animals that's that's very very cool I am not even sure where to start but we could like go over a little bit about that the negative side effects of positive reinforcement normally at least in the dog training world we as dog trainers are very like I'm I'm not sure but we are quite divided um there is a very there is trainers that are in the middle of things but there is two extremes to where we are talking about strictly positive reinforcement strictly is the way to go and is far more Superior than any other approach and I have like like you know when we talk about positive reinforcement of course there is there's some some downfalls and there is sometimes need for different approaches um but what what is your just general stand on do you believe that something that one's approach one par one reinforcement is superior than the other and of course there is always the like even if we go back to like I don't know what's at 1975 Michael studied to work he introduced the idea that reinforcement includes negative and positive and we should kind of abandon that idea and divide them um talk to me a little bit about this this is something that will be very interesting to the dog trainer audiences okay well um you know you you you've referenced uh the paper I wrote a number of years ago called negative effects of positive reinforcement which uh was actually my presidential address for The Association for Behavior Analysis International um and the purpose of that wasn't to argue that positive reinforcement is bad and negative reinforcement is good it was rather to point out that the common assumption which is that positive reinforcement is always good and negative reinforcement is always bad that that that that that assumption is is uh absolutely Incorrect and what I tried to do in the paper is show that first of all uh sometimes it's hard to tell the difference between positive and negative reinforcement and that's uh partly because of the arguments that Jack Michael made in the paper that you just mentioned um and it's it's partly because uh negative reinforcement is an inevitable uh element of the natural environment and it has many many positive effects so to uh just you know just dispose of negative reinforcement as if it were anathema as if it were you know something that that should be avoided at all costs just doesn't make sense you know you can you can trade up all kinds of maladaptive and self-defeating behaviors with positive reinforcement right right I'm sure you've you've seen novice dog trainers train dogs to do bad things yes in fact there is a saying that all bad behaviors have been most
7:28
언젠가는 정적 강화되었을 가능성이 높다는 말이죠. 맞아요, 정확해요. 정확합니다. 음, 제가 제안하려는 건 아니에요. 원하는 것을 하도록 개를 훈련하는 방법이 무엇인지 간에, 발에 전기를 가하고 옳은 행동을 할 때 전기를 끄라는 건 아니에요. 그건 야만적인 방식이겠죠. 하지만 제가 말씀드리고 싶은 건, 부적 강화는 당신이 길을 걸을 때 살피지 않고 연석 아래로 내려가지 않게 막아주는 역할을 한다는 겁니다. 이제 생각해 보세요. 도심 속에서 길을 걸을 때 차들이 쌩쌩 지나다니죠. 시속 25~30마일 정도일지 몰라도, 만약 그 앞을 가로막는다면 죽거나, 아니면 더 끔찍한 일을 당할 수도 있습니다. 그런데도 우리는 매일 거리를 오가죠. 치명적일 수 있는 거대한 기계들로부터 불과 몇 피트 떨어진 곳에서요. 그렇지만 우리는 전혀 걱정하지 않아요. 두려워하지도, 당황하지도, 불안해하지도 않죠. 왜 그럴까요? 우리가 불과 몇 피트 옆에 도사린 위험을 피하는 데 능숙하기 때문입니다. 우리는 또 다른 부적 강화의 우발적 상황을 겪고 있고, 그것이 매우 효과적이기 때문입니다. 그것들이 우리의 행동을 통제하고 전혀 우리를 귀찮게 하지 않습니다. 아마 우리가 과거를 충분히 되돌아볼 수 있다면 어린 시절, 아마도 부모님이 우리에게 길을 건너는 방법과 양옆을 살피는 방법 등을 가르치실 때 우리는 아마 약간 불안했을지도 모릅니다. 그러한 행동이 습득되는 과정에서 말이죠. 하지만 10살이나 11살 정도가 되면, 아마 그보다 더 일찍 우리는 전혀 두려움 없이 도시 곳곳을 다닐 수 있게 됩니다. 이것이 부적 강화입니다. 부적 강화는 우리가 추울 때 코트 단추를 채우는 이유입니다. 네, 부적 강화로 인해 일어나는 좋은 일들이 아주 많습니다. 그렇기에 부적 강화라는 것 그 자체를 나쁜 것이라고 말하는 것은 옳지 않습니다. 혹시 이런 일을 겪으시나요? 제가 앞서 언급했듯이 반려견 훈련 분야에서는 이것이 문제입니다. 우리는 꽤 분열되어 있고 어떤 이유에서인지 부적 강화, 정적 처벌, 혹은 혐오 자극의 사용이라는 단어를 언급하기만 해도 즉시 끔찍한 일이 일어난다는 것과 연결 지어 반응하게 됩니다. 혹시 이런 상황을 겪으시나요? 수업을 가르치실 때도 이런 문제를 다루셔야 하나요? 아니면 어떻게 하시는지 궁금합니다. 네, 물론이죠. 저는 확실히 그런 일을 겪고 있고, 어떤 상황인지 이해합니다. 질문해주셔서 감사합니다. 네, 정말 그렇습니다. 제가 발견한 해결책은 바로 학생들에게 이 분야의 기초 연구를 보여주고, 가능하면 그들이 이 분야의 연구에 직접 참여하게 하는 것입니다. 음, 수년간 저는 전기 충격 회피와 관련된 많은 연구를 수행해 왔습니다.
likely positively reinforced some at some point of time right exactly exactly exactly um now you know I'm not suggesting that the way to train a dog to do whatever it is that you want to do is to Electrify its feet and turn off the electricity when it does something right that would be that would be barbaric but I will point out that uh uh negative reinforcement is what keeps you from stepping off the curb without looking when you're walking down the street now consider uh when you walk down the street in an urban environment there are cars whizzing back and forth and maybe they're only going 25 or 30 miles an hour but if you step in front of one of them you're going to be dead or or maybe worse and uh and yet we walk down up and down the street every day within a few feet of these massive machines that could be deadly and we don't worry about it at all we're not scared we're not upset we're not anxious and why not because we are proficient in avoiding the danger that lurks just a few feet away we have another negative reinforcement contingencies are so effective that they control our behavior and they don't bother us in the least now maybe when if we could look back far enough into our childhood maybe when our parents were teaching us how to cross the street and to look both ways and all that maybe we were a little anxious as as that kind of behavior was being acquired but by the time one is I don't know 10 or 11 years old probably earlier than that we negotiate uh uh travel throughout cities without being scared at all and this is negative reinforcement negative reinforcement is why we uh button our coats when it's cold yes there's a there's a lot of there's a lot of good things that happen because of negative reinforcement and to suggest that negative reinforcement uh in itself is a bad thing would just be incorrect do you find that I mean like I as I mentioned in the dog training world this is this is a problem it's it's you know we're quite divided and and for whatever reason associating just just throwing the word negative reinforcement or positive punishment or any use of aversive immediately uh resonates with something horrible happening is that do you find anything like this do you have to deal with something like this when you when you even when you teach your in your classes or how yes oh yes uh I certainly do and uh what I find is that uh the the antidote is showing students the basic research in these areas and uh if possible getting the getting them to participate in the research in these areas um over the years I've uh I've had I've done lots of studies involving um avoidance of electric shock
11:02
매일같이 회피 스케줄에 노출된 쥐라면, 수개월 동안 계속해서 말이죠. 이것들은 정상 상태 실험인데, 쥐가 겁이 많거나 다루기 힘들 것이라고 생각할지도 모르지만 사실은 매우 친화적이고 유순합니다. 음, 아마도 그 이유 중 큰 부분은 그들이 빠르게 회피에 능숙해지기 때문일 겁니다. 그리고 제 실험실에서 쥐는 때때로 한두 시간 동안 단 한 번의 충격도 받지 않기도 합니다. 사실 그게 이 실험을 매혹적으로 만드는 이유죠, 아이반. 만약 당신이 제 실험실의 오퍼런트 챔버(조작 조건 형성 상자) 내부를 본다면 레버를 누르고 있는 쥐를 볼 수 있을 텐데 아무 일도 일어나지 않죠. 쥐가 레버를 누르기 전에도 아무 일 없고, 쥐가 레버를 누른 후에도 아무 일 없습니다. 그렇게 한두 시간 동안 계속 행동을 하니까 어떻게 이게 가능한지 의문이 들 겁니다. 음, 아무 일도 일어나지 않으니 소거(extinction) 스케줄처럼 보이지만, 사실은 충격 연기(shock postponement) 스케줄인 거죠. 그리고 어떻게든 동물의 행동이 이 스케줄의 통제 하에 들어오게 됩니다. 그 쥐 자체는 아주 침착하죠, 그렇죠? 톰은 꽤 느긋하고, 어, 조금도 겁이 없어요. 그리고 실험이 끝나면 다루기가 아주 쉬워요. 음, 여전히 우리는 알지 못해요. 우리는 여전히 의견 일치를 보지 못했어요. 무엇이 이 행동을 강화하는지에 대해서 말이죠. 분명히 강화 요인은 불분명해요. 우리는 일정이 무엇을 하는지 알고, 우리는 그 일정이 무엇인지 알고 있지만, 그 일정이 활성화하는 강화 메커니즘은 알지 못해요. 그게 이 연구를 매우 흥미롭게 만드는 점이죠. 하지만 제가 나쁜 사람은 아니에요. 이걸 연구한다고 해서요. 전 쥐들에게 못되게 굴지 않아요. 저는 쥐들을 사랑하고, 쥐들의 건강과 복지는 우리에게 결정적으로 중요해요. 왜냐하면 우리의 실험은 종종 동물의 수명 대부분 동안 지속되니까요. 그래서 우리는 쥐들이 건강하고 튼튼하기를 바라죠. 당연히, 우리가 어떤 형태의 혐오 자극을 사용할 때, 상황이 잘못될 수도 있죠. 물론이죠. 그리고 보통 무엇이 잘못되어야 이런 일이 일어날까요? 글쎄요, 음, 우리는 조건부와 비조건부에 대해 이야기하는데, 이게 매우 중요한 문제죠. 맞아요. 음, 하지만 당신이 쥐들과 작업할 때 무엇을 발견하나요? 글쎄요, 당신도 알다시피 음, 저는 말씀하신 대로 부적 강화와 정적 처벌 사이의 신중한 구분을 유지하는 것이 중요하다고 생각해요. 그 둘은 둘 다 혐오적 통제의 범주에 속한다는 점에서 통일되어 있습니다. 음 제 추측으로는요, 제 짐작입니다만 많은 개 훈련이 아마도 부적 강화보다는 처벌을 포함할 가능성이 더 높을 겁니다 음, 그리고 처벌에 관해서 실험실에서 매우 명확하게 나타나는 사실은 바로 아주 약한 자극도 효과적인 처벌인자로 작용할 수 있다는 점입니다 예를 하나 들어보겠습니다
and one might think that a rat that is exposed to an avoidance schedule day in and day out month after month these are steady state experiments would be a skittish or difficult animal to work with but in fact they're completely friendly and docile they um probably largely because they they vary quickly become proficient avoiders and um and in my laboratory a rat can sometimes go an hour or two without getting a single shock in fact that's what makes it so fascinating Ivan is if you were to look inside an operant chamber in my laboratory you might see a rat that's pressing a lever and nothing's happening nothing happens before the rat presses a lever nothing happens after the rat presses the lever and they go on and they do this for an hour or two at a time and you'd say how is this possible uh nothing's happening it appears to be extinct that the schedule's Extinction but it's not it's a shock postponement schedule and somehow the animal animal's behavior comes under the control of this schedule the rat itself is quite calm right the question Tom pretty laid back uh not the least bit skittish and when the experiment's over very easy to handle um it's uh it's it we still don't know we still don't have a uh a consensus about what is reinforcing this Behavior obviously the reinforcement is obscure we know what the schedule is doing we know that we know what the schedule is but we don't know the reinforcement mechanism that the schedule is activating and that's what makes it so interesting but I'm not a bad person because I because I study this I'm not mean to the rats I love the rats and their their health and well-being is of critical importance to us because our experiments often last uh most of the lifespan of the the animal so we need them to be hail and Hearty so obviously when we use any form of a versaf things can go sideways sure and normally what what what needs to go wrong for this to happen well um like we we talk about contingent and non-contention which is very big deal right um but what what do you find in in when you work with them well you know um I think it's important to uh uh maintain a uh a careful distinction between as you say negative reinforcement and positive punishment they're both they're they're unified in the sense that they both fall under the category of aversive control um my guess and I'm guessing uh is that a lot of dog training would would probably be more likely to involve punishment than negative reinforcement um and and the thing about punishment that's very clear in the laboratory is that very mild stimuli can function as effective punishers and I'll give you an example
14:40
음, 제 실험실에서 저는 처벌과 회피 모두를 연구해 왔고 전기 충격 외에 다른 자극들도 사용해 보았습니다만, 전기 충격만 놓고 봤을 때 쥐가 숙련된 회피 행동을 습득하고 유지하게 하려면 약 1밀리암페어의 전기 충격이 필요합니다 그런데 여러분이나 저에게 1밀리암페어는 아무것도 아닙니다. 거의 느껴지지도 않죠. 하지만 작은 쥐에게는 불쾌한 경험이 됩니다. 1밀리암페어는요 하지만 행동을 억제하려면 그보다 훨씬 적은 양으로도 충분합니다. 그래서 만약 쥐를 가령 변동 간격 계획에 따른 음식 강화로 훈련시킨 뒤 거기에 전기 충격 처벌을 덧씌우면 0.2나 0.3밀리암페어만으로도 행동을 억제할 수 있습니다. 자, 여기서 주목할 점은 0.2밀리암페어의 충격은 부적 강화를 유지할 수 없다는 것입니다. 즉 회피 행동을 유지할 수 없죠. 쥐는 그렇게 약한 충격을 피하려고 노력하지 않을 겁니다 하지만 처벌 패러다임에서는 그 정도의 약한 충격도 매우 효과적인 처벌인자로 작용합니다. 자, 이것은 쥐만의 특이한 점이 아닙니다. 여러분도 알 수 있듯이 이것은 보편적인 현상입니다 사람에게도 마찬가지인데 인간 조작적 조건 형성 실험이 여러 차례 있었고 아주 적은 액수의 돈조차도 상당히 효과적인 처벌로 기능한다는 점을 보여주었습니다 그래서 우리는 매우 민감합니다. 최소한 포유류는 처벌적 우연성에 매우 민감해 보입니다 그래서 효과를 보기에 그렇게 큰 처벌 자극은 필요하지 않습니다 따라서 가장 주의해야 할 위험은 만약 당신이 처벌 우연성을 사용하려 한다면 매우 부드럽게 할 수 있다는 사실입니다. 그렇습니다. 아주 부드러운 질책만으로도 효과적일 수 있습니다. 어쩌다 보니 인간들은 그리고 반려견 훈련사들, 아마도 인간 전반이 그렇겠지만, 우리는 오직 처벌이라면 강도를 높여야 한다고 생각합니다 그리고 그건 정확히 말하자면 아주 흥미로운 지점인데 분명히 많은 연구를 통해 잘 입증되어 있다고 생각합니다 처벌이 조건 강화제로 변질되지 않기 위해 지켜야 할 여러 규칙이 있습니다 물론입니다. 그래서 그 모든 규칙이 충족되면 당신이 말했듯이 아주 낮은 수준으로도 처벌의 역할을 다할 수 있고 반려견이 실제로 회피하거나 도망치게 설득할 수 있습니다 하지만 그것은 다른 이야기이고 상황에 따라 다릅니다 아주 흥미롭네요 글쎄요, 저는 수년간 반려동물을 키워왔습니다. 개와 고양이를 모두 키웠고 그 사실을 알고 있습니다. 저는 않을 겁니다 개에게 어떤 묘기를 가르치기 위해 벌을 주거나 행동을 처벌하는 방식으로 하지만 저는 어쩌면 바람직하지 않은 행동이나 주의를 분산시키는 행동을 부드럽게 처벌함으로써 행동 습득을 도울 수 있을지도 모릅니다. 그리고 저는 생각합니다. 반려견을 아끼는 모든 견주들은 일상적으로 처벌을 사용하고 있지만, 결코 그것을 처벌이라고 부르지는 않을 것입니다. 아마도 그냥
um in in my laboratory I've studied both punishment and and avoidance and I've used other stimuli besides electric shock but just with electric shock to to get a rat to acquire and maintain proficient avoidance requires an electric shock of about one milliamp now to you and me one milliamp is nothing you can barely feel it but to a little rat it's it's an unpleasant experience one million but to suppress responding you need much less than that so if you train a rat on let's say a variable interval schedule of food reinforcement and then you superimpose on that electric shock punishment you only need 0.2 or 0.3 milliamps to suppressed Behavior now notice that a 0.2 milliamp shock will not sustain negative reinforcement it will not it will not sustain avoidance a rat will not work to avoid a shock that mild but in a punishment Paradigm that mild of a shock will serve as a highly effective Punisher now this is this is not a peculiarity of rats you can see this in people too there have been a number of of human Opera experiments that have shown that a very tiny fraction of a penny will function as an effective Punisher and and so uh uh we're very sensitive we we mammals at least seem to be very sensitive to um punitive contingencies and it doesn't take much of a of a punishing stimulus to be effective so I think the danger that uh that is that you need to be aware of the most is that if you're going to use a punishment contingency you can be very gentle yes a very gentle rebuke can be effective somehow humans and dog trainers and probably just humans overall we think of oh this is punishment we need to be very uh Amplified the intensity and and that's exactly like this is very very interesting and I think it's for for sure very um well recorded in in a lot of research there is a lot of different rules that need to happen through punishment so it doesn't become a conditioned reinforcer and sure right so that but when all those roles are met as you said of very low level will do the job to punish to where to convince the dog to actually avoid or escape and and it's a different different story then it depends on yeah that's very very interesting well I don't I mean uh I've I've had pets over the years I've had dogs and cats both and I know this I'm not going to teach a dog to do any tricks by punishing it by punishing Its Behavior but I might be able to facilitate uh the acquisition of behavior by gently punishing conflicting Behavior or distracting behavior and I suspect that any dog owner that that cares about its dog engages in punishment on a frequent basis but they would never label it that so they might
18:28
안 돼, 안 돼, 안 돼라고 하거나 그래, 그거야 같은 말을 하겠죠. 음, 그게 바로 처벌 절차입니다. 사람들은 흔히 처벌을 화가 나서 동물에게 가하는 행동이라고 생각합니다. 좌절감을 느끼거나 화가 나서 개를 때리는 것 말이죠. 음, 그것은 효과적인 처벌 절차가 아닙니다. 네, 그건 복수일 뿐이죠. 특정 애견 훈련 업계에서는 안 돼라고 말하는 것과 같은 조건화된 처벌을 음 심지어는 처벌이라고 부르지도 않고 제지 수단이라고 부르죠. 왜냐하면 처벌이라는 단어는 워낙 부정적인 의미를 담고 있기 때문입니다. 어떤 이유에서인지 우리는 애견 훈련 업계에서 무언가를 처벌해야 한다고 말하기를 매우 두려워합니다. 그것이 나쁜 행동을 억제하는 데 도움이 되고, 그 결과 좋은 행동이 나타났을 때 강화할 수 있게 해주는데도 말이죠. 하지만 음 발견하는 것은 그것을 애견 훈련 분야는 매우 양극화되어 있습니다. 그리고 저는 어떻게든 현재 우리 사회 전체는 무엇에 대해 매우 무엇이 경계인지에 대해 비윤리적인 행동이 되는 것과 동물에게 어리석은 짓을 하는 것 사이의 경계는 무엇인가에 대해 반면 결과를 얻기 위해 지능적으로 혐오 자극을 사용하는 것과는 부적 강화에 있어서는 거의 같은 맥락입니다 음 제가 보는 부적 강화는 제대로 수행되었을 때 사실 매우 핫 핸즈(hot hands) 같은 단순한 놀이를 봐도 알 수 있듯이 말이죠 그건 부적 강화의 일종인 게임이에요 혐오 자극적인 요소가 포함되어 있지만 정말 흥미롭고 어린아이들은 그것을 한참 동안 즐기죠 그리고 네, 잘못될 수도 있습니다 하지만 우리가 부적 강화 부분에 집중해서 그 해결책을 설명해 줄 수 있다면 사실 저는 그것이 어느 정도의 자신감을 심어준다고 믿습니다 나는 이런 사람이고 나는 무엇을 피할 수 있는지 알며 어떻게 관리하고 어떻게 더 회복 탄력성을 가질지 알게 되죠 어 음 그냥 살아가면서 말이죠 많은 경우에, 우선 저는 개에게 특정 행동을 처벌하겠다고 말하는 대신 그 행동을 억제하겠다고 하는 것에 대해서는 개인적으로 별 문제가 없다고 생각합니다 제 연구에서 처벌이라는 단어를 사용할 때 그것은 말이죠 전문 용어로 사용하고 있는데, 안타깝게도 전문 문헌 밖에서 사용될 때는 그런 음 말씀하신 대로, 부정적인 의미가 담긴 용어가 되어버리죠 음, 그래서 어, 고객들이나 동료들과 일할 때, 만약 반려견 훈련 세계에서 새로운 종류의 전문 용어가 나온다면 그것은 저에게도 타당한 이야기이고, 어, 저는 절대 누구도 탓하지 않을 거예요. 제 생각엔 제 생각엔 음
just they might just say no no no yes you know something like that well that's that's that's a punishment procedure you know people tend to think of punishment as something that you do to the animal in anger you know you're frustrated you're angry and you whack the dog well that that's uh that's not an effective uh punishment procedure yes that's revenge in certain in certain dog training circles the that conditioned Punisher like like saying no um it's not even cult Punisher they they call it interrupter just because the word punishment has it's such a loaded word and for some reason we are in the dog training uh industry we are very afraid to to say oh we have to punish something because it will help us suppress and then therefore we can reinforce the good behaviors when they happen but um finding finding the it's very polarized the dog training is and I I believe somehow the whole society at the moment is is very polarized on on what is what where the boundaries boundaries between becoming unethical and and doing doing stupid things to to an animal versus intelligently applying aversive to to to accomplish to to get a result and with negative reinforcement it's it's almost the same thing um the way I look at negative reinforcement in when when done right is actually very like when we look at even a silly game as uh what a hot hands you know it's a it's a game of negative reinforcement it has an aversive element but it's so interesting and it's like little kids will play it forever and yes we can go wrong but as long as like if we if we focus on the negative reinforcement part we have able to explain the way out then actually I I believe that it becomes uh almost it builds certain level of confidence of I I'm somebody I can I can avoid things I know how to manage and how to be more resilient with uh um just just go on with life right in many cases well first I don't have a I personally have no problem if instead of saying we're going to punish certain behavior in the dog that we're going to discourage it for example I mean it it it's it's one thing you know when I use the word punishment in my research I'm using it as a technical term and it's unfortunate that when it's used outside the technical literature it has such a um as you say it has its loaded term um and so uh in working with clients uh and with colleagues if if there's a different kind of jargon that emerges in the world of of dog training that makes sense to me and uh I certainly wouldn't fault anybody for it you know I think in I think um
22:23
가장 흥미로운 어, 부적 강화(negative reinforcement)의 형태 중 하나는 흔히 연민이라고 불리는 것 같아요 제 말은, 한번 생각해 보세요. 만약 여러분이 어, 만약 텔레비전을 본다면, 어, 어, 여러분은 틀림없이 광고를 보게 될 거예요. 거기서는 학대받는 강아지나 고양이, 그리고 다른 귀여운 동물들을 보여주죠. 알다시피 쇠사슬에 묶여 있거나, 굶주리거나, 울고 있는 모습을 보여주고, 그러고 나서 '매달 19달러를 보내주시면 이런 일을 멈출 수 있습니다'라고 말하죠 글쎄요 사람들은 그걸 보고, 그리고 눈시울이 붉어집니다. 제 말은, 만약 그들에게 조금이라도 영혼이 있다면 말이죠. 그렇죠? 그들은 그걸 보고 눈시울이 붉어지며, 그들의 마음은 그 동물들을 향하게 되고, 그들은 19달러짜리 수표를 적거나 매달 수표를 보내기로 등록하는 것에 대해 꽤 기분 좋게 느낍니다 우리는 그걸 부적 강화라고 부르지 않지만 실제로는 정확히 부적 강화입니다 우리는 그것을 연민이라고 부르죠 그리고 여러분 아시다시피, 그런 것들이 너무나 많아요 그와 같은 호소가 많습니다, 그리고 저는 그것이 틀렸다고 말하는 것은 아닙니다, 어떤 사람들은 그것을 감정적 조작이라고 말할 수도 있겠지만 저는 그 정도까지는 생각하지 않습니다, 사람들은 세상의 상태를 보여주며 당신이 세상을 더 낫게 만들 수 있다고 말하지만, 그들은 나쁜 것을 보여줌으로써 당신이 더 낫게 만들도록 유도합니다 그리고 나쁜 상황이 사라지게 할 수 있다고 말하죠, 그것이 바로 부적 강화입니다, 여러 해 전에 인간 조건의 위대한 관찰자인 H.L. 멘켄이 에세이에서 이타주의에 대해 조금 언급했습니다 이타주의는 가장 순수한 형태의 자선이라고 여겨지죠 그런데 그는 이렇게 말합니다, 제가 의역하자면 이타주의는 가장 순수한 형태일지라도 불행한 사람들이 주변에 있는 것이 불편하기 때문에 일어나는 일이라고 말이죠 우리는 불행한 사람들을 보면 마음이 아픕니다 우는 아이들을 볼 때나 굶주린 동물들, 울고 있는 동료들, 속상해하는 학생들을 볼 때 우리의 마음은 그들을 향하고 우리는 돕기 위해 할 수 있는 일을 합니다, 그것이 부적 강화입니다 부적 강화는 우리가 세상을 더 좋은 곳으로 만드는 데 도움을 주는 일부입니다 와, 흥미롭네요, 정말 매우 흥미롭습니다 음 정말 깊이 생각해 볼 문제입니다 정말 사실입니다 이것은 부적 강화가 무엇인지에 대해 정말 많은 다른 견해가 존재합니다 저는 거의 이걸 꼭 기록해둬야겠네요 정말 그렇네요 음 그들도 그것에 대해 생각하고 있죠 이반, 요리를 할 때 당연히 좋아하죠. 또한 지금 당신은 주방에 있고 어, 뭔가를 집어 든다고 칩시다 아주 간단하게 예를 들어보죠, 당신이 끓는 물이 든 냄비를 집어 든다고요 무섭나요? 당황스럽나요? 불안한가요? 그걸 해야 한다는 사실에 식은땀을 흘리나요? 아니요, 그냥 자연스럽게 하죠. 저는 당신이 그런 생각을 한다고조차 의심스럽네요 하지만 당신은 숙련된 행동의 레퍼토리를 습득해왔고,
one of the one of the most interesting uh forms of negative reinforcement is often referred to as compassion I mean think about this if you if you uh watch television uh uh you will you will certainly see commercials where they show dogs and cats and other cute animals that are being abused you know they show them chained up they show them starving they show them crying and they and then they say why don't you send us 19 a month so we can stop this well people look at that and they and it brings it their eyes get moist I mean if they have any kind of Soul right they look at that their eyes get moist their heart goes out to these animals and they feel pretty good about writing a check for nineteen dollars or signing up to send a check every month We call we don't call that negative reinforcement but that's exactly what it is we call it compassion and and you know that's there are so many appeals like that uh and and I'm not saying they're wrong some people might say it's an emotional manipulation I I wouldn't go that far people are showing the state of the world and saying you can make it better but they they get you to make it better by showing you something bad and telling you that you can make the bad go away that's that's negative reinforcement many years ago the uh uh the great Observer of The Human Condition H.L Menken uh in an essay mentioned uh uh talked a little bit about altruism now you know altruism is supposed to be the purest form of charity and he says uh I'll paraphrase but he says altruism even in its purest form uh happens because it it's uncomfortable to have unhappy people around you know it we we don't our heart goes out to unhappy people to crying children to to to starving animals uh to crying colleagues you know to students who are upset our heart goes after them and we do what we can to help that's negative reinforcement negative reinforcement is part of what US helps us make the world a better place wow interesting very very interesting um that that's something to really yeah it's a I I almost need to but but it's so true this is uh there are so many different takes on on what negative reinforcement is um they're thinking about it too narrowly Ivan when you uh do you cook of course I love it well also you you uh here you are in your kitchen and you're uh you're picking up uh let's make it very simple you're picking up a pot of boiling water are you scared are you upset are you anxious do you break up into a into a sweat knowing that you're gonna have to do this no you just go about it I don't I doubt that you even think about it much but you've acquired a repertoire of very skilled behavior and you know
25:56
냄비를 어떻게 집어야 하는지, 그리고 어떻게 옮겨야 하는지 정확히 알고 있죠. 버너에서 버너로, 버너에서 조리대로 뜨거운 물을 쏟거나 스스로를 다치게 하지 않으면서 말이죠 음, 이것이 바로 회피 행동입니다 그게 어떻게 시작되었는지 알아내는 건 어렵지 않고 부적 강화 수반성에 의해 유지되죠 하지만 이건 좋은 일입니다 당신이 손가락을 데거나 그보다 더 심한 부상을 입는 것을 막아주니까요 그리고 그런 행동들은 정서적 부산물을 많이 만들어내지 않죠. 너무나 많은 사람들이 요리하는 것을 좋아합니다, 비록 그들이 심각한 상처를 입힐 수도 있는 열기로부터 불과 몇 인치 거리에 있는데도 말이죠 맞아요, 때로는 무섭기도 하죠 무서운 게 아니죠. 정확합니다, 정말 좋은 지적이에요. 그리고 가끔 부주의함에 대한 대가를 치르기도 하지만 당신은 자신이 무엇을 잘못했는지 정확히 알고 어떻게 개선해야 할지도 정확히 알고 있죠 적어도 향후 2년 동안은 다시 방심하기 전까지는 말이죠 그리고 네, 대략 2년이나 3년마다 저는 오븐에서 무언가를 태워 먹곤 합니다 네, 인정합니다. 그럴 때마다 항상 제 자신이 매우 멍청하게 느껴져요 그리고 당신이 지적했듯이 정말 흥미로운 점은 쥐가 실험실 환경을 탐색하는 과정에서 어떻게 길을 찾아가는가 하는 것입니다. 다른 예로 저는 반려견 훈련사 학교를 운영하는데요 제가 사용하는 비유는 고압 전신주에 올라가 작업하는 전기 기술자들에 대한 것입니다 그들은 매일 출근하면서 스트레스를 받아 땀을 흘리며 일하지 않습니다 그들에게는 답이 있으니까요 반려견 훈련 경험상 부적 강화(negative reinforcement)의 가장 결정적인 포인트는 개나 동물, 혹은 사람에게 즉각적으로 어떻게 탈출할 수 있는지 보여주는 것입니다 그리고 결과적으로는 혐오 자극을 완전히 회피하는 방법을 알려주는 것이죠. 설령 아주 잠깐 동안 오, 이거 좀 스트레스 받는데 싶은 요소가 있더라도 말입니다 이건 하나의 퍼즐 같은 것이고, 과연 해결할 수 있을까 싶지만 그 문제가 해결되는 순간 모든 두려움과 스트레스는 해소됩니다. 맞습니다 네, 맞아요 그리고 다시 말하자면, 최소한 반려견 훈련 분야에서는 우리는 퍼즐을 해결하는 그 순간에만 갇혀버리는 경향이 있습니다 그래서 그게 아주 나쁜 것이라고 지적하죠 자, 다시 한번 해보죠 그렇군요 하지만 이건 개가 실제로 해결책을 찾고 있는 순간이며 매우 일시적인 순간입니다. 우리가 적절한 일정을 가지고 있고 우리가 가장 먼저 도피하는 법, 어떻게 도피하고 다음에는 분명히 어떻게 회피하는지를 배우는 과정에 있다고 가정할 때 말이죠. 만약 동물이 그 단계에 갇혀 있다면 우리는 제대로 일을 수행하지 못한 것입니다. 그렇죠? 맞습니다. 그리고 아시다시피 개는 어 혐오적 우발 상황에 대처할 능력을 갖추고 있습니다. 우리는 이것을 알고 있습니다. 왜냐하면 어미가 어 혐오적 우발 상황을 유도하기 때문이죠. 그렇죠. 제 말은 어미는 어미가 강아지가 잘못된 행동을 할 때 낮게 으르렁거리거나 낮은 소리로 짖습니다.
exactly how to pick up the pan and how to move it from burner to stove from burner to burner burner to counter without spilling hot water and hurting yourself um this is this is avoidance behavior that's uh that was it's not too hard to figure out how it got started and it's maintained by negative reinforcement contingencies but these are good things they keep you from getting your fingers burned or worse and they don't generate a bunch of emotional byproducts too many people love to cook even even though they're they're inches away from uh heat that could do serious serious damage yes and sometimes I'm scared you're not scared exactly exactly very very good point and sometimes you pay the price for being negligent but you know exactly what you did and you know exactly how to change it for the next at least two years before you lose your guard again and and yeah about every every two or three years I I managed to burn burn something something out of the oven yes I'm I admit it and I always feel very stupid when I've done it and just like you pointed out which is super interesting how how a rat can be in in in the process of navigating in the the the laboratory in in another example I have I have a school for dog trainers and I bring a um my my analogy is with the electricians that go in this high voltage poles and they don't go to work every day stressed out sweating that they have an answer and it is the most critical point of negative reinforcement from what the eye from my experience in dog training is to to be able to show to the dog or animal or a human immediately how to escape and eventually how to completely avoid the aversive if we have that even if we have a a very brief element of oh wow this is a a little bit I'm stressed about it it's a puzzle can I solve it but the moment it gets solved all that fear and all that stress is resolved correct yes yeah and go again well I guess like in at least in the dog world we tend to get stuck on that moment of solving the puzzle and pointing out that okay it's very bad but this is at the moment when the dog is actually searching for solution and it's very temporary moment assuming that we have a a proper schedule and and we are on the way to learning how to escape first how to escape and next obviously how to avoid and and if the animal gets stuck in that place we have not done our job correctly right that's right and you know the the uh the dog is equipped to cope with uh aversive contingencies we know this because its mother uh engages in aversive contingencies right I mean the the mother will let let out a low growl or low pitch bark uh when the when the puppy does something
29:44
어 저는 지금 16살 된 고양이를 키우고 있는데 새끼 고양이와 함께 지내며 대처하고 있죠. 그 늙은 고양이는 새끼 고양이에게 경계를 설정합니다. 네, 그래서 어 부드럽지만 단호하게 앞발로 툭 치기도 하고 어 야옹 소리를 내면서 새끼 고양이가 선을 넘었다는 것을 알려주죠. 이건 누군가 훈련시켜야 하는 무언가가 아닙니다. 이것은 자연스러운 자연스러운 행동입니다. 그리고 어 만약 생명체들이 이런 일로 큰 충격을 받는다면 더 이상 개나 고양이는 존재하지 않을 겁니다. 왜냐하면 그들은 번식을 위해 같은 종들 주변에 있는 것을 견디지 못했을 테니까요. 이런 이야기를 할 때마다 저는 항상 꺼내 놓는 것이 있죠. 그것은 마치 근본적인 생물학 법칙과도 같아서 우리는 기분 좋은 것에는 다가가고 불쾌한 것은 피하게 되는데 이건 우리가 배워야 할 것이 아니라 마치 DNA에 내장된 프로그래밍과 같아서 우리는 이 모든 상황을 어떻게 헤쳐 나가야 하는지 정확히 알고 있습니다. 그래서 부정적인 비난을 하거나 혹은 부적 강화라는 말조차 쓰지 않더라도 혐오적인 경험 자체를 나쁜 것이라고 치부하는 것은 옳지 않습니다. 말씀하신 대로 그것은 매우 중요한 역할을 하니까요. 우리 모두 혐오 자극에 올바르게 반응하는 법을 모른다면 살아남지 못했을 것입니다. 그리고 그것이 항상 잘못된 방향으로 흐를 필요는 없습니다. 사실 모든 생명체에게 매우 유익한 것이기도 하죠. 심지어 단세포 생물조차도 위험을 알리는 것에서는 물러나니까요. 맞습니다. 네, 그리고 제 생각에 치료사들이나 아마도 동물 훈련사들 역시 제기할 반대 의견은 아마 이런 식일 것입니다. '자연계는 그렇다 치더라도, 우리는 훈련 과정의 일환으로 이런 우연한 상황들을 인위적으로 만들고 있지 않습니까? 그러니 우리의 학생이나 반려동물, 그리고 고객들에게 의도적으로 고통을 주는 행동은 하지 말아야 한다'는 것이겠죠. 음, 저도 그 의견에 동의하는 편입니다. 분명히 피해야 할 부분이라고 생각합니다. 불필요한 고통을 야기하지만, 많은 어, 부정적 강화나 처벌의 형태 중에는 사실 사람들이 말하는 그런 부작용을 일으키지 않거나, 설령 일으키더라도 경미하고 일시적인 것들이 많습니다. 일시적이라는 것이 핵심 단어죠. 맞아요, 네. 그 아시다시피, 어, 많은 놀라운 결과들이 있는데, 그것은 사람들이 종종 인지하지 못하는 범주에 속합니다. 정적 강화가 좋다는 범주에서요. 사람들은 정적 강화 스케줄도 혐오적인 성격을 띨 수 있다는 것을 보여주는 많은 실험실 연구 결과가 있다는 사실을 간과합니다. 그래서 훈련 세션 중 일부에서는 동물들이 확실하게 어, 실험을 중단하려는 반응을 보이기도 합니다. 실제로 그들은 정적 강화 스케줄로부터 벗어나려고 행동할 것입니다.
wrong and uh I have a I have a 16 year old cat right now that's coping with a kitten and the cat the old cat sets boundaries for the kitten yes and we'll we'll uh gently but firmly you know give it a bat with its paw and give it a little noise with its with its meow you know to let the kitten know that uh it's transgressed and this is not this isn't something that anybody had to train this is natural natural behaviors it and and uh if the organisms were devastated by this we wouldn't have cats or dogs anymore because they wouldn't tolerate being around their con specifics to procreate whenever we talk about this I always bring out the it's almost like a fundamental law of biology or whatever to where we we would approach something pleasant and we would avoid something unpleasant and we are it's not something that we need to learn it's it's like program it's baked in our dnas and we know exactly how to navigate through through all this correct and so taking like like blaming negatively or I shouldn't even say negative reinforcement but an abortive experience all together as something bad we it has such an important place as you said like we would all of us will will not be alive if we don't know how to respond correctly to aversive and it doesn't necessarily always need to go wrong it's actually very beneficial for for any living I mean any simple cell organism will back off from from something that it alarms to Danger correct yes and you know I think the uh I think the objection uh by uh therapists and I suppose uh animal trainers as well I think the objection would probably be go something like this well the natural world is one thing but we're contriving these contingencies as part of a training process and we shouldn't do anything to our our students our pets our clients um uh that would would cause intentionally cause them distress and you know I I I'm inclined to agree I think that you certainly don't want to create undue distress but there are many uh forms of negative reinforcement and punishment that don't really generate the kinds of byproducts that people are talking about or if they do they're mild and temporary temporary is a key key word yes yes yeah you know um uh there are many surprising results that that uh people are often unaware of in the category that positive reinforcement is good people Overlook the fact that we have lots of research in the laboratory that shows that schedules of positive reinforcement have aversive character and there are portions of sessions where animals will reliably uh engage in a response to turn off the experiment they will actually act to escape from schedules of positive
33:24
이것은 60년대부터 알려진 사실인데, 비록 오늘날에는 1960년대보다 그것을 유발하는 조건들에 대해 더 많이 알게 되었지만요. 어, 정적 강화 스케줄은, 음, 글쎄요, 여러분은 이렇게 생각하시겠죠. 만약 배고픈 비둘기를 데려다가 간헐적인 음식 강화 스케줄에 두면, 그것이 가장 좋은 것이며 처음부터 끝까지 긍정적인 상황일 것이라고요. 그리고 그 스케줄과 관련된 자극은 긍정적인 조건 강화물이 될 것이라고요. 하지만 꼭 그렇지만은 않습니다, 아이반. 맥락에 따라 달라지기 때문이죠. 예를 들어, VI 1분 스케줄을 예로 들어보죠. 잠시만요. 어, 아, 아. 그리고 그것은 소거와 교대로 나타납니다 1분 간격의 계획이 매우 긍정적인 것이 되고, 그것과 연관된 자극은 조건 강화제가 됩니다 하지만 만약 똑같은 VI 1분 계획을 VI 30초 계획과 교대로 한다면, 즉 보상 측면에서 두 배 더 밀도가 높은 계획이라면 그 VI 1분 계획은 부정적인 것이 되고, 그것과 연관된 자극은 조건 처벌제가 됩니다. 같은 계획, 같은 자극인데 그것이 포함된 맥락 때문에 강화제에서 처벌제로 바뀝니다 따라서 한 가지 요소만 보고 터널 시야에 빠져서 그것은 정적 강화이며 그와 연관된 자극은 좋을 것이라고 말할 수는 없습니다. 왜냐하면 맥락이 중요하기 때문입니다. 맥락을 놓치기란 너무 쉽죠. 흥미롭네요. 그리고 결국 뇌가 그것을 받아들이는 방식, 뇌가 그것을 인식하는 방식은 동일한 정서적이고 혐오적인 반응을 경험하게 되는 거죠. 맞습니다. 그렇죠. 그리고요. 만약 제가 당신에게 내년에 직장에 고용해서 500만 달러를 주겠다고 말한다면, 아마 당신은 기뻐할 겁니다. 하지만 전 세계의 프로 선수들이 있습니다. 축구 선수, 야구 선수, 미식축구 선수, 농구 선수들에게 만약 내년에 500만 달러를 벌게 될 것이라고 말한다면 그들은 매우 화를 낼 겁니다. 그들은 매우 불행하죠. 왜냐하면 그들의 세계에서는 그건, 그건 모욕이니까요. 같은 액수의 돈이라도 상황에 따라 환경적 맥락이 500만 달러라면 그게 좋은 건지 나쁜 건지는 상황에 달려 있죠. 맞아요, 맞아요. 그리고 또 다른 요소가 하나 있는데, 가끔은, 사실 가끔이 아니라 꽤 자주 만약 우리가 계속해서 만약 우리가 밖으로, 지금 밖으로 한 걸음 나가서 그 스키너 상자를 벗어나더라도, 만약 우리가 계속해서 어떤 행동을 보상만 한다면 그 행동 자체를 보람 있게 만들기 위해 어떻게 해야 할지에 집중하지 않고 무엇을 해야 하든 간에 행동 그 자체를 보람 있게 만드는 것에 집중하지 않으면 때로는 사람이나 동물이 원래 좋아하던 아주 자연스러운 행동도 우리가 빗나가게 할 수 있고, 완전히 망쳐버릴 수도 있어요. 하지만 정적 강화라는 것이 어떤 음, 간식 보상이나 돈을 기대하게 만들면서 그들이 원래 정말 좋아하던 일을 하게 만드는 것은
reinforcement this is something that's been known since the 60s although we know more about the conditions that generated today than we did in 1960 but a a schedule of positive reinforcement um well you would think that that if you take a Hungry pigeon and you put it on an intermittent schedule of food reinforcement that that would be the best thing and it would be a positive thing through and through and a stimulus correlated with the schedule would be a a positive condition reinforcer but it doesn't work that way Ivan it depends on the context so if you have let's say a VI one minute schedule and it alternates with Extinction one minute schedule becomes a very positive thing and a stimulus correlated with it becomes a condition reinforcer but if you take that same VI one minute schedule and you alternate it with a VI 30 second schedule a schedule that's twice as dense in terms of reinforcement then that VI one minute schedule becomes a negative thing and a stimulus correlated with it becomes a conditioned Punisher same schedule same stimulus goes from being a reinforcer to a Punisher because of the context in which it's embedded so you can't just get tunnel vision and look at just one element and say well that's positive reinforcement and a stimulus correlated with it's going to be good because the context matters and it's so easy to lose sight of the context interesting and ultimately the way that the the brain takes it the way the brain perceives it it experiences the same emotional and aversive response correct right and right you know uh if I told you that uh I was gonna I was gonna hire you for a job next year and I was going to pay you five million dollars I suspect you'd be happy but there are professional athletes all over the world soccer players baseball players football players basketball players if they if they were told they were going to make five million dollars next year they'd be very upset they'd be very unhappy yes because in their world that's uh that's a an insult same amount of money different environmental context is five million dollars good or bad it depends right right and and then there is that that other element that sometimes actually not sometimes quite often if we keep if we stay stay outside now step outside the the Skinner box but if we keep rewarding a behavior without trying to focus on how we can create whatever we need to do to make the behavior itself rewarding sometimes a very natural behavior that the person or the animal likes to do we can deflect it we can totally derail it but positive reinforcement expecting some some um treat reward or money from for doing what they originally actually
37:10
반려견 훈련에서 정말 큰 음 사람도 마찬가지일 거라고 확신해요. 아마도 보상을 기대하기 시작하는 순간, 좋아, 네가 돈을 줄 거야, 그럼 얼마를 줄 건데? 그러다가 말씀하신 것처럼 우리가 이렇게 되죠. 오, 돈을 충분히 안 주네? 원래 내가 아무런 보상 없이도 정말 좋아해서 하던 일이었다는 사실조차 잊어버리는 거죠. 맞아요. 네, 맞아요. 음 있잖아요, 때로는 외적 보상이 부적응적이라고 주장되기도 해요. 왜냐하면 그것들이 그렇지 않으면 내재적인 결과에 의해 강화되었을 행동에 대한 통제력 그럴 수도 있겠죠, 하지만 세심하게 설계된 훈련 프로그램에서는 외재적 강화제나 인위적인 강화제는 오직 그 행동이 자연스러운 결과의 통제하에 놓일 때까지만 사용됩니다. 스키너는 다음과 같이 지적합니다. 만약 읽기 학습을 자연스러운 결과에만 맡겨둔다면 아마 결코 습득되지 못할 것이라고요. 그러니 읽기는 명시적으로 가르치고 성적이나 칭찬, 그리고 그 외 무엇이든 동원해 강화해야 하는 거죠. 결국 이러한 인위적인 강화제와 인위적인 우발적 사건을 통해 아이는 충분히 탄탄한 읽기 레퍼토리를 구축하게 되고, 그 결과 스스로 만화책, 동화책, 소설책 또는 무엇이든 읽게 됩니다. 무언가를 만드는 법에 대한 설명서처럼요. 그런 것들은 읽는 행위 자체에서 자연스럽게 강화제가 주어지죠. 하지만 그 단계에 도달하려면 인위적인 우발적 사건과 인위적인 강화제를 사용해야 합니다. 요령은 어느 시점에 통제권을 한쪽에서 다른 쪽으로 넘겨야 하는지 아는 것입니다. 아이가 슈퍼맨 만화책을 읽기 시작하고 그것을 즐기게 되면 더 이상 읽기에 대한 보상을 주지 않잖아요. 그러니 그건 실수가 아니겠죠? 읽기나 파닉스 숙제를 할 때는 보상을 줬을지 몰라도, 아이가 일단 읽는 법을 배우고 나면 더 이상 그럴 필요는 없는 거죠. 그가 읽을 거리를 즐길 수 있을 정도로 충분히요 당신이 그를 내버려 두는 자료들을요 정말 맞는 말이에요 스키너에 대해 언급했으니 조금 더 깊이 들어가 볼 수 있을 것 같아요 왜냐하면 분명히 그가 제시한 것들 중에는 타당한 것이 정말 많지만 동시에 저는 그것이 많은 논란의 여지가 있다고 느껴지거든요 반려견 훈련에서 말이죠 이 부분에 대해 당신의 의견을 듣고 싶어요 음 특정 반려견 훈련사 집단에서는 오류 없는 학습(errorless learning)을 하는 것이 매우 흔하고 대중적이죠 그리고 물론 스키너가 당시에 그것을 제시했을 때 큰 화제였죠. 저는 그것에 대해 접해본 것이라고는 컴퓨터처럼 보이는 기계들이 나오는 유튜브 영상뿐이에요 기본적으로 그 과정을 거치지만 어떤 이유에서인지 그것은 실제로 크게 성공하지 못했어요 저희는 학교에서 분명히 그 시스템을 사용하고 있지 않지만요 여전히 그것에 대해 이야기하고 있고 여전히 그것이 어떤 자리가 있다고 믿고 있으며
really love doing in dog training it's a big um I I am sure which humans is the same probably the moment we start expecting okay you will pay me then how much you're gonna pay me and then you can as you said we become oh you're not paying me enough I I forget that I actually love doing this already without any reward at all correct yes yes well um you know uh sometimes it's argued that extrinsic rewards uh are maladaptive because they they take control over behavior that would otherwise be reinforced by its intrinsic consequences and I I guess that's possible but in a carefully crafted uh training program uh extrinsic reinforcers or contrived reinforcers are only used uh until the Behavior comes under the control of its natural consequences Skinner points out that um if you if you left the acquisition of reading to Natural consequences it would probably never be acquired so you know reading has to be taught explicitly and reinforced with grades and praise and who knows what else and eventually through these contrived reinforcers and contrived contingencies the child develops a strong enough reading repertoire that they will come to read comic books or story books or novels or what have you instructions on how to build something things that where the reinforcer comes naturally out of the act of reading but to get there you have to use contrived contingencies and artificial reinforcers and and uh the the trick is to know when to transfer control from one to the other you know once the kid starts reading Superman comic books and enjoys doing it you don't pay them for reading so that won't be a mistake right you might you might you might have paid him to do his reading and phonics homework but once he learns how to read well enough that he can enjoy reading material you you leave him alone very true since we mentioned skin Arab we we can go in a little bit further with it because I mean obviously there is so many things that he brought to the table that are so valid but at the same time I feel like there was it's there is a lot of controversial stuff in dog training this this is one one thing I want to hear your opinion on um it's in certain circles of dog trainers it's very common very popular to do errorless learning and of course kinner presented it at the time it was a big thing like I I've you know I the only exposure to it I have is YouTube videos of of these machines that look like computers but you know and and basically go through the process but for some reason it never really took off we obviously are not using this at school uh that that system
41:06
실제로 시행착오만큼이나 좋을 수 있다고 생각해요 여기에 대한 당신의 의견을 듣고 싶어요 음, 여기에는 적어도 두 가지 쟁점이 있어요 하나는 오류 없는 학습이 가능한가 하는 것이고 다른 하나는 만약 가능하다면 시행착오를 통한 학습보다 더 나은 장점이 있는가 하는 것이죠 원래의 혹은 고전적인 실험은 어, 커브 테라스의 박사 학위 논문이었고, 음, 그는 다양한 방법을 사용하여 변별, 단순 변별을 비둘기들에게 훈련시켰으며, 그는 그 비둘기들이, 그, 어 매우 적은 수의 오류만으로 변별을 습득했다는 것을 음 그가 처음에 그것이 가능하다는 것을 보여주었고, 두 번째로 어 그는 보여주었습니다 상관관계가 있는 자극, 그러니까 음, 이렇게 말해 보죠. 오류 없는 변별 학습에는 어떠한 부정적인 행동적 부산물도 없었다는 것을요. 자 나중에 어 미시간 주립대의 마크 릴링은 그와 그의 제자들이 보여주었습니다, 어 그는 테라스의 연구에서 혼입 변수 몇 가지를 지적했고 보여주었습니다. 음, 변별 학습에서 어 정서적 부산물이 있을지 여부를 결정하는 것은 오류 그 자체가 아니라 그것은 바로 부정적 변별 자극을 도입하는 시기였다는 것을요. 그래서 음 그것은, 오, 그것은, 그것은 지나친 일반화였습니다, 오류 없는 변별 학습이 본질적으로 우월하다고 주장하는 것은요. 이제 스키너는 스킨너는 그의, 그의, 음 교수 기계 개발에 있어서, 어 저는 그가 오류 없는 학습에 대해 그렇게 신경을 썼는지 잘 모르겠습니다 그는 오히려 행동을 형성하는 것에 더 관심이 있었습니다, 어떤 방식으로, 음 학문적인 종류의 행동을, 어떤 방식으로요 학생을 참여시키는 그런 방식으로 말이죠. 아주 작은 단계로 진행해서 학생이 실수를 할 가능성이 거의 없게 만들었죠. 그게 바로 스키너가 그의 프로그램 학습법을 통해 이루어낸 것입니다. 그가 제임스 홀랜드와 함께 집필했던 것 말이죠. 그리고 제 생각에는 다른 분들도 수고를 들여 양질의 프로그램 학습 자료를 만들어 낸 사람들은 비슷한 결과를 얻었을 겁니다. 하지만 제 생각에 이것을 잘 해내기는 매우 어렵습니다. 저는 교육 심리학자도 아니고 관련 문헌을 계속 쫓아가는 것도 아니라서 잘 모르겠지만, 그렇게 많은 자료가 있는 것 같지는 않습니다. 음, 제 생각에는 이게 워낙 노동 집약적인 작업이라서 그런 것 같아요. 최소한, 그러니까 생성하는 데 있어서 자료를 만드는 데는 많은 노동이 들어가죠. 일단 한번 만들어지면 모든 게 괜찮아지지만요. 저도 몇 번 읽어보긴 했는데, 확실하지 않아서 잘못된 참고 자료를 제시하고 싶지는 않네요. 하지만 분명 효과가 있는 지점이 있고, 어느 정도까지는 효과가 있어서 학생에게 자신감을 심어주죠. 학습자가 '나는 이걸 잘할 수 있어'라고 느끼게 만드는 겁니다. 왜냐하면 우리는 '아니야, 틀렸어'라고 비난하는 상황을 피하고 있으니까요.
but we are still talking about it and we're still believing that it's it's um there is a place and and actually it can it can be just as good as trial and error how I I want to hear your opinion on this well there's there's two at least two issues here one is um is uh is errorless learning um possible uh and if so does it have an advantage over learning with with errors the uh the the original or the classic experiment was uh curb Terraces doctoral dissertation and um he used different methods to train a discrimination a simple discrimination in pigeons and he showed that the pigeons that that uh acquired the Discrimination with a very very small number of Errors um were he first he showed it was possible to do that and secondly uh he showed that the stimulus that was correlated or well let me put it this way there were no negative behavioral byproducts of the errorless discrimination now later uh Mark rilling at Michigan State he and his students showed that uh he pointed out some of the confounded variables in Terrace's work and showed that um it wasn't the errors per se that that determined whether the uh there was going to be uh emotional byproducts in discrimination learning it was the timing of the introduction of the negative discriminative stimulus and so um it was an oh it was a it was an over generalization uh to argue that airless discrimination was inherently Superior now Skinner Skinner with his his um development of teaching machines uh I don't know if he was as concerned about errorless learning as he was about shaping up behavior in a in a way that um academic kind of behavior in a way that was uh that engaged the student and advanced in such small steps that the the uh student was unlikely to make an error and uh uh that's what Skinner accomplished with his uh programmed instruction uh that he writ that he wrote with um uh James Holland uh and I think other people who've who've gone to the trouble of of uh creating a high quality programmed instruction have had similar results uh but I think it's a very difficult thing to do well and uh I'm not I'm not an educational psychologist and I don't follow that literature but I I don't think there's a lot of it um and I think it's because it's it's uh it's pretty labor intensive yeah and at least I mean to generate it's labor intensive to generate the materials once they're generated everything's fine I I've like I've read some and I'm not sure like I I don't want to give the wrong references but there is a point where it does work to some point and it does build some confidence in the student that I can I am I'm good at this because we are avoiding the pointing the finger
45:01
하지만 난이도가 조금 더 올라가면 그때는 '앗, 잘못된 길로 들어섰네'라는 피드백이 필요해지죠. 음, 스키너의 프로그램 학습법을 보면 그는 그는 이미 풍부한 언어적 레퍼토리를 활용하고 있었던 거죠. 학생의 지식 기반을 활용하고 있는 것이지만, 하지만 개를 훈련시킬 때는 그런 종류의 활용할 레퍼토리가 없으며, 인간을 대상으로 한 작업에서 어, 기술을 창조하려 하지 않을 겁니다, 어 숙련된 기술을 수행하는 외과의사나 야구 투수, 조각가나 화가, 여기서 말하는 건 예술가를 말하는 건데, 그런 종류의 행동을 만들어내려면 수많은 시행착오 없이는 불가능하죠, 그렇죠? 아니요, 아니요, 만약 누가 자신이 아이를 렘브란트처럼 그림 그리게 훈련시킬 수 있다고 주장한다면, 그냥 뭐랄까, 단 한 번의 실패도 없이, 쓰레기 같은 결과물을 만들지 않고 말이죠, 그러면 저는 그 사람이 바보 같은 소리를 하고 있다고 말할 겁니다. 음, 그러니까, 숙련된 신체적 수행 능력은, 어, 학문적 성과와는 다른 방식으로 형성되어야 합니다. 학문적 성과는 이미 탄탄한 언어적, 학문적 행동을 기반으로 구축되니까요. 네, 맞아요. 그리고 분명한 차이점이 있죠. 인간과 동물의 차이점인데, 우리는 인간에게는 지시를 내릴 수 있지만 동물에게는 그게 정말 불가능하다는 점이죠. 바로 이 지점이 혐오 자극이 사용될 수 있는 이유입니다. 왜냐하면 결국 근본적인 법칙으로 돌아가기 때문이죠. 어떻게 생각하시나요? 어디까지, 얼마나, 생각하시나요? 혐오 자극이나 처벌, 또는 부적 강화의 사용이, 정확히 어디서부터 우리가 이것이 시작점이라고 말하게 된 걸까요? 좋은 생각이 아니고 도덕적으로도 사실 처음에는 그저 효과가 없었다고 생각해요 제가 맞는지는 모르겠지만요 글쎄요 제 생각에는 제가 그 부분에 대해 역사적 연구를 해본 건 아니지만 제 생각에는 이것이 대체로 스키너의 영향이 큽니다 스키너는 음 쾌락주의자였어요 아주 비판적이지 않은 의미로 말하는 겁니다 그는 사람들이 행복하기를 바랐고 사람들이 삶을 즐기기를 바랐으며 그를 위한 방법은 정적 강화라고 보았습니다 그는 정부가 의존하는 일부 혐오적 통제를 지적했고 이런 것들이 사람들이 정부에 불만을 갖게 하고 반란이나 다른 형태의 대항 통제를 야기한다고 느꼈죠 그래서 그는 사회 통제 수단으로서의 처벌을 적극적으로 반대했고 제 생각엔 그 연장선상에서 실험적 분석의 주제로서도 그랬던 것 같습니다 머레이 시드먼은 물론 자유 조작 회피 스케줄의 개발과 밀접하게 연관된 인물인데 나중에 '강압과 그 여파'라는 책을 썼고
of no you failed correct but the as it gets a little bit more complex then there is the need of Ops you you went the wrong path well in in if if you look at uh at Skinner's programmed instruction um he he he's he's taking advantage of the extensive verbal repertoire and knowledge base of the student he's taking advantage of that but when you when you train a dog there's you don't have that kind of repertoire to leverage and in human work uh you're not going to create uh skill perform the skilled performance of a surgeon or a baseball pitcher or a sculptor or a painter I'm talking about an artistic painter you're not going to shape up that kind of behavior without lots of Errors right no no you you show me somebody who claims that they can train a kid to paint like Rembrandt right out of the you know without ever making without ever producing junk and I and I will uh I will tell you that you're talking to a fool um so you know skilled skilled physical performances uh have to be shaped up in it in a different way than academic performances are already building on on strong uh verbal and academic Behavior yeah yes and then there is the the obvious difference between humans and animals that we can instruct and humans and it we really cannot do that with animals and then this is where diverse aversion has a place because it's a it really goes back to that fundamental law what what do you think about where where how how far down you think the use of aversive or punishment or negative enforcement where exactly was that starting point that we said this is not a good idea and it's not it didn't not really even morally I think originally it was like it just doesn't work I don't know if I'm right well I think uh I I haven't done a historical study of it but uh my I what I believe is that this is largely uh Skinner's influence uh Skinner was um a hedonist you know I say that in the most uh non-judgmental way he wanted he wanted people to be happy he wanted people to enjoy life and he saw the uh the route to that to be positive reinforcement and he pointed out some of the uh reliances that governments have on aversive control and he felt that these were the kinds of things that made people unhappy with their governments and generated rebellion and other forms of counter control so he he actively argued against punishment as a social control measure and I guess by extension as a topic of experimental analysis um Sidman Murray Sidman who of course is closely associated with the development of the free operative voidance schedule later wrote a book called The coercion since Fallout and
48:54
어떤 형태의 처벌이나 부적 강화도 정의상 강압적이라고 주장했습니다 전반적인 스키너의 영향과 이후 시드먼의 저술들이 이 분야에 대한 분석을 억제하는 데 큰 역할을 했다고 생각합니다 조작적 행동 분석가들에 의한 혐오적 통제 하지만 이에 관해서는 신경과학자들과 인지심리학자들이 연구한 수많은 연구 결과가 있습니다 그래서 우리는 그것을 그들에게 맡겨두었죠 그들이 이 주제를 연구하는 이유는 이것이 자연계의 분명한 일부이기 때문입니다 저는 스키너가 어쩌면 사회적 통제 메커니즘에 대해 타당한 지적을 할 수도 있다고 생각합니다만 기초적인 실험적 분석에 있어서는 우리는 세상의 모든 과정을 이해하고 싶어 합니다 그리고 만약 이것들이 자연계에서 일어나는 과정이라면 우리는 그것을 연구해야만 합니다 그렇죠, 물리학자가 사람들이 나무에서 떨어지면 다치기 때문에 중력을 무시하겠다고 말하지 않는 것과 같습니다 그들은 그것이 세상의 일부이기 때문에 연구하는 것이죠 정말 훌륭하고 뛰어난 비유입니다 정말 맞는 말이에요 와 우리가 매우 흥미로운 대화를 나누게 될 것이라는 걸 알고 있었어요 양육 본성 끝없는 논쟁이죠 음 아실지 모르겠지만 이제는 조금 오래된 책이지만 여전히 대중적이고 자주 언급되는 '탤런트 코드'라는 책이 있는데요 기본적으로 어떤 것을 충분히 오래 반복하면 매우 능숙해져서 유전적 성향을 극복할 수 있다는 내용을 다룹니다 그것과는 상관없이 음 그건 끝도 없고 영원히 계속되는 논쟁이죠. 어떻게 생각하시나요? 유전적 소인과 환경에 대해서요. 물론 둘은 겹치겠지만요. 하지만 저는 더 이상 말하지 않고 박사님께 말씀을 넘기겠습니다. 글쎄요. 음, 제 말은, 아시다시피 개에게 강화제 역할을 하는 것이 초등학교 1학년 아이에게는 강화제 역할을 하지 않을 수도 있죠. 그러니까, 분명히 유전적 유산의 결과여야만 하는 타고난 소인들은 따를 수밖에 없으며 그것들은 개발될 수 있는 행동 목록에 한계를 설정합니다. 제가 학생들에게 상기시키고 싶은 점은 유기체의 유전학 그 자체가 환경의 산물이라는 것입니다. 저는 학생들에게 이렇게 말합니다. 행동은 환경에 의해 통제되며, 환경에는 세 가지 측면이 있습니다. 현재 환경, 즉 지금 당장 일어나는 일이 있고, 역사적 환경, 즉 그 사람이 과거에 겪었던 경험들이 있으며, 그리고 조상 환경이 있습니다. 즉, 유기체의 조상들이 살았던 환경이 여러 세대에 걸쳐 유전자를 선택했고, 그것이 오늘날 이 개체가 가진 유전적 유산이 된 것이죠. 그래서 누군가에게, 예를 들어 아이에게
argued that any form of punishment or negative reinforcement was by definition coercive I think the influence of of of of Skinner in general and then sigmunds later writings in this area have done a lot to discourage the analysis of aversive control by operant psychologists by Behavior analysts but there's a whole bunch of research on it by neuroscientists and cognitive psychologists so we've left it to them and the reason they're studying it is because it's clearly a part of the natural world and I think that perhaps Skinner has a can make some valid points about social control mechanisms but when it comes to basic experimental analysis we want to understand all the world's processes and uh and and if these are processes that take place in the natural world we should be studying them yes you know what a a new a physicist doesn't say I'm not gonna I'm gonna ignore gravitation because when people fall out of trees they hurt themselves right you know they study it because it's part of the world that's that's excellent that's brilliant no it's very true yeah yeah yeah um wow I knew we're gonna have a very interesting conversation nurture nature the endless debate um there is a I don't know if you've it's kind of uh older book now popular book referred to a lot the Talent Code basically refers to if you if you repeat something long enough you will get very proficient and and you will overcome your genetic predisposition but irregardless of that um it's an endless never-ending debate and how how what is your take on genetic predisposition versus the environment of course they would overlap but I will I I'm not going to say anything more and let you talk well um I mean well as you know uh that functions as a reinforcer for a dog probably wouldn't function as a reinforcer for a uh first grade child you know so uh clearly uh uh the uh natural predispositions that must be uh the result of genetic Heritage have to have to be obeyed and they do set limits on on uh the kind of repertoire that can be developed the the thing that I uh I'd like to remind students of though is that the organisms uh genetics is itself a product of the environment you know I tell the students that uh behavior is controlled by the environment the environment has three aspects there's the Contemporary environment what's happening right now there's the historical environment the experiences that the person has had in the past and then there's the ancestral environment the environment of the organism's ancestors that over Generations selected the genetics the genetic Heritage that this individual has today So when you say to someone you say to a
52:49
어떤 물체를 들어 보이면서 이게 뭐냐고 물었을 때 그 아이가 사과라고 대답한다면, 음, 그건 무슨 일이 일어나는 건가요? 지금 이건 이건가요 대상과 질문인 이것은 무엇인가라는 질문이죠 하지만 아이가 사과라고 대답할 수 있으려면 특정한 경험을 해야 하는데 예를 들어 예를 들면 영어를 가르쳐 준 언어 공동체에서의 경험 그리고 사과를 경험했던 것들이 있어야 하죠 하지만 그 아이가 가져야만 했던 유전적 유산도 있죠 특정한 어떤 것을 가져야 하니까요 사과를 볼 수 있는 시각 체계와 질문을 들을 수 있는 청각적 자극 만약 그 유기체의 조상들이 바닷속에서 시간을 보냈다면 그런 종류의 감각 체계는 없었을 겁니다 그러니까 이 모든 것들이 함께 작용하지만 필연적으로 그것들은 모두 환경으로 거슬러 올라가게 됩니다 어떤 의미에서 본성 대 양육이라는 것은 없어요 전부 전부 양육인 거죠 맞아요 그건 이론적인 입장이지 아주 실용적이지는 않죠 도움이 안 돼요 개 훈련을 하는 사람에게는 도움이 안 된다고 생각해요 글쎄요 그게 어떤 틀을 제공해주기는 하죠 네 네 가끔 훈련사들이 갇히게 되는 상황을 말하자면 말하자면 부적절하다고 생각하는 행동을 교정하려고 할 때 그게 매우 정상적이고 유전적으로 프로그래밍된 행동이라는 점입니다 그리고 만약 우리가 그것을 계속 억제하려고 한다면 그리고 그것이 다시 그 유전적 입장으로 돌아가면 실제로 어느 정도 붕괴 수준에 이르게 될 겁니다 그것 대신에 이것이 존재한다는 것을 받아들일 방법을 찾는 것, 그리고 그것이 강하다는 것을요. 그리고 유전적 구성에 맞서기보다는 다른 무언가로 그저 자연스럽게 유도할 방법을 찾는 것이죠. 음, 이건 도그 트레이닝에서 아주 흔한,, 상당히 흔한 일입니다. 특히, 초보 트레이너들이라고 해야 할까요, 아니면 이제 막 시작하는 트레이너들에게서요. 우리는 도그 트레이너로서 우리는 믿습니다. 우리가 통제할 수 있고, 우리는 어떤 상황에서든 행동을 통제할 수 있다고 말이죠. 그리고 종종 우리는 정면으로 맞서는 것 같아요. 음, 선택적으로, 특히 개들의 경우, 우리는 개들을 특정한 일을 하도록 선택적으로 번식시켰으니까요. 예를 들어 우리가 사냥개인 포인터를 키우는데 우리가 '아, 우리는 이 개가 포인팅을 못 하게 할 거야'라고 말한다면 어떨까요? 이건 그냥 예시일 뿐이지만, 이건 그러니까, 뇌를 태워버리는 일이나 다름없어요. 왜냐하면 그 개는 심지어 어느 정도 수준에서는 의식조차 하지 않거든요. 그 행동을 하는 건 몸이 아니에요. 그건 그 모든 것이죠. 그 행동을 하도록 프로그램된 것이니까요, 그렇죠? 맞아요. 차라리 개가 보는 것을 멈추게 하거나 듣는 것을 멈추게 하겠다고 말하는 편이 나을 겁니다. 맞아요. 그렇죠. 그런 행동들, 아시겠지만 흥미로워요. 우리는 결코 보는 것이나 듣는 것에 대해서는 그런 식으로 말하지 않죠. 우리가 개가 듣지 못하도록 훈련하겠다거나 그런 식으로 말이죠. 하지만 그것도 행동입니다. 그저 너무나 깊이 뿌리박힌 행동일 뿐이죠. 생물학적 구성이라는 유기체들
child you hold up you hold up an object and you say what is this and they say apple well that what what's happening right now is the is this the this this object and the question what is this but the child can only respond Apple if it's had certain experiences including for example experiences in a verbal community that taught it English and in which it experienced apples but then there's also the genetic Heritage that it had to have it has a certain it has a visual system uh to see the apple and an auditory stimulus to hear the question if the if the organism's ancestors had spent their time on the ocean floor it wouldn't have that kind of uh sensory systems so all these things work together but inevitably they're all traced to the environment in a sense there is no nature versus nurture it's all it's all nurture right right that's not a it's a theoretical position it's not very practical doesn't help it doesn't help someone to train dogs I suppose well when it does give you a framework yes yes when it some sometimes where dog trainers can get into get stuck I should say trying to let's say correct some inappropriate in their mind behavior is that it's a very normal and genetically programmed Behavior and if we continuously try to suppress it and it goes again that generic press position it will actually reach certain level of breakdown instead of finding a way to to accept that this is there and it's strong and looking for a way to just guide it into something else instead of really go against the genetic makeup um so that's a that's a common quite common in dog training especially with I shouldn't say young trainers but trainers that are in in beginning and you know we as dog trainers we believe that we can control and we can control Behavior no matter what and very often I think we go head on against uh something that is selectively especially with dogs we we breed them selectively to do certain things like imagine if we decide that we have a hunting dog a pointer and we say oh we're gonna stop him from pointing it just is an example this will I mean this will burn the brain simply because the the dog's not even at a certain level not even consciously like it's not the the body that's doing it it's the it's everything program to do that behavior right right you might as well say we're going to stop the dog from seeing or stop the dog from hearing yeah those behave you know it's interesting we never talk about about seeing and hearing that way you know we're gonna we're gonna train the dog to stop hearing but that's Behavior it's just Behavior that's so deeply embedded into the uh organisms of biological makeup
56:38
그것을 바꾸려고 시도한다는 것은 우리에게 전혀 떠오르지 않았던 일입니다 그리고 그건 다행스러운 일이죠 잘 될 것 같지는 않거든요 맞아요 저에게 비어디드 콜리가 있었는데 그건 몰이 견종이거든요 네 그래서 저는 그 아이들이 두세 명의 아이들을 몰지 못하게 하는 데는 꽤 능숙했어요 두세 명의 아이들이 뒷마당에서 뛰어놀고 있을 때면 그 아이들은 몰려고 하지 않았지만 다섯이나 여섯 명이라면 그게 어쩔 수 없었죠 멈출 수가 없었어요 멈추게 할 수가 없었죠 네, 설령 그 순간에 우리가 그것을 멈추게 할 수 있다고 해도 그 행동을 그들의 레퍼토리에서 제거하는 것은 아니니까요 그렇죠, 동물에게... 제 생각에는 당신이 그 개를 훈련시키면서 뭐랄까 능동적으로 그 행동을 억제하도록 훈련시킨다고 볼 수 있겠네요 개들은 능동적으로 양립할 수 없는 어떤 행동을 하고 있는 거니까요 제 대학 동료 중에 한 분이 있는데 지금은 은퇴하셨지만 지질학자이신데 천부적인 반려견 훈련사예요 그분은 두 마리의 서로 다른 개에게 이걸 해냈어요 그러니 그게 단순한 우연은 아니었다는 걸 알죠 그분은 개를 데리고 와서 앉게 할 수 있었어요 앉으라고 말하면 개는 앉았고 그다음에는 치즈 한 조각을 가져와서 개 코 위에 올려두었죠 그러면 개는 가만히 앉아 있었어요 그 개가 조금 불안해하는 것을 볼 수 있었는데 그게 바로 그건 그런 거였죠 침을 흘리네요, 네, 침을 흘리고 있어요. 그러고 나서 그가 '가'라고 말하면 개는 치즈를 공중으로 홱 던져 올리고 치즈가 떨어질 때 바로 낚아채요. 그는 이것을 두 마리의 다른 개들로 했어요. 이건 이 개의 본능에 완전히 어긋나는 행동이지만, 개는 본능을 억제할 수 없었어요. 분명히 알 수 있듯이, 개가 가만히 있으려고 애쓰고 있고 침을 흘리고 있는 게 보이죠. 네, 하지만 그럼에도 불구하고 인상적인 묘기네요. 그렇긴 하죠. 우리는 자주 혼동하곤 합니다. 오퍼런트(조작적) 조건형성과 고전적 조건형성에 대해 이야기할 때 말이죠. 아주 기초적인 수준에서 가끔씩은 개 훈련사들에게 무엇을 말해줄 수 있을까요? 음, 그러니까, 두 가지에서 얻을 수 있는 다른 점들은 무엇인가요? 어떻게, 어떻게 그것들이 다른지 잘 모르겠어요. 어느 정도까지 말이죠. 음, 고전적 조건형성이 개 훈련에 얼마나 관여하는지 생각해보면, 가장 분명한 사례는 클리커 트레이닝이죠. 네, 제 생각에 우리가 고전적 조건형성에 대해 이야기할 때 종종 고전적 조건형성의 감정적 측면을 포함해서 이야기하곤 하죠. 그리고 기본적으로 행동을 특정 감정과 결합시키는 것을 이야기하죠. 그것이, 음, 제 생각에는, 우리가 고전적 조건형성에 대해 아는 것 중 하나는 어, 조건 자극과 무조건 자극은 시간적으로 매우 근접해서 발생해야 하고, 조건 자극이 먼저 나타나야 하며, 두 자극 모두 독특한 특징이 있어야 합니다. 그래서, 제 생각엔 만약
that had never occurs to us to to try to modify it and it's a good thing too I don't think it would work very well right I had bearded colleagues and uh you know that's a herding breed yeah and uh I was pretty good at getting them to stop hurting two or three children if two or three children were running around in the backyard they wouldn't try to Herb them but if there were five or six I have to I couldn't stop I couldn't stop yes and even even if we are able to to stop it at the moment we he we are not taking it out of their uh um repertoire right the animal you what I I would I would guess that you're training the dog to uh uh I guess you could say actively suppress the behavior they're they're they're actively doing something that's incompatible you know I have a colleague here at the University you know actually he's retired now he's a geologist but he's a natural dog trainer and he did this with two different dogs so I know it wasn't a pure accident but he could take this dog and have it sit he'd say sit and the dog would sit and then he would take a piece of cheese and he would put it on the dog's snout and the dog would sit there and he thought you could see the dog is kind of getting a little anxious and it's drooling yes it's drooling and then he'd say go and the dog would flip the cheese up in the air and as it came down snap it right up he did this with two different dogs this is this this goes against everything natural to this dog but he the dog couldn't suppress everything it's clearly you know it's you can see it's struggling to stay still and it's drooling yes but an impressive performance nevertheless we get confused quite often when we talk about Opera and then classical conditioning and to a very basic level sometimes What can you tell dog trainers um like like the what are the different takeaways from from both how how did they how how are they different I don't know to you know to what extent um classical conditioning is involved in dog training I guess the most obvious case is clicker training yeah I guess the way what we when we talk about classical conditioning a lot uh oftentimes we include it we include that emotional aspect of the classical conditioning and and you know basically pairing the behavior with certain emotions and then we talk that that is well I guess you know um in one of the things that we know about classical conditioning is that the uh the conditioned stimulus and the unconditioned stimulus have to occur in close temporal proximity in the uh conditions stimulus has to come first and there has to be something distinctive about about both stimuli so I guess if uh if
1:00:17
키가 크고 수염이 있으며 빨간 모자를 쓴 존 도(John Doe)라는 사람이 있다고 치죠. 그가 반려견 훈련 세션에 들어와서 개를 발로 차기 시작한다면, 음, 아마도 꽤 높은 확률로 다음번에 존 도가 나타났을 때 개는 그를 피해 달아날 것입니다. 즉, 그는 고전적 조건 형성의 원리에 의해 조건 혐오 자극이 되는 것입니다. 그런데 존 도의 가장 관련 있는 특징이 무엇인지는 명확하지 않을 수 있습니다. 모자, 수염 또는 그의 키와는 아무런 상관이 없을 수도 있습니다. 아마 그의 냄새일 수도 있고, 그의 소리일 수도 있겠죠. 하지만 음, 분명히 그것은 새로운 자극이 되어 독특한 자극으로 작용할 것이고, 새로운 고통과 그 자극이 결합하겠지만, 만약 존 도가 일반적인 개 훈련 과정의 일부였거나, 환경의 일부였다면 수개월 동안 말이죠. 그런데 평소와 다르게 그런 행동을 했다면 그가 조건 혐오 자극이 될 가능성은 훨씬 낮아집니다. 그러니까, 만약 당신이, 파블로프의 개 이야기를 예로 든다면요. 파블로프가, 파블로프가 관찰했을 때, 음, 먹이를 주는 사람이 개가 있는 방으로 들어올 때 말입니다. 개는 머물던 곳에서 침을 흘리기 시작할 것입니다. 왜냐하면 발소리는 항상 음식이 곧 나올 것이라는 신호였기 때문이지만, 개는 문이나 방의 문손잡이를 보고는 침을 흘리지 않았습니다. 네. 그것은 고정된 특징이었기 때문이죠. 음식이 오든 안 오든 항상 그 자리에 있었으니까요. 네. 존 도(John Doe)는 항상 그 자리에 있었고, 그러다 갑자기 존 도와 혐오 자극이 함께 나타나게 된다면, 어, 효과적인 고전적 조건 형성이 일어날 가능성은 낮아집니다. 그래서 제 생각에 이것이 의미하는 바는, 만약 개 훈련사가, 일을 잘하고 있고 주로 정적 강화와 연관되어 있다면, 가끔 혐오 자극이 포함될 수도 있는 실수를 한다고 해서 그 훈련사가 자격이 없다고 할 수는 없으며, 그것이, 어떤 정서적 반응을 만들어서는 안 되며 일시적이어야 한다는 것입니다. 그것이 바로, 기초 문헌들이 시사하는 바입니다. 그리고, 사실 그보다 더 나아가서, 처벌에 관한 임상 문헌은, 어, 존재하는 그대로, 처벌로 인한 심각한 정서적 부작용에 대한 증거를 거의 찾지 못합니다. 그 처벌이 강화 또한 제공하는 동일한 치료사에 의해 수행될 때 말이죠. 네, 제 생각에 이것은 매우 중요한, 구분이라고 생각합니다. 그, 그, 사람이나 동물이 그 존재가 되어서, 관계를 맺고 신뢰해야 한다는 점 말입니다. 그리고 어느 정도 관계를 형성해야 하죠 그 사람이나 훈련사와 처벌을 실행하는 것은 매우 중요합니다 그리고 물론 교육을 위한 장소나 필요하다면 화해의 과정이 필요한데
John Doe what you know who's a tall man with a beard and wears a red cap walks into the dog training session and starts kicking the dog uh I think there's a pretty good chance that the next time's John Doe shows up the dog is going to uh run away from him that is that he will become a conditioned averse of stimulus by virtue of classical conditioning now what the most relevant aspect of John Doe is might not be clear might not have anything to do with the cap the beard or his height it might be his scent uh it might be as sound uh but um uh clearly that would create you know that that would be a distinctive stimulus a new stimulus and a new con a new pairing of the pain with that stimulus but if if John Doe were part of the normal training of the dog part of the environment for months and then uncharacteristically engaged in this Behavior it would be much less likely that he would become a conditioned diverse of stimulus so you know look if you uh if if you if if you take the story of Pavlov's dog uh you know Pavlo when he when Pavlov observed is when the uh the feeder would come into the uh the room where the dog was housed the dog would start to salivate because the feet are always signaled that food was forthcoming but the dog didn't salivate to the door to the site of the doorknob to the room yes because it was a static feature was there whether food was coming or not yes it wasn't John Doe was there all the time and then and then suddenly there's a pairing of John Doe with with an aversive stimulus uh it's it's it's less likely that that effective classical conditioning will occur so I guess what that would mean is if if a dog trainer is doing you know a good job and uh is primarily associated with positive reinforcement uh an occasional mistake that might involve diversive stimuli shouldn't uh disqualify the train or it shouldn't be it shouldn't create the any emotional response should be temporary that's what that's what the uh basic literature would suggest and and and actually goes further than that the the clinical literature on punishment such as such as it exists uh finds very little evidence of serious uh emotional byproducts of punishment when it's carried out by the same therapists who also provide reinforcement yes it's very I think this is a very important uh distinction to be made that the the being the person or the animal they have to be they have to relate and trust and and have some some relationship with the the person or trainer that does implement the punishment it's critical and then there is a of course a a place for a school or reconciliation if needed to where it's
1:03:47
아니요 아니요 저는 정말로 당신을 미워하지 않아요 저는 당신을 해치려는 게 아니에요 단지 그건 하지 말라는 것이죠 알겠죠 음 제가 아기 고양이를 키운다고 말씀드렸었죠 그 아기 고양이는 어 얼마 전까지만 해도 정말 작았고 저를 졸졸 따라다녔어요 믿기 힘들 정도로요 그래서 제 발 밑으로 바로 들어오곤 했는데 한 번은 제가 뒷걸음질을 치다가 그 고양이 아기 고양이를 밟고 말았어요 그리고 이 아기 고양이는 정말 불행한 상태였죠 아팠고 당연히 도망갔어요 하지만 몇 분 뒤에 다시 돌아와서 제 무릎 위에 앉아 골골거렸죠 있잖아요 제 생각에 만약 그 고양이가 저와 겪은 경험의 전부가 그냥 제 삶에 나타나서 고양이를 밟는 것이었다면 네 다시 돌아오지 않았을 거예요 맞아요 하지만 그것은 그저 하나의 사건이었을 뿐이고 그 외에는 상당히 풍부한 강화가 이루어지는 관계였거든요 제가 항상 먹이를 줬으니까요 그렇죠 사랑을 주기도 하고요 등등 네 맞아요 이건 우리가 어떻게든 항상 반려견 훈련에서 이해하는 데 실패하는 중요성 같은 것인데 이런 식이죠 예를 들어볼게요 훈련사가 반려견을 데려왔다고 가정해 봅시다 위탁 훈련과 같은 교육 환경에서 행동 문제가 있을 때 그리고 매우 종종, 특히 경험이 부족한 훈련사들은 이렇게 하곤 하죠 개들이 오자마자 바로 그 문제를 해결하려고 서두릅니다 신뢰를 쌓지도 않은 채로요 그러니까 이봐, 나는 나는 너를 좋아해, 우리는 사실 너랑 알다시피, 그런 과정이요 비록 우리가 행동을 억제하고는 있지만 성공한 것처럼 보일지라도 개에게는 돌아갈 곳이 없습니다 음, 제가 찾으려는 단어가 무엇이냐면요 그러니까 안도감을 얻는 것이죠, 그러니까 제가 너는 나를 적대시하는 게 아니라, 단지 그 행동을 좋아하지 않는 것뿐이라는 안도감 말이죠 그리고 이것은 제가 생각하기에 매우 중요한 지점입니다 음, 처벌에는 정말 많은 규칙이 있지만 이것은 정말 중요한 하나죠 유대감을 갖는 것 어, 동물이나 사람이 당신을 특정 방식으로 바라보게 만드는 것 우리가 무언가를 시작하기 전에 말이죠 알다시피, 그의 책에서 강압과 그 후유증에 대해 시드먼은 말합니다 전기 충격을 사용하는 사람은 상어가 된다고요 네, 하지만 그건 사실이 아닙니다 오직 그 사람이 그것만 한다면 사실일 뿐이죠 만약 그게 사실이라면 우리는 모두 부모님을 미워했을 겁니다 알다시피, 적어도 전통적으로 제가 자란 50년대에는 잘못된 행동을 하면 체벌을 받았으니까요 하지만 저는 멈추지 않았죠 어머니가 나중에 아버지 오시면 두고 보자고 말씀하시던 게 기억나요. 정말 고전적이죠. 하지만 전 아버지를 미워하지 않았어요. 네, 아주 좋은 지적입니다. 그분이 하시는 또 다른 말씀이 있는데, 정확히 기억나지 않아서 제가 의역을 좀 할게요. 본질적으로 처벌하는 사람은 이런 강화 효과를 얻게 되고, 그게 만족감을 준다는 거예요.
like no no I I really I don't hate you I don't want to kill you it's just don't don't do this right well I mentioned to you that I have a kitten and uh the the not too long ago this kitten was very small and it followed me around like you wouldn't believe and was would get right under my feet and one time I took a step backwards and I stepped on the cat the kitten and this was a very unhappy kitten it hurt and and she ran away obviously but a few minutes later she was back and she was sitting on my lap and purring you know and uh I think if her entire experience was with me is that if I just walked into her life and stepped on her yeah she wouldn't have come back yes but that was just one uh incident in what otherwise was a pretty richly reinforcing uh relationship because I fed her all the time you know yes and they're in loved her and so on yes yeah this this is something that somehow we we always in in dog training we kind of failed to to understand the importance of this like like if you I'll give you some examples we let's say a trainer gets a dog in training in like you know board and train and there is a behavior problem and very often especially trainers that are an experience they would rush into tackling that problem today the moment the dog comes in without having to build that trust and like hey I I like you like we we actually you know and and then things even though we are suppressing Behavior it seems like we're successful the dog has no nowhere to to go back to um what is the word I'm looking forward to to get at comfort of hey I am you are not against me as a being you're just not liking that behavior and and this is a I think it's a critical point when we well there is many many rules with punishment but this is one big one to to have a report uh to where the animal or the human really looks at you in a certain way before before we start something to do like that you know uh in his book on coercion and it's Fallout Sidman says that a person who uses shock becomes a shark yes but it but it's not true it's it's only true if that's the only thing the person does um if it were true we would all hate our parents you know be uh or at least traditionally you know when I I grew up in the 50s and uh when you misbehaved you were spanked but I didn't stop you know I can remember my mother saying wait until your father gets home right classic but I didn't hate my father yes a very good point and another one he makes I I will paraphrase because I'm not sure but it's the the the person that's punishing basically gets this reinforcement and grad it's gratifying
1:07:44
그래서 처벌을 하면 할수록 더 처벌하고 싶어지죠. 그 부분도 저는 근거가 있다고 생각하지 않아요. 글쎄요, 어... 저라면... 만약... 제 생각에는 그럴 가능성이 가장 높은 경우는, 음... 어... 그러니까 개가 잘못된 행동을 할 때 말이죠. 당신이 전문 반려견 훈련사이고 개가 잘못된 행동을 하고 있을 때, 그 행동을 멈추기 위해 어떤 가벼운 처벌 절차를 사용해서 행동을 중단시키고, 그래서 다시 긍정적인 훈련 요소로 넘어갈 수 있게 한다면 그건... 어... 잘못된 행동을 멈추게 하는 것이 강화가 될 수는 있겠지만, 그것이 누군가가 어떤 상황에 매우 화가 나서 폭발했는데 그 상황이 사라져 버리는 경우만큼 강력한 강화제는 아니에요. 바로 그런 종류의 일이 너무나도 강력한 강화제가 되어서, 그게... 그런 일이 되는 거죠. 그래서 도구 상자에서 가장 먼저 꺼내는 도구가 처벌인 사람에 대해서는 걱정이 되긴 합니다. 네, 분명히 우리가 보고 싶어 하는 모습은 절대 아니죠. 맞습니다, 맞아요. 당신은 특정한 사람들이 특정 사람이 왜 그런 선택을 하는지 좋아하는 경향이 있습니다 그 길을 택하는 대신 그게 그 사람의 가정 교육 때문일까요 왜냐하면 우리는 종종 그 사람이 이런 환경에서 자랐기 때문에 이런 것들을 다 겪었으니 그래서 지금 그런 거라고 말하곤 하죠 음 네, 그니까 그거 있잖아요 아주 오래된 표현이 있죠 아마 18세기까지 거슬러 올라갈 텐데 매를 아끼면 아이를 망친다는 말이요 그 생각은 만약 당신이 잘못된 행동을 처벌하지 않으면 버릇없는 아이가 될 거라는 식이었죠 그런데 우리는 이제 더 이상 그렇게 생각하지 않습니다 우리는 처벌을 신중하게 선택적으로 사용해야 할 것이라고 보며 오직 적절한 행동을 강화하는 환경 안에서만 사용해야 한다고 봅니다 처벌 그 자체로는 멈추게 하는 것 외에는 별다른 효과가 없습니다 제 말은 봐요, 만약에 당신이 앞마당에 서 있는데 작은 아이가 갑자기 도로로 뛰어든다면 당신의 자연스러운 반응은 아이에게 꽤 위협적일 수 있겠죠 그리고 그건 자연스러운 반응일 겁니다 그게 아마도 아이가 다시 도로로 뛰어들 가능성을 크게 줄여줄 겁니다 음 그건 좋은 거죠 왜냐면 그런 행동은 정말 멈춰야 하니까요 하지만 그게 아이에게 길을 어떻게 건너야 하는지 가르쳐주지는 않잖아요 네 맞습니다, 하지만 그게 바로 우리가 그걸 가지고 있는 이유죠 강화, 맞습니다, 정확히 맞는 말씀입니다, 그래서 어 처벌 자체로는 어 무언가를 성취할 수 없을 겁니다 필요한 것을 달성하는 데는 말이죠, 그것은 특정한 일들을 이룰 수는 있겠지만 새로운 행동을 훈련시키지는 못합니다, 반면에 반면에 음 만약 만약 길거리로 자주 뛰어드는 아이가 있다면 그 아이에게 긍정적 강화를 할 기회는 많지 않을 것입니다 강화 지연 조건 형성 그게 효과가 있나요? 그리고 효과가 있다면 언제 어떨 때일까요? 글쎄요 음
so so the more you punish the more you want to punish and that also I don't think uh has any ground well you know uh I if if uh I would think that that would be most likely if the um well if if the dog is misbehaving if you're a professional dog trainer and the dog is misbehaving and you do some sort of mild punishment procedure to to stop the behavior so you can move on with the Positive elements of the training um it may be reinforcing to get the misbehavior to stop but that's not as potent of a reinforcer as it is when someone is becomes really angry at some situation lashes out and then the situation goes away that's uh that's the kind of thing that becomes such a potent reinforcer that you do worry about someone who the first tool that they pull out of the toolbox is punishment yeah that's certainly not what not not what we want to see correct correct do you think certain people tend to like what makes a certain person go that route instead of is it is it the upbringing of this person because a lot of times what we also would say well that person grew up in this kind of environment he he got all this so therefore he is now um you know yeah I mean you know there's there's an old expression uh uh it probably goes back to the 18th century spare the rod and spoil the child the uh the idea was that if you don't uh if you don't punish misbehavior you're going to have a rotten kid I guess and uh we don't see things quite that way anymore we see punishment as something that should be used uh cautiously and selectively and only in the context of a an environment that reinforces appropriate behavior punishment by itself doesn't accomplish a lot except to stop I mean look uh if if you were if you were if you were uh standing in your front yard and your small child suddenly darted out into the road your natural reaction might be quite frightening to the child and uh it would be a natural reaction and it would probably uh uh greatly reduce the likelihood that the kid runs out into the street again um that's a good thing right because you really have to have that kind of behavior stop but but in terms of but it doesn't teach the kid how to cross the street yes right but that's why we have reinforcement that's right that's exactly right so uh punishment by itself isn't going to accomplish uh what you need to have accomplished it might accomplish certain things but it doesn't it doesn't train new behavior on the other hand um if you if you have a kid that runs out into the street a lot you're not going to have many opportunities for positive reinforcement delayed conditioning does that work and if it works when would that be like like well um
1:11:32
실험실에서는 복도 끝에 있는 제 동료 앤디 라텔이 일련의 매우 정교한 실험을 통해 동물들도 학습한다는 것을 보여주었습니다, 그들은 꽤 긴 시간 지연된 강화 조건에서도 행동을 습득할 것입니다, 하지만 하지만 시간이 오래 걸립니다 그리고 그 어 그 결과는 즉각적인 강화제를 몇 번 사용하는 것과 비교하면 매우 미미합니다 강화제를 제공하는 것 말이죠 음, 단기적으로 일어나는 일이 무엇인지 아시죠 단기적인 효과는 장기적으로 일어나는 모든 것을 압도합니다 장기적으로 말입니다 그래서 저는 어 이렇게 말하는 것이 타당하다고 생각합니다, 지연 강화나 지연 강화 혹은 그 문제에 있어서 지연 처벌도 말이죠 네 아주 긴 지연 시간으로는 효과가 없습니다 어 두 가지를 연결하는 일종의 신호가 있지 않는 한 말이죠 두 가지를 연결하는 그게 바로 클리커 트레이닝이라고 생각합니다 그래서 만약 개가 어떤 행동을 했을 때 부저 소리가 10초 동안 독특한 부저음이 울리고 그다음에 음식이 제공된다면 아마 잘 작동하겠지만, 만약 개가 어떤 행동을 하고 10초 후에 음식을 받는다면 그건 아마 별로 효과적이지 않을 것입니다. 초기 습득 단계에서는 유지 단계에서는 괜찮을지 몰라도 초기 습득 단계에서처럼 행동이 매우 약할 때는 그렇지 않습니다. 즉각적인 강화는 음 작동하는 유일한 방법은 아닐지 몰라도 다른 어떤 방법보다도 더 효과적일 것입니다. 그럼 후행 조건 형성에 대해서는 어떤가요? 그것은 항상 논하기 매우 어려운 주제입니다. 적어도 개 훈련에서는요. 후행 조건 형성을 살펴보면 알다시피 고전적인 종류의 후행 조건 형성 실험들은 교과서에서 다루는 그런 부류의 것들인데요. 그런 것들을 보면 가끔씩 후행 조건 형성이 일부 조건 반응을 생성하기도 하지만 조건 반응으로 간주될 만한 것들이지만 그건 확실한 효과가 아니며 아마도 일종의 인위적인 결과물일 것입니다. 즉 소음 같은 것이죠. 만약 당신이 고전적 조건 반응을 만들어내고 싶다면 조건 자극은 무조건 자극보다 앞서야 합니다 설령 그것들이 동시에 발생하더라도 효과가 별로 좋지 않습니다 음 하나가 다른 하나보다 앞서야 합니다 수학적인 의미에서 예측해야 합니다 예를 들어 신호음이 음식을 예측해야 하죠 음 만약 동시에 온다면 신호음은 아무런 역할을 하지 못합니다 맞습니다 아무것도 예측하지 못하죠 사실은 동물이 그것을 알아차리지 못할 수도 있습니다 왜냐하면 그것은 음식에 의해 가려지기 때문인데 음식은 훨씬 더 두드러진 자극이기 때문입니다 몇몇 연구들이 있었는데 저는 저는 참고 문헌을 잘 기억하지 못하지만 그런 내용이 있었어요 지연 조건 형성(backward conditioning)이 실제로 아주 잘 작동할 수 있다는 것을 경우에 따라 어떤 포식 본능과 관련해서 이를테면 말이죠
in the laboratory my colleague down the hall Andy lattell has shown in a series of very elegant experiments that animals will learn they will acquire Behavior under conditions of Fairly long delayed reinforcement but but it takes a long time and the the uh the the end result is very small compared to what you would get with just a few immediate reinforcers um you know what happens in the short term triumphs or anything that happens in the long term so I would think that uh it would be fair to say that delayed reinforcement or delayed punishment for that matter yeah it doesn't work with with very long delays uh unless there's some sort of signal that connects the two I think that's what clicker training is about and you know if if uh if if if the dog does something and a buzzer sound a unique buzzer sounds for 10 seconds and then the food is delivered that probably will work fine but if the dog does something and 10 seconds later it gets the food that's probably not going to work very well during initial acquisition it might it might work okay in maintenance but what the behavior is is really weak uh as it is in the early stages of acquisition immediate reinforcement is uh well it might not be the only thing that works but it's going to work better than anything else how how about the backwards conditioning like that's a that's always very difficult one to discuss at least in dog training if you look at if you look at backwards conditioning uh the you know the classic kind of backward conditioning experiments that are of the sort that are reported in textbooks things like that you will see that occasionally backwards conditioning will generate some conditioned responses or what count as condition responses but it's it's not a reliable effect and it's probably an artifact of some kind it's uh it's noise it's it's the the uh if you if if you want to generate classically conditioned responses the the conditions stimulus has to precede the unconditioned stimulus even if the if they occur at the same time it doesn't work very well um one has to proceed the other it has to in in the mathematical sense predict like the tone has to predict the food um if it if they come at the same time the tone isn't doing anything correct it doesn't predict anything in fact it might not even be noticed by the animal because it will be overshadowed by the food which is much more Salient stimulus there were some studies and I I I'm so bad with references but there there was something to where they were pointing out the backward conditioning could actually be could could work very well in case of uh um some predatory instincts to to where let's say let's say
1:15:28
무언가 일이 일어나고 쥐에게 그런 다음 고양이가 걸어가는 경우죠 음 그리고 그 연합을 만들어낼 수 있는 유전적인 축적이 있습니다 그리고 저는 그 연구자들의 이름이 떠오르질 않네요 음 아주 흥미로웠고 네 저는 정말 잘 모르겠어요 대부분의 애견 훈련사들에게 지연 조건 형성은 우리가 학습을 위해 의존하거나 사용하는 그런 방식은 아닙니다 네 훈련 프로그램을 작성할 때 지연 조건 형성에 의존하지는 않겠죠 맞습니다 음 네 당신이 설명하는 것은 충분히 그럴 수 있겠다고 상상할 수 있어요 그럴 수 있죠 음 어 70년대에는 관심이 많았습니다 준비성에 대해, 그러니까 유기체가 준비되어 있다는 개념인데, 제가 이렇게 말하고 싶지는 않지만 제 입장에서는 조금 인지적인 측면이 있긴 하죠 하지만 특정 종류의 사건들을 연관 짓도록 준비되어 있다는 것이죠 그리고 이것은 여러 가지 방식으로 증명되었습니다 조작적 조건화의 경우를 보면요 만약 여러분이 쥐에게 전기 충격을 피하도록 가르치고 싶을 때 그 반응이 상자 밖으로 뛰어 나오는 것이라면 그것을 두세 번의 시도 안에 배울 겁니다 왜냐하면 동물에게 있어서 혐오 자극의 원인으로부터 물리적으로 멀어지려고 하는 것은 자연스러운 행동이기 때문이죠 만약 쥐가 제자리에 서서 레버를 누름으로써 전기 충격을 피하게 하려면 훨씬 오래 걸립니다 알겠죠? 핵심은 동물들이 한 가지 방식은 다른 방식보다 더 잘 배우도록 준비되어 있다는 것입니다 또 다른 예는 조건성 미각 혐오입니다 네, 존 가르시아의 실험이죠 쥐에게 이전에 맛본 적 없는 독특한 맛을 경험하게 한 뒤에 그것을 구토를 유발하는 것과 연결하면 단 한 번의 시도만으로도 그 맛을 피하게 될 겁니다 하지만 맛 대신 시각적 단서, 이를테면 번쩍이는 불빛이나 소리를 사용하면 그것들은 피하지 않죠. 그러니까 그것은 변별 자극의 선택은 조건 자극의 선택은 순전히 임의적인 것이 아니며 쥐가 실험실로 가져오는 어떤 것, 즉 유전적 유산에 관한 무언가에 달려 있습니다. 그러니 그런 관점에서 저는 이런 상황이 있을 수 있다고 생각합니다. 음, 매우 근접한 후행 조건화가 효과적일 수는 있겠지만, 여전히 선행 조건화만큼 효과적이지는 않을 것입니다. 맞죠? 다음 주제는 아주 큰 주제입니다. 학습된 무력감입니다. 그리고 그것은 아주 마치 아주 흔하게, 반려견 훈련 업계에서 아주 자주 거론되는 것으로 마치 우리가 그 상태에 도달하기가 매우 쉽고 빠져나갈 방법이 없으며, 그리고 그것이 영원히 지속된다는 식으로 말이죠. 저도 의견이 있고, 여러 가지 연구들을 읽어보았습니다. 음, 마이어(Maier)가 원래 연구 이후 2~3년 뒤에 후속 연구를 진행한 것도 알고 있습니다. 그들은 동물을 노출시키고 탈출하고 회피하는 방법을 가르쳤는데,
the something something happens to the rat and then the cat walks away um and that is there is that genetic buildup that can create that Association and I I cannot think of the name of the the researchers um that it was very interesting and yeah I really don't know much like like to to most of dog trainers backward conditioning is not something we we rely or search on uh for learning to to happen yeah you wouldn't program you wouldn't write a training program that relies on backwards condition right um yeah well the the thing that you're describing I can imagine that that could be the case um uh in the um in the 70s there was a lot of interest in um preparedness you know the idea that uh organisms are prepared to I wouldn't put it this way but it's a little bit cognitive for me but they're prepared to associate certain kinds of events and this was shown in a number of ways so just in the case of operant conditioning if you take if if you want to teach a rat to avoid shock and the response is jumping out of the box they'll learn it in two or three trials if uh because it's a natural thing for the animal to do is to try to move physically move away from the source of aversive stimulation if you require the rat to avoid Shock by stance staying in place and pressing a lever it takes a lot longer okay and the idea is that the animals prepared to learn One Way rather than the other another example is with a condition tasterversion yes John Garcia's experiments right you give the you give the rat uh a distinctive flavor something that's never tasted before and you follow that up with something that makes it sick well they'll after just one trial they will avoid that flavor but if instead of a flavor you use a visual space it sounds like a flashing light or a sound they don't avoid that so they it's the the choice of the of the discriminate the choice of the conditioned stimulus isn't purely arbitrary it depends on something that the the rat is bringing to the laboratory something about their genetic Heritage so you know in light of that I can imagine there could be situations where um very close backward conditioning might might be effective but it still wouldn't be as effective as this forward right so my next one is very big one learned helplessness and it it is a very like it's a very common very often brought up in the dog training industry to where like if we that it's very easy to reach that state and there is no way out and and it's there forever I have opinions I've I've read different uh um studies I know I know that even Maya right after like two or three years after the original study they they did a follow-up studies to where they
1:19:18
그들이 그런 정신 상태에 빠지는 것에 대해 훨씬 더 큰 저항력과 회복력을 보였습니다. 즉, 포기하는 것에 대해서 말이죠. 하지만 이 문제에 대한 당신의 견해를 정말 듣고 싶습니다. 이건 반려견 훈련 커뮤니티에서 매우 큰 이슈거든요. 네, 그게 큰 문제인가요? 사람들이 이 현상이 자신의 훈련을 방해한다고 보기 때문인가요? 아니요, 그것은 거의 허구에 가깝습니다. 거의 아니라고 할 수 있죠. 만약 만약 특정 행동을 하신다면 그러면 만약 훈련에 어떤 형태로든 혐오 자극을 사용한다면 개에게 학습된 무기력이 나타날 위험이 매우 큽니다. 따라서 무슨 수를 써서라도 피해야 합니다. 그게 바로 그 이야기였죠. 저는 셀리그먼과 마이어가 대학원에서 uh 그 초기 연구를 수행하고 있을 때였는데요. 그 연구에서 인상적이었던 점 중 하나는 기억하시겠지만 원래 연구는 개들을 대상으로 수행되었습니다. 그리고 uh 그들이 사용한 것은 제 기억이 맞다면 발바닥에 가해진 매우 강력한 전기 충격이었습니다. 네, 너무나 강력한 충격이라서 분출성 배변을 유발할 정도였죠. 그러니 아시다시피 이건 평범한 종류의 uh 혐오 자극과는 거리가 멀며, 전문적인 반려견 훈련사가 사용할 만한 것이 아니었습니다. 이건 정말 아주 극심한 강도의 uh 자극이었습니다. 저는 전문가가 아니고 그 연구를 살펴본 지도 아주 오래되어서 잘은 모르지만 그것이 바로 많은 사람들이 그 초기 연구에 대해 충격받았던 점입니다. 제 생각에는 uh 어, 흥미로운 점은 인간의 um uh 응용 문헌에서 uh 처벌이 더 더 널리 받아들여졌던 시절에는 문제는 결코 uh 학습된 무기력을 줄이는 것이 아니었습니다. 음, 문제는 학습된 무기력에 대한 우려가 아니라 바람직하지 않은 행동을 억제하는 데 효과적인 처벌인을 사용하는 것에 대한 우려는 효과적이었던 행동 [음악] 음 그건 뭐랄까, 우리가 조심하지 않으면 모든 행동이 멈출 것이라는 그런 게 아니었어요. 그런 일은 일어날 것 같지 않아요. 처벌이 기본적으로 강화 기반 절차의 아주 작은 일부로 존재하는 환경에서는 말이죠. 그 당시 실험들에서 일어났던 상황은 그런 방식이 아니었어요. 학습된 무기력 실험의 경우, 동물이 마치 이 절차가 치료적인 환경 속에 녹아들어 있다고 느끼는 그런 상황이 전혀 아니었습니다. 그렇죠. 네, 아시다시피 이건 중개 연구가 아니었어요. 맞아요, 이건 기초 연구였지, 정말 음 제 생각엔 그 실험들이 불러일으킨 응용적인 우려들은 아마도 과장되었던 것 같아요. 네. 이제는 그에 대해 별로 들어보지 못하셨겠지만요. 제 말은, 사실 우리 모두 정말 놀라실 수도 있겠지만, 네, 괜찮아요, 어쩌면 많이 들어보셨을 수도 있죠. 네, 반려견 산업에서는, 반려견 훈련 커뮤니티에서는 여전히 특정 분야에서 가장 중요하게 다뤄지는 주제거든요. 네, 맞아요, 아주 끊임없이 언급되는 내용이죠. 음, 그리고
exposed the animal and they taught it how to escape and avoid and they showed much more resistance and resilience to to be coming into that mental state to giving up but I I really want to hear your take on this one it's a very big one in the dog training Community yeah is it is it a big issue because people see this uh phenomenon interfering with their training no it's almost fictitious it's almost like no no if if you're doing certain things you will you will like if you use some form of aversion in in the training it you're really risking the dog to go into learning helplessness therefore you should avoid it at all costs that that's the The Narrative I was uh I was in graduate school when uh Seligman and Meyer were doing that original research and one of the Striking things about it remember the original research was done with dogs and uh they were using such severe electric shock applied is if I remember correctly to the pads yes that it was such a severe shock that It produced projectile defecation so you know this wasn't the run-of-the-mill kind of uh aversive stimuli that are likely to be used in by a professional dog trainer this was this was really very special shame intense uh stimulation now I'm not an expert it's been a long time since I've looked at that at that work but that's what struck many people about the original research I don't think that uh you know in in the uh in the it's it's it's interesting in the human um uh applied literature when uh punishment was was more was viewed as more acceptable the issue was never uh reducing learned helpnesses uh well the issue wasn't a concern with learned helplessness it was a concern with having a a Punisher that was effective in suppressing the bad behavior [Music] um it was it wasn't like oh if we're not careful all behavior is going to stop it doesn't seem likely in a in an environment where punishment is a very small component of an otherwise reinforcement-based procedure and that's not what was going on in those experiments and learned helplessness it's not like the animal was like this procedure was embedded in a uh in a therapeutic environment it wasn't that at all it was yes yes you know it was not translational research right it was it was basic research it wasn't it it really um I think I think the cons the the applied concerns that it generated were probably overstated okay you don't hear much about it anymore I mean we all actually we you're really surprised it's it's okay you may hear a lot of it yes in the dog industry the dog training in uh communities it's in the top of the list in certain really yes yes it's a very it's a it constantly is brought up um and and
1:23:09
네, 만약 우리가 방금 말씀하신 대로 한다면, 당연히 극단적인 결과들이 나타나겠죠. 음 네, 하지만 이건, 이건 사용될 수 있는 하나의 카드인 거죠. 설득해 보세요 아니요, 당신은 하게 될 겁니다, 이건 아주 잘 기록되어 있고, 바로 이 부분이 당신이 향하고 있는 곳입니다, 당신은 결국 어, 학습된 무력감을 만들게 될 것이고 개는 기본적으로 절대 회피하려 하지 않게 될 것이며, 어 그리고 포기하게 될 겁니다, 그리고 평생 동안 빠져나갈 방법은 없을 겁니다 글쎄요, 그것은 매우 무서운 전략입니다 제가 읽은 문헌에서는 그것에 대한 언급을 많이 보지 못했습니다 음, 아마 제가 말하는 기본 문헌, 어, 조작적 조건 형성 문헌에는 있을지 모르겠네요 음, 하지만 그것은 어, 신경과학 분야에서 흔히 사용되는 절차인 어, 강제 수영 실험을 떠올리게 합니다 이 실험의 작동 방식은 쥐나 생쥐, 설치류를 잡아서 물탱크에 넣는 것입니다 따뜻한 물이라서 충격 같은 것은 아니지만, 그럼 쥐는 무엇을 하나요? 개가 발을 저으며 물에 떠 있는 것처럼 헤엄을 칩니다 그러다가 잠시 후 멈추죠, 자, 멈췄을 때 익사하는 게 아니라 물에 뜹니다 그리고 이것은 보통 동물이 포기하는 데 걸리는 시간으로 여겨지는데, 그들이 표현하자면 포기하는 것이죠 이것은 동물이 얼마나 걸리는지, 어, 얼마나 우울해질 가능성이 있는지에 대한 징후로 간주되죠, 맞습니다, 그리고 포기한다는 것은 우울증의 징후로 여겨지지만, 물론 다른 가능성은 동물이 물에 뜨는 법을 배웠다는 것입니다, 훨씬 적은 노력으로 물 위에 머무르는 법을 배운 것이죠 그리고 많이 있어요, 아주 많은 제 주변의 보건과학 센터에서 대화하는 사람들 중 그렇게 생각하는 사람들이요. 그들은 말하죠. 연구비 지원 기관들이 요구하기 때문에 연구에 이것을 사용해야만 한다고 말이죠. 하지만 저는 그게 말도 안 된다고 생각해요. 왜냐하면 이건 우울증 모델로는 적합하지 않고, 학습 모델로는 적합하기 때문이에요. 학습을 위한 것이죠. 그러니까 그게, 그러니까 학습된 무기력 같은 건가요? 글쎄요, 아니에요. 그건 어떻게 어, 떠 있을지 배우는 거예요. 퍼즐에 대한 답을 찾는 거죠. 맞아요, 마이클. 혹시 하고 싶은 이야기가 있나요? 제가 질문을 너무 많이 한다는 건 알지만, 당신이 생각하기에 분명 무언가 있을 것 같아서요. 당신이 생각하기에 흥미롭고 개와 관련이 있으며, 당신의 경험을 통해 반려견 훈련사들에게 도움이 될 만한 것들이요. 글쎄요, 저는 잘 모르겠어요. 저는 반려견 훈련사들에게 가치 있는 말을 해줄 수 있는 그런 입장이 아닌 것 같아요. 왜냐하면 저는 그저 쥐 심리학자일 뿐이니까요. 저는 그저 어, 기본적인 원리만 연구하거든요. 대부분의 경우에 그렇지 않죠. 그럼 쥐 실험의 기본 원리에서 쥐와 실험을 계속 진행하기 위해 갖춰야 할 근본적인 요소들은 무엇인가요? 그리고 쥐가 무언가를 하는 것을 두려워하지 않게 하고, 잘못된 반응을 보이지 않게 하려면요? 음, 알다시피 어, 우리가 쥐에게 어떻게 회피하는지 가르칠 때,
yes if we do if we if we do what you just said of course we're gonna have some extreme results um but yes it is um it's one of the the cars that could be played to to to convince somebody no no you will this is well documented and and this is where you're headed you will you will create uh learn helplessness and the dog will just basically never try to avoid uh and and will give up and there will be no way out for for the rest of its life well it's a very scary tactic in the literature that I read I don't see much uh reference to it um maybe there are literature I'm talking about basic sure uh operate literature um but you know it does remind me of a procedure that's commonly used in uh um the neurosciences and that's uh the forced swim test and the way this works is you take a rat or a mouse you take a rodent and you put it in a tank of water warm water so it's not a shock or anything like that and what is what is the what Does The Rat do it does the dog paddle you know it Treads water and after a while it stops now when it stops it doesn't drown it floats and this is commonly the time it takes for the animals to give up the way they would put it is give up is is supposed to be uh a sign of how long it takes for the end how depressed the animals likely to be right and and giving up is a sign is is taken as a sign of depression but of course the alternative is that the animal has learned to float it's learned to stay on top of the water with much less effort and there's a lot there's a there's a lot of people around that I talk to in our Health Sciences Center that see it that way they say I have to use this in my research because the grant agencies expect it but I think it's ridiculous because it's not a good model for depression it's a good model for learning for learning so is it is it I mean there's it's like learned helplessness well no it's not it's learning how to uh float finding answers to the puzzle exactly right Michael anything that you want to talk about because I know that I I am asking a lot of questions but I know that there is probably something that you anything that you feel that it's interesting and related to dogs and it will benefit dog trainers from your experience of of I I don't I don't know I don't know that I have anything uh worthwhile to say to dog trainers uh because I'm just a rat psychologist you know I uh I just study basic principles and in most cases were not uh so in the basic principles with trots what are the fundamentals that need to be there in order to be able to continue with with experimenting with the rat and not not uh not have the rat being scared to do things for you and and have this
1:26:53
어, 글쎄요, 당신도 알다시피 어 전기 충격 우리가 하는 일 중 하나는 매우 신중하게 횟수를 조절하는 것입니다 전기 충격을 주는 그것은 독특한 자극이며 동물이 자연 상태에서 접할 수 있는 것이 아닙니다 그리고 음 분명히 동물들을 겁먹게 합니다 그들은 어 그들은 낑낑거리고 대소변을 보기도 합니다 그리고 어 그런 반응은 첫 번째나 두 번째 충격을 받을 때 나타납니다 그들은 그런 종류의 감정적인 반응을 보이죠 하지만 몇 시간의 훈련 후 2~3일이 지나면 동물이 효과적인 회피 행동을 습득하면서 그런 반응은 모두 사라집니다 그리고 어 2주 정도의 훈련을 거치고 나면 하루에 몇 시간씩 일주일에 5~6일 훈련하면 여러분이 보게 되는 것은 완전히 편안해진 동물입니다 그들은 어 그저 어 겉으로 드러나는 행동에서는 아무것도 보이지 않습니다 그들이 어 평생 동안 네 맞아요 네 오 전혀 그렇지 않아요 제 말은 그들은 어 음식 실험에 참여하는 쥐들만큼 오래 삽니다 저는 지금까지 충격 회피 실험을 하는 쥐에게 물린 적이 없습니다 하지만 음식 실험에서는 쥐에게 물린 적이 있죠 음식 강화 실험에서는 그런 일이 자주 일어나지 않아요 가끔 일어나는데 제 생각에는 그것은 동물들이 음 음식을 먹지 못한 상태이기 때문입니다 어 그런 이유로 어느 정도의 흥분을 유발합니다 그들을 다루기 시작할 때 그것이 음식에 대한 예고가 되어서 그들이 항의하기 시작하는 거죠 네 맞아요 그건 사실 공격성이라기보다는 음, 무엇을 먹어야 할지에 대한 혼란에 더 가깝습니다 음 하지만 충격이나 바람 연구에서 그런 동물들에게서는 그런 모습을 볼 수 없습니다 그들은 꽤 느긋하고 그들은 음 꽤 우호적입니다 네, 왜냐하면 가장 큰 주장 중 하나 그리고 제 말은 스키너는 처벌에 반대하는 많은 주장을 펼쳤지만 한 가지 큰 주장은 소위 반격이었죠 처벌에 대한 동물의 반응 어떻게, 타임아웃에 대해 연구하고 실험해 보신 적이 있는 걸로 아는데요 그거에 대해 어떻게 생각하시나요 타임아웃이 음 타임아웃은 어떻게 영향을 미치나요, 잠깐만요 좀 더 자세히 말씀드리자면 저희 반려견 훈련계에서는 일부 훈련사들이 타임아웃이 사실 매우 무해하다고 생각해서 거의 언제든 무엇에든 사용할 수 있고 주요 수단으로 쓰곤 합니다 정적 처벌과 타임아웃을 비교할 때 말이죠 그게 맞는 말인가요 음, 아시다시피 저와 제 학생들은 작년에 정적 강화로부터의 타임아웃에 관한 일곱 가지 실험을 담은 논문을 발표했습니다 모두 쥐를 대상으로 한 것이었죠 그리고 음 그 일곱 가지 실험 중에서 저희가 진행한 실험 하나를 포함했습니다 절차이지만 타임아웃 대신 처벌로 전기 충격을 사용했습니다. 처벌이죠.
wrong responses well you know uh when when we teach a rat how to avoid electric shock one of the things that we do is we are very judicious in the number of times we deliver a shock it's a it's a unique stimulus it's not something the animal Encounters in nature and um it clearly frightens them they uh they will uh they will whine they will defecate and urinate and uh that happens when they get their first first or second shock they will you'll see that kind of emotional reaction but within a few hours uh of training over a period of two or three days all that goes away as the animal acquires effective avoidance behavior and uh and after you know a couple weeks of training you know a couple hours a day five or six days a week what you see is a completely relaxed animal they uh they just don't uh there's there's nothing in their overt behavior that would suggest that there are uh for life yeah right yeah oh absolutely not I mean they live they live as long as uh as as the rats in the food experiments do you know I've never I've never been bitten by a rat in a shock avoidance experiment I have been bitten by rats in food experiments food reinforcement experiments doesn't happen very often so it happens occasionally and I think it's because the animals um our uh are food deprived and uh that generates a certain degree of agitation when they're when you start handling them and it it sort of predicts food so they start protesting yes yes it's it's not it's not really aggression as much as it is uh uh confusion about uh what it is that's out there to eat um but you don't see that with the animals and the shock of wind studies they're pretty laid back they're they're uh they're pretty friendly yes because one of the big one of the big arguments and then I mean Skinner had quite a few arguments against punishment but one big one was the the so-called counter-attack the response of the animal to to punishment how how what about I know you've done some work with uh experimenting and and studying time out um how how is Time Out affecting so so let me let me back up here we in dog training world certain trainers believe that timeout is actually very benign and we can just use it almost any time at anything and it's the go to approach if we have to do positive punishment versus timeout well is that correct statement well uh you know my students and I uh within the last year published a a paper that had seven experiments concerned with timeout from positive reinforcement all in all in rats and um we included a within those seven experiments one in which we ran the procedure but instead of using time out punishment we used electric shock punishment
1:30:47
그리고 쥐 실험실에서 타임아웃은 거의 전기 충격 처벌과 같은 방식으로 작동합니다. 음, 두 경우 모두 정서적 부작용은 나타나지 않지만, 둘 다 행동을 억제하는 데 효과적입니다. 실험실에서는 완전히 억제하거나 차단하지 않도록 적절히 약한 처벌을 사용해야 합니다. 그렇지 않으면 연구할 대상이 남지 않으니까요. 그래서 우리는 자극을 조절하려고 합니다. 어느 정도 중간 수준의 억제 효과를 얻기 위해서죠. 하지만 정적 강화로부터의 타임아웃은 어, 매우 효과적인 처벌 절차입니다. 그리고 어, 이건 처벌 절차입니다. 사람들이 뭐라고 부르든 상관없지만, 이건 부적 처벌(negative punishment)이 맞습니다. 그리고 어, 혐오 자극적이죠. 네, 혐오 자극적입니다. 그리고 어, 순수한 처벌 연구에서 확인하기는 어렵지만, 타임인 환경이 풍부할수록 그러니까 환경이 풍부할수록 타임아웃은 더욱 혐오스럽게 느껴져야 합니다. 물론 기술적인 이유로 처벌 연구에서 이를 확인하기는 어렵습니다. 그래서 우리는 이번 실험 세트에서 회피 실험을 몇 가지 수행했습니다. 동물이 타임아웃을 피하는 실험이었죠. 우리가 한 것은 쥐들을 챔버에 넣고, 가끔씩 가변 시간 간격으로 그 일이 일어나게 한 것입니다. 일정이 잡혀 무료로 사료를 제공받았습니다 [음악] 어 하지만 레버를 누르지 않고 30초가 지나면 타임아웃이 발생했습니다. 타임아웃 또한 30초 동안 지속되었죠. 따라서 그러한 상황에서 쥐는 타임아웃을 피하기 위해 레버를 누르는 법을 배웠습니다 그리고 우리가 그들이 사료를 받기 위해 누를 필요가 없도록 했습니다. 사료는 무료였죠 그들은 타임아웃을 받지 않기 위해 눌러야 했습니다 그 상태에 머무르기 위해서였죠 맞습니다. 우리가 한 일은 조건에 따라 무료 사료가 제공되는 비율을 변경한 것이었습니다 알고 보니 사료가 제공되는 비율이 매우 높으면, 즉 매우 밀도가 높아서 사료가 풍부할 때, 타이밍이 정말 좋을 때 쥐들은 사료 제공이 상대적으로 드물 때보다 타임아웃을 피하기 위해 더 열심히 노력하게 됩니다 상대적으로 어 드물게 말이죠. 따라서 세상이 좋을수록 타임아웃은 더욱 혐오스러운 것이 됩니다 음 흥미로운 점은 더 많은 사료를 제공하는 것이 세상을 더 좋은 곳으로 만든다는 것입니다. 실제로 그렇겠지만 그만큼 타임아웃을 훨씬 더 혐오스럽게 만들죠 풍부한 보상이 있는 환경에서의 타임아웃은 더 효과적인 처벌 수단이 되어야 합니다. 보상이 적은 환경에서의 타임아웃보다 말이죠. 그런데 정말 흥미로운 것은 수년 동안 제가 들어본 많은 사람들, 예를 들어 어린이집 같은 곳에서 우리는 처벌을 사용하지 않고 타임아웃을 사용한다고 말하는 것이 당연하게 들린다는 점입니다 그건 처벌이 맞습니다, 네, 그냥 노는 것이죠 노는 것과 같은 단어들, 네, 그리고 그건 부정적인 것이죠 처벌은 마치 이런 것과 같습니다 십 대 아들이 통금 시간을 어겼다고 자동차 열쇠를 빼앗는 것과 같죠
and time out in in our hands in the Rat Lab works pretty much like shock punishment does um we don't see emotional byproducts in either case but but they're both effective in suppressing Behavior now in the laboratory you don't want you want to use a mild enough punishment so that you don't completely suppress shut down do that you have nothing to study so you we look we try to calibrate the stimuli so we get sort of an intermediate level of suppression but time out from positive reinforcement is uh is a very effective punishment procedure and uh it's a punishment procedure you know people can call it whatever they want but it's negative punishment yes and uh it is aversive yes it's aversive and and and the uh uh it's hard to see it in pure punishment studies but the Richer the time in environment you know the Richer the the Richer the environment is then the more aversive a timeout should be now that's hard that's hard to see in a in a in a punishment study for technical reasons so we did some in this set of experiments we did some avoidance experiments where the animal is avoiding time out what we did was we we put the rats in the chamber and they just got every once in a while on a variable time scheduled they got free food [Music] uh but if 30 seconds passed without a lever press they got a timeout timeout also happened to last 30 seconds so under those circumstances the rat learned to press the lever to avoid the timeouts and what we were they didn't have to press to get the food the food's free they had depressed to avoid getting time to stay in that's right and what we did was across conditions we changed the rate of the free food deliveries and it turns out that if the rate of food deliveries is very high it's very dense there's a lot of food time the timing is really good then they will work harder to avoid the timeouts than if the food is relative the foods deliveries are relatively uh sparse so the better the time the world is the more aversive the timeout becomes um it's kind of interesting because giving more food makes the world a better place and I guess it does but it makes the time out much more aversive so time out from a rich a richly rewarding environment should be a more effective Punisher than Time Out From A A A less rewarding environment but it's it's really interesting the number of people who I've heard over the years like day care centers they'll say we don't use punishment we use time out of course that is punishment yes it's just playing words like playing yes and it's negative punishment it's just like it's like taking the car keys away from your adolescent son because he broke curfew
1:34:05
아시다시피 아들은 여러분에게 자신이 벌을 받았다고 말할 것이고, 실제로도 그렇습니다 하지만 이건 자동차 사용에 대한 타임아웃일 뿐입니다 그렇다면 타임아웃의 지속 시간은 어떨까요? 그리고 어떻게 해야 할까요? 언제 다시 시작해야 할까요? 왜냐하면 거기에는 세 가지 주요한 동물을 다시 복귀시키는 훈련 방식이 있기 때문입니다 그때 적절한 감정 상태가 아닐 때 말이죠 저희는 아직 연구하지 않았습니다 제 연구실에서는 오직 어... 저희는 30초 타임아웃을 사용했는데 왜냐하면 30초면 문헌상으로도 효과가 충분히 나타난다고 제안되었고 저희 경험상으로도 그랬기 때문에 다른 변수를 건드리지 않았습니다 그 변수에는 관심이 없었으니까요 하지만 그건 연구할 가치가 충분히 있는 변수입니다 아시다시피 그 분야에는 온갖 경험칙들이 떠돌고 있습니다 응용 문헌에서는 아이의 나이에 따라 타임아웃을 얼마나 길게 해야 하는지에 대해 다루지만, 그러한 주장을 뒷받침할 경험적 증거는 거의 없습니다 제 추측으로는, 그냥 추측입니다만, 타임아웃은 상대적으로 짧아도 여전히 꽤 효과적일 수 있다고 봅니다 그리고 몇 가지 다른 유형의 타임아웃이 있는데, 제 말은 분명 쥐들을 대상으로 할 때는 아마도 조금 다르긴 하지만 예를 들어 반려견 훈련에서 우리는 개를 데려다가 훈련장에서 멀리 떨어진 차 안의 크레이트에 둘 수도 있고 아니면 하키 경기처럼 현장에 그대로 둘 수도 있습니다. 그러면 페널티를 받긴 하지만 실제로는 벤치에 앉아서 지켜보는 것이죠. 경기에서 완전히 끌려 나가 제외되는 것과는 다릅니다. 네, 라커룸으로 쫓겨나는 건 아니죠. 그렇죠. 그리고 그것들은 확실히 다른 효과를 음, 제 생각에는 많은 경우에 그런 타임아웃이 현장에는 있지만 참여는 할 수 없는 상태가 허용되지 않는 것이 완전히 격리하는 것보다 더 효과적일 수 있습니다. 그러면 개나 어떤 동물이든 거의 잊어버릴 정도의 지점에 이르게 됩니다. 그냥 격리만 시키면 더 이상 왜 다시 경기장으로 돌아오는지에 대한 연관성을 찾지 못하는 것 아닐까요? 그냥 새로운 세션이 시작되는 건지 아니면 페널티 시간이 끝나서 돌아오는 건지 말이죠. 흥미롭네요. 그거 참 정말 흥미로운 관찰입니다. 전혀 생각지 못했던 부분인데요. 과연 실험실에서 그런 방식을 모방할 방법이 있을지 실험실에서 모델링할 수 있을지 생각해 봐야겠네요. 이 부분은 꼭 기록해 둬야겠어요. 끝나고 나서요. 그거 참 정말 흥미롭네요. 맞아요. 미처 생각지 못했던 부분이네요. 정말 좋은 지적입니다. 좋습니다, 마이클. 제 생각엔 우리가 제 말은, 아주 훌륭하다는 겁니다. 대화이고 저는 이것이 분명히 많은 대화를 불러일으킬 것이라 확신합니다. 반려견 훈련 커뮤니티 내에서요. 제가 우리가 이 대화를 어디로 이끌어 갈지 몰랐지만 어, 저는 우리가 다룬 모든 내용에 대해
you know that's the the son will tell you he's been punished and uh and he has but it's it's time out from the from access to the car what about the duration of time out and what about how when do we because there is top three School ways of bringing the animal back in when it's not the the proper emotional state right we have not we have not studied that in my lab we've only studied the uh uh We've Only We hit on a 30 second timeout because 30 seconds the literature suggested that that was long enough to be effective and uh and it was you know our experience so we didn't fiddle with it because we weren't interested in that variable but that is a variable that's well worth studying you know there's all these rules of thumb that are out there in the uh applied literature about how long a timeout should be depending on the age of the child but there's uh very little empirical uh evidence to support those claims my guess is that a time it's a guess uh is the timeout can be relatively brief and still be pretty effective and then we have the the the few different types of timeouts to where I I mean obviously probably with the rats is a little bit different but let's say in dog training we can take the dog and put it all the way away in the crate in the car out from the field or we can keep it present kind of like a hockey game and you get a penalty but you're actually on the bench watching versus being totally taken away and removed from the game yeah you're not sent to the locker room right and and they have they clearly have different effect on on um I think the in a lot of cases that that timeout that you you're still there but you're not allowed to participate um can be more effective than completely taking away and then the dog or or whatever the animal is it's almost at the point that it forgets I would just put away there is not really Association anymore of of why I'm coming back on the field now is it just another session or is it a we're back from penalty that's interesting uh you know that's uh that's a very interesting observation that had that had never occurred to me and I have to think about whether there's a a way to mimic that sort of thing in the laboratory see if we can model that in the laboratory I have to make a note of that after we're done that's uh that's very interesting indeed yes hadn't thought about that one that's that's a good one okay Michael I think we I mean I would say excellent conversation and I I'm sure this is gonna create a lot of conversations among the dog training Community like I I didn't know where we're gonna go with this but uh I'm I'm very happy of of
1:37:24
매우 기쁘게 생각합니다. 정말 흥미롭고 정말 감사의 말씀을 드립니다. 흔쾌히 응해주셔서요. 천만에요. 저도 정말 즐거운 시간을 보냈습니다. 그리고 당신 덕분에 제 학생들과 논의해 볼 어, 아이디어를 얻었습니다. 새로운 실험을 설계해 볼 수 있을지 말이죠. 네, 실험이나 아이디어를 위해 개가 필요하시다면 무엇이든 저희가 여기 있습니다. 개는 연구 대상으로 삼기에는 매우 비용이 많이 드는 동물이죠. 맞아요, 개는 그렇지만 모두가 '개'와 '연구'라는 말을 하면 관심이 즉각적으로 치솟습니다. 왜냐하면 개들은 사람들에게 아주 특별한 존재이기 때문이죠. 네, 음 그건 어 생각해 볼 만한 일이겠네요. 아시다시피 아마도 어, 대학 커뮤니티 안에 있는 반려견 보호자분들이 기꺼이 어, 몇 시간 동안 개를 빌려줄 의향이 있을 겁니다. 음, 저는 쥐를 대상으로 먼저 시도해 보고 싶네요. 맞아요, 그게 당신의 네, 네 물론이죠. 제 말은 무슨 이유로든 비둘기나 그런 거 있잖아요. 네, 어떤 이유에서든 그게 시작점이니까요, 그렇죠? 음, 그들은 실험하기에 아주 편리하거든요. 네, 바로 그곳에 사니까요. 복도만 걸어가면 바로 그곳에 있죠. 정말 감사합니다. 제가 정말 감사의 말씀을 다 드릴 수 없을 정도입니다. 팟캐스트에 출연해주셔서 감사합니다. 말씀드렸듯이 이 내용은 많은 반려견 훈련사들이 계속해서 다시 볼 영상이 될 거예요. 왜냐하면 저 자신도 직접 다시 볼 것이기 때문입니다. 제가 직접 이야기할 때와는 또 다른 느낌이라서요. 저는 그냥 편하게 앉아서 이런 내용들에 대해 말씀하시는 것을 듣기만 하면 되니까요. 정말 매우 흥미로웠습니다. 정말 감사합니다. 고맙습니다. 정말 감사드려요. 감사합니다. 외국어 [음악]
everything that we covered it's it's very interesting and I cannot thank you enough for for accepting that my pleasure I've I've thoroughly enjoyed myself and you've left me with a uh an idea to talk about with my students see if we can design some new experiments yeah any anything that you may need dogs for experiments or ideas we're here they are they are very expensive animals to have as subjects yes dogs are but everybody when you say dogs and research there is a the interest goes of the roof immediately because it it it's a special place with the dogs yeah well it would be uh dude would be something to think about it you know you there probably would be uh uh dog owners in the in the university community that would be happy to uh lend their dogs to us for a few hours well I like to try it with Rats first right that's your yes yes of course I mean for for whatever reasons or pigeons you know yes for whatever reason that's where it starts right well they're so convenient yes they live right there I just have to walk down the hall and there they are thank you I really like I cannot thank you enough for for coming to the podcast and as I said like it it will be this will be watched over and over from from a lot of dog trainers because I I will personally watch it myself because it's different when I'm talking and listening to where I can just sit back and listen to you talking about all this stuff it was very very interesting and then you're very thankful thank you thank you so much thank you foreign [Music]