Kay LaurenceKay Laurence (Learning About Dogs) · 글
The Fade-in Protocol에 관한 연구는 50여 년 전에 처음 기록되었으나, 널리 알려지거나 자주 선택되는 방식은 아님.successive discrimination task (연속 변별 과제) 방식operant chamber (조작적 조건 형성 상자) 벽면에 있는 key (디스크)를 쪼는 비둘기에게 음식으로 reinforcement (강화)를 제공.key light가 빨간색으로 켜졌을 때 쪼는 행동을 함.key에 대한 쪼기 행동이 확립되면, key의 색깔을 초록색으로 바꿈.reinforcement를 제공하지 않음.discriminative stimulus (SD): 빨간색 불빛. 음식을 얻기 위해 쪼아야 한다는 신호.stimulus delta (S ∆): 초록색 불빛. extinction (소거) 상태를 나타내는 신호로, 쪼아도 음식이 나오지 않음을 의미.key가 번갈아 제시됨.response generalization (반응 일반화)으로 인해 초기에는 많은 실수를 범함.key 색깔에 따른 적절한 differential response (변별 반응)가 나타남 (Pierce & Cheney, 2013).errorless discrimination training (오류 없는 변별 학습)discrimination training (변별 훈련)과는 다른 두 가지 절차를 도입함.S ∆ 상태인 초록색 key를 학습 초기에 도입함. (빨간색 조건에서의 쪼기 행동이 완벽히 확립되기 이전에 제시)fading (즉, fading in) 절차를 사용하여 초록색 key의 강도(밝기, 파장, 지속 시간)를 반복을 통해 점진적으로 증가시킴.discrimination 학습 속도가 훨씬 빠름.errorless discrimination procedures (오류 없는 변별 절차)로 훈련받은 비둘기: 약 25회의 오류 발생.standard procedures (표준 절차)로 훈련받은 비둘기: 2,000 ~ 5,000회의 오류 발생.T&E (Trial and Error, 시행착오) 방식의 비둘기: S ∆가 나타났을 때 감정적인 반응을 보임.errorless approach (오류 없는 접근법)로 훈련받은 비둘기: SD인 빨간색 디스크가 나타날 때까지 차분함을 유지함.errorless learning procedures를 사용한 경우, 표준 절차보다 학습 속도가 빠르고 오류가 적었으며, 학습 과정을 더 즐거워함.positive reinforcement (긍정적 강화)를 활용한 다양한 프로토콜이 있음에도 불구하고, 여전히 개가 '실수를 하도록 유도해야 한다'는 고정관념이 존재함.reinforcement를 제공한다 하더라도 그것을 진정한 positive learning (긍정적 학습)이라고 할 수 없음.Fading-In Protocol은 성공을 토대로 정교한 변별력(Discrimination skills)을 가르치는 우아한 과정임.cue(단서)로 행동을 지시할 때, 우리의 서 있는 방식(Way we are standing)이 그 정보를 뒷받침할 수도, 모순될 수도 있음.cue를 줄 때 강아지 쪽으로 한 걸음 다가가는 것은 강아지가 뒷걸음질 치게 만들 수 있음: 다가가는 행동이 언어적 cue와 모순됨(Contradicts the verbal cue).cue가 아닐 수도 있는 임시적인 cue)을 성공을 장려하고 오류를 방지하는 환경으로 둘러싸야 함.cue에 반응하고 있음.hand cue를 준다면 이는 혼란스러운 메시지(Mixed message)가 됨.cue를 인식하지 못하고 있다는 신호일 수 있음 (가르치는 사람이 너무 시끄러웠거나, 너무 빨랐거나, 모호했음).fade-in(점진적 도입)을 할 일련의 사건들을 준비해야 함.green:blink(테라스 실험에서의 빛의 밝기, 파장, 지속 시간 증가와 같은 개념)의 조건:a blink(눈 깜빡임 정도).green:blink는 반려견이 피하고 싶어 하는 이벤트가 절대 되어서는 안 됨.attractions(유인물/매력적 자극)로 훈련하며, aversions(혐오 자극)는 사용하지 않음.fade-in 방법과 fade-out(점진적 제거) 시점에 대해 완벽히 숙지해야 함.green:blink를 사용하고, 사이클 반복이 진행됨에 따라 반려견이 green:blink에 노출되는 정도를 스스로 조절해야 함.fade-in 적용 시작 전 단계: base behaviour(기본 행동)를 먼저 수행함.red light에 해당)임.Antecedent(선행 사건), Behaviour(행동) [Mark], Consequence(결과).cue(신호)를 찾고 있는지 확인한 후, cue(A)를 줌.mark하여 좋은 일이 생길 것임을 약속하고 reinforcement delivery process(강화물 제공 과정, C)를 시작함.cue를 찾게 됨.mark를 받은 후 보상을 찾으러 가는 시점에 최소한의 green:blink를 도입함.mark와 음식 섭취(collection of the food) 사이에 짧은 '눈 깜빡임'을 배치함.min green:blink 1: mark와 보상 사이에 배치.med green:blink 2: 1번보다 약간 더 강도가 높고 가까움.max green:blink 3: 가깝고 지속 시간도 김.mark를 받은 후, 자신의 경험상 이것이 곧 기대되는 reinforcement(강화물)로 이어진다는 것을 알고 있는 상태임.ABC cycle에서 가장 강력한 부분은 보상을 받으러 가는 순간임.green:blink 반응에 따른 해석:green:blink에 더 큰 관심을 보인다면: 사이클에서 유일하게 중단된 부분이 '간식 수거'임. 즉, 반려견이 green:blink를 선호하는 reinforcer(강화물)로 선택한 것임. 이는 향후 reinforcer 선택에 영향을 미칠 중요한 정보임.green:blink를 인지하면서도 계속해서 간식을 먹으러 간다면: green:blink와 treats(간식) 사이의 pairing(연합)이 시작된 것임.green:blink를 조사하거나 탐색하기 위해 멈춘다면:green:blink를 더 멀리, 더 짧게 조정하고 reinforcer의 가치를 높임.reinforcer(보상)를 추구할지 결정하는 과정임.WTF 모먼트가 사이클에 대한 집중을 깨뜨린다면, fade-in(자극의 점진적 투입)이 너무 강했음을 의미하므로, 강도를 조절하고 진행해야 함.Fade-in의 단계적 적용 및 관리fade-in이 cycle의 'C' 부분을 6~10회의 강도 변화 내에서 깨뜨리지 않는다면, 그 다음 단계로 넘어감.fade-in과 WTF 기회를 사이클의 더 앞부분으로 옮김.Green:blink 4: A와 B 사이 구간.Green:blink 5: 사이클 전반에 걸쳐 퍼뜨리기.reinforcement(강화)는 계속 제공되어야 하며, 만약 green:blink가 행동의 흔들림(wobble)을 유발했다면 그 흔들림마저 강화될 위험이 있음. (매우 위험한 상황임)green:blink가 사이클 내 다른 지점들에서 다양하게 시도되고, 주저함 없이 강도가 변화될 수 있을 때 비로소 행동 중간(5)에 green:blink를 적용할 수 있음.generalise).cosmetic behaviour)을 통해 이 기술을 가르치면, 개는 ABC cycle을 유지하는 능력을 다른 상황으로 전이함.proofing(검증/증명)이라고 부르지 않도록 주의해야 함.cycle에 집중하는 능력을 개발하는 '가르치는 과정'임.green:blink가 집중의 질을 떨어뜨린다면, 나쁜 습관(permanent poor association)을 형성하게 되며, 이는 종종 "어쨌든 대충 했으니까"라는 식으로 강화되기 쉬움.WTF 순간들을 웃어넘길 방법을 찾아야 함.green:blink를 치운 뒤 다시 시작할 방법을 마련해야 함.fade-in은 훈련 세션마다 수없이 많이 사용됨. 일상 속에서도 주의 깊게 관찰하면 하루 종일 사용 가능함.WTF를 연습함: 주인이 귀를 긁는 것은 그냥 green:blink라고 결정하거나, TV 광고가 나오는 것도 그냥 green:blink라고 판단함.dress-rehearse)할 필요가 있음.A Road to Nowhere: 익숙함이 사라지면 우리는 다시 편안함과 안전함으로 돌아가기 위해 알아볼 수 있는 이정표를 찾음. 이는 생존 본능이며, 우리를 살게 하는 것이기에 귀 기울일 가치가 있음.The Experienced Dog: 개가 충분한 준비를 마쳤다는 것은 모든 경우의 수를 안다는 뜻이 아님. 예상치 못한 일이 발생했을 때 자신의 기술을 활용해 문제를 해결할 수 있도록 다양한 조건(range of different conditions)을 경험했음을 의미함.Training이 아니라, 해당 학습자(learner)의 앞으로의 삶에 맞춰 속도를 조절한 신중하게 계획된 learning pathway(학습 경로)가 필요함.no-training day(훈련 없는 날)라고 해서 강아지가 햇볕 아래서 한가롭게 게으름을 피우는 날을 의미하는 것은 아님.talking head(말하는 사람)의 모습과 자신의 타고난 전문성을 납득시키려는 비건전한 혼합 방식을 보게 됨.teaching(가르치기) 하기 위한 계획을 세울 때 앞서 질문해야 할 것들.play(놀이)나 food delivery(음식 급여)를 하나의 ACTIVITY(활동)로 볼 때, 우리는 개와 동일한 사고방식을 공유하게 됨: "과연 경험할 만한 즐거움이 있는가?(is there pleasure to be experienced?)"training(훈련)이나 개와 함께 살아가는 친숙한 방식에 끌리기 마련임.prompts(단서)를 찾기.skills(기술)들을 다듬기.Border Collie(보더 콜리)였음.Klog(Kay Laurence의 블로그)임.errorless learning(오류 없는 학습)에 관한 강연을 듣고 Terrace’s paper(테라스의 논문)를 공부한 후, 작성자의 puzzle이나 WTF moments가 이것과 어떻게 부합하는지, 그리고 학습자가 오류를 범하게 만드는 훈련과는 어떻게 다른지 생각하게 됨.WTF moments는 개가 실패하도록 유도하지 않으며(그럴 의도도 없음), 매우 최소한으로 이루어져 개가 여전히 성공을 경험할 수 있다고 판단함 (기사에서 배운 faded-in 개념과 같음).WTF 수준에 머물렀으며, 더 심각한 Leave It! 단계로 확대되지 않았음.learning about dogs 과정을 통해 배우는 내용은 다음과 같음fading in(점진적 도입/희미하게 도입하기)의 일부임fading in 자체가 학습 과정의 필수적인 요소라는 점은 매우 놀라운 사실임fading in의 원리와 비유wait(기다려) 교육의 재해석:sheep balls(양몰이 공) 훈련을 시도하다 어려움을 겪는 경우, 해당 훈련을 시작하기 전에 fading in 개념을 먼저 완전히 숙지하고 적용해야 함.fading in: 학습 과정 중 자극을 점진적으로 도입하거나 강도를 조절하여 학습자가 자극에 익숙해지게 하는 기법.irrelevant stimulus: 학습 목표와 관련 없는 외부 자극.blink: 개가 자극을 인지하고 반응하는 짧은 순간.lure: 간식 등을 이용해 개의 행동을 강제로 유도하는 것.impulse control: 충동 조절.foundation: 기초 훈련.The Fade-in Protocol by Kay Laurence | Key Skills and Training , Seeing with New Eyes The origins of the protocol were first documented over 50 years ago, but it is not a well-known or frequently selected protocol. Susan Friedman explains the original experiment succinctly in her article Clearing the Path to Reinforcement with an Errorless Learning Mindset : Terrace (1963) researched errorless learning with a successive discrimination task. In the traditional successive discrimination procedure (different than Terrace’s procedure), a pigeon, for example, is reinforced with food for pecking a disk on the wall of an operant chamber (called a key light or key) when it’s illuminated red. After many repetitions, when the pecking behavior in the presence of the red key is well established the color of the key changes to green and pecking is no longer reinforced. With the standard protocol then, the red light is the discriminative stimulus (SD) that cues pecking for food reinforcement, and the green light is the stimulus delta (S ∆) that signals the extinction condition, i.e., pecking will not produce food reinforcement. The red and green keys are then alternately presented with the corresponding reinforcement and extinction conditions in effect. After initially making many errors (due to response generalization), the correct differential response to key color gradually occurs (Pierce & Cheney, 2013). Alternatively, Terrace used two procedures in his errorless discrimination training not typical of standard discrimination training. First, the S ∆ condition, the green key, was introduced very early in the program before pecking in the redlight condition was well established. Second, Terrace used a fading (i.e., fading in) procedure to present the green key at different values, gradually increasing brightness, wavelength and duration over the repetitions. These two procedures resulted in faster learning of the discrimination and very few errors. The pigeons trained with the errorless discrimination procedures made about 25 errors (i.e., pecking the green key light) compared to 2000 to 5000 errors made by the pigeons taught with standard procedures. Only those birds trained with T&E exhibited emotional responses in the presence of the S ∆. The pigeons trained with the errorless approach remained calm until the red disk, the SD, appeared. These findings have been widely replicated across species. Powers, Cheney, & Agostino (1970) found that preschool children taught a color discrimination with errorless learning procedures learned faster and with fewer errors, and
they enjoyed learning more than the children taught with standard procedures. Roth reported similar results with dolphins (as cited in Pierce & Cheney, 2013). Even though today we are surrounded by many available protocols for teaching with positive reinforcement, there is still a persistence that a dog should be set-up to make an error. The reasons for this are extensive: “the dog needs to know when they are wrong so that they can focus on what is required” “we are building resilience” “this teaches a dog impulse control” With this mindset underpinning the training, adding reinforcement to correct choices does not make it positive learning. Any time being wrong is deliberately included in the curriculum we are failing our learners. Sure, learning may be effective, unforgotten and lifelong but at a cost. The cost can be a mistrust in the teacher. When the learner is faced with a new lesson they may first ask “what is the trap?” This is a negative view and likely to cause hesitation and a resistance to discovery and exploration. If you think your teacher is only setting you up for success you approach with an enthusiasm and eagerness for the learning. All of us have been damaged by the error-trap. Whether it was deliberate on part of the teacher or not, it makes us cautious learners. In addition to this approach to learning we then experience frustration, that can escalate to rage, when we cannot solve the error. If you have ever used a computer connected to a printer you have experienced this. And it wasn’t good. Application of the Fading-In Protocol This is an elegant process that builds on success and teaches exquisite discrimination skills. When a stimulus is presented it is rarely in isolation. Reality surrounds our stimulus with relevant and non-relevant information. We cue a behaviour verbally the way we are standing can either support that information our contradict it. A step towards the dog as we cue a “sit” can force the dog to take a step back: the step towards contradicts the verbal cue. When teaching a behaviour we want to surround the stimulus (which may be the temporary cue, not the final one) with an environment that encourages success, not error. Under these conditions an error is simply information that tells you the environment needs changing. The dog is responding to a selected cue that does not contribute to our goal. The step towards causing the dog to step back. It is always the dog that get to choose what is relevant and what information they are receiving. If I have food in my hand and I have used food to hand-lure behaviour, the
n a step towards food in the hand is a correct response for that dog. If you have taught a nose-touch to hand, then a hand cue that asks the dog to wait, hold position, is a mixed message. To be able to teach without confusion, frustration or error we need to be clear in our communication and understanding of what information the dog is taking from the environment. An error is simply the difference between my expectation and the dog’s response. I consider errors valuable information: it shows me what the dog understands is likely to earn reinforcement it may tell me what the “hot” behaviour is at that moment (the most recently reinforced behaviour), it may tell me the dog is becoming mentally tired, particularly if the error rate has increased it may tell me a change in the environment is a stress overload. It could be too much information, a threat or opportunity the dog is struggling to resolve. it may tell me the dog cannot recognise the cue (I was too loud, too quick, or ambigious) Errors should be noted, and if possible logged with a clear description. The type of error, or response from the dog, can often give us important information, it should not be simply recorded as “error”. Before we begin the practical work we should prepare a range of events to fade-in. These are our green:blinks - the increasing brightness, wavelength and duration of the light in the Terrace experiment. They should be of short duration: a blink, of low intensity: quite far away, and have very little impact or relevance to the dog. Green:blinks should never be an event or something the dog wishes to avoid. We train with attractions, not aversions. If you have a training assistant they need to be fully aware of how to fade-in and when to fade-out. If you are training alone, use a fixed green:blink and as you progress through the repeats of the cycle you can control the exposure of the dog to the green:blink. You could use a pot of food on a chair outside the doorway. At first it is not on view, then on view for a moment, then increasingly in full view, then closer in proximity until the dog is rolling through their cycles alongside the chair. To begin the fade-in application we begin with the base behaviour. If it has been taught previously then it would benefit from refreshing. This is the behaviour the dog will default to when unable to make a choice. It is the red light in the Terrace experiment. We choose this behaviour in the cycle of Antecedent, Behaviour [Mark], Consequence. [Mark] is optional. We look to see if the dog is ready and seeking a cue, we give the cue (A),
the dog responds with (B), we mark this to promise good things and begin the reinforcement delivery process (C). As the dog consumes their treat they will be seeking another cue. More ..? After the dog has completed the behaviour, been marked and is seeking their reward, we begin to introduce the minimal green:blink. If it can be arranged it is a short “blink” between the mark and the collection of the food. Either as I reach, walk or open the pot, or the dog runs after a thrown treat or steps towards an offered treat. min green:blink
etween “A” & “B” Green:blink 5 Spilling around the cycle It is important that we keep the behaviour stable. Reinforcement is still occurring and if our green:blink has caused a wobble in the behaviour this will get reinforced. Dangerous ground. We should only progress to green:blinks during the behaviour (5) when green:blink has travelled many other places around the cycle and changed in intensity without causing hesitation. The delight in this process is many fold. Dogs generalise this very well. We can choose a cosmetic behaviour to teach it, not a life-important behaviour, and the dog will transfer the skills of maintaining the ABC cycle. Be careful not to label this process as “proofing”. We are not trying to test out the behaviour or the dog’s resilience. It is a teaching process that develops a dog’s capacity to focus on our training of cycles. If a green:blink produces a reduction in the quality of that focus, we are setting up a permanent poor association that often gets reinforced (“ because he sort of did it anyway ”). You will have to find a way to laugh off the exceedingly long WTFs that may occur, plan for it and make sure you find a way of letting the dog know you screwed up and you would like to start again (after having sent green:blink packing). I use fading-in many, many times during every training session. We use it all day long if we look for it and your dog practices many WTFs - deciding that when you scratch your ear it is just a green:blink, that when the adverts roll on the TV, it is just a green:blink. These are non-relevant events that do not impact of their comfortable snooze, chewing of a bone, or deep hand-massage. It is a quite natural survival skills of assessing what information the dog should pay attention to under specific circumstances. Often we need to dress-rehearse those circumstances to enable the dog to be confidence and fluent. A Road to Nowhere When familiarity is stripped away we seek recognisable signposts that will take us back to comfort and security. This is survival instinct. It is worth listening to as it keeps us alive. Read More The Experienced Dog Knowing your dog has receive sufficient preparation does not mean every eventuality, but a range of different conditions so that when the unexpected happens they will draw on their skills and solve the issue. Read More Be-toothed Learning Machines The thing they don’t tell you is that raising a puppy is DANGED HARD WORK. Biting everything, peeing everywhere, eating anything; not for the faint hearted. Read More Obnoxious Puppy The delight of your new puppy is
probably going to last a few weeks, maybe four if you are lucky. When 12 weeks old hits, and you will feel a slam, the Delight is going to demonstrate ungrateful, obnoxious traits. Read More It’s Not Training A carefully planned learning pathway, paced to suit that particular learner for their life ahead. Read More A Day of Learning A no-training day does not mean he gets a lazy day lying idly in the sun. Learning is still happening and this is significant and important for his development. Read More Wheat or Chaff? What is the purpose of this video? To sell a product, to instruct or to inspire? It should be clear from the first viewing. Often we are seeing an unhealthy blend of talking head, dripping treats into bored dog, convincing you of their innate expertise. Read More When it is not rewarding Just because it is our intent to reward does not make it always rewarding. Read More Ethical questions What to ask before when we make a plan to teach Read More Play Health Check When we look at play or food delivery as an ACTIVITY we share the same mindset as the dog: is there pleasure to be experienced? Read More Changing is growth We are naturally attracted to familiar ways of training or living with our dogs. We have often worked hard to learn those habits and there is a reluctance to make changes since this is hard work. It takes mentally energy to note what we are not doing well, recall what changes we need to make, find the prompts that can move us to the changes and then work on the skills those changes require. Read More Any Dog But a Collie After deciding I wanted to live with a dog, the only dog I ruled out was a Border Collie. Read More 7 Comments Michaela Hempen on 4th June 2018 at 4:17 pm What a great Klog! After listening to talks on errorless learning two weeks ago and studying Terrace’s paper, I thought about how your puzzle or WTF moments fit into this versus training that sets the learner up to errors. I figured that the WTF moments don’t set the dog up for failure (and don’t intend to) and are so minimal that the dog can still be successful (“faded-in” as I learnt reading the article) Many thanks for this and very well timed. My pondering remained a WTF and did not escalate into a Leave It! Reply Michaela on 4th June 2018 at 4:18 pm What a great Klog! After listening to talks on errorless learning two weeks ago and studying Terrace’s paper, I thought about how your puzzle or WTF moments fit into this versus training that sets the learner up to errors. I figured that the WTF moments don’t set the dog up for failure (and don’t intend to) and
are so minimal that the dog can still be successful (“faded-in” as I learnt reading the article) Many thanks for this and very well timed. My pondering remained a WTF and did not escalate into a Leave It! Reply Kay Laurence on 4th June 2018 at 9:44 pm Thanks Michaela, sometimes we don’t see the lines connecting the points until we wander around to a different view! Reply Sue McGuire on 4th June 2018 at 11:24 pm This klog (I am learning to go with the new title) does go to the heart of how incredibly perceptive our learners are during the process. It’s our attentiveness to this that provides real learning and real leaps of understanding. What I had missed is being attentive to “moving” the irrelevant stimulus into different sections of the learning cycle. For some reason I had attached myself to the “A” part. When in reality it’s all “part” of the fading in. It’s quite astounding when you consider “fading in” IS part of the learning process. Reply Kay Laurence on 5th June 2018 at 10:23 am It is strawberry season. Much anticipated (our version of the warm peaches direct from the tree). So if that strawberry, with some double cream is on the spoon and travelling to my mouth …… and the phone beeps ….. There is NO question, the strawberry goes in the mouth, then I decide whether to attend to the phone. But if I was just slicing the strawberries, I would most likely interrupt that moment to pick up the phone. Reply Julie Van Schie on 7th June 2018 at 6:13 am What I love about this training is that the dog still has choice. We are not waving extra extra high value food or toys in front of his nose and using those to lure away from the “blink”. Rather the dog is allowed to look, to process and to decide “not relevant”. That way we can truly move at the dog’s pace. You are always thought provoking Kay! Reply Kate Clark on 15th July 2025 at 12:27 pm I am working my way through the workbooks gradually and came to this section and the first time through is a bit mind blowing in a good way. The wait before you begin is a thousand times more complex than just teaching the good old wait. Then the mid starts spinning on the different possibilities and setups and applications to go onto. It is amazing. I need to get to grips with the fading in aspect and i will follow up with this with questions. Thank you Kay as i have struggled with trying to set up sheep balls with both of mine and this will come first before I go back to the sheepballs Reply Trackbacks/Pingbacks Impulse control – mindset matters pt. 2 of 2 - Tromplo - […] reinforcement to begin with. Once the altern
ate behaviors are strong enough, next I carefully “fade in” (credit Kay Laurence)… Submit a Comment Cancel reply Your email address will not be published. Required fields are marked * Comment * Name * Email * Website Δ