<![CDATA[Conditioned “NO” Beats Reward-Only Every Time]]>
Ivan Balabanov<![CDATA[Training Without Conflict® | Dog Training Podcast]]> · 팟캐스트 · 14분
본질적인 포식 본능을 가진 개는 스스로 보상을 얻기 위해 사냥과 살해를 수행하며, 이는 단순한 놀이와는 차원이 다른 강력한 내적 동기를 유발한다. 기존의 긍정 강화나 포식 대체 프로토콜은 포식 패턴이 활성화된 상태에서 개에게 인식되지 않으므로, 살상 이력이 있는 개를 통제하는 데 한계가 있다. 조건화된 처벌로서의 구두 신호는 이러한 사냥 의도를 즉각적으로 중단시키며, 행동에 대한 비용을 명확히 함으로써 개에게 실질적인 자유와 안전을 제공한다. 신호와 혐오적 결과 간의 명확한 연합은 정확한 타이밍과 고전적 조건화 과정을 거쳐 확립된다. 결과적으로 이 방법은 목줄 없는 산책 환경에서도 개를 확실하게 멈추게 하며, 유전적으로 깊게 프로그래밍된 포식 행동을 효과적으로 관리할 수 있게 만든다.
조건화된 "안 돼"는 언제나 보상 위주 방식보다 승리합니다
<![CDATA[Conditioned “NO” Beats Reward-Only Every Time]]>
0:00
자 여러분, 오늘 제 개들 중 한 마리를 데리고 흥미로운 것을 좀 보여드리겠습니다. 많은 개들이 다람쥐, 토끼, 자동차, 자전거 등 뭐든 쫓는 걸 좋아하지만 모든 개가 똑같지는 않습니다. 어떤 개들은 재미로 쫓습니다. 그건 일종의 달리는 스릴 같은 거죠. 완전한 포식 행동 순서 중 오직 추격 부분만을 하는 겁니다. 그래서 나무 앞에서 멈춰 흥분해서 짖기는 해도 결국 다시 주인에게 돌아오곤 하는데, 저는 이걸 '놀이 포식(play predation)'이라고 부릅니다. 왜냐하면 실제로 죽이려는 의도는 전혀 없기 때문이죠. 하지만 또 다른 유형의 개들이 있습니다. 제가 보여드릴 개가 바로 그런 경우인데, 죽이려는 목적으로 쫓습니다. 녀석들은 강하게 집중하고, 공격 기회를 노리며 한 시간 동안 은밀하게 멈춰 있거나, 아니면 두 번째 방식을 선택해 앞길에 있는 장애물은 아랑곳하지 않고 맹렬하게 쫓아갑니다. 자기 보존 같은 건 전혀 생각하지 않죠. 이것이 바로 완전한 포식 본능입니다. 동기가 더 깊고, 억제하기가 더 어렵습니다. 왜냐하면 포획하고 죽이는 것에 성공하는 것 자체가 엄청난 자기 강화 효과를 주기 때문입니다. 개는 말 그대로 아드레날린과 본능의 완성을 통해 스스로 보상을 받는 셈입니다. 이제 곧 보게 될 개에 대해 조금 말씀드리겠습니다. 이름은 칼린카이고 3살 된 암컷 말리노이즈입니다. 제가 직접 브리딩했고 2살 때 다시 데려왔으니, 녀석에게는 이미 과거가 있었죠. 다시 돌아오게 된 이유는 너무 다루기 힘들고 에너지가 넘쳤기 때문입니다. 도심의 콘도에서 살기에는 너무나 어리고 활동적인 워킹 말리노이즈였던 거죠. 하지만 녀석은 한 가지 기술을 완전히 마스터했는데, 그 기술은 바로 사냥하고, 잡고, 가끔은 다람쥐를 죽이는 것이었습니다. 제가 이런 배경 설명을 드리는 이유는 놀이로 쫓는 개와 아주 다른 의도를 가진 개라는 두 유형 사이의 차이를 알아야 하기 때문입니다. 그런 행동을 통제하는 처음 세 가지 방법은 매우 제한적이며 그리고 그들은 보통 몇 가지 프로토콜에 의존합니다. 탈감작, 역조건 형성, 차별 강화 그리고 소위 포식 대체 프로토콜이라고 부르는 것들입니다. 하지만 이 중 어느 것도 죽일 목적으로 추격하고 그런 이력이 있는 개들을 통제하는 데에는 효과적이지 않습니다. 예를 들어 탈감작을 살펴봅시다. 이것은 개를 유발 요인에 점진적으로 노출시키는 프로토콜로, 예를 들어 다람쥐를 더 먼 거리에서 혹은 낮은 강도로 노출하는 것입니다. 그래서 예를 들어 거리가 개가 아직 반응하지 않을 정도로 충분히 멀고 점차적으로 유발 요인에 더 가까이 다가가는 것입니다. 우리의 예시에서는 다람쥐가 되겠죠. 자, 만약 여러분이 경험담을 믿고 시도해보고 싶다면 그렇게 하십시오. 하지만
All right guys I'm gonna show you something interesting with one of my dogs today. Many dogs love chasing squirrels rabbits cars bikes you name it but not all dogs are the same. Some dogs chase for fun it's kind of the thrill of the run it's only the chase part alone of the complete predatory sequence so they may stop at the tree bark excitedly but ultimately they kind of bounce back to you and this is what I call play predation because there is really no intent to kill. Now there is this other type of dogs like the dog that I'm going to show you they chase with purpose to kill they lock in hard they stay frozen in a stealth mode for an hour waiting to strike or just choose option b and go in a wildest chases disregarding any obstacles along the way never really thinking about self-preservation or anything. This is the full predatory drive the motivation is deeper and harder to override because succeeding the the catching and the killing is hugely self-reinforcing. The dog literally pays itself with adrenaline and completion of the instinct. Now let me tell you a little bit about the dog that you're going to see shortly. Her name is Kalinka she's a three-year-old Malinois female and I got her back I bred her and I got her back when she was two years old so she had already some history it was the reason she came back was because she was just too much dog too high energy just a young and active working Malinois that doesn't really fit well living in a condo in a in the city however she mastered the skill and the skill is hunting and catching and sometimes killing squirrels. So the reason I'm giving you that background is because we need to make distinction between the two types of dogs the one that play chase and the one that have very different intention. The first three ways to control such behaviors are extremely limited and and they typically rely on few few protocols. Descentization, counter conditioning, differential reinforcement and what you would call predation substitute protocols. However none of this is effective in controlling dogs that chase with the purpose to kill and have history of doing so. Let's look at the desensitization for example it's a protocol that gradually exposes the dog to the trigger a squirrel in a bigger distance or with a low level of intensity. So the distance for example is big enough to where the dog doesn't react yet and gradually we start getting closer and closer to the trigger. A squirrel in our example. Now if you believe in anecdotal stories and want to give this a try go ahead but
3:48
탈감작은 매우 다른 목적을 위한 것이라는 점을 명심하세요. 그것은 일반적으로 공포 기반의 반응이나 공포증을 위한 것입니다. 포식 행위에서 개는 두려움을 느끼지 않습니다. 그들은 죽이려는 본능에 의해 매우 높은 동기 부여를 받습니다. 이것은 매우 깊게 뿌리 박힌 유전적 프로그래밍입니다. 누구나 200피트 떨어진 곳에 있는 다람쥐에 대해 개를 탈감작시킬 수 있습니다. 하지만 그 다람쥐가 개 코앞에서 갑자기 달아나는 순간 생물학적 임계치를 넘어서게 되고 상황은 끝납니다. 그 개는 이미 통제 불능 상태가 됩니다. 역조건 형성 접근법이란 무엇일까요? 익숙하지 않으신가요? 기본적으로 다람쥐에 대한 노출과 고가치 간식을 짝짓는 것입니다. 그 생각, 즉 의도는 개의 정서적 반응을 사냥에서 먹이를 받기 위해 주인을 바라보는 것으로 바꾸려는 것입니다. 다음은 차별 강화입니다. 차별 강화만으로는 역시 성공적이지 않습니다. 여기서의 아이디어는 다람쥐를 쫓는 것 이외의 다른 행동을 강화하는 것입니다. 앉아나 나를 봐 같은 행동들 말이죠. 생각할 수 있는 무엇이든 가능합니다. 이제 포식 대체 훈련도 있습니다. 여기서의 아이디어는 개가 살아있는 동물 대신 장난감을 쫓고 잡도록 장려하는 것입니다. 문제는 추격이 시작된 후에 강화물이 제공되면 개는 이미 그 행위에 완전히 몰입해 있다는 점입니다. 포식 본능 패턴 체계가 활성화됩니다. 그리고 이 시점에서는 다른 어떤 보상도 개에게 인식되지 않습니다. 개는 이미 몰입한 상태입니다. 끝난 거죠. 주머니에서 다른 다람쥐를 꺼내서 보여준다고 해도 소용없을 겁니다. 아무런 상관이 없습니다. 개는 이미 목표물에 완전히 고정되어 있어서 정신적으로는 물론, 어쩌면 물리적으로도 이미 그곳에 없는 상태니까요. 그러니 현실적으로 생각해 봅시다. 야생 동물, 자동차, 자전거는 여러분이 대처할 준비가 되었을 때 신호에 맞춰 나타나 주지 않습니다. 그건 현실이 아니죠. 제가 방금 언급한 몇 가지 접근 방식은 장난으로 쫓아다니는 개를 제어하는 데는 성공할 겁니다. 하지만 의도와 이력이 다른 개들에게는 확실하게 실패할 것입니다. 왜 효과가 없는지 말씀드리죠. 첫 번째 이유는 제가 '비용의 부재'라고 부르는 것입니다. 강압 없는 훈련(force-free)을 지향하는 트레이너들은 기본적으로 행동에 비용을 부과하는 것을 거부합니다. 그건 마치 근본적인 문제를 피하면서, 다람쥐를 죽이는 것보다 더 좋고 흥미로운 활동을 제공할 수 있다고 개를 설득하려는 것과 같습니다. 어떤 개들에게는 사냥의 아드레날린이 여러분이 제안하는 것보다 훨씬 더 큰 보상으로 느껴집니다. 그들은 단순히 여러분의 제안보다 더 높은 가치를 선택하고 다람쥐를 낚아챕니다. 이야기는 그걸로 끝이죠. 자, 이제 제가 복잡한 프로토콜이나 장난감, 음식 뇌물 없이, 그리고 개에게 칼라나 긴 리드줄을 사용하지 않고도 얼마나 확실하게 행동을 제어할 수 있는지 보여드리겠습니다. 그저 매번 개를 멈추게 하는 구두 신호 하나면 충분합니다.
keep in mind that desensitization is meant for very different purposes. It's typically for fear-based reactions and phobias. In predation the dog is not afraid. They are highly motivated by their instincts to kill. This is genetic programming that it's really deep. Anybody can desensitize a dog to a squirrel that is 200 feet away. However the moment that squirrel bolts under their nose the biological threshold is crossed and game over. That dog is gone. The counter conditioning approach. What does it mean if you are not familiar? Basically is pairing the exposure to the squirrel with high value treats. The idea, the intention is to change the dog's emotional response from hunt to look at the owner for food. Next we have the differential reinforcement. Differential reinforcement alone is also not successful. The idea here is to reinforce different behavior other than chasing the squirrel. Like sit or watch me. Whatever you can think of. Now there is also the predation substitute training. The idea here is to encourage the dog to chase and catch a toy instead of the live animal. The problem is that if the reinforcement is presented after the chase starts, the dog is already locked in. The predation motor pattern system is activated. And at this point, no other reinforcement are acknowledged by the dog. The dog is already locked in. Done. Like you can go as far as even to offer another squirrel from your pocket. It doesn't matter. The dog is already locked on the target and it's mentally, if not physically as well, gone. So let's be realistic. Wildlife, cars, bicycles, they don't show up on cue whenever you're ready to deal with them. That's not real life. Some of the approaches that I just mentioned, they will be successful in controlling a dog that is play chasing. But they will fail reliably with the dogs that have different intention and history. Let me tell you why it doesn't work. The first reason is what I call absence of cost. The force free trainers basically refuse to attach a cost to the behavior. It's like dancing around the real problem, trying to convince the dog that they have a better, more interesting activity to offer than killing the squirrel. Certain dogs find the adrenaline of the kill far more rewarding than your suggestion. They simply outbid you and take the squirrel. End of story. So what I'm going to show you now is how I can control behavior extremely reliably without the complicated protocols, without toys, food bribes, without any colors or long lines on the dog. Just a verbal cue that stops the dog every time.
7:27
제 경우 저는 '안 돼(no)'라는 단어를 사용하지만, 그건 제가 선택한 단어일 뿐입니다. 어떤 단어를 선택하든 상관없습니다. 그 신호는 교과서에서 '조건화된 처벌(Conditioned Punisher)' 또는 '이차적 처벌'이라고 부릅니다. 좋습니다, 말은 이 정도로 충분하겠군요. 밖으로 나가서 제가 조건화된 처벌이 어떻게 작동하는지 보여주기 위해 만든 환경을 보여드리겠습니다. 그런 다음 링커(Linker)를 데리고 다람쥐 사냥에 나설 겁니다. 그럼 시작해 보죠. 나탈리아가 우리를 도와줄 겁니다. 나탈리아가 우리를 도와줄 겁니다. 나탈리아가 우리를 도와줄 거예요. 나탈리아가 우리를 도와줄 거예요. 보통 다람쥐들이 여기 근처에 머물러요. 다람쥐를 약간 움직여 보세요. 네, 좋아요. 그러니까 기본적으로 그런 개념이에요. 개가 이게 설정된 상황이라고 생각할 가능성은 전혀 없어요. 우리가 탐색하게 내버려 두지 않는 이상요. 개는 여기에 항상 다람쥐가 있다는 걸 알아요. 그래서 여기 있거나 반대편 새 모이통 근처에 있죠. 아니면 물론 주변 어디든 있겠죠. 글쎄요, 잘 모르겠네요. 때로는 개가 다가와서 정말 빠르게 달려들 수도 있어요. 그리고 때로는 아주 은밀한 상태로 다가와서 몰래 뒤를 밟아 낚아채려고 할 수도 있죠. 그래서 계획은 개가 이 다람쥐를 발견하면 다람쥐를 쫓으려 할 것이고, 그때 제가 안 된다고 말하는 거예요. 그리고 제가 예상하는 것은 개가 쫓는 행동을 계속하지 않을 것이라는 점입니다. 제가 실제로 개를 멈출 수 있을 것이라는 거죠. 이건 의도가 다른 장난기 많은 개들과는 완전히 다른 경우입니다. 다람쥐를 낚아채서 죽이지 마. 제가 왜 안 된다고 말할 수 있는지, 그리고 어떻게 개를 통제할 수 있는지에 대해서는 나중에 다시 얘기할게요. 좋아요, 이제 개를 밖으로 내보내고 어떻게 되는지 보죠. 칼리나. 안 돼. 잘했어. 이리와. 이봐 칼린카. 나 간다. 이제 개는 다람쥐가 어디 있는지 알아요. 움직이지 않고 있죠. 하지만 다람쥐 쪽으로 가지는 않아요. 가지 않는 이유는 제가 이미 저곳은 가면 안 되는 곳이라고 가르쳤기 때문이에요. 이게 아주 중요해요. 이것이 바로 우리가 행동 통제에 대해 이야기할 때 언급하는 부분입니다. 칼리나. 이리와. 방금 저 회피하는 모습 봤죠? 칼린카. 하지만 이제는 자유롭게 풀어줄 것이고, 개는 다른 무엇이든 사냥할 것입니다. 다시 말하지만, 저는 개를 풀어주어 즐거운 시간을 보내게 할 수 있습니다. 개를 풀어주어 무언가를 잡게 할 수도 있고, 줄을 매지 않은 상태에서 말로 행동을 통제할 수도 있습니다. 그것은 정통 훈련 기술과 함께 옵니다. 무지개나 유니콘 같은 환상이 아닙니다. 긍정 강화나 차별 강화 프로그램만으로는 이를 달성할 수 없습니다. 칼리나. 이리 와, 얘야.
In my case I use the word no, but that's really the word I choose. It doesn't matter which word you will pick. That cue is called Conditioned or Secondary Punisher in the textbooks. Okay, enough of the talking. Let's step outside so I can show you the setup I have created to demonstrate how Conditioned Punisher works. And then I will take a linker out on a squirrel hunt. So let's do it. Natalia is going to help us. Natalia is going to help us. Natalia is going to help us. Natalia is going to help us. So typically this is where the squirrels hang around. Do a little movement on the squirrel. Yeah, good. So that's basically the idea. There is no chance that she will think this is a setup. Unless we let her explore. She knows that there are always squirrels here. So they're either here or by the bird feeders on the other side. Or of course anywhere around. So I don't know. Sometimes she may come and really make a very quick run. And sometimes she can come in that very stealthy mood and trying to stalk and catch. So the plan is that she sees this squirrel, she goes for it and I tell her no. And what I'm expecting to happen is that she will not continue. That I will actually be able to stop her. This is very different than again a playful dog that's you know having different intentions. Don't grab and kill the squirrel. So I will talk to you later why I can say no and I can control the dog. Okay so let's let her out and see how that goes. Kalina. No. Good. Come. Hey Kalinka. I'm going. So now she knows where it is. She isn't moving. However, she does not go. The reason she does not go is because I have already told her that this is off limits. And it's very important. This is what we are talking when we talk about behavior control. Kalina. Come. You see that avoidance right? Kalinka. However, now I'm going to let her be free and she will hunt anything else. Again, I can let her go and have fun. I can let her go and catch something or I can control the behavior verbally without anything on her. That comes with legit training techniques. Not rainbows and unicorns. No positive reinforcement, differential reinforcement programming can accomplish this. Kalina. Hey mama.
11:37
이쪽으로 와. 칼리나. 안 돼! 잘했어. 착하지. 이리 와. 아, 나의 다람쥐 사냥꾼. 이 일에 아주 능숙하죠. 네. 원리는 이렇습니다. '안 돼'라는 신호는 기본적으로 혐오적인 결과와 연결되어 있습니다. 시간이 지나면 신호만으로도 개를 멈추게 할 수 있습니다. 이를 기능의 전이라고 생각하면 됩니다. '안 돼'라는 신호는 단순히 소리에 그치지 않고 그 자체가 결과로서 기능하기 시작합니다. 정확한 타이밍과 고전적 조건화의 올바른 순서를 갖추어 공정하게 사용한다면, 가장 중요한 순간에 신속하고 확실하게 행동을 통제할 수 있게 됩니다. 이는 문제가 많은 개들에게도 진정한 자유를 주어, 지속적인 스트레스 관리 없이 목줄 없는 산책을 즐길 수 있게 합니다. 이제, 자신들이 그것을 할 수 있다고 믿는 모든 힘을 쓰지 않는(force-free) 훈련 옹호자들에게 말합니다. 그것은 불가능합니다. 실제로 작은 동물을 사냥하고 죽인 이력이 있는 개를 차별 강화, 역조건 형성, 탈감각화를 통해 그렇게 행동을 통제하는 것은 불가능합니다. 이것은 절대적으로 불가능합니다. 제가 보여드리고 싶었던 것은 이것이 전부입니다. 댓글을 남겨주시면 기꺼이 더 논의하겠습니다. 평정심을 유지하세요. 감사합니다.
Come this way. Kalina. No! Good. Good mama. Come. Oh. My squirrel hunter. She's very skilled at this. Yeah. Here is how it works. The work knob is basically paired with an aversive consequence. Over time, the signal alone stops the dog. You can think of this as a transfer of function. The no signal basically stops being just a sound and starts functioning as the consequence itself. When used fairly with clear timing and the proper order of classical conditioning, it will provide fast, reliable control of behavior at the most critical times when you need it. It gives even the most problematic dogs true freedom so they can enjoy off-leash walks without the constant stressful management. Now, for all force-free advocates that believe that they can do that, it's very much impossible to stop a dog that has a history of actually hunting and killing small animals through differential reinforcement, through counter conditioning, through descentization to have that kind of control of their behavior. This is absolutely not possible. That's all I wanted to show you. Leave comments and I'll be happy to discuss further. Stay calm. Thank you.