Michael Ellis on Creating Reward Events
Michael EllisLeerburg Dog Training Podcast · 팟캐스트 · 15분
성공적인 반려견 훈련은 핸들러와 반려견 간의 역동적이고 상호작용적인 보상 이벤트를 구축하는 것에서 시작한다. 전통적인 훈련이 정적인 간식이나 장난감 제공에 그쳤다면, 현대적인 접근 방식은 보상을 단순한 물건이 아닌 강도와 지속 시간이 변하는 활동으로 정의한다. 반려견은 특정 행동을 수행함으로써 자신이 원하는 보상 이벤트에 도달하고 이를 획득하는 법을 배운다. 핸들러는 자신의 신체 움직임을 통제하고, 터그 놀이, 추격, 먹이 획득 등 반려견의 사냥 본능을 자극하는 활동을 통합하여 동기 부여의 수준을 극대화한다. 각 반려견이 가진 유전적 기질, 소유욕, 에너지 발산 방식에 따라 훈련 계획을 개별화하며, 단순히 사료를 먹이는 행위보다 획득하는 과정 자체를 강화물로 활용한다. 훈련 초기 단계에서는 복종 동작의 완성도보다 핸들러에 대한 주의 집중과 상호작용에 대한 열망을 발전시키는 데 집중하여, 향후 더 높은 난이도의 훈련에서도 반려견이 방해 요소를 무시하고 참여할 수 있는 기반을 마련한다.
마이클 엘리스의 보상 이벤트 생성 방법
Michael Ellis on Creating Reward Events
0:00
보상 기반 훈련을 위한 동기 부여 구축 기법에 대해 이야기하겠습니다. 저희 훈련의 많은 부분은 저희가 이른바 좋은 보상 이벤트, 즉 좋은 상호작용적 보상 이벤트를 갖는 것을 전제로 합니다. 그리고 저희는 기초 훈련의 상당 부분을 특정 유형의 보상에 대한 욕구를 형성하는 데 할애합니다. 전통적으로 저희는 보상을 강아지에게 주는 물건으로 보았습니다. 간식으로 훈련하고, 공으로 훈련하고, 터그 놀이 장난감으로 훈련하는데, 그때는 그 물건을 통해 만들어지는 이벤트보다 물건 자체가 더 중요했습니다. 저희 훈련이 진화하고 발전함에 따라, 저희는 보상을 단순히 주어지는 물건이라기보다 일어나는 어떤 사건으로 바라보게 되었고, 실제로 강아지와 핸들러 사이에서 지속 시간과 강도가 변하는 상호작용적 이벤트로서 발생하는 것에 더 관심을 둡니다. 그래서 저희는 사용하는 물건 자체보다는 그 물건을 어떻게 사용하는지에 더 신경을 씁니다. 그래서 저는 간식을 장난감을 사용하는 방식과 거의 동일하게 사용할 수 있습니다. 터그를 사용하거나 터그 장난감을 사용하거나, 공을 사용하거나, 플라스틱 병을 사용하거나, 아니면 제가 직접 달려가 강아지를 밀치며 몸으로 놀아줄 수도 있습니다. 이처럼 강아지와 핸들러 사이에 만족스러운 상호작용을 만들어내는 다양한 방법이 존재합니다. 이 보상 이벤트가 구축되면, 저희는 이를 앞으로의 행동에 대한 강화물로 사용합니다. 그리고 저희 훈련의 거의 모든 것은 보상 기반 시스템 내에서 새로운 복종 행동을 가르치는 데 맞춰져 있습니다. 즉, 강아지는 특정 행동을 함으로써 자신이 원하는 것을 얻기 위해 일하는 법을 배우는 것입니다. 이 보상 이벤트가 가치 있고 강렬하며, 강아지가 이 보상 이벤트에 대해 동기 부여가 많이 될수록, 훈련은 훨씬 더 잘 진행될 것입니다. 집중력도 좋아지고 참여도도 높아지며, 강아지는 더 빨리 배우고, 보상을 얻기 위해 더 열심히 노력하며, 그 보상 이벤트를 얻기 위해 더 많은 방해 요소를 무시하게 될 것입니다. 오늘은 소위 말하는 동기 부여에 대해 이야기해 볼 겁니다. 어떻게 동기를 만들어낼까요? 동기를 형성하는 기본 원칙은 무엇일까요? 몇 가지 섹션으로 나누어 살펴보겠습니다. 우선 반려견 없이 우리가 어떻게 움직여야 하는지부터 이야기해 보죠. 반려견 훈련에서 반려견과 잘 상호작용하려면 신체적인 요소가 상당히 많습니다. 지적으로 접근하는 데는 한계가 있죠. 하지만 훈련을 제대로 하려면 신체적인 노력이 필요합니다. 몸을 움직여야 하고, 뛰어야 하며, 몸을 굽히고 비틀기도 해야 합니다. 자신의 움직임을 자각하고 신체를 통제할 수 있어야 하죠.
Our techniques for building motivation for reward-based training. So a lot of our training is predicated on having what we call a good reward event, a good interactive reward event. And we spend a significant part of our foundation training building desire for a certain type of reward. Traditionally we looked at rewards as things that we gave our dog. You're training with food, you're training with a ball, you're training with a tug toy, and it was the item that was more important than the event that was created with the item. As our training sort of evolved and progressed, we look at the reward more as something that happens and we're less concerned something that happens actually between the dog and the handler as an interactive event that varies in duration and intensity. And we care less about the item we're using and more about how we use that item. So I could use food very much the same way I use the toy. I would use a tug or I could use a tug toy or I could use a ball or I could use a plastic bottle or I could run and push my dog and roughhouse with them. So there's a variety of different ways that I can create a satisfying interaction between the dog and handler. Once we have this reward event constructed, we use it as reinforcement for behavior as we go forward. And almost all of our training is geared towards teaching the dog new obedience behaviors in a reward-based system, meaning they're learning to work to get something they want by doing specific behaviors. And this reward event, the more valuable and the more intense and the more motivated my dog is for this reward event, the better my training is going to go. The better my attention is going to be, the better my engagement, the faster the dog is going to learn, the harder they're going to work to get to it, and the more distractions they're going to ignore in order to get to that reward event. So today we're going to talk about what we call, this is basically motivation. How do we build it? What are the basic principles in building it? So we're going to break it down into sections. We're going to talk about how we move without the dog. So there's a lot of physicality to interacting well with your dog in dog training. We can intellectualize it as much as we want, but if you're going to do it well, it's a physical endeavor. You're going to have to move your body. You're going to have to run, you're going to have to bend, you're going to have to twist, and you're going to have to be aware of how you're moving your body and be in control of your body.
2:42
그래서 반려견과 상호작용할 때 필요한 핵심 기술과 동작들을 따로 분리해서, 반려견 없이 연습해 봅니다. 그 과정들을 하나씩 살펴보며 어떻게 하는지 이야기해 보겠습니다. 보상을 하나의 상호작용 이벤트로 만드는 것에 대해 깊이 다룰 겁니다. 그렇다면 어떻게 상호작용 이벤트를 만들어낼까요? 우리는 그 과정에서 어떻게 몰입할 수 있을까요? 그리고 어떻게 반려견이 우리와 그 활동을 함께하고 싶어 하게 만들까요? 먹이에만 동기가 부여된 반려견을 훈련할 때 제가 단순히 먹이 자판기 같은 존재가 되어도 훈련의 진척을 낼 수는 있습니다. 하지만 제가 직접 상호작용하면서 먹이를 받아먹든, 공을 쫓든, 터그 놀이를 하든, 반려견이 저와 그런 활동을 함께하고 싶어 할 때 훨씬 더 좋은 결과를 얻을 수 있습니다. 그 방법을 어떻게 실천하는지 자세히 이야기할 겁니다. 다양한 분야에서 차용해 온 기법들과 동기 부여를 높이는 여러 방법들도 다룰 예정입니다. 좌절감, 억제, 그리고 동기 부여를 높이기 위해 사용하는 다양한 동작 기술에 대해서도 이야기해 보겠습니다. 방호 훈련 스포츠에서 차용한 기술들에 대해서도 이야기할 예정입니다. 제가 수년간 방호 훈련 스포츠에 참여하면서 알게 된 점은, 방호 스포츠에서 반려견을 발달시키기 위해 사용하는 기술들이 다른 많은 분야에서도 유용하다는 사실입니다. 좌절과 절제를 통해 우리가 온라인으로 하는 활동은 반려견의 놀이, 터그 놀이, 추격, 콜(호출) 등 다양한 욕구를 개발하는 데에도 유용합니다. 그래서 우리는 보호 훈련 스포츠에서 차용한 기술들과 어떤 종류의 바이트 워크(입질 훈련)에 대해 이야기해 볼 것입니다. 실제로 많은 반려견과 보호 훈련을 하려는 것은 아니지만, 제가 구성하는 전반적인 보상 이벤트에 어떤 종류의 바이트 워크가 유익한지 다룰 것입니다. 그런 다음 반려견의 유전적 성향에 대해서도 이야기할 것입니다. 특정 반려견은 유전적으로 특정한 방식으로 놀고 상호작용하기를 원할 것입니다. 둥근 구멍에 네모난 못을 억지로 끼워 맞추려 하지 말고, 반려견의 강점을 살리는 것이 중요합니다. 어떤 반려견은 더 선천적으로 소유욕이 강하고, 어떤 반려견은 움직임에 더 관심이 많습니다. 어떤 반려견은 싸우거나 거칠게 노는 것, 그리고 입을 더 공격적으로 사용하는 것에 더 관심이 있습니다. 만약 제가 움직임에 더 관심이 많은 반려견을 데려다가 터그 놀이를 좋아하게 만들고 거칠게 놀아주려 한다면, 그 반려견은 그것을 만족스럽게 느끼지 않을 수도 있습니다. 그리고 제가 원하는 방식으로 반려견이 놀게 하려는 시도 때문에, 반려견이 그 활동 자체를 싫어하게 만들 수도 있습니다.
So we break out some of the core skills and the core movements that you use in interacting with your dog, and we do them without the dog. So we're going to go through some of those steps and talk about how we do that. We're going to talk a lot about the reward as an interactive event. So how do we build that interactive event? How are we present in it? And how do we get the dog to want to do that activity with us? I can have a dog that's motivated strictly for food, and I can be kind of a human food dispenser, and I can make progress training that dog. But I get much better results from a dog that wants to do that activity, whether it's take a piece of food, chase a ball, play tug, with me specifically in an interactive way. And you're going to see we'll talk a lot about how we do that. We're going to talk about some of the techniques that we've borrowed from different disciplines and things to increase motivation. We're going to talk about frustration, restraint, a variety of different movement techniques that we use to increase motivation. We're going to talk about techniques that we've borrowed from protection sports. So one of the things that I found over the years is when I got involved in doing protection sports is that there are techniques to develop the dog in protection sports that are useful for many other disciplines. What we do online through frustration and restraint also has use in developing your dog's desire to play, tug, chase, recalls, a variety of different things. So we're going to talk about what techniques we've borrowed from protection sports and what kind of bite work. We're not really doing protection work with a lot of the dogs, but what kind of bite work is beneficial for my overall reward event that I'm constructing. And then we're going to talk about genetic propensities in dogs as well. So certain dogs genetically are going to want to play and interact in a certain way. And it's important that we don't try to fit a square peg in a round hole, that we go with the dog's strengths. Some dogs are more naturally possessive. Some dogs are more interested in movement. Some dogs are more interested in kind of fighting and rough housing and using their mouths more aggressively. And if I'm trying to take a dog that's more interested in movement and turn that dog into a kind of tugger and I want to rough house with that dog, that dog might not find that satisfying. And in my attempts to make the dog play the way I want to play, I make the dog not like the activity.
5:16
따라서 반려견을 특정 방향으로 이끌고 싶지만, 동시에 반려견이 유전적으로 타고난 기질에 맞춰가야 합니다. 반려견이 선천적으로 타고난 특정한 성향들이 있으며, 우리가 그것을 형성할 수 있는 범위는 제한적입니다. 또한, 반려견마다 에너지를 발산하는 방식이 다릅니다. 그래서 우리의 훈련 계획과 반려견과 상호작용하는 방식도 그것에 따라 달라져야 합니다. 여러분은 우리가 소위 '내향적' 반려견과 '외향적' 반려견이라고 부르는 것에 대해 자주 듣게 될 것입니다. 그중 일부는 주어진 반려견이 에너지를 어떻게 발산하느냐에 관한 것입니다. 내향적인 반려견은 매우 집중할 수는 있지만, 주변을 많이 움직이지는 않습니다. 그들은 가만히 앉아 있거나 자세를 잘 유지하는 경향이 있는 반면, 외향적인 반려견은 움직임이 많습니다. 그들이 내면에서 느끼는 모든 것이 신체를 통해 밖으로 드러나는 것을 볼 수 있습니다. 꼬리를 흔들고, 방방 뛰고, 몸을 떨며, 가만히 앉아 있는 것을 힘들어하죠. 따라서 우리는 각 개의 기질 유형에 따라 서로 다른 훈련 계획을 세워야 합니다. 저는 내향적인 개와 훈련할 때는 훈련의 한 측면에, 외향적인 개와 훈련할 때는 또 다른 측면에 더 집중해야 합니다. 이것은 우리가 개에게 보상을 주거나 함께 훈련할 때 무엇을 할지 결정하게 만드는 유전적 형질의 한 예라고 할 수 있습니다. 저희의 다른 영상들을 보셨거나 훈련 과정을 보셨다면, 저희가 '참여(engagement)'라는 개념에 대해 자주 이야기하는 것을 들으셨을 겁니다. 다시 강조할 필요가 있다고 생각하는데, 여기서 '참여'란 단순히 개가 주인에게 지속적인 관심을 기울이는 것을 의미합니다. 또한 간식, 장난감, 놀이, 또는 어떤 형태의 상호작용이든 보상으로 제공되는 것을 개가 원한다는 뜻이기도 합니다. 개는 주인에게 원하는 것이 있고, 그것을 얻기 위해 주인에게 계속 주의를 집중하게 됩니다. 이것이 저희 전체 훈련 시스템의 초석입니다. 만약 개가 당신에게 주의를 기울이지 않는다면, 개에게 행동을 가르치려는 시도는 의미가 없습니다. 그래서 저희는 초기 훈련의 상당 부분을 참여를 이끌어내는 데 할애합니다. 이는 보상 이벤트의 가치를 높이는 것과 밀접한 관련이 있습니다. 이 두 가지는 동시에 일어납니다. 그래서 저는 개를 데리고 나와 마커(marker)를 각인시키기 시작합니다. 개가 음식을 따라오도록 가르치기 시작하죠. 개가 음식을 쫓도록 가르치고, 장난감을 가지고 놀고, 붙잡아 두었다가 부르는 훈련(restrained recalls)을 시작합니다. 우리가 이야기해 온 모든 것들 말이죠. 이러한 활동들과 저와의 상호작용에 대한 욕구를 형성하는 과정에서, 개는 자연스럽게 저에게 주의를 기울이게 됩니다. 그리고 우리가 개발하고자 하는 것은 지속적으로 관심을 보이며 우리에게 무언가를 원하는 개입니다. 그런 개는 훈련하기가 매우 쉽습니다. 그래서 저희 모든 훈련 영상에서 이 부분을 반복해서 강조하는 것입니다.
So I want to steer the dog in certain directions, but I also have to kind of go with what the dog gives me genetically. And there are certain types of propensities that a dog will have naturally and we can shape them only so much. Also, different dogs manifest energy differently. So our training plans and how we interact with them has to be a function of that as well. You'll hear us talk a lot about what we call internal versus external dogs. And part of that is just how that given dog manifests energy. An internal dog can be very focused, but doesn't move around very much. They tend to sit still well and hold themselves well. And an external dog moves a lot. Everything that they're feeling inside you see coming out through their body. Their tails wag, they bounce, they vibrate, they have a hard time sitting still. So we have to have a different training plan for each of those dogs based on their temperament type. I need to focus more on one aspect of the training with an internal dog and more on another with an external dog. This is kind of an example of a genetic trait that steers what we do when we're rewarding and working with the dog. If you've seen any of our other videos or watched any of our training, you hear us talk a lot about the concept of engagement. And I think it bears repeating again that engagement simply means that the dog will pay sustained attention to you and wants what you have to offer in terms of a reward event. Whether it's food, toys, play, some kind of interaction. The dog wants something from you and will hold their attention on you in an attempt to get it from you. This is the cornerstone of our entire training system. If your dog is not paying attention to you, then we have no business trying to teach the dog's behaviors. And so we spend a significant part of our early training developing engagement. It kind of goes hand in hand with developing value in the reward event. These two things happen together. So I bring my dog out. I start charging my markers. I start teaching my dog to follow food. I start teaching my dog to chase food. I start playing with toys. I start doing restrained recalls. All the things that we've talked about. And in the course of building the desire for these activities and these interactions with me, the dog naturally pays attention to me. And what we want to develop is a dog that pays continuous attention and wants something from us. It's a very easy dog to train. So you'll hear us harp on this over and over again in all of our training videos
7:44
이 모든 것이 퍼즐의 매우 중요한 조각이기 때문입니다. 에너지를 더 쏟아서 반려견과의 교감(engagement)을 발전시키고 보상 상황에 대한 열정과 에너지를 쌓는 것은 여러분의 훈련 인생에 큰 도움이 될 것입니다. 그러므로 훈련 초기 단계에서는 이러한 것들을 발전시키는 데 집중하는 것이 더 좋습니다. 즉, 반려견이 어떤 행동을 하는지 걱정하는 것보다 보상 시스템, 의사소통 시스템, 그리고 교감에 신경을 쓰는 것이 더 낫습니다. 반려견이 앉거나 눕거나 가져오기를 하거나 그런 것들은 중요하지 않습니다. 여러분이 에너지를 쏟는다면 훈련의 이러한 다른 측면들, 즉 교감, 의사소통, 그리고 탄탄한 보상 이벤트 구축에 필요한, 상호작용하는 보상 이벤트에 에너지를 쏟는다면, 나머지 모든 것은 쉽게 자리를 잡게 될 것입니다. 이번 섹션에서는 보상을 상호작용적 이벤트로서 다루고 그와 관련된 몇 가지 원칙에 대해 이야기하겠습니다. 저희 훈련을 오랫동안 접해오신 분들은 저희가 보상의 상호작용적 특성에 대해 얼마나 많이 이야기하는지 아실 겁니다. 줄다리기는 반려견과 정기적으로 하는 놀이입니다. 왜냐하면 특정 반려견들에게는 매우 강력한 동기부여가 되기 때문입니다. 물론 모든 개에게 해당되는 것은 아니지만, 이 놀이는 반려견과 보호자 사이의 직접적인 상호작용이기 때문입니다. 반려견 스스로는 할 수 없는 것이죠. 그래서 우리는 이러한 상호작용적 특성을 모든 보상 시스템에 적용하고 싶습니다. 제가 음식을 사용하든, 공이나 줄 장난감을 사용하든, 아니면 단순히 반려견과 신체적으로 놀아주며 저를 쫓아오게 하거나, 반려견을 살짝 밀치며 장난을 치든 상관없습니다. 저는 그것이 반려견과 제가 함께하는 활동이 되길 바랍니다. 그래서 우리는 그것의 상호작용적 특성에 대해 많이 이야기할 것입니다. 그리고 이번 섹션에서 다룰 몇 가지 개념은 첫째, 움직임은 강화가 된다는 점입니다. 이는 매우 중요합니다. 가만히 앉아서 반려견에게 보상을 건네주는 것은 동기부여가 크게 되지 않습니다. 움직임입니다. 보상의 움직임과 움직임의 핸들러, 이 두 가지 모두입니다. 그래서 우리는 그것을 더 흥미롭고 재미있게 만들기 위해 우리 스스로가 어떻게 움직여야 하는지에 대해 많이 집중할 것입니다. 보상 이벤트와 그 전후에 일어나는 일 혹은 일반적인 행동 간의 대비에 대해 이야기할 것입니다. 단순히 우리가 움직이는 것만이 아니라, 보상 이벤트가 발생하기 전의 상황과 보상 이벤트 자체 사이에 강한 대비가 존재하기 때문입니다. 그러므로 제가 가만히 서 있다가 갑자기 점프해서 무언가를 한다면, 그것은 개에게 매우 강력한 보상이 됩니다. 일반적인 행동과 보상 중에 일어나는 일 사이에는 큰 대비가 있습니다. 평소 제 행동과 보상이 비슷해 보이기 시작할수록, 둘 사이의 차이가 적을수록 개에게 주는 동기부여는 줄어듭니다. 그래서 우리는 훈련을 기복이 심한 상태로 만들기 위해 노력합니다.
and in all because it is such an important piece of the puzzle. And extra energy devoted to developing engagement in your dog and to building power and passion for the reward event will carry you a long way in your training life. So you are better focusing your attentions in the early part of your training on developing these things, the reward system, our communication system, and engagement than you are worrying about what behaviors your dog does. It doesn't matter if your dog sits or downs or retrieves or any of those things. If you devote the energy necessary to these other aspects of your training, engagement, communication, and building solid reward events, interactive reward events, then all the rest of the stuff will come into place easily. In this section, we're going to talk about the reward as an interactive event and some of the principles that surround that. So for those of you that have been around much of our training, you realize how much we talk about the interactive nature of a reward. Tug of War is a game that we play with our dogs on a regular basis because it's highly motivating for certain dogs, not all dogs, but also because it's directly interactive between the dog and the handler. The dog can't do that by themselves. And so we want to take that interactive nature and apply it to all of our reward systems, whether I'm using food, a ball, or a tug, or just playing with my dog physically, having the dog chase me around and pushing my dog off of me and playing games with my dog. I want it to be something that the dog and I do together. So we're going to talk a lot about the interactive nature of that. And some of the concepts that we're going to cover in this section are one, movement is reinforcing. This is huge. Sitting still and handing your dog's rewards is not very motivating. Movement. Movement of the reward and movement of the handler, both of those things. So we're going to focus a lot on how we move ourselves to make that more interesting and exciting for the dog. We're going to talk about the contrast between the reward event and what's happening around the reward event or normal behavior. Because it's not simply that we move, it's that there's a strong contrast between what was happening before the reward event occurs and the reward event itself. So if I'm standing still and suddenly I jump and go and we do something, that's highly reinforcing to a dog. There's a big contrast between normal behavior and what happens during the reward. The more than my normal behavior and my reward start to look alike, the less difference
10:23
매우 활기찬 보상 이벤트 뒤에 차분함이 이어지고, 다시 활기찬 보상 이벤트가 이어지는 식인데, 이러한 대비가 매우 높은 동기부여를 제공합니다. 이제 배고픔에 의한 본능과 움직임에 기반한 사냥 본능 시퀀스의 차이에 대해 이야기해 보겠습니다. 배고픔에 의한 본능의 한 가지 특징은, 음식으로 훈련할 때 많은 사람이 단지 음식을 사용해 개에게 동기를 부여한다는 점입니다. 개를 배고프게 만들고 개가 음식을 원하면, 개가 무언가를 하고 제가 음식을 한 조각 주는데, 훌륭합니다, 효과가 있죠. 하지만 그 시점에서 개는 오로지 배고픔이라는 본능에 의해서만 동기부여를 받고 있는 것입니다. 그저 얼마나 배가 고픈가의 문제입니다. 그리고 배고픔에 의한 본능은 반복을 통해 강도가 더해지지 않습니다. 개가 배고플 때 음식을 먹으면 포만감이 들어 본능이 줄어듭니다. 따라서 그런 방식으로 훈련할수록 실제로 개는 음식에 대한 동기가 낮아집니다. 반면 사냥 본능이나 음식 혹은 장난감을 이용한 움직임 기반의 보상 체계는 반복할수록 강도가 강해집니다. 그러니 어린 강아지가 공을 쫓는다면, 공을 더 많이 던져주고 그것을 더 많이 반복하게 할수록 강아지는 더욱 열중하게 됩니다. 그렇게 하면 개의 동기부여가 점점 더 강화됩니다. 이것이 바로 제가 '재생산'이라고 부르는 과정입니다. 그들은 그것에 대해 알게 됩니다. 따라서 리허설은 먹이 사냥 본능에 기반한 행동이나 추격 행동을 강화합니다. 그러므로 우리의 보상은 음식을 사용하든 장난감을 사용하든 그런 방식으로 움직이는 것에 좌우될 것입니다. 왜냐하면 그것이 훨씬 더 강력한 동기를 부여하기 때문입니다. 이제 보상 이벤트의 가변적인 지속 시간과 그것이 소위 '강화 후 휴지기' 또는 주의 산만에 미치는 영향에 대해 이야기해 보겠습니다. 많은 사람들이 훈련 중에 반려견에게 보상을 줄 때 그 보상은 항상 일정한 길이로 정해져 있습니다. 개가 앉으면 간식 하나를 줍니다. 그래서 그 보상은 간식 하나로, 아주 짧은 시간 동안만 지속됩니다. 제가 보상을 주는 방식이 예측 가능하다면, 보상은 항상 특정 시간 동안만 지속되고, 그러면 개는 보상 직후 잠시 동안 주의를 다른 곳으로 돌리기 시작합니다. 다음 보상이 또 몇 초 동안은 나오지 않을 것을 알기 때문입니다. 그래서 개가 하는 행동은, 강화 후 휴지기라고 부르는 것인데, 개가 '좋아, 보상을 받았어'라고 생각하고는 잠시 딴청을 피우다가 다시 집중하게 됩니다. 왜냐하면 당장은 다음 보상이 없다는 것을 알기 때문입니다. 그래서 우리는 보상 이벤트의 지속 시간을 다양하게 합니다. 때로는 간식 하나일 수도 있고, 때로는 간식 세 개를 10초 동안 쫓아가게 할 수도 있고, 때로는 30초 동안 터그 놀이를 할 수도 있으며, 때로는 5초 동안 터그 놀이를 하기도 합니다. 저는 그 지속 시간을 다양하게 바꿉니다. 그래서 개는 보상이, 보상 이벤트가 얼마나 지속될지 알 수 없습니다. 따라서 개들은 매번 보상을 받은 후 딴청을 피우지 않습니다.
between the two, the less motivating that is to the dog. So we try to make our training very much full of peaks and valleys. Really high active reward events followed by calm, high active reward events, and those contrasts are highly motivating. We're going to talk about the difference between hunger drive and movement-based prey sequences. So one of the things about hunger drive, so if we're training with food, lots of people will use just the food to motivate the dog. You get your dog hungry, your dog wants food, he does something, I hand him a piece of food, great, it works. But he's being completely motivated by his hunger drive at that point. It's simply how hungry he is. Where, and hunger drive does not intensify through rehearsal. Your dog can be hungry and when he eats it satiates that and that drive drops. So the more you train that way, actually the less motivated the dog becomes for the food. Whereas prey or movement-based reward systems with food or toys intensify through rehearsal. So if your dog chases a ball as a young dog, the more you throw the ball for your dog and the more they rehearse that, the more intense they get about it. So rehearsal intensifies prey-based or chasing behaviors. So our rewards are going to be predicated on moving that way, whether we're using food or a toy, because it's much more motivating. We're going to talk about variable duration of the reward event and its effect on what we call post-reinforcement pause or inattention. Lots of people, when they reward their dog in training, the reward is always of a fixed length. My dog sits, I give him a piece of food. So that reward is a single piece of food, it lasted X seconds, a very short period of time. If I'm predictable in how I do that, my reward always lasts a certain amount of time, then my dog starts to check out right after the reward for a brief period, knowing that there's not another reward coming for X seconds again. And so what that dog does, they call that post-reinforcement pause, where your dog says, okay, I've got my reward. Now he checks out for a few seconds and then he checks back in because he knows that he's not getting another reward. So we vary the duration of our reward event. So sometimes it's one piece of food, sometimes it's you chase three pieces of food for 10 seconds, sometimes I play tug with you for 30 seconds, sometimes I play tug with you for five seconds. I vary that duration. So the dog never knows how long the reward is going to go on, the reward event is going to go on. And so they don't check out after each
12:57
왜냐하면 보상이 계속될 것이라고 기대하기 때문입니다. 그런 가변성이 개를 계속 참여하게 만들고, 잠시 집중력을 잃는 순간들을 없애줍니다. 이제 청각적 자극에 대해 이야기하겠습니다. 우리는 강아지 물기 훈련 영상에서 이에 대해 이야기했지만, 복종 훈련에서도 동일하게 적용합니다. 그래서 장난감이나 음식으로 강아지를 유도할 때 우리는 작은 소리를 냅니다. 그리고 이러한 소리가 의미하는 바를 개에게 조건화합니다. 자극적이고 흥미롭습니다. 그래서 제 반려견이 제 손에 있는 간식을 쫓아올 때, 저는 반려견이 그것을 쫓아올 때 작은 소리를 냅니다. 원하시는 대로 어떤 소리든 내셔도 됩니다. 하지만 그 청각적 자극은 고전적 조건 형성을 통해 반려견을 자극하고 그들을 흥분하게 만듭니다. 그래서 나중에 보상 과정에서 반려견을 자극하고 그 활동에 대한 전반적인 동기 부여와 흥분을 높이는 데 사용할 수 있습니다. 우리는 이제 이야기할 것입니다. 개의 뇌는 보상을 얻는 행위 자체가 보상을 가지고 있는 것보다 더 강화가 되도록 설계되어 있습니다. 즉, 개는 쫓고 탐색하는 행동을 좋아하도록 생물학적으로 설계되어 있습니다. 예를 들어, 야생에 있는 개과 동물은 무언가를 사냥합니다. 그리고 대부분의 사냥 시도는 성공하지 못하죠. 다들 자연 다큐멘터리 보신 적 있죠? 80%의 경우 개나 동물은 자신이 쫓던 것을 얻지 못합니다. 그런데 만약 80%가 실패한다면 당신은 계속해서 어떤 행동을 할까요? 그런 경우는 거의 없겠죠. 그래서 생물학적으로 그 행동을 할 때 기분이 좋게 만들어 계속하게끔 설계된 것입니다. 그러니 쫓는 행위 와 탐색하는 행위는 개의 뇌에서 기분을 좋게 만드는 부분을 활성화합니다. 그리고 그것을 많이 반복할수록, 그들은 그것을 더 좋아하게 됩니다. 따라서 우리의 보상에는 탐색과 추격이 포함되어야 합니다. 실제로 음식 조각을 먹는 것보다는 음식을 얻어내는 과정이 더 중요합니다. 그리고 개에게는 획득 단계에 있는 것이 훨씬 더 큰 동기 부여가 됩니다. 그렇다면 이것을 어떻게 보상 과정에 통합할까요? 그리고 나서 우리는 상호작용 퍼즐의 기본 세 가지 요소인 터그 놀이, 추격, 그리고 먹는 것에 대해 이야기할 것입니다. 터그 놀이는 당연히 보호자와 반려견 사이에서 직접 이루어집니다. 우리는 터그 놀이를 하고 추격을 합니다. 반려견은 저를 쫓거나, 제 손에 있는 것을 쫓거나, 혹은 제가 던지는 것을 쫓을 수 있습니다. 그리고 물론 먹는 것이 있죠! 우리는 반려견의 먹이 본능을 이용해 동기를 부여합니다. 음식을 위해서요. 그리고 우리가 어떻게 음식을 제공하느냐에 따라 그 보상 사건이 얼마나 가치 있게 될지가 결정됩니다.
reward, because they expect it to kind of keep going. That variability keeps them engaged and you don't have those little moments of checking out. We're going to talk about our auditory stimulator. We talk about this in our puppy bite work video, but we also do it in our obedience. So when we're teasing a dog with toy or food, we make little noises. And we condition the dog that these noises are stimulating and exciting. So if my dog's chasing a piece of food that's in my hand, I go and I make little noises as the dog chases it, whatever you want, you can make any kind of noise you like. But that auditory noise becomes classically conditioned to stimulate the dog and get them excited. So later on, we can use that to stimulate our dog during the reward event and bring up their overall motivation and excitement for that activity. We're going to talk about a dog's brain is wired in such a way that the acquisition of rewards is more in reinforcing than having them. So dogs are wired to like chasing and searching behavior. It's biological. So for instance, if there's a cane in the wild, they're hunting for something. And most hunting attempts are unsuccessful. Everyone watch the nature documentaries, right? 80% of the time, the dog or the animal does not get what they're going after. And what behavior would you continue to do if you were 80% unsuccessful? Very few of them. And so biology has conspired to keep them doing it by making it feel good. So chasing and searching activates a part of the dog's brain that feels good. And the more they rehearse it, the more they like it. So our rewards should include searching and chasing. It's less about the actual getting the piece of food and more about acquiring the piece of food. And it's much more motivating to the dog to be in the acquisition phase. So how do we incorporate that into our reward event? And then we're going to talk about tugging, chasing and eating, which are the basic three pieces of our interactive puzzle. So tugging is obviously directly between the handler and the dog. We're tugging, chasing. The dog can chase me, can chase what's in my hand or can chase something that I throw. And then eating, of course, right? So we're using the dog's hunger drive to motivate them for food. And then how we deliver the food predicts how valuable that reward event is going to be.