5 Simple Hacks to Help Your Dog Learn Faster (Reinforcement Process) #31
Susan GarrettShaped by Dog with Susan Garrett · 팟캐스트 · 20분
반려견 훈련은 보호자가 제공하는 보상의 가치를 반려견에게 온전히 전달하는 과정이다. 로스트 비프와 같은 고가치의 보상이 훈련 과정을 거치며 메커니즘의 오류로 인해 가치가 손실되지 않도록 관리하는 것이 핵심이다. 보호자는 강화가 단일 사건이 아닌 하나의 연속적인 과정임을 인지하고, 보상 전달 과정에서 발생하는 시간적 공백이나 부적절한 움직임이 반려견의 행동에 미치는 영향을 최소화해야 한다. 구체적인 훈련 단계인 기준 설정, 마킹, 보상 전달, 위치 선정, 릴리스 신호 사용을 체계적으로 적용할 때 반려견은 더 빠르고 효율적으로 목표 행동을 습득한다. 최종적으로 훈련의 결과는 보호자의 의도가 아닌 반려견의 인식에 의해 결정되므로, 보상 배치는 보호자의 목표와 일치해야 한다.
반려견의 학습 속도를 높여주는 5가지 간단한 팁 (강화 프로세스) #31
5 Simple Hacks to Help Your Dog Learn Faster (Reinforcement Process) #31
0:00
안녕하세요 여러분, Shaped by Dog에 오신 것을 환영합니다. 저는 수잔 개럿이고 오늘 제가 나눌 이야기는 여러분의 반려견 훈련을 훨씬 더 쉽고, 훨씬 더 효과적이며, 훨씬 더 효율적으로 만드는 방법입니다. 반려견 훈련에 강력한 힘을 더하는 방법을 알려드리고 싶습니다. 그리고 이는 여러분이 반려견 훈련 여정 어디에 있든 동일하게 적용됩니다. 저와 같은 전문가이든, 스포츠 분야의 경쟁자이든, 특히 복종 스포츠 대회에 출전하는 분들을 위한 특별한 내용입니다. 여러분이 초보자이든, 심지어 아직 반려견이 없어서, 그저 기다리는 중이라 해도 상관없습니다. 이 내용을 잘 기록해 두셨다가 나중에 새 반려견이나 강아지를 데려왔을 때 훈련에 엄청난 차이를 만들어 줄 것입니다. 시작하기 위해 기본 개념부터 잡죠. 모든 훈련은 가치의 전달입니다. 무엇을 배우든, 어떤 행동을 하든 상관없이, 개이든 사람이든 상관없이 가치의 전달이 일어납니다. 예를 들어, 여러분은 지금 이 팟캐스트를 듣고 계십니다. 아마 처음에는 단순히 구글에 '개'라는 단어를 검색했다가 우연히 들어왔을 수도 있지만, 다시 찾게 될 것입니다. 이상적으로는 여러분이 이 팟캐스트를 다시 찾는 이유는 그것이 여러분에게 어떤 가치를 제공했기 때문일 것입니다. 제 목표는 정보, 교육, 그리고 즐거움이라는 가치를 드리는 것입니다. 저는 두 가지 E를 잡았죠. 그게 제가 추구하는 것입니다. 여러분에게 즐거움을 드리면서도 훌륭한 반려견 훈련 교육을 제공하고 싶습니다. 여러분이 직장에 가는 것도 가치의 전달입니다. 여러분이 어떤 행동을 하면 보수를 받죠. 만약 보수를 받지 못한다면 계속 직장에 다니시겠습니까? 만약 다른 사람들을 위해 일하는 것에서 기쁨을 느끼는 가치를 얻는다면 그럴 수도 있겠죠. 모든 행동은 가치의 전달입니다. 우리 반려견들이 우리와 함께 어질리티를 하거나, 우리 곁을 걸어주거나, 장난감을 가져오는 것은, 우리가 그런 행동을 훈련시킬 때 반려견들에게 우리가 제공한 가치 때문입니다. 그래서 모든 훈련은 가치의 전달입니다. 좋습니다. 모든 훈련은 가치의 전달입니다. 어떻게 하면 여러분의 훈련을 강화할 수 있을까요? 어떻게 하면 훈련을 더 효과적이고 효율적으로 만들 수 있을까요? 바로 가치를 살펴보는 것입니다. 우리가 로스트 비프 한 조각으로 시작한다고 해봅시다. 이건 반려견에게 가치가 있죠. 그러니 반려견에게 엄청난 가치가 있는 것을 선택하세요. 반려견이 '아, 정말 좋아'라고 느끼는 것이요. 이제 우리는 그 가치를 가져와 행동으로 전달해야 합니다. 하지만 반려견 훈련 과정에서 그 가치의 일부가 손실됩니다. 그 가치는 어떻게 손실될까요? 바로 반려견 훈련사로서 우리의 메커니즘 때문에 그렇습니다. 따라서 우리가 해야 할 일은 가치를 유지하기 위해 메커니즘을 개선하는 방법을 이해하는 것입니다. 만약 로스트 비프가 반려견에게 10만큼의 가치라면,
Hey everybody, welcome to Shaped by Dog. I am Susan Garrett and today I'm going to share with you how you can make your dog training a lot easier, a lot more effective, a lot more efficient. I want to teach you how to supercharge your dog training. And it's going to be the same regardless of where you are in your dog training journey. If you're a professional like me, if you're a competitor in sports, something special for those of you who compete in the sport of obedience, it doesn't matter if you're brand new or if you don't even have a dog, you're just waiting. This is something you can take note of and that when you get your new dog or puppy, it will make a massive difference when you train. To start, let's get a base level. All training is a transfer of value. It doesn't matter what you're learning or what behavior you are doing, whether you are a dog or a human being, there's a transfer of value that occurs. For example, you're listening to this podcast. You might happen on it the first time, maybe just because you Googled the word dog, but you would return to it. Ideally, you're returning to it because it brought you some value. My goal is to bring the value of information, education, and entertainment. I got the double E. That's what I'm looking for. I'm hoping to entertain you, but to give you some great dog training education. You go to work, there's a transfer of value. You do a behavior, you get a paycheck. If you didn't get a paycheck, would you still be going to work? You might if you got the value of doing something for others that filled you with joy. Every behavior is a transfer of value. Our dogs will do agility with us, or they will walk by our side, or they will retrieve a toy for us because of the value we brought to them when we were training them to do those behaviors. All right. All training is a transfer of value. How are we going to supercharge your training? How are we going to make your training more effective and efficient? By looking at the value. Let's say we start with a piece of roast beef and it has a value to your dog. So, pick something that's got a massive value to your dog. The dog says, this is my, oh, I love that. Now, we need to take that value and transfer it into the behavior. But the process of dog training loses some of that value. And how is the value lost? By our mechanics as dog trainers. And so, all that we have to do is improve our understanding of how to improve our mechanics in order to maintain the value. If roast beef is a value of 10 to your dog,
3:23
우리는 그 10이라는 가치를 그대로 전달하여 훈련하는 행동의 끝에서도 10이라는 가치를 얻어야 합니다. 앉아, 엎드려, 장난감 쫓기, 손님에게 발 올리지 않기, 식탁에서 구걸하지 않기 등 무엇이든, 혼자 있을 때 짖지 않기 등 어떤 행동이든 우리는 그 전달을 원합니다. 훈련할 때 가치가 손실되는 것을 원치 않습니다. 그 가치의 전달이요. 우리의 훈련은 로스트 비프가 가진 10의 가치를 끝까지 유지할 수 있을 때 가장 효율적이고 효과적입니다. 그것이 저와 같은 전문가, 즉 도그 스포츠 최고 수준에서 뛰어난 성과를 거두고 성공한 사람과의 차이입니다. 왜냐하면 제 가치 전달이 더 효과적이고 효율적이기 때문이죠. 이제 여러분도 그렇게 할 수 있도록 돕겠습니다. 네 가지가 있습니다. 먼저 제 멘토인 밥 베일리가 강조한 내용을 공유하겠습니다. 그는 강화가 하나의 과정이지, 이벤트가 아니라는 훌륭한 만트라를 가지고 있습니다. 즉, 반려견이 앉았을 때 '착하다'라고 말하며 쿠키를 주는 상황에서, 강화는 쿠키가 개의 입으로 들어가는 것만이 아닙니다. 강화 과정이 존재합니다. 만약 여러분이 이를 이해하고 존중하기 시작한다면, 그 과정이 바로 로스트 비프의 10이라는 가치를 유지하며 전달하는 방법입니다. 가치를 잃지 않고, 행동을 유도하는 과정에서 흥미가 떨어지지 않게 하는 것이죠. 왜 그렇게 유지하고 싶어 할까요? 왜냐하면 만약 반려견이 로스트 비프를 정말 좋아한다면, 우리가 그 로스트 비프에 대한 애정을 제가 말했을 때 엎드리는 것과 같은 행동으로 전이시킬 수 있다면, 반려견은 빠르게 엎드릴 것이고 신나서 엎드릴 겁니다. 매번 그렇게 할 거예요. 그것이 바로 가치 전이가 완벽하게 이루어졌다는 증거입니다. 자, 이제 훈련 그 자체를 살펴봅시다. 우선, 우리는 반려견이 보여주는 반응, 스스로 하는 행동을 평가합니다. 그러다가 마음에 드는 행동을 발견하죠. 아까 이야기하던 엎드려 동작을 예로 들어보겠습니다. 엎드려의 기준이 무엇인지 알아야 합니다. 무엇이 엎드려를 정의할까요? 개가 더 이상 서 있거나 앉아 있지 않을 때라고 할 수 있을까요? 좋아요. 배는 땅에 닿았지만, 신이 나서 팔꿈치는 떠 있는 상태라면요. 그게 엎드려일까요? 고민해 봐야 할 문제죠. 사실 고민할 필요조차 없어야 합니다. 여러분은 기준을 알고 있어야 해요. 기준을 모르면 훈련할 수 없기 때문입니다. 훈련에 들어가기 전에 여러분의 기준이 무엇인지 확실히 정하세요. 이제 훈련은 반려견이 하는 행동을 평가하다가 보상하고 싶은 행동이 나올 때까지 기다리는 과정입니다. 이제 반려견이 엎드려를 이해했고 여러분은 더 빠르게 엎드리는 행동 등에 보상하고 싶다고 가정해 봅시다. 첫 번째 단계는 반려견의 모든 반응을 살펴보고 그중 마음에 드는 것을 포착할 수 있어야 한다는 것입니다.
we want to transfer that 10 and get a value of 10 at the end of the behavior that we're training. Regardless if it is sit down, chase a toy, keep your feet off my guests, don't beg at the table, don't bark when you're alone. Whatever the behavior is, we want the transfer. We don't want to lose the value by when we're training. That transfer of value. Our training is most efficient and effective if we can hold that 10 value for the roast beef all the way through. That's the difference between a professional like me, somebody who has excelled at the highest level of dog sport and had success because my transfer of value is more effective and efficient. I'm going to help you get yours there right now. There's four things. First of all, I'm going to share with you. My mentor, Bob Bailey, he has this great mantra and it is that reinforcement is a process. It's not an event. Meaning when your dog sits and you say, good boy, here's a cookie. The reinforcement isn't the cookie getting to the dog's mouth. There is a reinforcement process. And if you understand it and you start respecting it, that process is how we get the transfer of value from the roast beef of 10 to stay and not lose any value, not go down on the way to getting into the behavior. Why do we want to keep it that way? Because if your dog loves roast beef and we can transfer that love for roast beef into something like lying down when I say, your dog's going to lie down fast and they're going to lie down excited. They're going to do it every single time. That's how you know you've got a great transfer of value. And so, let's look at training itself. First, we're just evaluating the responses that our dogs get, our dogs offering, our dogs are doing. And then we see the one we like. Let's pick the down since we were talking about that. You have to know what is the criteria of a down. What makes it a down? Is it when the dogs, you can say, well, the dog's no longer in a stand or a sit. All right. So, their stomach's on the ground, but their elbows are off because they're excited. Is that a down? Let me think about that. You shouldn't have to think about it. You know your criteria because you can't train it if you don't know it. Know what your criteria is before you get into that. Now your training is evaluating what your dog is doing until you see what it is that you want to reward. Now let's say your dog understands a down and you want to reward, you know, faster downs or whatever. The first part is that you are able to look at all the responses and see the one you like.
6:26
그것이 강화 과정의 첫 번째 단계입니다. 기억하시나요? 강화는 이벤트가 아니라 과정입니다. 첫 번째는 마음에 드는 행동을 골라내는 것이고, 그다음은 마킹하는 것입니다. 항상 모든 행동을 마킹할 필요는 없지만, 대부분은 마킹한다고 가정해 보죠. 클릭기로 마킹할 수도 있습니다. '좋아' 같은 말로 마킹할 수도 있고, 'Yes'와 같은 다른 단어로 마킹할 수도 있습니다. 제 말은, 단어는 짧을수록 좋습니다. 그것이 강화 과정의 일부입니다. 좋습니다. 여기까지가, 첫 번째 단계는 여러분이 무엇을 좋아하는지 인지하거나 파악하는 것입니다. 쾅. 두 번째는 마킹입니다. 그리고 간격이 있죠. 좋아하는 것을 보고, 좋아하는 것을 골라내고, 그것을 마킹하는 것 사이에는 간격이 있습니다. 그것이 강화 과정의 일부입니다. 세 번째는 보상을 제공하는 것입니다. 이것도 강화 과정의 일부입니다. 여러분이 좋아하는 것을 마킹한 시점과 보상을 제공하는 시점 사이의 간격이죠. 그리고 마지막은 보상의 위치 선정입니다. 마지막이라고 해서는 안 되겠네요. 한 가지 더 있습니다. 바로 행동으로부터의 해방, 즉 개가 다음 행동으로 넘어가도록 허락을 주는 것입니다. 자, 다섯 가지가 있다고 칩시다. 좋습니다. 좋아하는 것을 보는 것. 1단계는 기준을 알고 그것을 평가하는 것입니다. 그리고 좋아하는 것을 보고 인지하는 사이의 간격이죠. 이때 여러분은 감정을 배제해야 합니다. 알겠죠? 왜냐하면 팟캐스트 16화 '그것 이전의 그것'을 기억하시나요? 음, 강화 과정에서 개들은 '그것 이전의 그것'을 아주 빠르게 포착합니다. 그래서 만약 여러분이 훈련 중에 지나치게 열중해서 고개를 갸웃거리고 있다가, 개가 제대로 이해하기 시작할 때 간식을 주러 움직이기 시작하면, 개는 여러분이 집중하던 상태에서 강화물을 향해 움직이는 것으로 변했다는 것을 알아차리고 행동을 멈추거나 행동을 짧게 끊어버릴 것입니다. 세기 전환기에 '클레버 한스'라는 말에 대한 오래된 행동 연구가 하나 있습니다. 독일에서 있었던 일인데, 이 사람은 말에게 어떤 수학 문제든 낼 수 있었습니다. 나눗셈, 산수, 곱셈, 덧셈, 뺄셈, 무엇이든 온갖 수학 문제를 냈죠. 심지어 문장제 문제도 있었던 것으로 기억합니다. 말은 80% 이상의 확률로 정답을 맞혔습니다. 제 기억에 87% 정도였던 것 같은데, 정답을 맞혔죠. 그래서 행동주의자들은 매우 놀랐습니다. 말 주인이 아닌 다른 사람들도 질문을 할 수 있었거든요. 다른 사람들이 질문을 던져도 말은 정답을 맞혔습니다. 간단히 말해서 그들이 발견한 것은 말이 포커를 할 때처럼 상대의 단서를 포착했다는 점입니다. 제 추측으로는 처음에 한스의 주인은
So, that's step number one in the reinforcement process. Remember? Reinforcement is a process, not an event. The first is picking out what you like. The next is marking it. Now you don't always mark all behaviors, but let's assume most of them you do. It could be marking with a clicker. I like it. It could be marking with a word. Good. Could be marking with a different word like, yes. I mean, the shorter the word, the better. That's part of the reinforcement process. All right. So, we've, first step is acknowledging or recognizing what you like. Boom. Second is marking. And there's a gap between seeing what you like, picking out what you like, and then marking it. That's part of the reinforcement process. The third is delivering the reward. That's part of the reinforcement process. The gap between you marking what you like and you delivering the reward. And the final thing is the placement of the reward. And I shouldn't say the final thing. There is one other thing, and that is the release from the behavior, giving the dog permission to move on. So, let's say there's five things. All right. Seeing what you like. Step number one, know your criteria and then evaluate it. And the gap between seeing what you like and acknowledging it, and you need to be non-emotional. All right. Because remember podcast episode number 16, the thing before the thing. Well, the reinforcement process, the dog picks up the thing before the thing very quickly. So, if you are training and you're like all intense and your head's all cocked to the side, and then when the dog's starting to get it, you start moving because you're going to move towards your cookie. The dog recognizes you went from intense to moving towards reinforcement and they will stop the behavior or they will cut it short. There is an old turn of the century behavior study with a horse called Clever Hans. And it was in Germany, this fellow could give the horse any kind of math problem, division, arithmetic, multiplication or addition or subtraction, anything, all kinds of math problems. I think there was actually even word problems. And the horse like over 80% of the time got it right. I think it was 87% of the time got it right. And so, that was astonishing to these behaviorists. Other people like could ask the question, not only the fellow who owned the horse, other people could ask the question, the horse got it right. What they, long story short, they found was the horse picked up tells, like poker tells, right? So, my hallucination is at first Hans's owner
9:22
말의 발을 아주 주의 깊게 보고 있었을지도 모릅니다. 말이 정답을 맞히면 주인은 고개를 들었고 말은 '아, 저 사람이 고개를 움직이니까 보상을 받겠구나'라는 걸 알게 된 거죠. 그게 바로 주인이 보상을 주기 직전에 하는 첫 번째 행동이니까요. 결국 말은 사람들의 눈을 보기 시작했습니다. 그래서 계속 발로 땅을 굴렀습니다. 예를 들어 '5 더하기 5는 뭐야?'라고 물으면 땅을 구르기 시작하는 식이죠. 그러다 사람들이 고개를 들기 시작하는 걸 보면 땅을 구르는 속도를 늦췄습니다. 그리고 사람들이 고개를 특정 방향으로 돌리면 땅 구르기를 멈췄죠. 그게 바로 '행동 이전의 행동'입니다. 그것이 훈련에 영향을 미치게 되고 효율성을 떨어뜨릴 것입니다. 따라서 여러분이 첫 번째 단계를 평가할 때, 눈에 보이는 기준들을 평가할 때는 감정을 배제하고 마킹을 하기 전까지는 움직이지 않아야 합니다. 이제 행동을 확인하고 마킹하는 사이의 간격은 중요합니다. 왜냐하면 예를 들어 강아지가 어떤 행동을 해서 여러분이 '아, 방금 그거 좋았어'라고 생각하는데 내가 원하는 게 바로 그거야. 맞아. 클리커를 꺼내서 클릭을 해야지 하는데, 어, 클리커가 거꾸로 있네. 바로잡아서 클릭해야지 하고 클릭하는 경우처럼요. 자, 그 간격 동안 여러분이 보상하는 것은 원래의 행동이 아니라 지속 시간입니다. 그러니까 강아지가 엎드려서 그 자세를 유지했다고 가정해 봅시다. 여러분은 이제 엎드린 행동 자체가 아니라 그 엎드린 상태를 마킹하게 되는 겁니다. '그래, 잘했어. 팔꿈치가 바닥에 닿았네. 마음에 들어'가 아니라요. 아니면 또 무슨 일이 일어날 수 있냐면, 강아지가 엎드린 뒤 여러분이 클리커를 만지작거리는 동안 강아지는 엎드린 자세에서 일어나 앉아 버릴 수도 있습니다. 그러면 여러분은 엉뚱한 행동을 클릭하게 되는 거죠. 훈련을 비효율적으로 만듭니다. 그래서 1단계는 감정을 배제하고 원하는 것을 골라내는 것입니다. 2단계는 마킹하는 것입니다. 3단계는 보상 전달입니다. 즉, 마킹을 했죠. 좋습니다. 마음에 듭니다. 그러고 나서 주머니에 손을 넣고 쿠키를 하나 꺼내는데, 아, 이건 너무 크네. 좋아, 반으로 쪼개자. 그다음에 나머지 쿠키들은 주머니에 다시 넣고 쿠키를 주려고 하는데, 손에 다른 쿠키 몇 개가 더 잡혀서 그걸 다시 집어넣으려고 하죠. 그럼 전달 방식이 엉망이 되는 겁니다, 그렇죠? 주머니에 손을 넣어 뒤적거리고 쿠키가 너무 커서 쪼개는 데 걸린 시간 때문에요. 그만큼 간격 사이의 시간이 낭비된 거죠. 매우 비효율적으로 변하고 있다는 걸 기억하세요. 그러고 나서 손안에 다른 쿠키들이 있는 상태로 쿠키를 전달하게 되면, 개는 하나를 받으면서 동시에 열 개가 멀어지는 걸 보게 됩니다.
might've been looking at his feet really intently. And when the horse got it right, he put his head up and then the horse knew, oh, I'll get a reward because his head's moving. And that's the first thing he does before he rewards me. Eventually the horse was just looking at people's eyes. So, he would just keep pawing the ground. Like what's five plus five? He'd start pawing the ground. And when he'd see that people were starting to lift their head, he'd slow the pawing the ground down. And then when they turned their head a certain way, he stopped pawing the ground. That is the thing before the thing. That is going to influence your training and it's going to make it less efficient. So, when you are evaluating step one, when you're evaluating the criteria you're seeing, you need to be unemotional and not be moving until you've marked it. Now the gap between seeing it and marking it is important because let's say your dog is doing something and you go, oh, that, that was good. I think that's the one I like. Yeah. I got to get my clicker and then I'm going to, I'm going to click, oh, the clicker's upside down. I'm going to turn it up right side up and then I'm going to click it. Well, that gap, what you're rewarding now is not the behavior. You're rewarding the duration. So, let's assume the dog went into a down and stayed into a down. You are now marking the stay of the down, not the actual, oh good, your elbows hit the ground. I like that. Or what might else have happened, your dog might've gone into a down and while you were fumbling with your clicker, he might've got up out of the down and then gone into a sit. And then you click the wrong thing, making your training inefficient. So, step number one, going through and unemotionally picking out what you want. Step number two, marking it. Step number three is the delivery of the reinforcement. So, you've marked it. Good. I like that. And then you go into your pocket and you pick out a cookie and then it's like, oh, that one's too big. Okay. I'll break that one in half. And then I'll put these ones back in my pocket and then I'm going to give you the cookie, but maybe I got a couple other in my hand and I'm going to pull that away. So, that delivery was crap, right? The time it took to reach into your pocket and to sift through what you have and to break them up because they were too big. So, that took time away from the gap. Remember you're getting very inefficient. And then you deliver the cookie with other cookies in the small of your hand, in the pocket of your hand. So, the dog gets one and sees 10 going away.
11:37
잠깐만요. 나 계산할 줄 아는데. 이건 나한테 최악이야. 그러니 보상을 전달할 때는 빠르고 효율적으로 전달하세요. 쿠키 하나를 집어서 그 하나만 줘야 합니다. 많은 개들이 보상의 가치가 줄어드는 것을 보기 때문입니다. 그리고 당신이 주는 그 작은 쿠키 하나가 오히려 처벌처럼 느껴질 수도 있습니다. '어, 나머지는 다 어디 갔지?' 하면서요. 훈련에 불안을 야기할 수 있죠. 어떤 개들은 그런 상황에 익숙해지기도 하지만, 테리어 종류를 훈련해 본 입장에서 보면 가치를 앗아갈 때 개들은 전혀 좋아하지 않습니다. 그러니 전달이 중요합니다. 보상을 미리 준비하세요. 아마 손에 쥐고 있거나, 저는 훈련할 때 항상 손에 쿠키 하나를 들고 있습니다. 그래서, 쾅, 쾅, 쾅, 쾅 할 수 있죠. 그 간격을 줄여야 합니다. 원하는 것을 보고, 마킹하고, 전달하세요. 쾅, 쾅, 쾅, 쾅. 아주 빠르게. 조금 더 쾅, 쾅 하세요. 개 훈련에 조금 더 쾅, 쾅을 더하세요. 오, 이게 바로 이 팟캐스트의 이름이 되어야겠네요. 에피소드일지도 모르겠네요. 이야기가 샜네요. 네 번째 요소는 보상의 위치입니다. 음, 수잔, 방금 당신이 전달이라고 했잖아요. 같은 거 아닌가요? 오, 전혀 아니에요. 위치 선정은 비법입니다, 여러분. 정말 비법이죠. 자, 예를 들어 여러분의 개가 엎드려 자세를 취하길 원한다고 해보죠. 마커를 찍고, 굿이라고 말한 뒤, 간식을 아주 빠르게 전달하려고 합니다. 그런데 간식을 전달할 때 개가 여러분에게 닿으려고 몸을 일으킨다면요. 저는 단지 만나려는 것뿐이에요. 그래서 여러분이 저에게 쿠키를 주는 걸 돕기 위해 제가 중간쯤 다가가려는 거죠. 그러고 나서 쿠키를 받아먹고 바로 다시 엎드려 자세로 돌아가요. 사실 여러분은 그들이 기준을 어긴 것에 대해 보상한 셈이 된 겁니다. 그래서 강화물의 위치는 여러분의 목표에 기여해야 합니다, 그 목표가 무엇이든 간에요. 자, 복종 훈련을 하시는 분들에게 말씀드리자면, 복종 훈련을 하는 많은 사람들은, 물어오기(retrieve) 훈련을 하는데 개가 아주 구체적인 것을 물어와야 합니다. 나무 조각이나 끝에 방울이 달린 플라스틱 막대 같은 거죠. 이걸 덤벨이라고 부릅니다. 그래서 개가 물어오면, 입에서 물건이 나올 때 보상을 받게 되죠. 그래서 많은 개들이 물고 있기를 원하지 않아요. 자, 제가 개들에게 이 훈련을 시킬 때는, 실제로 덤벨을 물고 있는 동안에 쿠키를 전달합니다. 개들은 입에 쿠키를 물게 되고 그러고 나서 제가 덤벨을 빼내죠. 그건 그 자체로 하나의 과정이라, 이 팟캐스트의 범위를 벗어나긴 하지만요. 제 요점은 강화물의 위치가, 여러분의 궁극적인 목표에 기여하는가 하는 것입니다. 제 궁극적인 목표는 개가 덤벨을 물고 있게 하는 거예요. 저는 개가 입에 덤벨을 물고 있는 상태에서 보상을 줄 겁니다. 도그 어질리티 훈련을 할 때, 여러분은 개가 시소 위에 계속 머물길 원합니다. 개가 시소 위에 있을 때 보상을 주고 있나요? 아니면
Wait a minute. I can do math. That sucked for me. So, when you're delivering the reward, deliver it fast and efficient. One cookie goes in and you give the one cookie because a lot of dogs are going to see the value being taken away. And that little one you give may become more of a punishment than a reward because of, Hey, what about the other ones? It could cause some angst in your training. Now some dogs will learn to live with it, but I know training terriers, they are not happy when you take away the value. So, the delivery is important. Have your reinforcement ready. Maybe have it in your hand or have it in a, I always have one cookie in my hand when I'm training. So, I could be boom, boom, boom, boom. You want to cut down those gaps. See what you want, mark it, deliver it. Boom, boom, boom, boom. Super quick. A little more boom, boom. Get a little more boom, boom into your dog training. Oh, it should be the name of this episode maybe. I digress. The fourth element is the placement. Well, Susan, you just said the delivery. Isn't that the same thing? Oh, nay, nay. The placement is secret sauce, guys. It's secret sauce. So, let's say you want your dog to go into a down. You mark it, you say good, you go to deliver it really quickly. And as you deliver it, your dog kind of comes up to reach you. I'm just going to meet you. So, I'm going to help facilitate you giving me that cookie by meeting you part way. And then they take the cookie and they go right back into the down. You actually have rewarded them for breaking criteria. So, the placement of the reinforcement needs to contribute to your goal, whatever your goal is. So, for you obedience people out there, a lot of people in obedience, we have a retrieve and the dog has to retrieve something very specific. It's like a piece of wood or plastic dowel with bells on the end. It's called a dumbbell. And so, the dog retrieves it, they get rewarded when it comes out of their mouth. So, a lot of dogs don't want to hold it. Now, when I train my dogs to do this, I actually deliver the cookies to them while they're holding it. They actually get cookies in their mouth and then I take it out. That's a process all on its own, not really in the scope of this podcast. But my point is the placement of the reinforcement, does it contribute to your ultimate goal? My ultimate goal is to have my dog hold the dumbbell. I will reward him with his mouth still holding the dumbbell. When you're training in dog agility, you want your dog to stay on the seesaw. Are you rewarding the dog when they are on the seesaw? Or
14:19
오, 착하다라고 말하며 개가 여러분 쪽으로 걸어오게 해서 걸어오는 도중에 쿠키를 주고 있나요? 강화물의 위치가 여러분의 궁극적인 목표에 기여하고 있나요? 정말 중요합니다. 강화물 배치와 관련된 다른 부분은 첫째, 그것이 당신의 목표에 기여하는가? 둘째, 당신과의 근접성이 동일한가? 예를 들어, 당신은 개가 저녁을 먹는 동안 식탁에서 구걸하지 않도록 자기 침대에 가만히 있기를 원합니다. 그래서 당신은 가끔 일어나서 침대에 있는 개에게 간식을 줍니다. 그리고 매일 그렇게 반복합니다. 당신은 개가 침대에 머무는 법을 배우고 있다고 믿겠죠. 행동에 대해 한 가지 알아야 할 점은, 우리가 무엇을 훈련하고 있는지 확실히 아는 것은 오직 개뿐이라는 것입니다. 우리는 우리가 무엇을 훈련하고 있는지 안다고 생각할지 모르지만, 사실 모릅니다. 항상 최종 결정권은 개에게 있습니다. 그러니 어쩌면 개의 마음속에서는 간식을 얻기 위해 당신 곁에 머무는 것이 전부일지도 모릅니다. 그래서 만약 식탁이 너무 멀어지면 개는 당신 쪽으로 더 가까이 다가오기 시작할 것입니다. 어떻게 하면 좋을까요? 예를 들어, 제가 개에게 저로부터 직선으로 멀어지게 달리는 법을 가르치고 싶다면, 보상을 주기 위해 개를 다시 제게로 부르지는 않을 것입니다. 저 멀리 원격 급식기를 사용하거나, 직선으로 달려 나가는 것에 대한 보상으로 장난감을 던져줄 것입니다. 강화물의 배치는 중요하며, 당신과의 근접성도 중요합니다. 이게 어떤 모습일까요? 만약 개가 방을 나갈 때 짖지 않도록 가르치고 싶다면, 간식을 주려고 계속 방으로 다시 들어가서는 안 됩니다. 다른 사람에게 방으로 들어가 간식을 주게 하거나 원격 급식기를 구매하세요. 좋습니다. 이런 것들을 알고 있어야 합니다. 무엇을 평가해야 할지 빠르게 파악하고 결정을 내릴 수 있어야 합니다. 그리고 원하는 행동을 확인하고 결정을 내린 바로 그 순간, 즉시 들어가서 강화하거나 행동을 표시할 수 있어야 합니다, 그게 당신이 하고 있는 훈련이라면 말이죠. 행동을 표시하고 나서 빠르게 효과적으로 강화물을 전달하세요. 강화물을 전달할 때는 강화물의 배치와 그것이 내 최종 목표와 나로부터의 최종 근접성에 어떻게 기여하는지 고려하세요. 마지막으로, 개가 보상을 받은 자세를 유지하고 있는 동안 릴리스 신호를 주어야 합니다. 그것은 강화물을 더욱 강력하게 만드는 것과 같습니다. 제가 여기 Shaped by Dog에서 허락의 힘에 대해 이야기했던 것을 기억하시나요? 그 에피소드에서 저는 우리의 말, 즉 릴리스 신호가 어떤 개들에게는 우리가 훈련에 사용하는 간식보다 더 강력한 강화물이 된다고 말씀드렸습니다. 따라서 가치를 전이하고 음식이나 장난감의 높은 가치를 유지하는 마지막 단계는,
are you saying, oh, good boy. And he walks towards you and gets his cookie walking towards you. Does the placement of the reinforcement contribute to your ultimate goal? Super important. The other part of placement of reinforcement, number one, does it contribute to your goal? Number two, is it the same proximity to you? So, for example, you want your dog to just hang out in his bed while you eat dinner so he's not begging at the table. And every now and again, you get up and you feed him in his bed. And you keep doing that day after day. And you believe the dog's learning to stay in his bed. Here's the thing about behavior. Only the dog really knows for sure what we're training. We may think we know what we're training, but we don't know. The dog always has the last word. So, maybe in the dog's mind, it's all about staying close to you to get cookies. And if the table gets too far away, the dog's going to start moving closer. What could you do? If, for example, I wanted to teach my dog to run in a straight line away from me, I wouldn't call him back to me to reward him. I'd either use a remote feeder out there, or I would throw a toy to reward him for running out in a straight line. The placement of the reinforcement is important and the proximity to you. What does that look like? If you want a dog to learn to not bark when you leave the room, then you can't keep going back to the room to feed them. Either have somebody else go back to the room to feed them or invest in a remote feeder. All right. So, you've got to know those things. Be able to see what you want to evaluate quickly, make the decision. And the moment you see what you want and you've made that decision, you've got to be able to get in there and reinforce or mark the behavior, if that's what you're doing. Mark the behavior and then deliver the reinforcement fast and effectively. And when you deliver the reinforcement, then consider what is the placement of the reinforcement and how does that contribute to my end goal and the end proximity away from me. And finally, you're going to give your release while the dog is still holding the position you rewarded. That is like supercharging the supercharge. Remember, I talked about the power of permissions here right on Shaped by Dog. In that episode, I said our words, our releases are as reinforcing into some dogs, more reinforcing than the cookies we are using in our dog training. And so, the final piece to this transfer of value in maintaining that high value of your food or your
17:15
장난감을 사용하는 경우에도 그 가치를 유지하면서, 마지막으로 중요한 점은 릴리스 단어를 말할 때 개가 설정한 기준을 계속 유지하도록 하는 것입니다. 물론, 여러분은 항상, 언제나 릴리스 단어를 사용해야 합니다. 개에게 엎드려라고 한 뒤 바로 다른 일을 하러 가면 개에게 혼란을 줄 수 있기 때문입니다. 항상 그래야 합니다. 개에게 앉아, 엎드려, 서 등의 자세 신호를 주거나, 어질리티에서 타겟을 수행하고 끝에서 멈추거나 출발선에서 기다리게 할 때, 무엇을 하든, 개에게 통제된 자세를 유지하도록 요구하는 모든 것 뒤에는 반드시 릴리스 단어가 따라와야 합니다. 제가 제안하는 것은 릴리스 단어를 주기 전에 개가 기준을 잘 유지하고 있는지 확인하라는 것입니다. 예를 들어 개가 앉아 있는 상태에서 앞으로 몸을 기울이기 시작해 엉덩이가 바닥에서 조금씩, 아주 조금씩 들리기 시작한다면, 개가 그런 행동을 할 때 릴리스 단어를 준다면, 도대체 무엇을 위해 릴리스를 하는 걸까요? 맞습니다. 차를 운전하며 듣고 계신 여러분의 대답이 들리는 것 같네요. 엉덩이를 바닥에 붙이고 있어야 한다는 기준을 어기는 행동에 대해 릴리스를 하고 있는 셈이 되는 거죠. 이제 막 새 반려견을 맞이할 준비를 하는 분들에게 너무 어렵게 느껴지지 않았기를 바랍니다. 이러한 기초적인 행동 훈련에서 우리 전문가들만큼 성공적인 결과를 내는 것은 매우 쉬운 일입니다. 반려견 훈련사 여러분, 여러분이 해야 할 일은 그저 사용 중인 음식이나 장난감의 가치를 인식하는 것뿐입니다. 그리고 그 가치를 유지하고, 더 나아가 다음 다섯 가지 사항을 염두에 둠으로써 훈련 효과를 극대화하고 계신가요? 보상할 행동을 확인하고, 마킹하고, 보상을 전달하고, 보상의 위치를 정하고, 마지막으로 릴리즈 신호를 주는 것입니다. 이 팟캐스트가 도움이 되셨다면 부탁 하나만 드려도 될까요? 강아지를 사랑하는 다른 사람 한 명 이상에게 공유해 주실 수 있나요? 여러분의 소셜 미디어 페이지에 공유해 주신다면 제가 정말 감사할 것 같습니다. 제 목표는 전 세계의 반려견 보호자들이 자신의 반려견을 더 잘 이해하도록 돕고, 전 세계의 반려견들이 가능한 한 최고의 삶을 살 수 있도록 돕는 것이기 때문입니다. 그 목표에 동참해 주셔서 미리 감사드립니다. 그럼 다음 Shaped by Dog 시간에 뵙겠습니다.
toys, if you're using toys, maintaining that value, the final piece is making sure that the dog is maintaining the criteria you've established when you deliver the release word. And of course, you are always, always giving a release word. So, you're not going to tell your dog down and then go to work because then you're going to confuse your dog. Always. If we give our dogs a positional cue, sit down, stand, or in agility doing a target and stopping at the bottom of something or waiting at a start line, whatever we do, anything that requires a dog to hold a control position must always be followed up with a release word. And in what I'm suggesting is make sure the dog is maintaining criteria before you give that release word. So, if you want your dog to sit and he starts leaning forward and it turns into like his butt just starts lifting off the ground a little bit and a little bit and a little bit, you actually, when you, if you gave a release word when the dog was doing that, you would be releasing them for what? That's right. I heard you say that while you're driving in your car. You would be releasing the dog for breaking your criteria of keeping your butt on the ground. Now, I hope this wasn't too overwhelming for those of you who are just getting ready to get a new dog. It's super easy to be as successful in these foundational behaviors as any one of us professional dog trainers. All that you need to do is recognize the value of the food or the toy that you're using. And are you maintaining that value and ideally supercharging your training by being mindful of those five points? See what you're going to reward, mark it, deliver it, placement of reinforcement, and give a release. Hey, if you're finding value in this podcast, would you do me a favor? Would you share it with one other dog loving person or maybe more? Share it on your social media pages? I would forever be indebted to you because my goal is to help dog owners worldwide better understand their dogs and to help dogs worldwide to have the best life possible. And I thank you in advance for contributing to that goal. I'll see you next time on Shaped by Dog.