How the Best Professional Dog Trainers Use Reinforcement #94
Susan GarrettShaped by Dog with Susan Garrett · 팟캐스트 · 18분
반려견과의 파트너십은 보호자가 제공하는 강화물의 체계적인 관리와 연결을 통해 구축된다. 엘리트 수준의 훈련에서 개가 보상 없이 작업하는 것처럼 보이는 이유는 작업 자체를 가장 높은 가치의 강화물로 인식하기 때문이며, 이는 보호자가 사전에 차곡차곡 쌓아 올린 강화의 결과이다. 보호자는 관심, 음식, 기존의 신호, 장난감, 그리고 개가 본능적으로 좋아하는 활동이라는 다섯 가지 범주의 강화물을 개에게 제공한다. 보호자를 거치지 않고 개가 스스로 강화를 획득하는 상황은 거래적 관계를 형성하여 통제력을 약화시킨다. 따라서 모든 강화물은 보호자의 허락을 통해 전달되어야 하며, 낮은 가치의 강화물을 높은 가치의 강화물과 연결함으로써 강화물의 계층 구조를 확장해야 한다. 개가 선호하는 보상의 위계를 파악하여, 낮은 보상도 높은 보상으로 이어지는 매개체로 활용할 때 진정한 파트너십이 완성된다.
최고의 전문 개 훈련사들이 보상을 사용하는 방법 #94
How the Best Professional Dog Trainers Use Reinforcement #94
0:00
경찰견이나 세계 수준의 어질리티 챔피언 독, 서비스 독처럼 아주 훈련이 잘 된 개들이 핸들러로부터 음식이나 장난감 보상을 전혀 받지 않고도 일하는 것처럼 보이는 모습을 본 적 있나요? 특히 스스로 시도해 보았을 때 반려견이 '아니요, 사양할게요'라며 거부하는 상황을 겪어보셨다면요. 만약 그렇다면 오늘 팟캐스트가 아주 마음에 드실 겁니다. 안녕하세요, 수잔 개럿입니다. Shaped by Dog에 오신 것을 환영합니다. 오늘 저는 왜 일부 엘리트견들이 그렇게 훌륭하게 일하는데도 강화물을 받지 않는 것처럼 보이는지에 대한 미스터리에 대해 이야기하려고 합니다. 그 미스터리가 풀립니다. 이번 내용은 이제 막 강아지를 입양했거나 유기견을 새로 데려오신 분들에게 매우 중요합니다. 그리고 반려견과 함께 해결하기 어려운 문제로 고군분투하시는 분들에게도 매우 중요할 것입니다. 특히 반려견의 행동을 이끌어내기 위해 음식이나 장난감 유인물을 사용해야 한다고 배워오신 모든 분들께 말씀드리고 싶습니다. 오늘 팟캐스트에 특별히 더 집중해 주셨으면 합니다. 이번 에피소드는 많은 분들에게 전환점이 될 것이라고 생각합니다. 어떻게 이런 내용을 다루게 되었는지 말씀드리죠. 최근 제가 소셜 미디어에 글을 하나 올렸는데, 세계 최고의 훈련사들을 위대하게 만드는 것은 그들이 사용하는 음식 보상 자체가 아니라는 내용이었습니다. 그것은 바로 그들이 음식 보상을 어떻게 사용하는가에 대한 것입니다. 그랬더니 많은 분들이 댓글을 달아주셨습니다. 세계 최고의 훈련사라면 음식을 사용해서는 안 된다는 것이었죠. 마치 세계 최고의 훈련사들은 음식이나 장난감을 전혀 쓰지 않는다는 것처럼 말입니다. 경찰견이나 세계 수준의 어질리티 챔피언 독, 서비스 독처럼 아주 훈련이 잘 된 개들이 핸들러로부터 음식이나 장난감 보상을 전혀 받지 않고도 일하는 것처럼 보이는 모습을 본 적 있나요? 특히 스스로 시도해 보았을 때 반려견이 '아니요, 사양할게요'라며 거부하는 상황을 겪어보셨다면요. 만약 그렇다면 오늘 팟캐스트가 아주 마음에 드실 겁니다. 안녕하세요, 수잔 개럿입니다. Shaped by Dog에 오신 것을 환영합니다. 오늘 저는 왜 일부 엘리트견들이 그렇게 훌륭하게 일하는데도 강화물을 받지 않는 것처럼 보이는지에 대한 미스터리에 대해 이야기하려고 합니다. 그 미스터리가 풀립니다. 대부분의 경우 이런 사람들은 그저 트롤링을 하려는 사람들이었다는 걸 알고 있습니다. 어떻게든 소란을 피워보려는 거였죠. 하지만 우려스러웠던 건 그 게시물들 중 일부가 실제로 30명에서 50명 정도의 사람들로부터 좋아요나 추천을 받았다는 점입니다. 그래서 생각했죠. 와, 강화가 어떻게 사용되는지에 대해 오해가 좀 있구나. 분명히 말씀드리자면, 모든 강화 훈련사들이 다 같은 건 아니거든요. 강화 기반 훈련사라고 해서 모두 같은 방식으로 음식을 사용하는 것도 아닙니다. 오늘 팟캐스트에서 무엇이 다른지 알려드리려고 합니다. 그 게시물에 대한 대응으로 제 친구를 초대하기로 했습니다. 시애틀 경찰청에서 46년간 근무한 베테랑이자 훈련 시 확실하게 음식과 장난감을 사용하는 놀라운 반려견 훈련사입니다. 그는 자신의 훈련 방식에 음식과 장난감을 활용하죠. 스티브 화이트를 초대해 엘리트 수준의 훈련에서 음식을 사용하는 것에 대해 이야기를 나누기로 했습니다. 엘리트 수준의 훈련 말이죠. 쇼 노트에 링크를 공유해 드릴 테니 스티브와 제가 나눈 대화를 들어보시기 바랍니다. 스티브와 제가 나눈 대화 말입니다. 정말 흥미로운 내용이 될 거라고 생각합니다.
Have you ever watched a really well-trained dog, like a police dog, a high-level world champion agility dog, a service dog, that they seem to be working without getting any kind of food or toy reward from their handler? And especially if this is something that you've tried on your own and your dog says, no, thank you, I'm opting out. Well, if you have, you are going to love today's podcast. Hi, I'm Susan Garrett. Welcome to Shaped by Dog. Today, I'm going to be talking about that mystery, that why it seems that some of the elite dogs work so brilliantly and it doesn't look like they're getting reinforcement. The mystery is solved. This is going to be critically important to those of you who just have a new puppy or a new rescue dog, of course. It's going to be critically important for those of you that are struggling with some challenge with your own dog that you just can't seem to fix. And I really want all of you who have been taught that you should be using food lures or toy lures to create behavior with your dog. I would like all of you to pay particular attention to today's podcast. I think this is going to be a pivotal podcast for many people. And let me tell you how it came about. Recently, I put up a post on social media where I shared that it isn't the food rewards that the best dog trainers in the world use that make them so great. It's how they use their food rewards. And there was a number of comments from people who said, well, the best dog trainers in the world shouldn't use food. Like the best dog trainers in the world don't use food or toys. And I recognize for the most part, these were just trolls that were, you know, looking to try and stir some action. But what concerned me was that some of these posts actually got 30 or 50 people giving them a like or a thumbs up. And I thought, wow, there is some confusion here in how reinforcement is used. Because let me tell you, all reinforcement trainers are not the same. All reinforcement-based trainers do not use food in the same way. I'm going to share how things look different in today's podcast. Now, in response to that post, I decided I would invite a friend of mine, 46-year veteran of the Seattle Police Force and a phenomenal dog trainer who definitely uses food and toys in his training. I'd invite Steve White to join me for a discussion on the use of food in training at the elite level. And I'm going to share a link in the show notes so that you can listen into that conversation that Steve and I had. I think you're going to find it a really interesting one.
2:58
우리 모든 엘리트 반려견 훈련사들이 공통적으로 이해하는 사실은, 일단 개들에게 '작업' 자체가 가장 큰 강화물이라는 것을 납득시키고 나면, 음식이나 장난감 없이도 개들이 즐겁게 일한다는 것입니다. 모든 품종의 모든 개들에게 가능한 일일까요? 저는 가능할 수도 있다고 믿지만, 어떤 개들에게는 더 쉽다는 것도 알고 있습니다. 왜 그런지에 대해서는 오늘 나중에 말씀드리겠습니다. 하지만 많은 개들에게 작업은 엄청난 가치를 지니게 되는데, 이는 강화가 차곡차곡 쌓여왔기 때문입니다. 분명히 말씀드리자면, 여러분이 엘리트 수준의 개를 보면서 '와, 음식이나 장난감 없이도 일을 하네'라고 생각한다면 그것은 착각입니다. 왜냐하면 결국 개들은 핸들러로부터 어떤 형태로든 강화를 받게 되기 때문입니다. 그 과정은 길어질 수 있습니다. 30초 간격으로 일어나는 일이 아닐 수도 있습니다. 그들은 더 오래 작업하도록 훈련받았지만, 여전히 수행하는 작업으로부터 많은 강화물을 얻고 있습니다. 그들이 하고 있는 작업에서 말이죠. 왜냐하면 엘리트 반려견 훈련사들은 강화물을 쌓는 법을 이해하고 있기 때문입니다. 그래서 개에게 줄 수 있는 가장 큰 보상이 무엇인지(개마다 다를 수 있죠) 그 가치를 개에게 줄 수 있는 가장 낮은 수준의 보상으로 연결할 수 있는 것입니다. 그 가치가 개에게 줄 수 있는 가장 낮은 수준의 강화로 이어질 수 있습니다. 이에 대해서는 팟캐스트 16화에서 '일이 일어나기 전의 일'이라는 주제로 말씀드렸습니다. 그 '전의 일'에 대해서요. 만약 개가 카운터 위로 올라갔을 때 내려오라고 하고, 개가 내려왔을 때 '잘했어'라고 하며 간식을 준다면, 당신은 실제로 개에게 착한 행동을 하려면 먼저 나쁜 행동을 해야 한다고 가르치고 있는 셈입니다. 착해지기 위해 나빠져야 한다는 것을요. 그것은 강화물을 쌓는 것이지만, 정말로 원치 않는 방식으로 쌓는 것입니다. 그런 방식은 피해야 합니다. 또한 80화도 확인해보시면 좋습니다. 그곳에서 여러 관계 유형에 대해 이야기했는데, 오늘 제가 하려는 이야기가 바로 그것이기 때문입니다. 음식 때문에 일하는 것처럼 보이는 개들과, 보호자와 함께하는 것을 좋아하는 마음 자체로 일하는 것처럼 보이는 개들, 그리고 오직 음식이나 장난감 때문에 일하는 개들 사이의 차이점 말이죠. 그 차이는 정말로 당신과 개 사이의 관계에 달려 있습니다. 그 관계는 당신과 개 사이의 관계에 관한 것입니다. 그리고 관계는 우리가 주고받는 강화를 통해 구축됩니다. 부모와의 관계든, 자녀와의 관계, 파트너나 배우자와의 관계, 혹은 직장 동료나 상사와의 관계에서도 마찬가지입니다. 모든 관계는 강화 또는 그 부족함을 통해 구축됩니다. 실화입니다. 자, 이제 당신의 반려견들에 대해 이야기해 봅시다. 제가 반려견에게 사용하는 강화의 유형은 다섯 가지가 있습니다. 반려견을 강화하기 위해 사용하는 다양한 도구들을 늘려 나가는 것을 고려하는 것이 정말 중요합니다. 한 가지에만 매몰되어서는 안 됩니다. 한 가지에만 고정되지 않는 것이 좋습니다. 다섯 가지 유형입니다. 첫 번째는 관심입니다. 그것은 눈맞춤일 수도 있습니다. 제가 반려견을 쳐다보는 것만으로도 그들의 관심을 끌 수 있습니다.
What all of us elite dog trainers understand is that dogs will work happily without the use of food or toys once you've been able to convince them that the work is the most reinforcing thing. Is that possible with all dogs of all breeds? I believe it may be, but I do believe it's easier for some dogs. And I will share with you why later on today. For many dogs though, the work becomes massively valuable because the reinforcement has been stacked. And let me tell you, when you look at an elite level dog and say, wow, they're working without food or toys, it's an illusion because what's happened is they will eventually get some reinforcement from their handler. It may be drawn out. It may not be happening, you know, in 30 second intervals. They've been trained to work for longer, but they still are getting a lot of reinforcement from the work they're doing. Because elite dog trainers understand that you stack reinforcement so that the highest reinforcement that you could ever give a dog, and that could be different for any dog, that value can go into the lowest level reinforcement that you can give your dog. I talked about this in podcast episode number 16, where I talked about the thing before the thing. So, if your dog gets on the counter and you tell them off and they get off and you say, good off and give them a cookie, you actually are teaching the dog you have to be bad in order to be good. That's stacking reinforcement, but that's stacking it in a way you really shouldn't want to be doing it. Also, you might want to check out episode number 80, where I talk about the different types of a relationship, because that is really what I'm talking about today. And the difference between dogs who seemingly work for food and dogs who seemingly work just for the love of what they're doing with their owner versus dogs who only work for food or toys. That really is about the relationship that you have with your dog. And the relationship is built through the reinforcement we have, whether it's a relationship with a parent, relationship with a child, your relationship with your partner, your spouse, relationship with your coworkers or your boss. All relationships are built through reinforcement or lack thereof. True story. So, let's talk about your dogs. There are five categories of reinforcement that I use with my dog. And it's really important that you start thinking about growing the different things you use to reinforce your dog. You don't want to get stuck on just one thing. Five categories. Number one, attention. That could be a look. I could look at my dog and that will get their attention.
5:44
반려견은 즉시 제가 자신들과 상호작용하고 있다는 것을 알게 됩니다. 제 호흡 방식일 수도 있습니다. 제가 깊게 숨을 들이마시면, 반려견들은 '오, 뭔가 시작되려나 보다'라고 알아차리는 신호가 됩니다. 그것은 제가 반려견에게 관심을 주는 방법입니다. 관심은 조금 더 명확한 행동일 수도 있습니다. 쓰다듬거나 칭찬하는 것일 수 있는데, 이 둘은 별개의 것이죠. 하지만 여전히 관심이라는 범주에 속합니다. 관심은 야단치는 것일 수도 있습니다. 그래서 짖어서 야단을 맞는 개들은 종종 그 야단이 짖는 행동을 강화하는 원인이 되기도 합니다. 짖는 행동을 강화하는 것이죠. 결국 관심 때문에 더 짖게 되는 것입니다. 때로는 스스로 무엇을 하고 있는지 잘 모를 때가 있습니다. 자, 첫 번째 범주는 관심입니다. 두 번째 범주는 음식입니다. 분명한 것들이 있죠. 훈련 시 사용할 수 있는 다양한 간식들입니다. 매일 반려견에게 급여하는 사료도 있습니다. 어떻게 급여하고 계신가요? 음식을 줄 때 반려견은 무엇을 하고 있나요? 주위를 맴돌며 짖거나, 몸을 부딪치고 있지는 않나요? 서로 먼저 먹겠다고 으르렁거리고 있지는 않나요? 음식을 주는 행위 자체가 음식을 받을 때 반려견이 하는 모든 행동을 강화하는 것입니다. 세 번째 범주는 반려견이 이미 교육을 통해 알고 있는 신호입니다. 강화입니다. 예를 들어, 반려견에게 오라고 해서 달려왔는데 그 뒤에 앉으라고 지시한다면, 달려오는 행동이 앉기라는 행동에 의해 보상받는 것입니다. 물론 앉기 행동은 장난감이나 칭찬, 혹은 음식으로 보상받을 수 있지만, 그것은 반려견이 달려오기로 한 결정 또한 강화하게 됩니다. 큐(신호)는 강력한 강화제입니다. 우리는 최고 수준의 스포츠에서도 항상 이를 사용합니다. 그렇기 때문에 반려견이 나쁜 행동을 할 때 '내려와'라고 말하면, 사실 그 큐를 준 순간 반려견이 하고 있던 행동을 보상해 주는 셈이 됩니다. 자, 첫 번째는 주의 집중입니다. 두 번째는 음식입니다. 세 번째는 강화를 통해 구축해 온 잘 알려진 큐입니다. 네 번째 범주는 장난감입니다. 프리스비, 공, 혹은 가장 좋아하는 회수용 장난감처럼 던져주는 장난감이 있을 수 있습니다. 터그 놀이처럼 상호작용하거나 가볍게 던져주는 방식의 장난감도 있죠. 그것들은 상호작용형 장난감입니다. 또한 강아지들이 가지고 혼자 멀리 달려가서 노는 장난감들도 있습니다. 저는 이 장난감들에 별표를 표시할 텐데, 제 분류법상 실제로는 다섯 번째 범주에 속하기 때문입니다. 그것은 바로 반려견이 강화물로 여기는 활동들입니다. 산책하기, 차 타기, 수영하기 같은 당연한 활동들부터, 다람쥐 쫓기, 다른 개 쫓기, 혹은 옆집 개와 울타리를 사이에 두고 짖거나 공격적인 행동을 하는 것 등이 해당합니다.
They will instantly know I'm engaging with them. It could be the way I breathe. If I take a deep breath, like that is a trigger for my dogs to know, ooh, something's on. That is me giving them my attention. Attention could be something more obvious. It could be your patting or your praise, which are two separate things. They still fall under the category of attention. Attention could also be scolding. So, dogs who get scolded for barking, quite often that scolding is what's reinforcing the barking. So, you get more attention barking. Sometimes you're not quite aware of what you're doing. So, category number one, attention. Category number two is food. So, there's the obvious, all the different training treats that you could use. There's the food you deliver to give your dog their meals every day. How are you delivering it? What is that dog doing when you're delivering that food? Are they running around circles around you, barking, bouncing off of you? Are they growling at each other because they want their food first? The delivery of the food is rewarding all the things that are going on when you deliver that food. So, the third category are cues that your dog knows because you've built them through reinforcement. So, for example, if you ask your dog to come and they come running and then you ask the dog to sit, the behavior of coming running gets rewarded by the sit. Now, the sit could be rewarded with a toy or your praise or food, but that also reinforces the decision of the dog to come running. Cues are powerful reinforcers. There are things that we use at the highest level of sport all the time, which is why when your dog's doing something naughty and you say off, you actually are rewarding them for what they're doing at the time you've given them that cue. So, we've got attention. We've got food. We've got known cues that you've built through reinforcement. The fourth category, toys. Now, there could be toys that you throw like a flying disc or a ball or a favorite retrieve toy. They could be toys that you interact with in a way of like a game of tug or a quick toss. So, those are interactive toys. Now, there's also toys that dogs are given that they run off and play with by themselves. I'm going to put an asterisk beside those toys because they really belong in my books in the fifth category, and that would be activities that your dog finds reinforcing. Activities like the obvious things like going for a walk or going for a car ride, going for a swim, chasing squirrels, chasing other dogs, barking or aggressing at the fence with the other next-door neighbor's dog.
8:32
스포츠를 예로 들면, 어질리티의 다양한 장애물들은 제 반려견들에게 강력한 강화제입니다. 많은 개들에게 있어 터널을 통과하는 것은 어질리티에서 가장 강력한 강화물이 되는 장애물입니다. 제 반려견인 모멘텀의 경우, 시소에 아주 열광하며 이는 큰 강화제가 됩니다. IPO나 경찰견 훈련 같은 활동을 하는 개들에게는 무는 행동, 즉 소매를 무는 것이 엄청난 강화제가 됩니다. 그래서 그것은 아마도 그런 스포츠를 하는 사람들이 개에게 줄 수 있는 최고의 강화 수단일 것입니다. 이제 이러한 모든 활동들은 여러분이 개에게 허락을 구하도록 가르치기 시작하면, 여러분을 위해 작용하기 시작합니다. 여러분의 허락 없이 이루어지는 모든 활동은 손실된 강화일 뿐입니다. 우리가 개에게 주는 모든 것은 우리를 거쳐서 전달되어야 합니다. 우리를 거치지 않게 되면, 결국 제가 '거래적 관계'라고 부르는 상황이 발생합니다. 거래적 관계는 인간과 인간 사이, 그리고 인간과 동물 사이에서도 일어날 수 있습니다. 예를 들어, 부모가 '침대 정리하면 25센트를 줄게'라고 말할 수 있죠. 혹은 아이가 '오늘 밤 아이패드 스크린 타임을 더 주면 침대를 정리할게요'라고 말하는 식입니다. 그게 바로 거래적인 관계입니다. 흔히 반려견 훈련을 처음 시작하는 사람들은 애견 훈련소를 찾아가고, 그곳에서 개와의 관계를 거래적으로 만드는 법을 배웁니다. 가장 먼저 배우는 것은 개가 앉거나, 엎드리거나, 따라오거나, 옆에서 걷게 하기 위해 개 코앞에 간식을 대는 방법입니다. 즉, 개에게 '간식이 보일 때만 내 말을 들어'라고 가르치는 셈이죠. 그리고 여기서 사람들이 오해를 합니다. 내가 가진 간식은 고양이를 쫓고 싶은 욕구만큼의 가치가 없는데, 어떻게 하면 개가 고양이를 쫓지 않게 할 수 있을까요? 글쎄요, 그것은 거래적인 사고방식입니다. 그리고 여러분의 생각이 맞습니다. 거래적인 관계를 맺고 있다면, 어떤 종류의 처벌을 사용하지 않고서는 개가 고양이를 쫓지 않게 만들 방법은 없습니다. 하지만 굳이 거래적인 관계를 가질 필요는 없습니다. 여러분도 제가 가진 것처럼 개와 동반자 관계를 맺을 수 있습니다. 그것은 개를 집으로 데려오는 즉시 시작됩니다. 저는 강아지가 내리는 결정이나 선택에 따라 5가지 강화 범주가 저를 통해 어떻게 전달되는지 공유하기 시작합니다. 자, 제가 제 강아지를 칭찬하게 될까요? 안타깝게도 많은 분들이 이렇게 합니다. 강아지가 장난감을 물고 반대 방향으로 달려갈 때 말이죠. 이리 와, 착하지, 잘했어. 그거 가져와, 가져와, 착하지, 잘했어. 다시는 반복해서 보고 싶지 않은 행동을 칭찬으로 강화하고 있는 것입니다. 정말 재미있는 건, 그 말투와 아주 높은 톤의 칭찬 때문에 제 강아지가 지금 막 위층으로 올라왔다는 점이죠, 그렇죠? 그건 매우 강력한 강화제입니다.
Playing sports, agility is the different obstacles in dog agility are powerful reinforcers for my dogs. For many dogs, a tunnel, going into a tunnel is the most reinforcing obstacle in agility. For my dog, Momentum, she's crazy about a seesaw, a big reinforcer. For dogs that do things like IPO or police work, the opportunity to bite, bite a sleeve is massively reinforcing. And so, that is probably the number one reinforcing thing that people that do those sports could ever give their dog. Now, all of these activities, they start to work for you as a dog owner once you start to put permission before them. All of those activities without permission from you are reinforcement that's lost. Everything we give our dogs, we need to have it to come through us. When it doesn't come through us, what ends up happening is what I call a transactional relationship. A transactional relationship could happen between humans and humans and humans and animals. So, you know, a parent could say, make your bed and I'll give you a quarter. Or the child says, I'll make my bed if I get extra screen time on my iPad tonight. That's transactional. So often, people starting out their dog training career will go to a school who will teach them to be transactional for their dog. The very first things you'll be taught is how to put a cookie on your dog's nose to get them to sit or to down or to come with you or to walk beside you. So, you're teaching the dog, do things for me when you see the big cookie. And that's where people misunderstand. How can I get my dog to not chase a cat because the cookie I've got doesn't have the value of the cat? Well, that's transactional thinking. And you're right. You will never get a dog to not chase the cat without using punishment of some kind if you have a transactional relationship. But you don't need to have a transactional relationship. You could have what I have and that's a partnership with my dogs. It starts as soon as I get them home. I start to share with them how the category of five reinforcements come through me based on decisions or choices those puppies make. So, am I going to praise my puppy? Unfortunately, a lot of people do this. When the puppy is running in the opposite direction with a toy. Come on, good baby, good baby. Bring me that, bring me that, good baby, good baby. Reinforcing through your praise something you don't ever want to see repeated. What's really funny is that talking, that very high-pitched praise just caused my puppy to just come upstairs, right? It's very reinforcing.
11:37
칭찬이든, 관심이든, 주는 사료든, 주는 신호든, 선택해서 사용하는 장난감이든, 물론 허락하는 활동이든 간에, 여러분이 반려견에게 제공하는 모든 강화에 대해 생각해야 합니다. 다른 개와 노는 것은 강력한 강화제입니다. 저는 강아지에게 '좋아, 가서 놀아'라고 말하기 전에 앉아, 핸드 타겟, 또는 엎드려를 시킵니다. 그건 엄청난 강화제이기 때문이죠. 왜 그냥 공짜로 주나요? 강화 효과를 낭비하는 셈입니다. 우리는 그 모든 강화가 여러분을 통해 이루어지게 만들어야 합니다. 그것이 바로 강화 기반 훈련이 작동하는 방식입니다. 그럴 때 반려견은 여러분이 요청했기 때문에 고양이를 쫓는 것을 멈추게 됩니다. 반려견이 '음, 주인님이 미트볼을 가지고 있네'라고 생각해서가 아닙니다. 저는 고양이를 쫓아갈 거예요. 그냥 한번 생각해보세요. 여러분이 소파에 앉아 있는데 반려견을 보고 말을 걸고 쓰다듬으며 칭찬하고 있다고 상상해보세요. 그리고 만약 개를 여러 마리 키운다면, 저희 집처럼 한 마리에게 말을 걸고 쓰다듬으며 칭찬하기 시작하자마자 다른 개가 '오, 그래, 내 차례야, 나야 나'라고 생각하며 달려오겠죠. 안 돼, 안 돼, 쟤 쓰다듬지 마. 나를 쓰다듬어, 나야 나. 집에도 그런 상황이 있나요? 네. 왜냐하면 녀석들은 모두 당신의 관심을 원하기 때문이죠. 당신은 관심을 나눠주고 있고요. 저도 그 관심을 조금 가져갈게요. 이제 제가 모두에게 관심을 주고 있을 때, 뒷방에서 킴이 건조대에서 강아지 식기를 꺼내 카운터에 올려놓는다고 상상해보세요. 사료를 줄 준비를 하기 위해서죠. 저한테 '엄마, 관심을 주세요'라고 하던 청중들이 어떻게 제로가 되는지 상상해보세요. 모두가 '오, 음, 평소엔 당신의 관심이 최고야'라고 생각하죠. 하지만 지금은 더 좋은 강화물이 곧 나타날 상황인 거예요. 킴이 사료를 나눠주면서 아침이나 점심, 혹은 저녁을 주려 할 때(참고로 녀석들은 점심을 먹지 않지만요). 만약 제가 문으로 가서 '이봐, 수영하러 갈 사람 누구야?'라고 말한다면 어떨까요? 어떻게 될까요? 제 개 다섯 마리 중 네 마리는 '아침은 나중에 먹을게요'라고 할 겁니다. 우리는 수영하러 갈 거예요. 자, 불독인 테이터 샐러드의 경우라면 '이봐, 다람쥐나 칩멍크 쫓을 사람 누구야?' 같은 말을 해야 할 겁니다. 그러면 녀석도 그 자리를 떠나게 만들 수 있겠죠. 제가 무슨 말을 하는지 아시겠나요? 개들에게는 보상에 대한 계층, 즉 자연스러운 서열이 존재합니다. 우리는 반려견의 보상을 확장하고 그 모든 보상이 우리를 통해 전달되도록 해야 합니다. 그래야 원래는 더 가치가 높은 보상이 있었더라도, 당신과 함께 무언가를 할 수 있는 기회가 그 모든 것을 능가하게 됩니다. 왜냐하면 그 모든 것이 당신을 통해 오기 때문이죠. 그렇다면 이 파트너십은 어떻게 작동할까요? 첫째, 무엇이 언제 당신의 반려견을 강화하는지 알아야 합니다. 무엇이 무엇을 능가하는지 말이죠. 그리고 그것을 바꾸기 위해 노력해야 합니다.
You need to think about all of the reinforcement you give your dogs, whether it's praise or the attention, the food you give them, the cues you're giving them, the toys you decide to use, and of course the activities that you allow them to engage with. Playing with another dog is a high reinforcer. I'll ask my puppy to sit or hand target or down before I say, all right, go play. Because that's a massive reinforcer. Why just give it away? It's lost reinforcement. We need to make all of that reinforcement come through you. And that's how reinforcement-based training works. That's when a dog will stop chasing a cat because you've asked them. And it's not because they're saying, well, she's got a meatball. I think I'm going to chase this cat. Because just think about it. Imagine that you're sitting on your couch and you see your dog and you start talking to them and patting them and you're praising them. And then if you have more than one dog, if it's anything like my house, as soon as I start talking and patting and praising one dog, then another dog goes, oh, yeah, I think it's me, me, me, me. No, no, don't pat them. Pat me, me, me. Do you have that at your home? Yeah. Because they all want your attention. You're doling out attention. I'll take a little bit of that. Now, while I'm giving everyone attention, imagine if in the back room, Kim takes the dog dishes from the draining rack and puts them on the counter because she's going to start dishing out food. Imagine how my audience of I want your attention, mama, goes to zero. Everyone goes, oh, well, your attention's awesome most of the time. But right now, there's a better reinforcer about to happen. And as Kim's doling out that food and about to give them their breakfast or lunch or dinner, they don't get lunch. What if I went to the door and said, hey, who wants to go swimming? Guess what? Four of my five dogs would say, we'll get breakfast later. We are going for a swim. Now, with tater salad, the bulldog, I would have to say something like, hey, who wants to chase this squirrel or chipmunk? That would get him to leave. Do you see what I'm saying? There's a hierarchy, a natural hierarchy for reinforcement with our dogs. We need to make sure that we expand our dog's reinforcement, have it all come through us so that even though there originally might have been a high value reward, the chance to work with you trumps all of it because it all has come through you. So, how does this partnership work? Number one, you need to know what reinforces your dog and when. What trumps something? And you need to work to change that.
14:19
그러니 첫 번째로, 다섯 가지 범주에서 무엇이 반려견을 강화하는지 파악하고 각각의 범주를 키워나가세요. 단순히 '아, 우리 개는 음식을 정말 좋아해'라고만 하지 마세요. 녀석은 그냥 음식 때문에 일할 뿐일 겁니다. 아니면, 이 방송을 듣고 계신 분들 중에 "우리 개는 오직 물기(bite)를 위해서만 일해요"라고 말하는 아주 뛰어난 엘리트 개를 키우시는 분들에게도 말씀드리고 싶네요. 그걸 키워나가야 합니다. 물기 훈련을 하러 가도록 허락하기 전에 이 터그 놀잇감으로 먼저 터그 놀이를 해야 한다고 가르쳐야 합니다. 그렇게 해야 터그 놀이의 가치를 높일 수 있습니다. 제 보더콜리들은 일단 어질리티를 시작하면 간식이나 장난감을 전혀 받지 않아도 행복해합니다. 그저 어질리티를 할 기회만을 원하거든요. 하지만 제가 강화 기반의 훈련사라면, 어질리티 자체가 강화물인 상황에서 어질리티 중에 발생하는 문제를 어떻게 해결할 수 있을까요? 강화의 위계를 높게 유지해야 합니다. 그러니 시소 훈련을 하고 싶다면, 먼저 이 터그 놀잇감으로 놀이를 해야 합니다. 개들이 간식은 안 먹으려 하겠지만 터그 놀이는 할 수 있다는 걸 알고 있습니다. 때로는 터그 놀이를 하기 위해 시소로부터 더 멀리 떨어져야 할 수도 있겠지만, 결국 해냅니다. 그런 다음 같은 방식으로 간식을 활용해 훈련합니다. 우리의 목표는 개의 강화 피라미드를 뒤집는 것입니다. 개가 쫓는 가치가 매우 높은 보상을 한두 가지만 두는 대신, 최상단에는 다수의 고가치 보상을 두고 최하단에는 가치가 낮은 보상을 아주 조금만 남기는 것입니다. 그리고 강화의 위계를 쌓기 시작하세요. 그러니 여러분의 개가 어떤 음식을 좋아하는지 파악하세요. 개가 언제 어떤 음식을 먹지 않는지 파악하고 그것부터 쌓아가기 시작하세요. 예를 들어, 몇 주 전에 저는 테이터 샐러드를 데리고 산책을 나갔고 킴도 함께였습니다. 저는 이런 게임들을 합니다. 제가 그 개와 하는 게임 몇 가지를 그녀에게 보여주고 있었죠. 한번은 리콜(부르기)을 했습니다. 녀석은 간식을 기대하며 전속력으로 돌아왔는데, 제가 대신 터그 놀잇감을 주었습니다. 그러자 녀석은 '어, 이건 계약 조건이 아닌데'라는 반응을 보였습니다. '난 너랑 터그 놀이 안 해.' '내가 돌아오면 넌 나한테 간식을 줘야지.' 말도 안 돼요. 그래서 생각했죠. 그냥 물어오기만 할 수 있다면, 굳이 터그 놀이를 할 필요는 없겠구나 하고요. 아니요, 절대 아니죠. 우리 방식은 그게 아니니까요. 그래서 그다음 2주 동안 저녁 식사 때마다 테이터를 데리고 나가서 마당 곳곳의 여러 위치에서 물건을 가져오도록 훈련시켰습니다. 하지만 그 2주 동안은 그렇게 해서 저녁 식사를 얻게 했죠. 이제는 저랑 터그 놀이를 하자는 제안을 거절할 리가 없어요. 식사 시간이라는 아주 가치 높은 보상과 저와의 터그 놀이를 연결했으니까요. 그러면 아마 사료 한 숟가락을 더 얻게 될지도 모르죠. 터그 놀이를 두 번 할 수도 있고요. 세 번 할 수도 있죠. 네 번 할 수도 있고요. 물론 처음에는 물건에 살짝 닿기만 해도 성공이라고 생각했습니다. 그러니 여러분의 강화물이 무엇인지 파악하세요. 무엇이 강화물이 아닌지도 알아야 합니다.
So, number one, what reinforces your dog in those five categories and grow each of those. Don't just say, oh, my dog loves food. He'll just work for food. Or even, hey, you guys who are listening to this who have very high achieving elite dogs who say, my dog just works for the bite. You've got to grow that. You've got to say, you need to tug with this tug toy before I give you permission to go and do that bite work. That's how you build up the value for the tug. My border collies, once they get going in agility, they will happily never take another cookie or a toy. They just want the chance to do agility. But if I am a reinforcement-based dog trainer, how am I ever going to fix a problem I have in agility if agility is a reinforcer? I have to keep that hierarchy of reinforcement high. So, if you want to do a seesaw first, you have to tug with this. I know they won't take food, but they can tug. And I may have to get further away from the seesaw before I get that tug, but I get it. And then I work in food the same way. Our goal here is to reverse our dog's reinforcement pyramid. Rather than having only one or two really high value rewards that the dog will go for, we want to have a multitude of high value rewards at the top and only a very small number of low value rewards at the bottom. And start stacking your hierarchy of reinforcement. So, know what food your dog loves. Know what food your dog won't take when and start building that. For example, a few weeks ago, I was walking tater salad and Kim was with me. And I play these games. I was showing her some of the games I play with him. And one time I did a recall. He came barreling back expecting his cookie, but instead I gave him a tug toy. And he went, oh, that's not an arrangement. Yeah, I don't tug with you. When I come back, you give me a cookie. There is no way. And I thought, well, just if you could just pick it up, you don't actually have to tug. Oh, nay, nay, not our deal. And so, for the next two weeks, I took tater out at dinnertime and I got him to retrieve different things all over the yard, different locations. But that's how he earned his dinner for two weeks. Now, there is no way he's going to say no to retrieving for me because I've taken the value of something that's super high, mealtime, and I've put tug with me. And then you possibly might earn another scoop of your food. I might do two tugs. I might do three tugs. I might do four tugs. Now, of course, at first, I only wanted to get that little bit of a touch. So, know what your reinforcements are. Know what your reinforcements aren't.
16:59
개가 거들떠보지도 않는 것도 가치를 높여서 언제든 받아들이게끔 만들어 보세요. 그 목표는 강화물을 쌓아가는 방식으로 달성할 수 있습니다. 개에게 미소를 짓고, 말을 걸고, 쓰다듬어 준 뒤에 터그 놀이를 할 수도 있겠죠. 터그 놀이를 한 후에 다른 개와 놀게 하거나, 수영을 하게 하거나, 다른 무언가를 할 수 있도록 허락해 줄 수도 있습니다. 강화물을 쌓으세요. 그 모든 것은 여러분을 통해 전달됩니다. 그 모든 것의 가치는 엄청나게 커질 겁니다. 마지막으로, 개에게 관심을 줄 때 무엇을 보상하고 있는지 주의 깊게 살펴보세요. 그게 단순히 '야, 그만해'라고 하는 것이라도, 여러분은 무언가를 보상하고 있는 셈입니다. 강화는 정말 놀라운 것입니다. 강화는 행동을 형성하고, 반려견과 놀라운 관계를 구축할 수 있게 해주죠. 파트너십으로 발전하는 관계 말입니다. 진정한 파트너십이 형성되면, 훈련 자체가 동물에게 본질적인 보상이 됩니다. 여러분이 인내심을 갖고 반려견을 강화하는 섬세한 기술에 주의를 기울인다면, 여러분과 여러분의 반려견에게도 정확히 그런 일이 일어날 수 있습니다. 다음 시간, Shaped by Dog에서 다시 뵙겠습니다. 무슨 일이야? 무슨 일이야?
Work at growing what your dog won't take so that they'll take it at any time. And you're going to get there by stacking your reinforcements. So, you might smile at your dog, talk to them, pat them, and then play a game of tug. You might play a game of tug and then give them permission to go and play with another dog or go for a swim or do something else. Stack reinforcements. They all come through you. They all grow massive in value. And the final thing is pay attention to what you're rewarding when you're giving your dog your attention. Even if it's to say, hey, knock it off, you're rewarding something. Reinforcement is amazing. Reinforcement builds behavior and it can build an amazing relationship with your dog, a relationship that becomes a partnership. And with a true partnership, the work becomes intrinsically reinforcing for the animal. And that's exactly how it can happen for you and your dog if you're patient and you pay attention to the fine art of reinforcing your dog. I'll see you next time here on Shaped by Dog. What's up? What's up?