How To Use Everyday Reinforcement To Grow Lasting Skills
Susan GarrettDogs That · 영상 · 8분
일상적인 강화를 사용하여 지속적인 기술을 키우는 방법
How To Use Everyday Reinforcement To Grow Lasting Skills
0:00
반려견과의 현재 관계와 강화(reinforcement)를 어떻게 발전시킬 수 있을지 이야기해 봅시다. 그리고 그것을 전략적인 강화 활용으로 발전시켜야 합니다. 첫 번째 단계는 반려견이 음식을 사랑하게 만드는 것입니다. 음식에 관심이 없는 반려견을 키우는 분들도 계실 텐데, 그런 경우에는 그 사랑을 만들어내야 합니다. 하지만 반려견이 좋아하는 음식을 하나 찾았다면, 두 번째 음식을 사용해 보세요. 아마 첫 번째 음식을 먼저 주고, 그다음에 더 높은 가치의 음식을 주는 방식일 겁니다. 이렇게 강화 전략을 발전시키는 첫 번째 단계부터 시작합니다. 반려견이 음식을 사랑해야 한다는 것이죠. 만약 반려견이 음식을 좋아하지 않는다면, 훌륭한 행동을 만들어낼 수 없습니다. 훌륭한 행동은 훌륭한 강화에서 비롯되기 때문입니다. 그러니 첫 번째는 반려견이 음식을 좋아하게 만드는 것입니다. 많은 분들이 반려견이 음식 없이는 아무것도 하지 않는 상황에 처해 계시죠. 그래서 우리는 1단계부터 5단계까지 나아가야 합니다. 2단계는 음식을 얻는 과정에 조건(contingencies)을 추가하는 것입니다. 제 경우, 반려견과 가장 먼저 하는 게임 중 하나가 '네 선택이야(It's Your Choice)' 게임입니다. 팟캐스트에서 여러 번 언급한 적이 있죠. 쇼 노트에 링크를 남겨둘게요. 강화에 조건을 건다는 것은 주머니에 간식을 넣어두는 상황을 의미합니다. '네 선택이야' 게임을 하면 반려견은 주머니 속의 간식을 무시하게 될 텐데, 왜냐하면 다가와서 엉덩이를 툭 치는 것으로는 간식을 얻을 수 없다는 걸 알기 때문이죠. 혹은 저처럼 집 안 카운터 곳곳에 음식 그릇을 두는 사람도 있을 겁니다. 만약 반려견이 놀라운 행동을 하면 바로 신호를 줍니다. 예를 들어 '옳지(good dog)'라고 말하거나, 특정 위치 보상 신호인 '스포트라이트(spotlight)'를 사용해 '쿡(cook)'이라고 한 뒤, 걸어가서 간식을 주는 식이죠. 그렇다면 이것이 어떤 모습일까요? 제가 식사할 때 반려견이 부엌에 들어오는 것을 좋아하지 않는다고 가정해 봅시다. 제가 요리할 때 주방 가장자리에 있는 강아지 방석에 개가 있으면 간식을 줄 거예요. 특히 강아지일 때는 더더욱요. 의도적으로 그 방석 위에 있고 싶게 만들 거예요. 그러니 그것이 우발적 보상이 될 수 있죠. 강아지는 관계 은행에 예기치 못한 예금을 받게 되는 셈입니다. 네가 그 방석 위에 있는 게 정말 좋아. 누군가 문으로 들어올 때 짖지 않는 네 모습이 정말 좋아. 당신이 좋아하는 행동은 무엇이든 강화하게 될 거예요. 강아지들은 음식이 올 줄 모르죠. 무언가 하고 있는 중에 당신이 다가가 '정말 잘했어'라고 말하며 강화하는 겁니다. 거기에 우발적 상황이 있는 것이죠. 훈련 게임으로 개를 훈련할 수도 있고, 일련의 훈련 게임을 활용할 수도 있습니다. 저희 Recallers 프로그램처럼 전략적으로 단계를 짠 게임을 제공하는 훈련 계획을 통해서요.
Let's talk about how we can evolve the current relationship with your dog and reinforcement, and evolve it into a strategic use of that reinforcement. So, the first step is the dog loving food. Some of you don't have a dog that loves food. You need to create that love. But when you find one food the dog loves, then you're going to use a second food, maybe give it to them first and give them the higher value food second. So, we start with step number one as we evolve your reinforcement strategy. That is the dog's got to love food because if they don't love food, you can't grow brilliant behaviors. Brilliant behaviors are grown from brilliant reinforcement. So, number one, dog loves food. That's where a lot of you are, that the dog won't do anything unless they see the food. And so, we've got to go from step one all the way down to step five. So, step two is we've got to add some contingencies to them getting their food. And that's where for me, one of the first games I play with my dogs is the It's Your Choice game. I've talked about it a lot on the podcast. I'm going to leave a link in the show notes. Contingencies on that reinforcement means you might have some cookies in your pocket. Now, if you play It's Your Choice, your dog will eventually ignore those cookies in the pocket because they know coming up and bopping you in the hip is not going to get cookies. Or you might be like me, I have like food bowls around the house on the counters and they're there. If I see something amazing happening, then I will mark it. I might say, good dog. I might give what we call a spotlight, a location-specific reinforcement marker, cook, and I'll walk over and I'll give the dog a cookie. So, what would that look like? Let's say I don't like my dogs in the kitchen when I'm eating, or I don't like my dogs in the kitchen when I'm cooking. If I see a dog in a dog bed on the peripheral of the kitchen, I'll give them a cookie, especially when they're puppies. I will intentionally make them want to be in those beds. So, that could be a contingency. The dog gets the surprise deposit into the relationship bank. I love when you're in that bed. I love when you aren't barking when someone comes through the door. Whatever you love, you're going to reinforce. So, they don't see the food coming. They are in the midst of something and then you go, that's mighty good. And you reinforce them. There's a contingency on that. It could be you're training the dog with a training game, with a series of training games, with a training plan like our Recallers program, where we give you strategically layered games that
2:38
강아지와의 강화 전략을 발전시킬 수 있도록 돕습니다. 그리고 또 하나 해야 할 일은 강화에 대한 우발적 상황을 만드는 동안 요구하는 개(Demandica dog)를 무시하는 것입니다. 다가와서 손을 툭 치며 '나한테 먹을 걸 줘야겠어, 뒤에 간식이 있는 것 같아'라고 말하는 개, 간식 통 앞에서 앉아 짖는 개, 저녁 먹을 시간이 되면 낑낑거리는 개를 말하죠. 요구하는 개를 무시해야 합니다. 음식, 장난감, 활동은 모두 요구해서 얻는 것이 아니라 제공받는 것이라는 점을 이해하도록 도와야 합니다. 자, 그럼 음식을 좋아하게 만들고 그 음식에 대한 우발적 상황을 만들어야 합니다. 이제 그 음식의 가치를 전이시켜야 합니다. 강아지가 무언가를 하고 보상을 받는 게임을 하는 것, 그것이 가치를 전이시키는 방법입니다. 그렇게 음식의 가치가 크레이트 게임 같은 게임으로 옮겨가는 것이죠. 가치 전이가 일어납니다. 크레이트 게임을 해본 분들은 아시겠지만, 강아지들이 그 안에서 보상을 받은 경험이 많아서 크레이트로 날아 들어가기를 좋아하게 될 거예요. 가치가 전이된 거죠. 다른 것으로도 전이될 수 있습니다. 터그 놀이로 전환하세요. 이제 또 하나의 가치 있는 강화물을 얻게 된 것입니다. 항상 간식만 사용하는 게 아니죠. 때로는 터그 놀이를 활용합니다. 이는 핸드 타겟이나 다른 놀이처럼 당신과 함께하는 행동으로 전이될 것이며 이제 활용할 수 있게 됩니다. 산책 중에 개에게 가까이 오라고 할 때, 아마 누군가를 지나치고 있을지도 모릅니다. 저는 개에게 강화 구역으로 들어오라고 할 겁니다. 제 옆은 또 다른 강화된 행동이죠. 그들이 그곳에 있을 때 간식이 없다면, 그냥 핸드 터치를 요구할 겁니다. 산책 나갈 때 간식을 안 가져가는 경우는 매우 드물지만, 그런 일이 생기기도 합니다. 정말로 생기죠. 그래서 우리는 음식을 사랑하는 단계에서 음식에 의존하는 단계로, 이제는 음식의 가치를 장난감, 게임, 활동으로 전이시키는 단계까지 왔습니다. 이제 행동의 지속 시간을 늘려볼 겁니다. 즉, 방석에 올라가자마자 보상을 주지는 않을 겁니다. 개가 주방 근처 방석에 올라가면 당신이 주방을 몇 번 왔다 갔다 한 뒤에 보상을 줄 수도 있죠. 우리는 보상 사이의 간격을 늘려가는 중입니다. 그래서 결국 당신이 원했던 것처럼 당신이 식사하는 동안 개가 방석에 머물게 하고 싶다면, 우선 개가 방석에 있는다는 것의 의미를 완벽하게 이해하게 만들어야 합니다. 식탁에 앉아 있지 않은 상태에서 그 지속 시간을 늘려가야 하죠. 그렇게 개가 방석 위에서 편히 쉬는 멋진 행동이 완성되면, 이제는 식탁에 앉아 있는 동안에도 똑같이 합니다. 처음에는 간식 몇 개를 때때로 방석 쪽으로 던져주는 것부터 시작할 수 있지만, 결국에는 일어나서 식탁을 정리하고 개에게
help you evolve the reinforcement strategies with your dog. And the other thing that you need to do while you're growing contingencies on your reinforcement is ignore the Demandica dog. The dog that comes up and flips your hand and says, I think you need to feed me. I think there's cookies back here. The dog that sits and barks at the cookie jar. The dog that whines when it's near dinner o'clock. You need to ignore the Demandica dog and you need to help grow their understanding that food, toys, activities are all reinforcement that are delivered, not demanded. Okay. So, we've got, got to love food, got to have some contingencies on that food. Now we've got to transfer the value of that food. Playing games where the dog does something and then they're reinforced, that's a way to transfer value. So, the value goes from the food into a game like crate games. The transfer of value happens. So, any of you who've ever played crate games, I know that your dogs will love to fly into the crate because they've got such a history of reinforcement there. The value has been transferred. It could be transferred into a game of tug. Now you've got another valuable reinforcer. You're not always using food. You're sometimes using tug. It will be transferred into acts with you like the hand target or other games that now you can use. If I'm out on a walk and I ask my dog to come near me, maybe I'm passing somebody. I will ask them to come into reinforcement zone, another reinforced behavior at my side. While they're there, if I don't have any cookies, I'll just ask them for a hand touch. Highly unlikely I'm out for a walk without cookies, but it does happen. It does happen. So, we've gone from the, I love food stage to contingencies on that food to now we're transferring the value of the food to food, toys, games, activities. Now we're going to grow the duration of the behavior. So, they're not going to get a reinforcement the moment they get in the bed. They might get in the bed around the kitchen and you might walk around the kitchen a few times and then you're going to reinforce them. We're stretching out the duration between reinforcements. So, that eventually, let's say you wanted your dog in the bed while you're eating. Well, first, you're going to have an amazing understanding of what being in that bed means. You're going to grow that duration without you being at the dinner table. So, once you have this amazing behavior of your dog hanging out on a bed, now you're going to do it while you're at the dinner table. It might start with you having a couple of cookies that you throw
5:13
간식을 준 뒤 설거지를 하러 갈 수도 있습니다. 나중에는 방석에 있는 것만으로 간식을 주지 않아도 개가 방석에 머물게 될 겁니다. 왜냐하면 그곳은 여전히 무작위로 보상을 받을 수 있는 매우 강화된 장소이기 때문이죠. 예를 들어 제가 주방으로 걸어 들어갈 때, 새로 키우는 강아지가 있다면 저는 그저 강아지가 방석에 있는 것을 보고 보상을 줄 것입니다. 강아지가 개 침대에 있을 때 보상을 주고, 만약 다른 강아지도 그 개 침대에 있다면 그 강아지들도 함께 보상을 받습니다. 이렇게 우리는 특정 행동의 지속 시간을 늘리고 있습니다. 또한 행동들을 다른 행동들로 확장하며 행동 연쇄를 만들어가고 있죠. 여러분도 이미 행동 연쇄를 가지고 있을 겁니다. 분명히 그래요. 여러분은 깨닫지 못할 수도 있지만, 분명 하나는 가지고 있죠. 처음 강아지를 데려왔을 때, 아침에 일어나자마자 가장 먼저 한 일이 무엇이었는지 내기를 해도 좋습니다. 만약 강아지가 크레이트에 있었다면, 강아지를 꺼내서 밖으로 데리고 나가 배변을 하게 했을 겁니다. 이제 시간이 지나 어느 정도 익숙해지면, 여러분은 먼저 일어나서 화장실을 다녀온 뒤 강아지를 데리고 밖으로 나가는 단계에 이르게 됩니다. 그러니 밖에서 배변하도록 하는 보상이 이제는 하나의 연쇄 행동으로 구축된 셈이죠. 이제는 '침대에서 일어난다, 밖으로 나간다'는 단순한 과정이 아닙니다. '일어나서 내가 화장실을 가고, 그다음에 밖으로 나간다'는 과정이 된 것입니다. 이 과정은 더 길어질 수도 있습니다. 예를 들어 제가 호텔에 머물고 있을 때, 아마 수백만 층쯤 되는 높은 곳에 있을지도 모릅니다. 일어나서 화장실에 가고 샤워를 하겠죠. 다 자란 반려견들을 데리고 밖으로 나가기까지 30분이 걸릴 수도 있습니다. 물론 강아지라면 그렇게 하지 않겠지만 말입니다. 이렇게 우리는 행동 연쇄를 확장해 나갑니다. 예를 들어, 제가 언급했던 리인포스먼트 존(보상 구역)은 제 옆에 있는 공간으로, 리콜 게임을 통해 그 가치를 키워왔기 때문에 반려견들이 가치를 느끼는 장소입니다. 처음에는 제자리에 멈춰서 시작할 수도 있죠. 반려견이 스스로 리인포스먼트 존을 찾아가면 그곳에 머무는 대가로 간식을 줍니다. 결과적으로, 저는 '정말 잘했어'라고 말하며 리인포스먼트 존에 도달하게 합니다. 그러고는 제가 한 걸음 내딛고, 여러분에게 간식을 줍니다. 이제 강아지는 리인포스먼트 존을 찾을 수 있게 되었죠. 저는 한 걸음, 두 걸음 옮긴 뒤에 보상을 줄 수 있습니다. 저는 이 과정을 반복할 수도 있어요. 강아지가 리인포스먼트 존을 찾으면 간식을 주고, 다음 간식을 주기 전까지 세 걸음을 걷는 식으로요. 이것이 바로 행동 연쇄를 확장하는 것입니다. 그리고 그 과정은... 강아지 훈련에서 강화의 사용이 어떻게 진화했는가입니다.
back into the bed from time to time, but eventually you might get up, clear the table, give the dog a cookie and go and do the dishes. And eventually it'll go to, I'm not even feeding the dog for being in the bed, but they're being in their bed because it's a highly reinforced position that will still randomly get reinforced. Like when I'm walking into the kitchen, when I have a new puppy, if I reinforce a puppy that's in a dog bed and I see another dog in the dog bed, they're getting reinforced as well. So, we're growing duration of a specific behavior. We're also growing behaviors into other behaviors, a behavior chain. So, you already have a behavior chain. I'm sure of it. You may not realize it, but you have one. When you first got your puppy, I'd be willing to bet the first thing you did when you got up in the morning, if your puppy was in a crate, was to get them up and run outside with them so that they can eliminate outside. Now, eventually you got to the point where you would get up and you would go to the bathroom yourself, and then you would come and get your dog and get them outside. So, the reinforcement of relieving themselves outside now has been built into a chain. It doesn't go from, I'm out of bed, we're going outside. It's, I'm out of bed, I'll go to the bathroom, then we'll go outside. And it can grow even more when I'm in a hotel and I might be on the trillionth floor. I get up, I will go to the bathroom, I'll shower. It might be 30 minutes before I go outside with my adult dogs. I wouldn't do this with a puppy, of course. So, we're growing a behavior chain. For example, I mentioned reinforcement zone, a place on my side where my dogs find value because we've grown it from recaller games. Now, I might start stationary. The dog finds reinforcement zone on their own. They get a cookie for being there. Eventually, I'm going to say, you are so good at that. We find reinforcement zone. I'm going to take a step and I'm going to feed you. So, now you can find reinforcement zone. I can take one, two steps before I feed you. I might ping pong that. You find reinforcement zone. I give you a cookie. We take three steps before you get your next cookie. That is growing a behavior chain. And that is how the use of reinforcement is evolved in dog training.