Dog Training Outside The Box: Transfer Of Value Case Study #177
Susan GarrettShaped by Dog with Susan Garrett · 팟캐스트 · 19분
반려견 훈련은 고통이나 강압 없이 강화 중심의 프로그램을 통해 수행된다. 반려견은 언제 강화가 제공되는지, 어떤 행동이 강화를 유도하는지를 스스로 이해하는 과정이 필요하며, 보호자는 반려견의 열광적인 반응을 이끌어낼 강화물의 위계와 환경에 따른 가치 변동을 파악해야 한다. 강화의 위치와 전달 방식은 목표 행동의 강화에 결정적인 영향을 미치며, 낮은 가치의 보상에 대한 흥미가 부족한 개에게는 고가치 보상을 활용한 가치 전이를 통해 훈련 동기를 부여한다. 특정 상황에서만 활용되던 고가치 보상을 점진적으로 일상적인 보상과 연결함으로써, 반려견은 점차 일반적인 보상에도 높은 참여도를 유지하게 된다. 훈련의 핵심은 틀을 깨는 사고를 통해 개마다 다른 최적의 강화 경로를 찾고, 훈련 환경과 방해 요소를 고려한 세심한 훈련 계획을 세우는 데 있다.
기존 틀을 벗어난 개 훈련: 가치 이전 사례 연구 #177
Dog Training Outside The Box: Transfer Of Value Case Study #177
0:00
지난 몇 주 동안 저는 여러분의 질문에서 자연스럽게 파생된 시리즈를 선보이고 있습니다. 그 질문들은 제가 개를 훈련하는 방법론과 학생들을 위한 레슨 계획을 짜는 방식에 관한 모든 것이었습니다. 그리고 종종 들어오는 질문은 "수잔, 유도물(lure) 없이 이걸 어떻게 훈련하나요?"입니다. 간식 유도물 없이 힐링(healing)을 어떻게 가르치는지와 같은 간단한 것일 수도 있고, 아니면 어떻게 간식 유도물 유무와 관계없이 개에게 뒷다리를 가리키도록 가르치는지, 또는 어떻게 개가 물건을 가져오도록 셰이핑(shaping)할 수 있는지에 대한 질문일 수도 있습니다. 제 대답은 사실 틀을 깨는 사고를 요구합니다. 오늘 밤에는 깊이 있게 파고들어 그 과정이 어떤 모습인지 살짝 보여드리려고 합니다. 안녕하세요, 수잔 가렛입니다. Shape My Dog에 오신 것을 환영합니다. 지난주 팟캐스트에서 저는 애틀랜타에서 가르쳤던 저먼 셰퍼드에 대해 언급했습니다. 이 일은 적어도 20년도 더 넘은 일이네요. 제가 세미나를 진행하고 있었는데, 한 여성이 아침 일찍 찾아왔습니다. 그녀는 이렇게 말했죠. 이 아이가 제 개입니다. 이름을 벨라라고 하죠. 솔직히 제 기억력이 그리 좋지 않아서 그 여성의 이름이나 개 이름은 기억나지 않습니다. 그녀는 "복종 훈련을 위해 벨라에게 덤벨을 가져오도록 가르쳐야 해요"라고 말했습니다. 그리고 제 강사가 강제로 물게 하는(force fetch) 훈련을 하고 싶다고 했다더군요. 그건 개가 입을 열 때까지 귀를 꼬집는 것을 의미합니다. 그런 다음 덤벨을 입에 물리는 것인데, 이를 부적 강화(negative reinforcement)라고 합니다. 개가 원하는 행동을 하면 고통을 제거하는 방식이죠. 여러분은 굳이 그럴 필요 없습니다. 만약 이 팟캐스트를 듣고 계신다면, 아마도 여러분은 개에게 절대로 고통을 주고 싶지 않으실 겁니다. 하지만 제가 어떤 방식으로든 이 여성분이 그 행동을 하도록 돕지 못했다면, 벨라는 그런 운명에 처할 상황이었습니다. 잠시 물러나서 우리가 강화 기반 프로그램으로 훈련할 때에 대해 이야기해 보죠. 이는 신체적 교정이나 언어적 위협을 사용하지 않고 개를 훈련하기로 결정했다는 것을 의미합니다. 여러분은 "안 돼"라고 말하고 싶지 않으시겠죠. 이봐요, 안 돼요. 아. 개에게 칼라 팝(collar pop)을 주어서는 안 됩니다. 귀를 꼬집어서는 덤벨을 입에 물게 하려고 절대 해서는 안 되죠. 그런 식으로 훈련하기로 결정했다면, 당신의 가장 중요한 도구는 강화(reinforcement)입니다. 따라서 당신은 강화의 사용과 이해에 있어 아주 능숙해져야 합니다. 자, 이에 대해 깊이 파고들어 봅시다. 저는 이를 개가 이해해야 할 것과 인간이 강화에 대해 정말로 이해해야 할 것으로 나누어 설명합니다. 개에게는 꽤 단순한 문제이니 개부터 이야기해보죠. 첫째, 개는 언제 강화가 가능한지 이해해야 합니다. 즉, 지금 강화를 얻게 될 것이라는 뜻이죠. 그렇지 않으면 개는 끊임없이
Over the last few weeks, I have been rolling out a series that kind of grew organically from your questions all about the methodology of how I train my dogs, how I come up with the lesson plans for my students. And often it's how do you train this without a lure, Susan? It could be something as simple like, how do you teach healing without a food lure? Or how do you teach your dog to point his back leg with or without a food lure? Or how could you shape your dog to retrieve? And my answer to that really requires some out of the box thinking. And tonight I'm going to do a deep dive and give you a glimpse of what that looks like. Hi, I'm Susan Garrett. Welcome to Shape My Dog. In a podcast last week, I mentioned a German Shepherd that I was teaching in Atlanta. And this was at least or more than 20 years ago. I was giving a seminar. This woman showed up right at the beginning of the day. She said, this is my dog. Let's call her Bella. Honestly, my memory is not that good that I can remember either the woman's name or the dog's name. She said, I need to teach Bella how to retrieve the dumbbell for obedience. And my instructor said that they would like to do a force fetch, which means they were going to ear pinch the dog until the dog opened its mouth. Then they would put the dumbbell in and that's called negative reinforcement. They remove the pain when the dog does what they want. You guys don't really need to know that. If you are listening to this podcast, you probably don't ever want to cause your dog pain. But this was what was going to be Bella's fate if I couldn't help this woman somehow do that behavior. Let's just take a step back and talk about when we are training in a reinforcement-based program, meaning you've decided you want to train your dog without the use of physical corrections or verbal intimidation. You don't want to say, hey, no, ah. You don't want to give them a collar pop. You certainly don't want to give them an ear pinch in order to get them to open their mouth for a dumbbell. When you decide to train that way, then your number one tool is reinforcement. And so, you've got to become brilliant at the use of and the understanding of reinforcement. So, let's just do a deep dive into that. And I break it up into what we need the dog to understand and what the human really needs to understand all about reinforcement. So, let's talk about the dog since it's pretty simple for them. Number one, the dog needs to understand when reinforcement is available, meaning now you're going to earn your
2:44
강화를 얻으려고 할 테니까요, 특히 유도(luring) 훈련을 받아왔다면 더욱 그렇습니다. 그래서 우리가 온라인 강의실에서 수강생들에게 가장 먼저 가르치는 것은 개에게 언제 강화가 오는지 알려주는 법입니다. 강화가 오기 전까지는 그냥 편안하게 있어도 된다는 것을요. 여러분 모두 마커(marker)에 대해 잘 아실 겁니다. 마커는 개에게 방금 한 선택이 옳았다고 알려주는 것입니다. 그 선택은 옳았어요. 그리고 팟캐스트 174편에서 배웠듯이, 개는 이제 올바른 선택을 했다는 사실에 도파민이 급증할 것입니다. 그들은 이제 강화가 올 것이라고 기대할 수 있죠. 바로 마커가 우리 개들에게 그것을 가르치는 것입니다. 마커는 클리커(clicker) 같은 것이 될 수 있어요. 그래서 제 개는 클리커 소리를 들을 때마다 항상 강화를 받습니다. 실수로 클릭했더라도 저는 항상 강화합니다. 그것이 바로 제가 개와 맺은 약속이기 때문입니다. 제가 사용하는 단어와 같은 다른 마커들도 있습니다. 팟캐스트 151편을 참조하시면, 제가 위치 지정 마커(location specific markers)에 대해 이야기하는 것을 들을 수 있는데, 이는 개에게 강화를 여기서 줄 것이라거나, 저기로 가서 강화를 기대해도 좋다는 것을 알려줍니다. 만약 제 개가 특정 자세를 유지하길 원한다면 말이죠. 아니면 특정 위치일 때 저는 '쿠키'라고 말하는데, 이는 제가 개에게 직접 간식을 주겠다는 뜻입니다. 그래서, 정지 상태에서는 '움직이지 않아도 된다'고 알려주는 쉬운 신호가 됩니다. 저는 개 입으로 간식을 직접 전달합니다. 힐 자세에서는 '쿠키'라고 말하고 제 엉덩이에 닿게 한 뒤 간식을 줍니다. 그러면 엉덩이가 개가 머물러야 할 일반적인 위치를 알려주는 타겟 구역이 됩니다. 좋습니다. 그래서 제가 언급했었죠. 151번으로 가시면 제가 사용하는 모든 단어를 확인하실 수 있는데, 그 단어들은 보상이 곧 주어질 것이니 여기서 기다리라는 의미입니다. 또한 개에게 보상을 받을 수도 있지만 지금 하던 것을 계속하라는 의미의 단어들도 있습니다. 예를 들어 '좋아', '슈퍼', '훌륭해' 같은 단어들은 그냥 계속하라는 뜻이며, 잘하고 있다는 것을 의미합니다. 개는 주변 환경에 있는 보상을 쫓아가지 말아야 한다는 점을 이해해야 합니다. 이제 게임이 시작되었다거나, 보상이 곧 오거나, 아니면 지금 가서 가져와도 된다는 마커 단어를 듣기 전까지는 말이죠. 개가 알아야 할 두 번째는 보상을 얻는 방법입니다. 즉, 개가 쉐이핑을 받을 때 스스로 행동을 제안하는 법을 알아야 합니다. 평생 간식 유도 훈련만 받은 개들은 유도 물체가 나오기 전까지는 아무것도 하지 않아도 보상을 받아왔습니다. 그래서 많은 개들이 스스로 행동을 시작하지 못합니다. 보세요, 손을 흔들 수도 있고, 뒤로 물러나거나, 눈을 굴리거나, 스웨터를 짜는 행동을 보여줄 수도 있겠죠. 그래서 사람들은 '아, 이 개는 안 돼, 클릭커 트레이닝 같은 방법은 우리 개에겐 안 통할 거야'라고 생각하곤 합니다.
reinforcement because otherwise they're going to be constantly trying to get the reinforcement, especially if they've been lured. So, the first thing that we teach our students in our online classrooms is how to teach the dog to know when the reinforcement is coming. You can just chillax until it's coming. And you guys know all about markers. So, a marker is what tells the dog that choice was correct. The choice was correct. And you know from podcast episode number 174, that now they're going to get a dopamine spike, that they've made the correct choice. They can expect a reinforcement is on the way. And so, a marker is what teaches our dogs that. It could be a marker like a clicker. So, every single time my dog hears a clicker, they will get a reinforcement. Even if I click by mistake, I will always reinforce because that's the bond that I've made with my dog. There are other things like words that I'll use. And if you refer to podcast episode number 151, I talk about location specific markers, which tells my dog that they're going to get a reinforcement delivered here, or they can go here and expect this reinforcement. If I want my dog to hold a position or a location, I will say cookie, which means I'm going to deliver the cookie to their mouth. So, stationary positions, that's an easy one that tells them you don't need to move. I will deliver the cookie into your mouth. In heel position, I will say cookie, touch my hip and deliver the cookie. And the hip becomes a targeted area for the dog to stay in that general location. Okay. So, I've spoke about them. You can go to 151 and look at all the words that I use that mean your reinforcement is coming. Expect it here. I do also have a couple words that mean you possibly might earn reinforcement, but keep doing what you're doing. And things like good or super or excellent, like just words that mean keep going, you're on the right track. The dog needs to understand, don't go after whatever reinforcement is out in your environment unless you hear a marker word that says, game on now, it's coming, or now you can go get it. The second thing that the dog needs to know is how to earn that reinforcement. Meaning when they are being shaped, they need to know how to offer a behavior. And for dogs who've been food lured their whole life, they have been reinforced for doing nothing until the lure comes out. So, a lot of dogs won't start offering, look, I can wave, I can back up, I can roll my eyes, I can knit you a sweater. So, people will assume, oh, this dog, yeah, clicker training, that method won't work for my dog.
5:21
네. 그거 아세요? 클릭커 트레이닝은 하나의 방법이 아니라 과학을 적용하는 것입니다. 그래서, 지구상의 모든 종에게 효과가 있습니다. 여러분의 개에게도 효과가 있을 겁니다. 개는 학습의 법칙을 따를 것입니다. 약속합니다. 하지만 첫 번째 행동을 스스로 제안할 수 있도록 아주 단순하게 만들어 줄 필요는 있습니다. 좋습니다. 그러니까 가능합니다. 우리 개들에게 필요한 건 그저 이해하고 우리가 그들을 교육함으로써 그들은 마커 단어를 이해하게 될 것이고, 이는 곧 여러분의 강화가 가능하다는 것을 의미하며, 그들은 무언가를 어떻게 제시해야 하는지 이해해야 합니다. 이제 여러분의 차례입니다. 무엇을 알아야 할까요? 우선 여러분은 반려견에게 최고의 강화물이 무엇인지 알아야 합니다. 그것이 첫 번째입니다. 무엇이 최고로 좋아서 정신을 못 차릴 정도로 열광하게 만드는지, 그런 강화물이 나오면 너무 좋아서 어쩔 줄 모르는지 알아야 합니다. 또한 아주 좋은 강화물이 무엇인지도 알아야 합니다. 그리고 허용 가능한 강화물이 무엇인지도 알아야 합니다. 자, 이런 것들은 반려견의 훈련 경험 연령과 훈련하는 장소에 따라 달라질 것입니다. 그러니 현재 훈련 단계와 환경에서 무엇이 열광적인지, 무엇이 좋은지, 무엇이 허용 가능한지 알아야 합니다. 좋습니다. 그럼 이런 방해 요소가 있는 상황에서는 어떨까요? 왜냐하면 주변 환경에 방해 요소가 많다면, 더 높은 수준의 강화물이 필요할 것이기 때문입니다. 그래서 이전에 허용 가능했던 것들이 주변에 방해 요소가 많을 때는 허용되지 않을 것입니다. 반려견을 훈련할 때마다 여러분은 어디서 훈련하는지, 무엇이 열광적인지, 무엇이 좋은지, 무엇이 허용 가능한지를 고려해야 합니다. 왜냐하면 그것들은 계속 변할 것이기 때문입니다. 또한 훈련하는 환경에서 어떤 제약이 있는지도 알아야 합니다. 예를 들어, 수업 시간에 훈련한다면 삑삑이 장난감은 사용할 수 없습니다. 왜냐하면 다른 모든 개들의 주의를 분산시킬 것이기 때문입니다. 마찬가지로 여러분 반려견의 열광적인 첫 번째 강화물이 수영인데 근처에 물이 없다면, 그 환경에서 훈련할 때 그것을 반려견의 첫 번째 열광적이고 뛰어난 강화물로 사용할 수는 없습니다. 즉, 환경은 단순히 방해 요소뿐만 아니라 환경의 구조, 즉 어떤 문제나 자산이 있어서 강화물로 활용할 수 있는지에 대한 것입니다. 예를 들어, 제가 보더 콜리를 훈련할 때 이 실내에서 이 개가 가장 좋아하는 최고의 보상은 저나 다른 사람이 장난감을 가지고 도망치기 시작해서 그 개가 그것을 잡고 터그 놀이를 할 수 있게 하는 것입니다. 밖에 있을 때 이 개가 가장 좋아하는 최고의 보상은 엄마를 쫓아갈 기회입니다. 그러니 필요할 때 바로 실행할 수 있도록 알고 있어야 합니다.
Yeah. You know what? Clicker training isn't a method. It's just applying science. So, it works for every species on the planet. It will work for your dog. Your dog will abide by the laws of learning. I promise. However, you may need to make that very simple for them to offer their first behavior. All right. So, it's possible. That's all our dogs have to do is understand and through our educating them, they will understand their marker words, which means your reinforcement is available and they have to understand how to offer something. Now it's your work. What do you need to know? First of all, you have to know what are your dog's best reinforcements. That is number one, what is like off the charts, crazy, outrageous, oh my gosh, I'm losing my mind because this reinforcement is coming out. You also need to know what is a really good reinforcement. And then you've got to know what is an acceptable reinforcement. Now, those things are going to change based on the age of the dog's experience in training and the location where you're training. So, what's outrageous, what's good, what's acceptable in this environment at this stage of your training. All right. And what about under these distractions? Because if there's a lot of environmental distractions, you're going to have to go even higher. So, what was previously acceptable won't be acceptable when there's a lot of distractions around. Every time you train your dog, you have to consider where you're training and what is outrageous, what is good, what is acceptable because those are going to keep changing. You also have to know what are you limited by the environment you're training. So, for example, if you're training in a class, you can't use a squeaky toy because you're going to be just distracting all of the other dogs. Likewise, if your dog's outrageous number one reinforcement is swimming and there's no body of water nearby, that can't be your dog's number one outrageous outstanding reinforcement when you're training in this environment. So, the environment, it's not just the distractions, it's the structure of the environment, what problems or what assets does it present that you can use as a reinforcement. For example, when I am training my border collie this in the building, her number one most outrageous reinforcement is me or somebody else taking off running with a toy that she can catch and play tug with. When we are outside, her number one outrageous reinforcement is the chance to chase her mother. So, it changes knowing that you want to go to that
7:47
이제 고려해야 할 다른 점은 여러분이 훈련하려는 행동에 대해 가장 훌륭하고, 적절하며 허용 가능한 보상이 무엇인지 파악하는 것입니다. 예를 들어, 정밀한 힐(Heel) 자세를 훈련하고 있다면, 그 정밀한 힐 자세에서 개의 입으로 바로 전달할 수 있는 것을 선택하는 것이 좋습니다. 수영을 시켜주는 것을 보상으로 활용할 수는 있지만, 가끔씩만 가능하기 때문에 제약이 따를 것입니다, 반면 가치 높은 음식을 전달하는 것은 더 빠르게 수행할 수 있죠. 그래서 훈련하려는 행동이 무엇인지, 그리고 평상시 상황에서 개가 가장 좋아하는 것을 활용하는 것이 합리적인지를 따져봐야 합니다. 좋습니다. 여러분이 훈련 중인 행동에 효과가 있을까요? 다음으로 고려할 점은 그 보상을 어떻게 전달하느냐입니다. 이는 전적으로 여러분의 목표에 달려 있습니다. 훈련 목표가 무엇인가요? 만약 제가 반려견이 발톱 다듬는 것을 받아들이도록 만들고 싶다면, 고려할 점이 정말 많지만 기본 자세는 옆으로 눕는 것이겠죠. 저는 마커 단어를 사용하여 그 자세에서 벗어나도 된다는 신호를 주거나, 더 가능성 높은 방법으로는 보상을 반려견에게 직접 가져다줄 것입니다. 보상의 위치가 바로 누워 있는 행동을 강화하기 때문이죠. 보상은 하나의 과정입니다. 마커를 사용하더라도 보상의 위치를 여전히 고려해야 합니다. 예를 들어, 개에게 멀리 달려가는 것을 가르치고 마커를 줬다면, 그건 좋은 방법이죠. 그리고 개가 돌아와 여러분 앞에 앉는 동안 여러분이 주머니에서 쿠키를 꺼낸다면 주머니요. 네, 당신은 그들이 당신을 떠나는 것을 강화하는 셈이 되겠지만, 동시에 그들이 당신에게 다시 돌아와 앞에 앉는 것을 강화하는 것이기도 해서, 당신에게서 도망치는 것과는 반대되는 행동이 될 것입니다. 그래서 당신에게서 도망치는 행동을 훈련하는 데는 훨씬 더 오랜 시간이 걸릴 것입니다. 정말 중요합니다. 강화의 위치는 매우 중요합니다. 왜냐하면 그 간격이 존재하기 때문이죠. 다음은 강화의 전달입니다. 수잔, 그럼 강화의 전달과 위치의 차이점은 무엇인가요? 위치는 개가 무엇을 하고 있는지를 의미합니다. 전달은 당신이 강화를 그들에게 가져갈 때 얼마나 신속하고 의도적으로 하는지를 의미합니다. 좋아요. 강화 과정입니다. 이건 과정이에요. 그래서 사람은 가치, 전달, 그리고 위치에 신경 써야 합니다. 개는 좀 쉬운 편이죠. 이제 다시 우리의 친근한 저먼 셰퍼드 이야기로 돌아가서 이 모든 고려 사항을 적용해 봅시다. 제가 이분에게 처음 물은 것은 '무엇이 가치 있는가?'였습니다. 그러자 그녀는 놀라운 뷔페를 열었습니다. 작은 플라스틱 통들을
when you need it. Now, the other thing you need to consider, what is the most outrageous, good, and acceptable reinforcement for the behavior that you are training. So, if you were going to be training, let's say precision heel position, you may choose to use something that can be delivered to the mouth of the dog in that precision heel position. So, you could use like go for a swim to reinforce that, but it would be more limiting because it could only happen once every once in a while, where delivering high value food could happen more quickly. So, what is the behavior and is it reasonable to consider using what the dog loves most of all in a normal situation? All right. Does it work for the behavior that you are training? The next consideration is the delivery of that reinforcement. And that is really dependent upon your goal. What is your goal of training? So, if I was wanting to shape my dog to accept their nails being trimmed, there's a lot of things that come up to this, but the position would be they're lying on their side. And I would either give them a marker word, which meant you could get out of that position, or more likely what I would do is bring the reinforcement to them. So, the placement of the reinforcement actually reinforces them lying down. Reinforcement is a process. If you use a marker, you still need to be considerate of the placement of the reinforcement. So, if you wanted to teach a dog to run far away from you and you marked, that's good. And then they came back and sat in front of you while you dug a cookie out of your pocket. Yeah, you would be reinforcing them leaving you, but you would also be reinforcing them for coming back and sitting in front of you, which would be in opposition to them running away from you. So, the behavior of running away from you would take a lot longer for you to train. Super important. The placement of reinforcement is critical because you've got that gap. Next is the delivery of reinforcement. Well, what's the difference, Susan, the delivery and the placement? The placement is what the dog is doing. The delivery is how swift and deliberate you are when you are bringing that reinforcement to them. Okay. Reinforcement process. It's a process. So, the human's got to be concerned with the value, the delivery, and the placement. The dog, they've got it kind of easy. Now let's get back to our friendly German Shepherd and take all these considerations in. The first thing I said to this lady is, what is of value? And she opened this amazing buffet. She had all these little Tupperware
10:18
여러 개 가져왔는데, 하나에는 잘게 썬 로스트 비프, 다른 하나에는 잘게 썬 닭고기, 또 다른 하나에는 잘게 썬 삶은 달걀이 들어 있었고, 잘게 썬 치즈와 작은 동결 건조 빙어도 있었습니다. 마치 어제 일처럼 생생하게 기억나네요. 와, 정말 운 좋은 개네요. 이것들은 모두 굉장히 가치가 높은 보상들로 보였습니다. 그래서 제가 말했죠. 좋아요. 클리커는 어떤가요? 클리커를 알고 있나요? 네. 개는 클리커 소리가 나면 보상을 받는다는 걸 알고 있어요. 좋습니다. 두 번째로 그녀는 개에게 물어오기(retrieve)를 가르치고 싶다고 했습니다. 하지만 지난 팟캐스트 에피소드에서 제가 훈련의 감정에 대해 이야기했던 것을 기억하시나요? 무엇이 중요한지, 즉 물어오기 훈련 과정의 어떤 부분으로 넘어가기 전에, 저는 이 개가 훈련에 어떤 감정을 가지고 올지 알아야 합니다. 그래서 저는 훈련을 무엇이 중요한지에서부터 잠시 떼어놓습니다. 당신이 알아두는 것이 매우 중요합니다. 당신이 가르치고 싶은 것을 바로 훈련하기 시작해서는 안 됩니다. 중요한 것에서 분리하세요. 예를 들어, 어질리티 훈련에서 개가 도그 워크 끝에서 멈추도록 가르치고 있다면, 저는 그 훈련을 실제 훈련 환경에서 분리해서 계단 끝에서 멈추는 법을 가르칠 겁니다. 개가 엄청나게 좋아할 때까지 반복한 다음, 중요한 훈련 상황으로 가져갈 거예요. 그래서 우리는 훈련을 중요한 환경에서 분리할 것이고, 거기에는 타겟 스틱이 사용되었습니다. 저는 개가 그저 코로 타겟 스틱을 건드리기만 하면 됩니다. 클릭하고 보상을 줄 겁니다. 좋아요. 우리는 그룹으로 이 훈련을 하고 있었어요. 그녀가 타겟을 제시했고 개가 건드렸죠. 그런데 만약 곰돌이 푸의 이요르를 의인화한 개가 있다면, 바로 그 개였어요. 그녀가 로스트 비프 한 조각을 주자 아주 느릿느릿 씹어 먹더군요. 그러고 나서 다시 타겟을 제시했죠. 개는 그녀를 쳐다보는 듯하더니, 매우 느리고 신중하게 고개를 돌려 타겟을 건드렸어요. 그녀가 클릭했고, 개는 다시 그녀를 쳐다봤죠. 그녀가 닭고기를 줬는데도 반응은 똑같았어요. 아무리 엄청난 고가치 보상을 줘도 말이죠. 그래서 세션이 끝날 때 개들을 들여보내고 제가 말했어요. 만약 이게 당신의 최고의 고가치 보상이라면, 내일 당신의 강사에게 귀 꼬집기(ear pinch) 훈련을 받게 될 거라고요. 왜냐하면 개가 강화물에 전혀 관심이 없는 상태에서는 절대로 리트리브(물어오기) 셰이핑을 할 수 없기 때문입니다. 셰이핑은 가치의 이전이니까요. 당신이 사용하는 강화물에 개가 느끼는 가치가 훈련 중인 행동으로 옮겨가는 것이죠. 그래서 물어봤어요. 터그 놀잇감을 좋아하나요? 아니요. 테니스 공은요? 별로요. 알겠습니다. 당신의 개가 이제껏 보고, 경험하고, 먹거나 냄새 맡았던 것 중에 귀가 쫑긋 서고 발끝으로 서게 만들며, 꼬리가 흔들리기 시작하는 무언가가 있나요? 아, 있어요! 그럼요!
containers, chopped up roast beef in this one, chopped up chicken in this one, a chopped up hard-boiled egg in this one, chopped up cheese, and some little freeze-dried minnows. I remember it like it was like, wow, this is one lucky dog. These seem like all pretty outrageously high value rewards. And I said, okay. And what about a clicker? Does you know a clicker? Yeah. He knows a clicker means it's going to get a reward. All right. Number two, she said, I want to teach him to retrieve. But remember in the last podcast episode, I talked about the emotion of training. Before I go to what's important, any part of the shaping process for the retrieve, I need to know what kind of emotion this dog is going to be bringing into the training. And so, I take my training away from what's important. Super important for you to know. You don't just start training what you want to train. Take it away from what's important. So, if I was teaching in agility a dog to stop at the end of a dog walk, I would take that away from training and teach them to stop at the end of a set of stairs until I get a high level of, oh my gosh, this is so good. Then I would take it into what's important. So, we're going to take the training away from what's important and that involved a target stick. So, I want the dog to just touch his nose to the target stick. We're going to click and reward that. Okay. So, we're doing this in a group. She presented the target. The dog touched it. And if ever there was a dog personifying Eeyore from Winnie the Pooh, he did it. She gave him a piece of roast beef and he chewed it up so slowly. Then she presented it. And then the dog kind of was looking at her, looked away very slowly and deliberately touched the target. She clicked. He looked back at her. She gave him chicken. Same response, no matter what kind of outrageous high value reward she gave. So, at the end of the session, we put the dogs away and I said, I'm very confident if this is your outrageous high value reward, your dog will be getting an ear pinch from your instructor tomorrow because there is no way we are going to shape a retrieve when he really doesn't care about the reinforcement. Because shaping is about the transfer of value. The value the dog has for the reinforcement you're using goes into the behavior you're training. So, I said, does he like tug toys? No. Does he like tennis balls? Not really. Okay. Is there anything that your dog has ever seen or done or eaten or smelled that makes his ears come up and he's on the balls of his feet and the tail
12:48
그녀가 말하길, 아이들이 뒷마당에서 비눗방울 놀이를 할 때, 비눗방울 아시죠? 불면 방울이 나오는 거요. 제가 말했죠. 좋아요, 가서 가져와 보세요. 다이소에서 산 비눗방울입니다. 비눗방울을 준비했죠. 자, 여기서 핵심은 최고의 강화제를 준비했다는 겁니다. 이제 분리도 했고, 마커도 준비됐습니다. 이제 환경을 조작해야 합니다. 그래서, 저는 의자 위에 비눗방울을 두고 이렇게 말했죠. '여기서 행동을 형성해 볼 겁니다.' 그리고 말했습니다. 이 개가 비눗방울을 어떻게 대하는지 보고 싶었거든요. 그래서 타겟을 제시했습니다. 그리고 개가 타겟을 건드렸을 때, 제가 말했죠. '이제 강화 과정을 만들어 볼 겁니다.' 그래서 개가 건드리면, '버블(bubbles)'이라고 말해주세요. 개에게는 아무런 의미가 없는 단어였죠. 비눗방울인지도 몰랐을 테니까요. 개가 그분 쪽을 쳐다봤습니다. 그분은 25피트 떨어진 의자까지 달려가서 비눗방울을 집어 들고 몇 번 불기 시작했습니다. 그 개는 제가 본 적 없는 모습으로 돌변했습니다. 그분 머리 위로 뛰어넘고, 악어처럼 입을 쩍 벌리며 비눗방울을 낚아채려 했죠. 그분이 비눗방울을 다시 준비하는 동안, 꼬리를 헬리콥터처럼 휘둘렀습니다. '세상에, 비눗방울이잖아.' 그렇게 비눗방울을 두 번 불어줬습니다. 됐습니다. 다시 제자리에 두세요. 자, 이리로 다시 돌아오세요. 보통 사람들은 이렇게 말하겠죠. 세상에, 개가 정말 좋아하는 걸 찾았네요. 이제 평생 비눗방울로만 훈련하면 되겠어요. 여기에는 두 가지, 아니 어쩌면 세 가지 문제가 있습니다. 첫째, 집안 전체가 끈적거리는 비누 거품으로 뒤범벅이 될 겁니다. 누가 그걸 원하겠어요? 둘째, 결국 개도 지치게 됩니다. 아까 한 행동은 정말 힘을 많이 쓰는 활동이었거든요. 그리고 셋째, 항상 비눗방울을 들고 다니는 게 얼마나 불편하겠습니까? 그래서 우리는 이 엄청난 고가치의 강화제의 가치를 다른 강화제로 옮겨야 합니다. 소위 말하는 '가치 전이(transfer of value)'죠. 그 과정은 이렇습니다. 그분이 타겟을 제시했습니다. 개가 타겟을 건드렸고, 클릭하고 로스트 비프 한 점을 줬습니다. 그리고 개는 천천히 씹어 먹었죠. 두 번째로 타겟을 제시했습니다. '버블'이라고 말하고 바닥을 가로질러 뛰어갑니다. 또 다른 비눗방울이 나왔습니다. 이 개는 두 번 만에 '그래, 저 공을 건드려볼게'라고 생각했죠. 그래서, 그 개는 공을 건드리기 시작했고 스스로 비눗방울 쪽으로 달려가려고 방향을 틀었습니다. 하지만 아무도 비눗방울이라고 말하지 않았죠. 마커가 없으면 보상도 없습니다. 그렇기 때문에 우리가 개가 스스로 보상을 얻게 내버려 두지 않는 것이 매우 중요합니다. 이리 와, 벨라. 다시 건드렸고 약간 움직이긴 했지만, 클릭은 음식을 먹을 수 있다는 신호죠. 그러고 나서 다시 클릭하고 음식을 줬고, 또 한 번 건드리게 한 뒤 비눗방울을 보여줬습니다. 우리는 달려갔고
just starts going? Oh yeah. Oh yeah. She said, when the kids play with the bubbles in the backyard, you know those soap bubbles, you blow and the bubbles come out? I said, all right, go get me some dollar store bubbles. We got the bubbles. Now, here's the key. We've got our best reinforcement. We've got a split. We've got our marker. And now we've got to manipulate the environment. So, I put the bubbles on a chair and I said, we are going to go and shape behavior over here. And I said, I just want to see what this dog's like with bubbles. So, we presented the target. And when he touched it, I said, we're going to create a reinforcement process. So, when he touches it, I want you to say the word bubbles, which is meaningless to this dog. He didn't know they were bubbles. He turned to her. She started running the 25 feet to the chair, picked up the bubbles and doing some streams. This dog turned into something the likes I've never seen before. He was jumping over the woman's head, snapping like an alligator, grabbing those bubbles. And in between when she was loading, his tail was going like a helicopter. Oh my gosh, there's bubbles. So, two streams of bubbles. That's it. Put it back. Come on back over here. Now, what most people would say, oh my gosh, we found something that the dog loves. We're just going to train with bubbles for the rest of his life. There's two problems. Actually, probably three problems with that. Number one, you get that icky soap scum all over your house. Who wants that? Number two, eventually the dog's getting tired. Like that was a pretty exhausting exhibition. And number three, how inconvenient is it to always be packing bubbles? So, what we have to do is take the value of the high outrageous reinforcer and put it into our other reinforcers, aka the transfer of value. And so, this is the process. She presented the target. The dog touched it, click, give a piece of roast beef. And he did his slow chew. Presented the target the second time. We say bubbles, run across the floor, another stream of bubbles. It only took two before this dog said, yeah, I'll touch that ball. So, she started going like hit the ball and then she was turning to sprint to the bubbles on her own. But no one said bubbles. Without a marker, there's no reinforcement. That's why it's just super important that we're not letting a dog just grab reinforcement on their own. Come on back here, Bella. She hit it again and she was kind of moved, but click means you're getting food. Then we did a click and food again, and then another touch and bubbles. Off we go running
15:19
비눗방울이 더 나왔죠. 그게 세션의 끝이었습니다. 나머지 세션에서는 코 터치 5번에 음식 보상을 하나, 그다음엔 비눗방울을 줬습니다. 어쩌면 코 터치 1번에 비눗방울을 줄 수도 있었죠. 그러고 나서 비눗방울을 주기 전에 코 터치를 10번 하게 했습니다. 그러고 나서는 3번 하고 비눗방울을 다시 한 번 줬을지도 모르겠네요. 그다음엔 20번까지 늘렸을 수도 있고요. 결과가 어땠을까요? 하루가 끝날 때쯤엔, 우리가 원했다면 벨라가 나무로 무엇인가를 깎아왔을지도 모른다고 생각합니다. 벨라는 클릭하고 음식을 받는 것에 매우 신이 났습니다. 왜냐하면 음식을 더 많이 받을수록 비눗방울에 더 가까워지기 때문이죠. 아주 곧, 비눗방울은 세션당 한 번, 혹은 일주일에 한 번 정도만 나오게 될 겁니다. 그리고 결국에는 굳이 사용할 필요가 없는 특별한 보상이 될 겁니다. 왜냐하면 음식은 이제 간신히 참을 만한 수준에서 정말 좋은 것으로 바뀌었기 때문입니다. 비눗방울에 더 가까워질 기회를 의미하게 되었으니까요. 따라서 틀을 깨는 사고는 모든 개에게 선형적인 경로가 존재하지 않는다는 것을 의미합니다. 여러분은 보상 과정을 살펴봐야 합니다. 개에게 무엇이 엄청난 자극이 되는지 살펴보고, 비눗방울을 25피트 멀리 배치함으로써 어떻게 흥분을 고조시킬 수 있을지 고민해야 합니다. 우리는 개가 뛰게 만들었고, 생리학적 변화를 일으켰으며, 게임 안에 또 다른 게임을 만들어 저와 함께 일하는 것이 신난다는 것을 가르쳤습니다. 저와 함께 움직이는 것이 신난다는 것을요. 그런 개를 훈련하는 것은 물건을 가져오게 하거나 뒷다리를 가리키게 하는 등 무엇이든 가능했을 겁니다. 이제 필요한 것은 훌륭한 훈련 계획을 세우는 것과 자신의 환경이 어떤지, 방해 요소가 무엇인지 파악하는 것뿐입니다. 원하는 행동이 무엇인지, 그리고 가장 적절한 강화물은 무엇인지 알아야 합니다. 왜냐하면 지금 시점에서는 선택지가 아주 많기 때문이죠. 좋습니다. 긴 시리즈였는데, 질문을 보내주신 여러분께 감사드립니다. 덕분에 제가 반려견 훈련을 시작하게 된 계기와 제가 모든 수강생을 어떻게 가르치는지 명확히 정리할 수 있었습니다. 그리고 이 기회를 빌려 여러분께 이번 10월에 아주 특별한 것을 공개할 예정이라는 소식을 전하고 싶습니다. 여러분은 좋아하는 영화를 볼 수 있는 비디오 스트리밍 서비스에 익숙하실 겁니다. 그리고 매달 멤버십에 가입해서 시리즈나 영화, 다큐멘터리를 시청할 수 있죠. 저와 저희 팀은 반려견 훈련에 관심이 있는 분들을 위한 서비스를 준비 중이며, 이름은 '도기 플릭스(Doggy Flicks)'입니다. 도기 플릭스는 10월에 100% 무료로 공개될 예정입니다. 제 팟캐스트를 계속 들어오셨다면 아시겠지만, 저는 제가 아는 것을 공유하는 것을 좋아합니다. 저는 반려견과 보호자들을 위해 더 나은 삶을 만들어가는 것을 좋아하니까요. 그래서 저는 최소 일 년에 한 번씩, 도기 플릭스의 모든 멤버에게 교육용 영상 시리즈를 제공하기로 약속했습니다. 그 첫 번째 시리즈는 10월에 공개됩니다.
and more bubbles. So, that was the end of the session. The rest of the sessions is we started going five nose touches with food and one bubbles. And then maybe we do one nose touch and bubbles. And then we went to 10 nose touches before we got the bubbles. And then we might've done three and then one again with bubbles. And then we might go to 20. And guess what? By the end of the day, I think Bella would have whittled us something out of wood if we wanted. She was so excited about the click and getting food because the more she got of that, the closer she got to bubbles. Very soon, the bubbles would only have to come out, I don't know, once a session or once every week. And eventually it would be just a special reinforcement that you wouldn't have to use because food had gone from barely tolerable to really good because they meant the chance to get closer to bubbles. So, outside the box thinking means there isn't a linear path for all dogs. You look at the reinforcement process. You look at what is outrageous to the dog and you look at how can you grow excitement by putting the bubbles 25 feet away. We got the dog to run, changing the physiology, building in a game within a game that working with me is exciting. Moving with me is exciting. Shaping that dog to retrieve an object or point at her back leg, anything would have been possible now. All that would be required is a great training plan and knowing what's your environment, what are the distractions, what is the behavior you want to create, and what's the most appropriate reinforcement because you have so many at this point. Okay. This has been a long series and I thank you guys for your questions that has led me to put together really defining my origin in dog training, defining how I train all of my students. And I want to take this opportunity to share with you this October, I'm going to be rolling out something very special. You're familiar with video streaming, like your favorite movies. And every month there's a membership that you can join and just watch series or movies, some documentaries. Well, my team and I are rolling out one for people who have an interest in training their dogs and it's called Doggy Flicks. So, Doggy Flicks is going to be available in October, 100% free. Because if you've been following my podcast, you know, I love to share what I know. I love to create a better life for dogs and their humans. And so, I'm making a commitment that a minimum of once a year, all the members of Doggy Flicks will receive an educational video series. And the first one will be coming out in October.
18:03
지금 바로 doggyflicks.com에 접속하세요. D-O-G-G-Y-F-L-I-X.com입니다. 대기 명단에 등록하실 수 있습니다. 대기 명단에 등록하신 모든 분께는 도기 플릭스가 출시될 때 알림이 발송될 겁니다. 멤버십을 생성하게 될 텐데, 기억하세요. 지금 당장은 여기서 제가 이야기한 모든 내용을 계속 배우고 싶은 분이라면 누구나 100% 무료로 이용할 수 있습니다. 이 내용을 실제 반려견 훈련에 어떻게 적용할까요? 그것이 바로 도기 플릭스의 첫 번째 시리즈 주제가 될 것입니다. 그러니 지금 링크를 클릭하세요. 지금 바로 쇼 노트를 확인하시면 대기 명단으로 안내해 드릴게요. 저희 팀과 제가 무엇을 준비하고 있는지 정말 기대가 큽니다. 전 세계 반려인들에게 놀라운 교육을 제공하기 위해 'Doggy Flicks'라는 새로운 플랫폼을 만들고 있습니다. 오늘 팟캐스트를 통해 어떤 소중한 정보나 깨달음, 유익한 내용을 얻으셨는지 정말 궁금합니다. 유튜브에 오셔서 댓글을 남겨주세요. 여러분의 피드백을 받는 것을 정말 좋아하거든요. 여러분은 정말 최고예요. 다음 주 'Shaped by Dog'에서 다시 뵙겠습니다.
You can go to doggyflicks.com right now. That's D-O-G-G-Y-F-L-I-X.com. And there's a waiting list. Everybody who joins that waiting list will get a notification when Doggy Flicks is available. You will create a membership. Remember, right now, this is 100% free for anybody who wants to continue to learn everything that I've talked about here in the podcast. How do we apply this to actual dog training? So, that's going to be the topic of the first series in Doggy Flicks. So, go ahead, click the link that's in the show notes right now, take you over to the waiting list. I'm just super excited for what my team and I are putting together to create this brand new venue of Doggy Flicks to bring amazing education to dog lovers everywhere. I'd love to know what gems, aha moments, and nuggets you received from today's podcast. Please jump on over to YouTube and leave a comment for me because I love getting your feedback. You guys are amazing. I'll see you next time right here on Shaped by Dog.