Common Misconceptions Around Shaping: Why You May Find Dog Training Frustrating #261 #podcast
Susan GarrettDogs That · 영상 · 22분
셰이핑에 관한 흔한 오해들: 반려견 훈련이 좌절스럽게 느껴지는 이유 #261 #팟캐스트
Common Misconceptions Around Shaping: Why You May Find Dog Training Frustrating #261 #podcast
0:00
Shaped by Dog 259회에서 저는 셰이핑에 관해 모든 것을 이야기했습니다. 그리고 문제 해결 에피소드를 원하시면 질문을 남겨달라고 했었죠. 정말 많은 질문을 받았습니다. 정말 많은 질문을 받았습니다. 그래서 그 모든 것들을 포함해 더 많은 내용을 다루려고 합니다. 그 질문들을 보면서 Shaped by Dog에서 셰이핑에 관한 에피소드를 하나 더 할 필요가 없다는 걸 깨달았습니다. 여러 편을 다뤄야겠더군요. 자, 이제 그 첫 번째 시간으로 셰이핑에 관한 여러분의 질문에 답해 드리겠습니다. 첫 번째 시간으로 셰이핑에 관한 여러분의 질문에 답해 드리겠습니다. 안녕하세요, 수잔 개럿입니다. Shaped by Dog에 오신 것을 환영합니다. 와, 정말 훌륭하고 많은 질문을 받았습니다. 그런데 셰이핑에 대해 오해하고 계신 부분이 있다는 걸 알게 되었습니다. 그 이유는 대부분의 팟캐스터나 블로그, 그리고 아마도 저 자신조차도 과거에는 셰이핑을 설명할 때 '점진적 접근(successive approximations)'이라는 표현을 쓰기 때문일 겁니다. 즉, 개가 작은 행동 하나를 하면 보상해주고, 그다음 또 다른 작은 행동으로 이어지고, 또 그다음 행동으로 이어지도록 하는 것이죠. 그래서 결국 개가 우리가 원하는 행동을 하도록 만드는 것입니다. 하지만 트레이너로서 제가 보기에 이 방식은 개에게 엄청난 좌절감을 줄 수 있습니다. 우리는 항상 더 많은 것을 기대하기 때문이죠. 그러면 개는 스트레스를 받게 되고 많은 불안감과 압박을 느끼게 됩니다. 또 다른 문제는, 우리가 정말로 원하는 것이 무엇인지 우리 자신과 개 모두에게 명확하지 않을 때 조금 더, 조금 더, 조금 더를 반복하면서 이른바 '속임수 행동(cheat behaviors)'이 많이 만들어진다는 것입니다. 그래서 셰이핑이 약간 엉망이 되어버리곤 합니다. 제 개들을 예로 들어볼게요. 90년대 초반에 저는 제 잭 러셀 테리어와 보더 콜리에게 똑같은 행동을 셰이핑하고 있었습니다. 제 목표는 복종 훈련을 위한 '고 아웃(go out)', 유럽 친구들이 부르는 방식으로는 '센드 어웨이(send away)'를 가르치는 것이었습니다. 그래서 저는 그들에게 셰이핑을 시작했습니다. 클리커와 간식을 한 움큼 들고, 제가 벽에 아주 가까이 서서 점진적 접근법을 선형적으로 사용했습니다. 벽을 쳐다볼 때마다, 더 오래 쳐다볼 때마다, 혹은 벽에 더 가까이 다가갈 때마다 클리커를 눌러 보상을 주었죠. 결국 두 마리 모두 벽에 앞발을 올리게 만들었습니다. 속도는 더뎠고 두 개 모두 일에 대한 의욕이 강한 상태였습니다. 그래서 행동에서 불안함과 허둥대는 모습이 많이 보였지만, 어쨌든 목표는 달성했습니다. 두 마리 다 벽에 앞발 두 개를 올리게 했으니까요. 하지만 제 훈련 계획의 다음 단계는, 이른바 '투퍼(twofer)'라고 불리는 것을 원했습니다. 즉, 개가 행동을 완전히 이해했다는 것을 보여주기 위해, 첫 번째 행동에는 클릭을 하지 않고 두 번째 행동에만 클릭을 하는 것이었죠. 그래서 제 잭 러셀 테리어인, 트위스터에게 먼저 시도했습니다. 투퍼를 도입할 당시 저는 개로부터 한 걸음 정도 떨어져 있었고, 개들은,
In Shaped by Dog episode 259, I talked all about shaping. And I said, if you wanted me to do a troubleshooting episode, just leave me some questions. And I got a lot of questions. So, all of that and more. And from those questions, I realized I don't need to do another shaping episode here on Shaped by Dog. I need to do several. And here we go with the first one, your questions answered all about shaping. Hi, I'm Susan Garrett. Welcome to Shaped by Dog. And wow, did we get a lot of really, really great questions. And I realize that there's a misconception about shaping. And I think it's because most podcasters, most blog posts, most people, and probably including myself at one point, they describe shaping with the phrase successive approximations, meaning you reward the dog for doing one little behavior. And that leads you to another little behavior and then another little behavior and another. So, the dog will eventually do what you want them to do. But from where I see things as a trainer, that can potentially lead to a ton of frustration for the dog because we're always expecting more. And that stresses and creates a lot of anxiety for dogs and a lot of pressure. The other problem with that is we get a lot of what's called cheat behaviors built into that a little bit more and a little bit more and a little bit more when what we really, really want isn't super clear both in our minds and in the dog's mind. So, shaping kind of gets a little bit screwy. I'm going to give you an example with my own dogs. Back in the early 90s, I was shaping both my Jack Russell Terrier and my Border Collie to do the exact same behavior. And my goal was to teach them to do a go out for obedience, a send away as my European friends might call it. So, I started shaping them linearly using successive approximations with a clicker and a handful of cookies, me standing very, very close to the wall, clicking them for looking at it, for looking at it longer, for maybe getting closer to the wall. Eventually I got them to put their paws on the wall. It was slow and both those dogs were driven to work. So, I got a lot of anxiousness in their behaviors, a lot of franticness, but I got the job done. I got both of them to hit their two paws on the wall. But the next step in my training plan was I wanted what was referred to as a twofer, meaning I want the dog to show me they understand the behavior by I'm not going to click the first one and I'll click the second one. And so, my Jack Russell, I did first Twister. Now, when I was adding a twofer, I was maybe a stride away from the dog and they were
2:51
저를 떠나 벽을 터치하러 가던 중이었습니다. 별도의 신호는 주지 않았고, 아직 훈련 초기 단계였습니다. 트위스터는 저에게서 떨어져 벽으로 가서 앞발을 올린 뒤, 간식을 받으러 다시 돌아섰습니다. 마치, 녀석의 머릿속에서 생각이 돌아가는 게 눈에 보이는 것 같았습니다. 녀석은 '어, 클릭 소리가 안 났네? 나를 못 봤나?' 하는 듯했죠. 그래서 다시 돌아서서는 더 높은 벽 위치를 더 강하게 짚으며 마치 이렇게 말하는 것 같았습니다. '보세요, 제가 제대로 했잖아요.' 저는 웃음이 터져 굴러버렸고, 바로 클릭하고 보상을 주었습니다. 그래서 그, 이후로는 제가 클릭을 안 하면 벽에서 떨어지려 하지 않았습니다. 계속 벽을 쾅쾅 두드렸죠. 제 말은, 그게 바로 제가 찾던 반응이었다는 겁니다. 그래서 무엇이든 셰이핑하는 데 뛰어난 제 보더 콜리에게도, 다음으로 투퍼를 시도하려고 했습니다. 그런데 녀석은 벽을 터치했을 때 제가, 클릭하지 않았는데도 트위스터처럼 행동하지 않았습니다. 벽에서 떨어져 저에게 걸어오지 않은 거죠. 녀석은 앞발을 벽에 둔 채 그대로 머물며 마치 이런 표정으로 저를 돌아보더군요. 당신이 요청한 대로 하고 있어요. 당신이 여기서 무슨 짓을 했는지 안 보여요? 말하자면, 언제 클릭할 건가요? 그걸요? 더 오래 기다려 줄까요? 그런데 녀석은 그냥 내가 포기하기를 기다렸어요. 그래서 나는 '알겠어, 알겠어, 알겠어'라고 생각했죠. 내가 무슨 짓을 했는지 알겠더라고요. 매번 똑같은 지점에서 클릭했거든요. 녀석은 그 게임이 벽에 닿는 것이라고 생각하는 거죠. 그래서 머릿속으로 이 모든 것을 고민하고 있었어요. 그리고 나는 말했죠. 이제 내가 할 일은 녀석이 벽에서 떨어질 때 클릭하는 것이라고요. 그러면 녀석이 '아니, 아니, 아니, 닿는 게 아니라 떨어지는 거야'라고 이해할 테니까요. 그래서 나는 기다렸고, 준비하고 있었어요. 그런데 녀석이 한 행동은 벽 위쪽을 더 쳐다보더니 움직여서 자신의 팔꿈치와 앞발을 벽에 대고 어깨너머로 나를 쳐다보며 '음, 내 앞발만으로는 부족하다면, 내 몸을 더 벽에 대길 원할지도 몰라'라고 생각하는 것 같았어요. 그래서 나는 '그래, 그래' 했죠. 녀석은 다른 행동을 제시할 수 있어요. 이것이 바로 점진적 접근법(successive approximations)의 핵심이죠. 녀석들이 클릭할 수 있는 무언가를 제시할 때까지 기다리는 거예요. 그래서 녀석이 그때 했던 행동은 아래를 내려다보고 뒷발을 최대한 벽 쪽으로 바짝 끌어당기는 것이었어요. 그래서 마치 프리즈(freeze) 자세를 취한 것처럼 온몸을 벽에 밀착시켰고, 그 순간 나는 너무 웃겨서 뒤로 넘어졌어요. 정말 웃겼거든요. 마치 어제 일처럼 생생하게 기억나는데 아마 1993년쯤이었을 거예요. 그래서 점진적 접근법을 사용할 때는 잘못될 수 있는 수많은 변수가 있어요. 만약 일에 대한 의욕이 강한 개라면 포기하지는 않겠지만, 조금 좌절감을 느낄 수도 있어요. 그리고 그런 의욕이 높은 개들에게서 느끼는 좌절감 때문에
leaving me and going and touching the wall. There was no cue given. It was still early in the process. So, Twister left me, went up, hit the wall, turned around to come back to get the cookie. And it was almost as if I could see the wheels turning in her head. And she went, wait, I didn't get the click. Did she not see me? And she turned around and went like higher up on the wall with more force as if to say, look, I did it. And of course, I fell over laughing, clicked her and gave her her reinforcement. So, from then on, if I didn't click her, she wouldn't leave the wall. She would just keep pounding on the wall. I mean, that was exactly what I was looking for. And so, my Border Collie who was brilliant at shaping anything, I was going to try her with a twofer next. And here's what she did is she touched the wall and I didn't click, but she didn't do like Twister did. She didn't come off and start walking towards me. She stayed there with her paws on the wall. And she kind of looked over her shoulder at me like, I'm doing what you asked. Do you not see what you've done here? Like, when are you going to click that? Do you want me to hold longer? And she just waited me out. And I thought, okay, okay, okay. I see what I've done. I clicked at the exact same spot every time. And she thinks the game is be in contact with the wall. So, I'm thinking through this all in my brain. And I said, what I'm going to do now is I'm going to click her when she comes off the wall. So, she understands, no, no, no, it's touching, come off. So, I'm waiting, I'm ready. But what she did is she looked up the wall further and then she moved so that she could put her elbows and her paws on the wall and looked over her shoulder at me thinking, well, if my paws aren't enough, maybe you want more of my body on the wall. And I'm like, okay, okay. She can offer something else. This is what successive approximations are about. You wait until they offer something that you can click. And so, what she did then was she looked down and she scooted her rear paws as close to the wall as she could. So, she got her entire body laying on the wall as if she was in the freeze position, at which point I fell over laughing because it was hysterical. I remember it like it was yesterday and it was probably somewhere in, I don't know, 1993 that that happened. And so, with successive approximations, there's just like a gamut of things that could go wrong. And if you have a dog who's driven to work, they may not give up, but they may get a little bit frustrated. And with that frustration with the
5:23
짖는 소리가 나올 수도 있어요. 아니면 다른 방식으로 불안함을 드러낼 수도 있죠. 예를 들어 제자리에서 돌다가 행동을 하거나, 살짝 깨무는 시늉을 하고 나서 행동을 할 수도 있고요. 물론 모든 개가 다 그런 것은 아니죠, 그렇죠? 그래서 그런 의욕이 없는 개들의 경우에는, 개들은 '어, 무슨 말인지 모르겠네' 하고 그냥 가서 눕거나, 혹은 귀를 긁거나 여러분을 빤히 쳐다볼지도 모릅니다. 그래서 많은 사람들이 셰이핑을 연속적 근사법(successive approximations)이라고 생각합니다. 하지만 그 생각은 이제 버리시길 바랍니다. 인터넷 세상으로 보내서 자유롭게 풀어주세요. 왜냐하면 제가 여러분께 생각해보시길 바라는 것은, 지난 에피소드에서 설명해 드린 틀을 깨는 셰이핑이기 때문입니다. 이는 여러분의 반려견이 이미 알고 있는 행동 블록을 사용하는 셰이핑입니다. 자, 지금 이 방송을 듣고 계신 여러분 모두, 제가 알기로는 여러분이 반려견에게 보상을 주었던 행동들이 분명히 있을 겁니다. 간식을 주었거나, 무언가를 할 수 있도록 허락해주었거나, 장난감을 주었을 겁니다. 그러니 여러분께 부탁드리고 싶은 것은, 운전 중이시라면 그냥 생각만 해보시고, 나중에 여유가 될 때 적어보세요. 반려견의 행동 블록은 무엇인가요? 반려견이 보상을 받았던 행동들을 말이죠. 예를 들면 앉기, 엎드리기, 서기와 같은 것들입니다. 많은 강아지들이 이 세 가지 행동으로 보상을 받았을 거라 확신합니다. 그런 것들이 유용하게 쓰일 수 있는 행동 블록입니다. 예를 들어, 만약 여러분이 강아지에게 배로 기어가는 법을 가르치려는데, 그 강아지가 엎드려서 간식을 받은 적이 단 한 번도 없다면, 과연 어떻게 될까요? 전에 한 번도 보상받은 적 없는 엎드린 자세를 강아지가 스스로 취하게 만드는 것이 얼마나 어려울지 상상이 가시나요? 그러니 어딘가 일지에 여러분의 반려견이 가진 모든 행동 블록을 적어두세요. 그래야 새로운 행동을 셰이핑할 때 무엇을 활용할 수 있는지 알 수 있으니까요. 자, 이제 질문을 받아보겠습니다. 그전에 먼저 지난 에피소드에서 다루었던 내용 몇 가지를 상기시켜 드릴게요. 성공적인 셰이핑 세션을 갖는 방법이죠. 그 내용을 조금 더 구체적으로 보충해 드리겠습니다. 우선, 강아지가 정말 환장할 정도로 좋아하는 간식, 즉 보상의 우선순위(hierarchy of reinforcement)를 갖추는 것이 매우 중요합니다. 강아지가 '오 세상에!' 할 만큼 반응하는 보상을 가지고 계신가요? 네. 그리고 보상 수단이 있나요? 제가 위계라고 말할 때, 예를 들어 이런 개를 키운다면, 테이터 샐러드(Tater Salad)의 경우, 제가 가장 좋아하는 보상을 사용하면 너무 흥분할 수도 있는데, 이게 좀 재밌는 점입니다. 그 녀석은 15개월 된 구조견으로 여기 처음 왔을 때는 정말 훈련하기 싫어하던 차분한 개였거든요. 그런데 지금은 너무 흥분해요. 그래서 저는 더 정교한 행동을 만들어갈 때 그 녀석에게는 가치가 낮은 보상을 사용하곤 합니다. 이것이 보상의 위계입니다. 그리고 '우리 개는 정말 훈련하기 싫어해요'라고
higher drive dogs, you might get vocalization. You might get them showing anxiety in other ways, like you might get a spin and then a behavior, or you might get a little nip and then a behavior. Now, not all dogs are like that, are they? So, with dogs that aren't driven like that, they're going to go, Oh, I don't get it. And they might just go and lay down or they might like scratch their ear or stare at you. And so, many people think of shaping as successive approximations. And that is something I want you to release. Send that out into the interwebs and let it be free to roam. Because what I want you to think about it, what I described in the last episode is outside the box shaping. It's shaping by using behavioral blocks that your dog knows. Now, every single one of you listening to this, I know there are behaviors that you have reinforced in your dog that you have either given them a cookie, you've given them permission to do something, you've given them a toy. So, I'd like you, if you're driving in your car, just think about this. But when you get a chance, write down, what are your dog's behavioral blocks? Things that they have earned reinforcement for. It could be something like a sit, a down, a stand. I'm sure many dogs have received reinforcement for those three things. Those are behavioral blocks that could come in handy. For example, if you were trying to teach a dog to crawl on their belly, yet they'd never ever received a cookie for lying down. Can you see how difficult it would be to get that dog to offer a down position when it's never been a previously reinforced position? So, list all of the behavioral blocks in a journal somewhere so you'll know what you have to work with when you're trying to shape a behavior. Okay. I'm going to get your questions. I'm going to first remind you of some of the things we talked about in the last episode, how to have a successful shaping session. And I'm going to flesh that out a little bit more. So, super important that you have a hierarchy of reinforcement, meaning food that your dog goes cuckoo for Cocoa Puffs about. Do you have reinforcement that your dog goes, oh my gosh, yes. And do you have reinforcements? When I say a hierarchy, if you have a dog like, for example, Tater Salad, if I use his number one reinforcement, he might get too crazy, which is kind of funny because he was a pretty laid back dog who really didn't want to work when he arrived here as a 15 month old rescue dog, but he gets so jacked up. So, I would use like lesser value reinforcements for him when
7:51
말씀하시는 분들이 바로 거기서부터 시작해야 합니다. 강아지가 음식을 먹도록 보상의 위계를 만드는 것이죠. 그 점에 대해서는 에피소드 259에서 이야기했습니다. 둘째로, 위치 지정 보상 마커가 필요합니다. '쿡(cook)', '서치(search)', 그리고 저는 '차우(chow)'도 사용하시길 적극 권장합니다. '차우'는 그릇 안에 간식이 있다는 뜻입니다. 그릇에서 간식을 꺼내 먹을 수 있죠. 혹은 생식을 하시는 분들이라면 저는 그릇에 한 숟가락 정도의 음식을 담아둡니다. 이제 그릇 옆에 간식을 던져주고 '서치'라고 하면, 그들은 그릇에 있는 음식을 먹으러 가선 안 됩니다. 오직 '차우'나 여러분이 정한 다른 위치 지정 보상 마커를 들었을 때만 그릇에서 음식을 먹어야 합니다. '쿡'은 입으로 직접 받아먹는 것, '서치'는 바닥에서 찾는 것을 의미합니다. 그리고 '서치'는 타겟팅 훈련을 할 때 아주 좋은 신호입니다. 기억하시나요? 에피소드 259에서 담요 위로 강아지를 유도하는 훈련을 했었죠. 그 담요를 점점 작게 줄여서 이제는 앞발 타겟팅이 되었습니다. 자, 강아지가 타겟 위에 앞발을 올리는 것에 더 큰 가치를 느끼게 하려면 '리셋 쿠키(reset cookie)'라는 것을 활용할 겁니다. 그건 이런 식입니다. 여러분이 신호를 주죠. '서치'라고 말하면 강아지에게 타겟에서 내려와 바닥에 있는 간식을 찾아도 된다고 알려주는 겁니다. 저는 그 간식을 강아지 뒤쪽으로 살짝 던져서 강아지가 다시 돌아와서 다시 할 수 있게 만듭니다. 그렇게 하면 강아지는 보상을 얻고 다시 제자리로 돌아오게 됩니다. 훈련의 흐름을 끊지 않으면서도 강아지가 계속해서 자발적으로 훈련에 참여하도록 만드는 아주 좋은 방법이죠. 이런 식으로 보상의 위계와 위치 지정 마커를 결합하면 훨씬 더 효과적인 훈련이 가능합니다. 여러분의 개가 어떤 보상에 어떻게 반응하는지 잘 관찰해 보세요. 그리고 그 데이터를 바탕으로 훈련 계획을 조정해 나가시는 겁니다. 타겟을 매우 쉽게 찾게 하세요. 이제 경험이 더 많은 반려견에게는 정말 도전적인 과제를 주고 싶습니다. 저는 쿠키를 제 뒤로 던질 수도 있어요. 그러면 반려견은 다시 돌아와서 타겟을 향해 저를 마주 보는 방법을 알아내야 하죠. 이제 중간 단계의 반려견이라면, 그 두 지점 사이 어딘가에 던져주면 됩니다. 그러니 리셋 쿠키는 반려견이 위치 기반 강화 마커 'chow(차우)'를 이해하고 있을 때만 가능합니다. 정말 중요합니다. 'It's your choice(네 선택이야)' 말이죠. 만약 'It's your choice'를 이해하지 못하는 반려견을, 'It's your choice'를 통해 훈련시키려 한다면, 반려견은 손에 든 간식에 너무 집착해서, 무엇을 시도할지 생각할 수 없게 될 겁니다. 또한, 반려견이 'It's your choice'를 잘하지 못한다면, 위치 기반 강화 마커인 'chow'를 사용할 수 없어요. 왜냐하면 제가 만약 반려견이 직선으로 뒷걸음질 치길 원한다면, 제가 할 수 있는 방법 중 하나는 쿠키가 든 그릇을 뒤쪽 어딘가에 두는 거예요. 그래서 몇 걸음 뒤로 갔을 때, 강화하는 방법 중 하나로 앞다리 사이로 쿠키를 굴려주는 거죠. 하지만 그냥 'chow'라고 말해서 반려견이 뒤로 돌아
I'm shaping a more precision behavior. So, hierarchy of reinforcement. And for those of you who say, my dog really doesn't want to work, that is where you're starting. You're creating a hierarchy of reinforcement so that your dog will take the food. And I spoke about that in episode number 259. Second, you're going to have those location specific reinforcement markers. Cook, search, and I really encourage you to use chow as well, which is there's a cookie in the bowl. You can take the cookie out of the bowl. Or for those of us who are raw feeders, I put a spoonful of the food in the bowl. Now, if I throw a cookie beside the bowl and I say search, they aren't to take their food from the bowl. Only take the food from the bowl if they hear chow or whatever location specific reinforcement marker you're going to use. So, cook, come into your mouth, search, look for it on the floor. And search is a great cue when we're working to create a targeting. Remember in episode number 259, I had you shape the dog onto a blanket. We made that blanket smaller and smaller. Now we have a paw target. Well, the way we're going to build in more and more reinforcement for that dog finding value in putting their paws on a target is we're going to use what's called a reset cookie. So, that's where you're cute. You would say search, which tells the dog you can get off and look for a cookie on the floor. And I'll throw that cookie a little bit behind my dog so that they come back up and they find that target super easy. Now for a dog that's more experienced, I really want to challenge them. I might throw the cookie behind me. So, then they have to figure out how to come around and face me again on that target. Now your dog in between, you're going to throw somewhere in between those two spots. So, the reset cookie is only possible if you have a dog who understands that location-specific reinforcement marker search. Super important. It's your choice. If you are trying to shape a dog who doesn't understand it's your choice, the dog is going to be so obsessed with the food in your hand, they're not going to be able to think about what to offer. Also, if your dog doesn't have really good, it's your choice. You can't use the location-specific reinforcement marker chow because if I say, want my dog to back up in a straight line, one of the things I might do is put a bowl with a cookie in it somewhere behind so that as I get a couple steps, one place I might reinforce that dog is by rolling a cookie between their front legs. But I also might just say chow so they turn around and get the cookie
10:22
바로 뒤에 있는 쿠키를 먹게 할 수도 있습니다. 그게 뒷걸음질할 때 비뚤어지지 않게 도와줄 수 있죠. 그래서 'chow'의 사용은, 제가 지금 강아지를 훈련할 때 거의 매일 사용합니다. 훈련 중에 어딘가에 'chow'를 사용하곤 하죠. 'It's your choice' 없이는 불가능한 일입니다. 자, 크레이트 게임에 대해 이야기해 보죠. 크레이트 게임과 핫 존(Hot Zone)이 있는데, 둘 중 하나만 해도 된다고 말씀드렸지만, 사실 크레이트 게임의 모든 단계를 거치면 훨씬 더 많은 행동 블록을 갖게 됩니다. 즉, 크레이트를 보거나 켄넬로 들어가라는 신호를 주면 크레이트로 달려가는 반려견을 만들 수 있다는 뜻이죠. 그리고 그것은 반려견이 보호자 곁을 떠나 이동하는 것에 대해 강화받았다는 것을 의미하는 행동 블록입니다. 나중에 거리감이 필요한 다른 훈련을 할 때 유용하게 쓰일 수 있는 아주 긴 거리의 훈련입니다. 그러니 크레이트 게임과 핫 존을 활용하세요. 저는 개인적으로 크레이트 게임에 집중하는 것을 추천하지만, 물론 둘 다 병행하셔도 좋습니다. 왜냐하면 핫 존을 이용하면 한 번에 여러 마리의 개를 동시에 훈련할 수 있기 때문입니다. 짧은 세션으로 진행하세요. 자, 오늘 이후에 여러분이 해주셨으면 하는 일이 있습니다. 1분 이하의 짧은 세션 5번을 꼭 실천해 보세요. 그 세션을 영상으로 촬영하고 나중에 무엇을 배웠는지 제게 알려주세요. 그리고 기억하세요. 제가 드린 숙제는 그것뿐만이 아닙니다. 지금 여러분의 반려견이 가지고 있는 모든 행동적 결함을 모두 적어보라고 했습니다. 6번 항목은 5번 항목과 관련이 있습니다. 쉐이핑 훈련을 꼭 영상으로 찍어서 다시 확인해 보세요. 제가 여러분께 확인하라고 하고 싶은 것은 두 가지입니다. 첫 번째는 여러분이 기대했던 것과 실제 개가 한 행동이 무엇인지 확인하는 것입니다. 이제 영상을 다시 볼 때는 개를 무시하고 여러분 자신을 살펴보세요. 여러분은 무엇을 하고 있었나요? 왜냐하면 개가 했던 행동과 여러분이 보상을 제공한 방식, 혹은 여러분이 서 있던 자세나 시선이 향하던 곳 사이의 연관성을 발견하게 될 것이기 때문입니다. 그러니 여러분의 테크닉을 비판적으로 바라보세요. 친구 여러분, 그것이 개가 성공하느냐 마느냐에 가장 큰 영향을 미치기 때문입니다. 선행 조건을 어떻게 배치하셨나요? 지난 에피소드에서 이야기했지만 정말 중요한 부분입니다. 예를 들어, 개에게 후진을 가르친다면 저는 아마 바닥에 무릎을 꿇고 시작할 것입니다. 만약 서 있는 상태에서 개에게 후진을 시킨다면, 개는 위쪽의 무언가를 생각할지도 모릅니다. 하지만 여러분이 바닥에 있으면 개는 조금 더 낮은 위치의 무언가를 생각하게 될 것입니다. 게다가 보상을 주는 위치를 개의 다리 사이로 더 정확하게 조절할 수 있습니다. 그렇게 하면 개가 더 적극적으로 후진하도록 유도할 수 있죠. 즉, 선행 조건 배치는 단순히
right behind them. That might help them to not go in a crooked line when they're backing up. So, the use of chow, I use it pretty much every day when I'm training my puppy right now. I will use chow somewhere in our training. And that's just not possible without it's your choice. All right, crate games. Crate games and hot zone. Now I said you could do one or the other, but I got to tell you, if you work through all the stages of crate game, you have so many more of those behavioral blocks. Meaning, you have a dog who will run to their crate when they see the crate or when you give them the cue to go in their kennel. And that is a behavioral block that means your dog has been reinforced for leaving you and traveling a great distance to do something that will come in handy when we want to do other things at a distance. So, crate games and hot zone. I would really focus on crate games, but you can do them both for sure. Because hot zone allows you to train more than one dog at one time. Short sessions. Okay. Here's what I'd like you to do after today. I want you all to commit to doing five sessions that are one minute or less. Video those sessions and come back and tell me what you learned. And remember, that's not the only piece of homework I gave you. I also asked you to write down all the behavioral blocks you know your dog has right now. And number six is related to number five. Please video your shaping and review the shaping. And there's two things I want you to look for. Number one, you're going to look at what did you expect and what did the dog do? Now you're going to go back and you're going to ignore the dog when you look at the video the next time. What did you do? Because you're going to see a connection in between what the dog did and how you delivered the reinforcement or how you were standing or where you were looking. So, really be critical of your mechanics because that my friend has the biggest impact on whether the dog has success or not. How did you arrange your antecedents? I spoke about this in the last episode, but it's just so important. Now, if I was teaching the dog to back up, I would probably start with myself kneeling on the ground. If you're getting a dog to back up and you're standing up, then the dog might be thinking of things up here. When you're on the ground, the dog is going to be thinking a little bit lower. Plus your placement of reinforcement can be more exact right between your dog's legs. So, that's encouraging them to back up more. So, the antecedent arrangements, it's not just what
12:46
주변의 다른 산만 요소들이 무엇인지, 보상을 어디에 배치하는지, 혹은 도구를 어떻게 잡고 있는지에 관한 것입니다. 아니면 당신이 앉아 있거나 서 있거나 무릎을 꿇고 있는 자세가 어떤지 등이요. 그런 요소들이 개가 지금 당신이 요구하는 행동의 구성 요소를 이해하는 능력에 어떤 영향을 미치고 있을까요? 왜냐하면 올바르게만 한다면, 개에게 올바른 반응은 아주 명확해야 하기 때문입니다. 그리고 개가 무언가를 스스로 해보려고 애쓰거나 마치 안드로메다에서 온 것처럼 엉뚱한 행동을 하길 기다리는 대신, 점진적 근사치 과정이 때로는 그렇게 보이기도 하죠. 당신과 개 모두 그 과정에 몰입하게 되는데, 빠르게 진행되기 때문입니다. 물론 훈련 중에 잠시 소강상태가 있을 수 있지만, 그런 순간은 아주 드뭅니다. 또한 만약 당신의 셰이핑(shaping) 훈련이 혼란스럽고 정신없어 보이는데, 갑자기 개가 클릭 소리와 간식을 받는다면, 그것은 결코 올바른 방향으로 나아가는 것이 아닙니다. 왜냐하면 당신이 만들고 있는 그 행동 안에 그런 정신없는 불안함까지 함께 심어주게 될 것이기 때문입니다. 누가 그런 걸 원하겠어요? 우리 중 누구도 개가 정신없거나 불안해하는 것을 원치 않습니다. 8번, 개에게 말을 걸거나 도움을 주고 싶은 충동을 참으세요. 만약 개가 멈칫한다면, 개가 생각할 시간을 주세요. 그래도 안 된다면, 한 30초에서 1분 정도 지나서도 개가 아무것도 하지 않는 것 같다면, 그냥 중단하고 핫 존(hot zone)으로 돌아가서 선행 조건을 다시 조정하세요. 개가 멈췄을 때 당신이 말로 신호를 주거나 몸짓으로 유도하거나 손가락으로 가리키거나 정식 명령을 내려서 도와준다면, 그것은 셰이핑을 하는 것이 아니라 지시를 내리는 것입니다. 그리고 그와 함께, 당신의 감정을 정말 의식해야 합니다. 한숨을 쉬거나 신음 소리를 내지 마세요. 개를 마킹하고 보상할 때는 기뻐해도 됩니다. 물론 그런 감정은 드러내도 좋지만, 실망감은 여러분의 반려견이 고스란히 느끼기 때문입니다. 여러분은 중립적인 태도를 유지해야 합니다. 네, 저도 어쩔 수 없이 반려견이 기대한 행동을 하는 것을 보면 신이 납니다. 그러니 반려견에게 간식을 주면서 함께 축하해 주는 것은 괜찮지만, 간식을 준 뒤에는 다시 행동을 관찰하고 마킹 및 보상할 대상을 찾는 중립적인 상태로 돌아가야 합니다. 아홉 번째, 1분 평가 시간을 기억하세요. 모든 세션은 1분 이내의 평가 세션으로 시작해야 합니다. 그리고 마지막으로 다시 한 번 강조하자면, 셰이핑은 연속적 근사치에만 의존해서는 안 됩니다. 반려견이 이전에 보상을 받았던 행동들을 성공적으로 수행할 수 있도록 행동 블록을 배치하는 것이어야 합니다.
are the other distractions around you. It's your placement of reinforcement or how you're holding your tools or how you are sitting, standing, or kneeling yourself. How is that impacting that dog's ability to grasp the concept of what behavioral building block you are looking for them to offer you right now? Because if you do this right, the correct response should be so obvious to the dog. And rather than waiting for the dog to offer something and the dog like trying to grasp something from outer space, like successive approximations sometimes look, both you and the dog are engaged in the process that you're a part of because it's happening fast. Now, there may be some lulls in your training, but they're very, very few and far between. Also, if your training, if your shaping looks chaotic and frantic, and all of a sudden the dog gets a click and a cookie, then that also is not taking you in the path in the right direction because you will be building in all that frantic anxiety into the behavior that you're building. And who wants that? None of us wants frantic or anxious in our dogs. Number eight, resist the urge to talk to your dog or help the dog. If the dog stalls out, give them a moment to process. And if it doesn't, maybe after like 30 seconds or a minute, the dog doesn't look like they're moving towards anything, then just break it off, have them hop it up in the hot zone and rearrange your antecedents. Because if the dog stalls out and you help them by giving them verbal prompts or prompting them with your body or giving them a finger point or giving them a formal command, then you're not really shaping, you're telling. Now, also along with that, I want you to be really conscious of your emotions. Don't sigh like you're ever going to get, don't groan. You can be happy when you are marking and reinforcing the dog. Sure, show that kind of emotion, but don't show disappointment because your dogs are going to feel that. You're going to be neutral. And yes, I can't help but get excited when I see my dogs doing what I expected. So, it's okay to celebrate with the dog as you're giving them the cookie, but go back into the place of neutrality as you are just an observer of behavior looking for something to mark and reinforce. Number nine, remember that evaluation minute. Every single session has to start with an evaluation session that's one minute or less. And finally, I'm going to remind you one more time, shaping shouldn't be about successive approximations. It should be about arranging behavioral blocks so that the dog can successfully move through things
15:26
여러분은 반려견에게 코로 찬장 문을 닫도록 가르치고 싶을지 모릅니다. 하지만 '오, 수잔, 저는 반려견에게 코로 찬장 문을 닫으라고 가르치며 보상한 적이 없는데요.'라고 생각할 수 있죠. 그럼 셰이핑을 할 수 없다는 뜻인가요? 아닙니다. 반려견에게 코 터치를 가르치셨나요? 그렇다면, 그게 첫 번째 행동 블록입니다. 이번에는 손에 테이프를 붙여보면 어떨까요? 아, 두 번째 행동 블록인 테이프 타겟팅을 할 수 있게 되죠. 이제 그 테이프를 파리채 같은 곳에 붙여서 손에서 멀어지게 해봅시다. 그러면 타겟팅을 할 수 있을까요? 네, 할 수 있습니다. 이제 그걸 찬장 문에 붙이고 파리채를 확장해서 여러분의 손을 그곳에서 빼보세요. 그러면 반려견이 그곳을 타겟팅할 수 있을까요? 바로 그거죠. 우리는 매우 빠르게 반려견이 우리가 원하는 행동을 하도록 유도하는 행동 블록들을 갖추게 된 것입니다. 자, 마지막으로, 네, 여러분의 질문에 답변할 차례입니다. 첫 번째 질문, 셰이핑을 할 때 비보상 마커를 사용해야 할까요? 아니요. 선행 조건 설정이 여러분이 기대하는 반응이 명백하게 나오도록 준비되어야 합니다. 그러니 여러분의 도움은 필요 없습니다. 모든 것은 여러분의 계획과 환경 설정에 달려 있습니다. 개와 그 개가 이전에 강화받았던 반응들을 기반으로 합니다. 개가 자발적인 반응을 보이는 것을 편안하게 배우도록 돕는 좋은 훈련은 무엇일까요? 음, 에피소드 259에서 말씀드렸던 '서치(search)'라는 장소 특정적 강화 마커와 담요를 사용하는 것만큼 간단한 방법이 있습니다. 그 훈련은 모든 개가 앞발 타겟팅(paw targeting)을 스스로 하도록 동기를 부여할 것입니다. 만약 그렇지 않다면, 당신이 사용하는 강화물의 가치가 개에게 충분히 높지 않을 가능성이 큽니다. 저는 '행동에 클릭하고 위치에 보상하라'는 의견을 들은 적이 있습니다. 동의하시나요? 그것은 당신이 정지된 행동을 셰이핑(shaping)하고 있는지, 아니면 움직임이 있는 행동을 강화하고 있는지 구분하는 데 아주 좋은 경험칙입니다. 너무나도 많은 사람들이 개가 자신에게서 멀어지게 한 뒤 클릭하고 다시 자신에게 돌아오게끔 보상하려 합니다. 하지만 저는 무언가를 던져서 개에게 보상할 것이고, 솔직히 클릭조차 하지 않을 것입니다. 저는 그냥 '굿(good)' 같은 언어적 마커를 사용하고 강화물을 개에게 던져줄 것입니다. 이제, '행동에 클릭하고 위치에 보상하라'는 원칙이 항상 100% 우리가 하는 방식은 아닙니다. 왜냐하면 제가 이미 리셋 쿠키(reset cookies)에 대해 언급했으니까요, 맞죠? '서치'라고 말함으로써, 우리는 위치에 대해 보상하는 것이 아니라 의도적으로 리셋을 만들어내는 보상을 하는 것입니다. 그 리셋은 개가 우리가 원하는 것을 할 수 있도록 해줍니다. 저는 셰이핑을 시도하지만 제 메커닉(mechanics)이 너무 서툴러서 개와 저 모두 매우 좌절합니다. 셰이핑은 초보자를 위한 것이 아니라는 말을 들었는데, 사실인가요? 제 생각에 셰이핑은 누구에게나 적합합니다. 왜냐하면 더 많이 할수록 더 잘하게 되기 때문입니다.
that they've previously been reinforced before. Now, you might want to teach your dog to close the cupboard door with her nose, but, oh, Susan, I've never reinforced my dog for closing the cupboard with her nose before. So, that means I can't shape it. No. Have you taught your dog to nose target? Well, that's behavior block one. What if we put a piece of tape on your hand now? Ah, behavior block two, we can target tape. Now let's put that tape say on a fly swatter or something that you can get away from your hand. Can they target then? Yeah, they can. Now let's put that on a cupboard door and extend the fly swatter to get your hand out of there. Can they target it in there? Boom. We've got behavior blocks that in a very speedy way has led our dog to offer the behavior we were wanting. Okay. And finally, yes, I'm getting to your questions. So, question number one, should I be using a non-reward marker when I'm shaping? No. Your antecedent arrangements are arranged in a way that the response you're looking for is the obvious response. So, no help from you. It is all on your planning, your setting up of the dog, and the dog's offered previously reinforced responses. What is a good exercise to help a dog to learn to be okay with offering responses? Well, something as simple as the location-specific reinforcement marker of search and the blanket, as I spoke about in episode number 259, that exercise will get every dog motivated to offer paw targeting. And if it doesn't, chances are your reinforcement isn't high enough value to the dog. I've heard the comment, click for action and reward for position. Do you agree? You know, that is a great little rule of thumb to help differentiate between are you shaping a stationary behavior or are you reinforcing a behavior of motion? So, so often people want their dogs to run away from them and then they click and reward them back at them. But I would click and reward the dog by throwing something and I wouldn't even click, honestly. I would just use a verbal marker, like, good and throw their reinforcement out there to them. Now, click for action, reward for position isn't always 100% what we do because I've already mentioned reset cookies, right? By saying search. We aren't really reinforcing for position, but we're intentionally reinforcing to create a reset that allows a dog to do what we're looking for. I try to shape, but my mechanics suck and my dog and I get very frustrated. So, I've heard shaping isn't for novices. Is this true? So, shaping is for everybody
17:58
만약 당신과 반려견이 좌절하고 있다면, 그것은 당신의 선행 조건(antecedent) 설정 문제로 돌아가야 합니다. 그리고 저는 앞발 타겟팅으로 다시 돌아갈 것입니다. 간단한 것으로 시작해서 거기서부터 발전시키세요. 다시 말하지만, 점진적 근사치(successive approximations)는 행동 블록을 사용하는 셰이핑보다 당신과 반려견을 더 좌절시킬 가능성이 큽니다. 그리고 앞서 언급했듯이, 개가 좌절하면 당신은 질 낮은 행동을 보게 될 것입니다. 그와 같은 행동들 말이죠. 짖거나 낑낑거리고, 여러분을 앞발로 긁는 행동들이요. 여러분이 원하지 않는 것들이 그 행동 안에 구축될 거예요. 그러면 여러분은 더 좌절하게 될 겁니다. 수잔, 셰이핑 세션에 별도의 신호를 주나요? 아뇨, 저에게는 그냥 개 훈련 세션일 뿐이니까요. 그래서 제가 선행 조건이나 환경을 어떻게 구성했는지가 우리 개들에게는 우리가 곧 새로운 것을 배우거나 과거에 작업했던 것을 이어서 작업할 것이라는 꽤 큰 신호가 됩니다. 개가 계속해서 틀리면 어떻게 하나요? 저는 세션을 끝내고, 핫존(hot zone)으로 점프하게 한 뒤 그 행동에 대한 보상을 줄 거예요. 그리고 영상을 확인해서 선행 조건 배치 중 어떤 부분이 제가 개에게 정말로 원했던 행동과 반대되는지 평가할 거예요. 자, 만약 개가 어떤 묘기를 하도록 셰이핑 되었는데 그 묘기를 계속해서 반복적으로만 한다면, 아마 목줄을 잡거나 자리를 이동시키는 방식으로 중단시킬 수 있죠. 하지만 다시 말하지만, 저는 개가 스스로 문제를 해결하는 것을 정말 좋아해요. 단, 아무도 좌절하지 않는 방식으로요. 그래서 선행 조건을 재배열해서 정답이 매우 명확한 환경을 만들 수 있다면, 그것이 저의 첫 번째 선택이 될 것입니다. 셰이핑은 가르칠 수 있는 모든 행동과 묘기에 효과가 있나요, 아니면 특정 상황에서만 셰이핑을 쓰나요? 저는 모든 것에 셰이핑을 사용합니다. 그래서 셰이핑으로 가르칠 수 없는 것은 딱히 생각나지 않네요. 어떤 것들은 제가 어떻게 셰이핑해야 할지 상상조차 할 수 없긴 하겠네요. 알겠습니다. 댓글을 남겨주세요. 무엇이었는지 알려주세요. 이게 모든 견종에게 효과가 있나요, 심지어 지능이 낮은 견종에게도요? 으악. 저는 개인적으로 지능이 낮은 견종은 없다고 생각합니다. 저는 어떤 견종은 다른 견종보다 특정 기술에 더 적합하다고 믿습니다. 그리고 네, 셰이핑은 개뿐만 아니라 앵무새, 햄스터, 쥐, 뒷마당의 까마귀와 다람쥐에게도 효과가 있습니다. 제 말은, 정말 많은 것들이 있죠. 야생 동물에게 먹이를 주라고 부추기고 싶지는 않지만, 모든 동물은 학습을 합니다 셰이핑도 그렇고요. 네, 우리 사람들도요. 좋아요. 여기서 받아들여야 할 내용이 참 많죠. 저는 여러분이 유튜브로 넘어가서 댓글을 남겨주셨으면 해요. 여러분 반려견의 행동 구성 요소가 무엇인지 알려주세요.
in my opinion, because the more you do it, the better you get at it. If you and your dog are getting frustrated, that comes back to your antecedent arrangements. And I would go back to the paw targeting. Start with something simple and then grow from that. Again, successive approximations are probably going to frustrate you and your dog more than shaping with behavioral blocks. And as I mentioned earlier, that when the dog gets frustrated, you'll get cheap behaviors like barking, whining, pawing at you. Things you don't want are going to get built into that behavior. And that's going to be even more frustrating for you. Susan, do you cue shaping sessions? No, because to me, they're just dog training sessions. And so, you know, how I've arranged the antecedents or my environment is a pretty big cue to my dogs that we're about to learn something new or work on something that we've been working on in the past. What do you do if the dog keeps getting it wrong? I would end the session, have them jump in the hot zone, give them a reinforcement for that. And then I would go to my video and evaluate what part of the antecedent arrangements were in opposition to what I really wanted my dog to do. Now, if you have a dog that's been shaped a certain trick and they just keep offering that trick over and over and over again, you can interrupt it by maybe doing a collar grab and moving them. But again, I really like the dog to figure things out for themselves, but in a way that doesn't frustrate anybody. And so, if I can rearrange the antecedents and create an environment where the correct is super obvious, that would be my first choice. Does shaping work with all behaviors and tricks that can be taught or shaping only for specific things? I use shaping for everything. So, I can't think of something that it can't be taught with. Some things I just can't fathom how to shape. Okay. Leave me a comment. Let me know what it was. Does this work with all breeds, even unintelligent breeds? Yikes. I personally don't think there are unintelligent breeds. I believe that there are breeds that are better suited for some skills than others. And yes, shaping works for all, not just dogs, but parrots and hamsters and rats and your backyard crows and squirrel. I mean, there's so many things. I don't want to encourage you to feed wildlife, but all animals learn by shaping even yes, us people. Okay. A lot of things to take on board here. I want you to jump over to YouTube and leave me a comment. Let me know what are your dog's behavioral building blocks,
20:24
반려견이 이전에 보상받았던 행동 단위들 중에서 셰이핑을 할 때 활용할 수 있는 것들이 무엇인지요. 그리고 제가 구체적으로 어떤 행동을 어떻게 셰이핑하면 좋을지 단계별로 설명해주길 원하신다면, 기꺼이 그렇게 해드리겠습니다. 그리고 다음 에피소드에서는 너무 흥분해서 셰이핑하기 힘든 강아지들을 위해 우리가 무엇을 할 수 있는지 다뤄볼 예정입니다. 하지만 이 팟캐스트를 다시 듣고 184번 에피소드로 돌아가서 진 도널드슨의 '푸시, 스틱, 드롭(push, stick, drop)'에 대해 들으신다면, 여러분의 셰이핑 세션에서 발생했던 문제들을 확실히 해결할 모든 도구를 얻으실 수 있을 거라 확신합니다. 제 유튜브 채널로 와주세요. 그리고 거기에 가신 김에 이번 에피소드에 '좋아요'를 눌러주시고 댓글도 남겨주세요. 셰이핑에 대해 더 알고 싶은 점이 있다면 알려주세요. 아직 두 개의 에피소드가 더 계획되어 있으니까요. 다음 시간에 Shaped by Dog에서 뵙겠습니다. 이 불쌍한 강아지가 구독 버튼을 눌러달라고 빙글빙글 돌고 있네요. 구독 한번 해주실 수 있을까요? 거기에 가신 김에 알림 설정도 함께 눌러주세요. 이미 구독 중이시라면, 그건 여러분을 위한 거예요. 스스로에게 멋진 보상을 하나 챙겨주세요. 정말 엉뚱해요. 참 엉뚱한 녀석이죠. 상호작용 상호작용 상호작용 상호작용 상호작용 상호작용 상호작용 상호작용 상호작용 상호작용 상호작용
the behavioral units that your dog has been previously reinforced for that you can use in shaping another behavior. And if there's a specific behavior you'd like me to walk you through what that looks like, then I'm happy to do that. And in an upcoming episode, I am going to talk about what we can do for those dogs that are just so frantic, it's hard to shape them. But I'm pretty sure that if you re-listen to this podcast and go back and listen to podcast episode number 184, where I talk about Jean Donaldson's push, stick, drop, that you will have all the tools to absolutely fix what's been going on in your shaping session. But jump over to my YouTube channel. And while you're over there, please give this episode a thumbs up, but leave me a comment. Let me know what more you want to know about shaping because I have two more episodes planned for you. I'll see you next time right here on Shaped by Dog. This poor dog's turning herself in circles trying to get you to hit that subscribe button. Could you please do it? And while you're there, go ahead, hit the notification bell as well. And if you're already a subscriber, that's for you. Go grab yourself a great reinforcement. Such a goof. interact interact interact interact interact interact interact interact interact interact interact