Common Misconceptions Around Shaping: Why You May Find Dog Training Frustrating #261
Susan GarrettShaped by Dog with Susan Garrett · 팟캐스트 · 21분
셰이핑은 점진적 근사치에만 의존하는 방식에서 벗어나 반려견이 이미 학습하여 보상을 받은 행동 블록을 활용하는 과정입니다. 기존의 점진적 근사치 방식은 세밀한 행동 단위로 보상을 나누어 진행하므로 의욕이 강한 개에게는 스트레스와 불안을, 의욕이 낮은 개에게는 무관심이나 좌절을 유발할 수 있습니다. 반려견의 행동 블록을 일지에 기록하여 새로운 기술을 가르칠 때 무엇을 활용할지 체계적으로 계획합니다. 성공적인 셰이핑을 위해 강화물의 위계를 설정하고, 위치 특화 강화 마커를 활용하며, 1분 이내의 짧은 세션을 반복하여 집중력을 유지합니다. 훈련사는 자신의 선행 조건 배치와 보상 타이밍을 객관적으로 평가하여 반려견이 문제 해결 과정을 명확히 이해하도록 돕습니다. 반려견이 동작을 스스로 제안하도록 유도하고, 정체된 상황에서는 핫 존을 이용해 휴식하며 재설정을 시도합니다. 실망감을 드러내는 대신 중립적인 태도를 유지하고 반려견의 성공적인 시도에 기뻐하며 훈련의 흐름을 조절합니다.
셰이핑 주변의 흔한 오해: 개 훈련이 좌절스럽게 느껴지는 이유 #261
Common Misconceptions Around Shaping: Why You May Find Dog Training Frustrating #261
0:00
Shaped by Dog 259화에서 저는 셰이핑에 대해 전부 다루면서, 만약 문제 해결 편을 원하신다면 질문을 남겨달라고 했습니다. 정말 많은 질문을 받았습니다. 그래서, 이 모든 것과 그 이상을 다룰 예정입니다. 질문들을 보면서, Shaped by Dog에서 셰이핑에 대한 에피소드를 하나 더 만드는 것만으로는 부족하다는 것을 깨달았습니다. 여러 편을 만들어야겠더군요. 자, 이제 첫 번째 시간으로 여러분의 셰이핑 관련 질문에 답해 드리겠습니다. 셰이핑에 관한 모든 질문들에 대한 답변입니다. 안녕하세요, 수잔 가렛입니다. Shaped by Dog에 오신 것을 환영합니다. 와, 정말 정말 훌륭한 질문들을 많이 받았습니다. 셰이핑에 대한 오해가 있다는 것을 깨달았습니다. 제 생각에는 아마도 대부분의 팟캐스트 진행자, 블로그 게시물, 일반인, 그리고 아마 저 자신도 한때는 그랬을 텐데, 셰이핑을 '점진적 근사치(successive approximations)'라는 용어로 설명하기 때문인 것 같습니다. 즉, 개에게 작은 행동 하나를 할 때마다 보상을 주고, 그것이 또 다른 작은 행동으로, 그다음에는 또 다른 작은 행동으로 이어지게 하는 것이죠. 그렇게 하면 개는 결국 우리가 원하는 행동을 하게 됩니다. 하지만 훈련사의 입장에서 보면 이 방식은 개에게 엄청난 좌절감을 줄 수 있습니다. 왜냐하면 우리는 항상 더 많은 것을 기대하게 되고, 그것이 개에게 많은 스트레스와 불안, 그리고 압박감을 주기 때문입니다. 또 다른 문제는 '꼼수 행동(cheat behaviors)'이라고 불리는 것들이 많이 생겨난다는 점입니다. 조금 더, 조금 더, 계속해서 더 많이 요구하다 보면 정작 우리가 정말로 원하는 것이 무엇인지 우리 머릿속에서도, 그리고 개의 머릿속에서도 명확하지 않게 됩니다. 그래서 셰이핑이 조금 엉망이 되어버리죠. 제 개들을 예로 들어보겠습니다. 90년대 초반, 저는 제 잭 러셀 테리어와 보더 콜리에게 정확히 같은 행동을 셰이핑하고 있었습니다. 제 목표는 복종 훈련에서 '고 아웃(go out)', 유럽 친구들이 부르는 '센드 어웨이(send away)'를 가르치는 것이었습니다. 그래서 저는 셰이핑을 시작했죠. 제 목표는 복종 훈련에서 '고 아웃(go out)', 유럽 친구들이 부르는 '센드 어웨이(send away)'를 가르치는 것이었습니다. 그래서 저는 셰이핑을 시작했죠. 클리커와 쿠키 한 줌을 들고 연속적인 근사치를 사용하여 직선적으로 훈련했습니다. 저는 벽에 아주, 아주 가까이 서서, 개들이 벽을 쳐다볼 때, 더 오래 쳐다볼 때마다 클릭해 주었습니다. 아마 벽에 더 가까이 다가가면 클릭했죠. 결국 저는 개들이 벽에 발을 올리게 만들었습니다. 과정은 느렸고 두 마리 모두 일에 대한 의욕이 넘쳤습니다. 그래서 행동에서 불안한 모습이 많이 보였고, 정신없이 서두르는 모습도 있었지만, 결국 해냈습니다. 둘 다 앞발 두 개를 벽에 대게 만들었죠. 그런데 제 훈련 계획의 다음 단계는 이른바 '투퍼(twofer)'라고 불리는 것을 하는 것이었습니다. 즉, 개가 행동을 이해했다는 것을 보여주길 바라는 마음에서 첫 번째 행동에는 클릭하지 않고, 두 번째 행동에 클릭을 하려는 것이었죠. 그래서 제 잭 러셀 테리어인 트위스터에게 먼저 시도했습니다. 투퍼를 추가할 때,
In Shaped by Dog episode 259, I talked all about shaping and I said, if you wanted me to do a troubleshooting episode, just leave me some questions. And I got a lot of questions. So, all of that and more. And from those questions, I realized I don't need to do another shaping episode here on Shaped by Dog. I need to do several. And here we go with the first one, your questions answered all about shaping. Hi, I'm Susan Garrett. Welcome to Shaped by Dog. And wow, did we get a lot of really, really great questions. And I realize that there's a misconception about shaping. And I think it's because most podcasters, most blog posts, most people, and probably including myself at one point, they describe shaping with the phrase successive approximations, meaning you reward the dog for doing one little behavior and that leads you to another little behavior and then another little behavior and another. So, the dog will eventually do what you want them to do. But from where I see things as a trainer, that can potentially lead to a ton of frustration for the dog because we're always expecting more and that stresses and creates a lot of anxiety for dogs and a lot of pressure. The other problem with that is we get a lot of what's called cheat behaviors built into that a little bit more and a little bit more and a little bit more when what we really, really want isn't super clear both in our minds and in the dog's mind. So, shaping kind of gets a little bit screwy. I'm going to give you an example with my own dogs. Back in the early 90s, I was shaping both my Jack Russell Terrier and my Border Collie to do the exact same behavior. And my goal was to teach them to do a go out for obedience, a send away as my European friends might call it. So, I started shaping them linearly using successive approximations with a clicker and a handful of cookies, me standing very, very close to the wall, clicking them for looking at it, for looking at it longer, for maybe getting closer to the wall. Eventually I got them to put their paws on the wall. It was slow and both those dogs were driven to work. So, I got a lot of anxiousness in their behaviors, a lot of franticness, but I got the job done. I got both of them to hit their two paws on the wall. But the next step in my training plan was I wanted what was referred to as a twofer, meaning I want the dog to show me they understand the behavior by I'm not going to click the first one and I'll click the second one. And so, my Jack Russell, I did first, Twister. Now, when I was adding a twofer,
2:49
저는 개와 한 걸음 정도 떨어져 있었고 개들은 저에게서 떠나 벽으로 다가가고 있었습니다. 아직 큐(cue)는 주지 않았습니다. 여전히 초기 단계였으니까요. 그래서 트위스터는 저를 떠나 벽으로 가서 벽을 치고, 쿠키를 받으러 다시 돌아오려고 몸을 돌렸습니다. 마치 머릿속에서 톱니바퀴가 돌아가는 게 보이는 것 같았죠. 녀석은 '잠깐, 클릭 소리가 안 났는데? 나를 못 봤나?' 하는 듯했습니다. 그리고 다시 돌아보더니 '봐, 내가 했잖아'라고 말하듯이 더 높은 벽 쪽을 더 세게 쳤습니다. 물론 저는 웃음을 터뜨리며 그 순간 클릭하고 보상을 주었습니다. 그 후부터는 클릭을 안 하면, 녀석은 벽을 떠나지 않았습니다. 계속 벽을 쿵쿵거렸죠. 제 말은, 그게 바로 제가 원하던 모습이었다는 겁니다. 그래서 무엇이든 셰이핑(shaping)하는 데 천재적인 제 보더 콜리에게도 다음으로 투퍼를 시도해 보려고 했습니다. 그 개는 벽을 터치했고 저는 클릭하지 않았는데, 트위스터처럼 행동하지 않았습니다. 벽에서 떨어져 저에게 걸어오지도 않았고요. 녀석은 그냥 벽에 발을 올린 채로 제 쪽을 어깨너머로 쳐다보며 '나는 지금 하고 있는데' 하는 눈빛을 보냈습니다. 당신이 물었죠. 지금 당신이 무슨 짓을 했는지 안 보이나요? 이를테면, 언제 클릭할 건가요? 제가 더 오래 버티길 원하나요? 그리고 그녀는 그저 제가 어떻게 나오나 기다렸죠. 저는 생각했습니다. 그래, 알겠어. 내가 무슨 짓을 했는지 알겠어. 저는 매번 똑같은 지점에서 클릭했거든요. 그래서 그녀는 벽에 닿아 있는 것이 게임이라고 생각한 거죠. 그래서 저는 머릿속으로 이 모든 것을 생각했습니다. 그리고 말했죠. 이제 제가 할 일은 그녀가 벽에서 떨어질 때 클릭하는 것입니다. 그래야 그녀가 '아니, 벽에 닿는 게 아니라 벽에서 떨어져야 하는구나'라고 이해하겠죠. 그래서 저는 기다리며 준비하고 있었습니다. 그런데 그녀는 벽 위쪽을 더 올려다보더니, 움직여서 팔꿈치와 앞발을 벽에 대고 저를 어깨 너머로 쳐다봤습니다. '내 앞발로는 부족하다면, 내 몸을 벽에 더 많이 대길 원하는 건가?'라고 생각한 거겠죠. 저는 생각했습니다. 그래, 알겠어. 그녀는 다른 행동을 제안할 수 있는 거구나. 이것이 바로 점진적 접근법(successive approximations)이라는 겁니다. 클릭할 수 있는 행동을 보일 때까지 기다리는 거죠. 그래서 그녀가 그때 한 행동은 아래를 내려다보고 뒷발을 최대한 벽 쪽으로 바짝 붙여서, 마치 프리즈(freeze) 자세를 취한 것처럼 온몸을 벽에 밀착시킨 것이었습니다. 그 순간 저는 너무 웃겨서 뒤로 넘어질 뻔했습니다. 정말 웃겼거든요. 마치 어제 일처럼 생생하게 기억나는데, 아마 1993년쯤이었을 겁니다. 그래서 점진적 접근법을 사용할 때는 잘못될 수 있는 변수가 너무나 많습니다. 만약 일에 대한 의욕이 강한 개라면, 쉽게 포기하지 않겠지만, 약간 좌절감을 느낄 수도 있습니다. 그리고 의욕이 높은 개들에게서 그런 좌절감이 나타나면,
I was maybe a stride away from the dog and they were leaving me and going and touching the wall. There was no cue given. It was still early in the process. So, Twister left me, went up, hit the wall, turned around to come back to get the cookie. And it was almost as if I could see the wheels turning in her head. And she went, wait, I didn't get the click. Did she not see me? And she turned around and went like higher up on the wall with more force as if to say, look, I did it. And of course, I fell over laughing, clicked her and gave her her reinforcement. So, from then on, if I didn't click her, she wouldn't leave the wall. She would just keep pounding on the wall. I mean, that was exactly what I was looking for. And so, my border collie who was brilliant at shaping anything, I was going to try her with a twofer next. And here's what she did is she touched the wall and I didn't click, but she didn't do like Twister did. She didn't come off and start walking towards me. She stayed there with her paws on the wall. And she kind of looked over her shoulder at me like, I'm doing what you asked. Do you not see what you've done here? Like, when are you going to click that? Do you want me to hold longer? And she just waited me out. And I thought, okay, okay, okay. I see what I've done. I clicked at the exact same spot every time. And she thinks the game is be in contact with the wall. So, I'm thinking through this all in my brain. And I said, what I'm going to do now is I'm going to click her when she comes off the wall. So, she understands, no, no, no, it's touching, come off. So, I'm waiting, I'm ready. But what she did is she looked up the wall further, and then she moved so that she could put her elbows and her paws on the wall and looked over her shoulder at me thinking, well, if my paws aren't enough, maybe you want more of my body on the wall. And I'm like, okay, okay. She can offer something else. This is what successive approximations are about. You wait until they offer something that you can click. And so, what she did then was she looked down and she scooted her rear paws as close to the wall as she could, so she got her entire body laying on the wall as if she was in the freeze position. At which point I fell over laughing because it was hysterical. I remember it like it was yesterday and it was probably somewhere in, I don't know, 1993 that that happened. And so, with successive approximations, there's just like a gamut of things that could go wrong. And if you have a dog who's driven to work,
5:16
짖는 소리가 날 수도 있습니다. 혹은 불안감을 다른 방식으로 표출할 수도 있는데, 예를 들어 제자리에서 돌다가 행동을 하거나, 살짝 깨물고 나서 행동을 할 수도 있죠. 물론 모든 개가 그런 건 아니죠, 그렇죠? 그래서 그렇게 의욕적이지 않은 개들의 경우에는, 어떻게 접근해야 할까요? 강아지들은 '아, 잘 모르겠는데'라고 생각할 거예요. 그냥 가서 누워버릴 수도 있죠. 아니면 귀를 긁거나 여러분을 빤히 쳐다볼 수도 있고요. 그래서 많은 분들이 셰이핑을 점진적 근사치라고 생각하시는데요. 이제 그 생각은 버리셨으면 좋겠어요. 그 개념은 인터넷 세상으로 보내서 자유롭게 놓아주세요. 왜냐하면 제가 여러분께 생각해보라고 하고 싶은 것은, 지난 에피소드에서 설명해 드린 틀을 벗어난 셰이핑이기 때문이에요. 이건 여러분의 반려견이 이미 알고 있는 행동 블록을 사용하는 셰이핑이죠. 자, 지금 이 방송을 듣고 계신 여러분 모두, 여러분이 반려견에게 강화물을 주었거나, 어떤 행동을 하도록 허락했거나, 장난감을 주어서 보상을 한 행동들이 분명 있을 거예요. 그래서 운전 중이라면 그냥 생각만 해보시고요, 기회가 될 때 여러분 반려견의 행동 블록이 무엇인지 적어보세요. 즉, 반려견이 보상을 받은 적이 있는 행동들이죠. 앉기, 엎드리기, 서기 같은 것일 수도 있어요. 분명 많은 강아지들이 이 세 가지 행동에 대해 보상을 받은 경험이 있을 거예요. 그런 것들이 유용하게 쓰일 수 있는 행동 블록들이죠. 예를 들어, 만약 여러분이 강아지에게 배를 대고 기어가는 행동을 가르치려 하는데, 강아지가 이전에 엎드리기 자세로 간식을 받아본 적이 한 번도 없다면 어떨까요? 이전에 강화된 적이 없는 자세를 취하도록 유도하는 것이 얼마나 어려울지 아시겠죠? 그러니 모든 행동 블록을 어딘가 일지에 기록해 두세요. 그래야 새로운 행동을 셰이핑할 때 무엇을 활용할 수 있을지 알 수 있으니까요. 좋아요. 이제 여러분의 질문을 살펴볼게요. 우선 지난 에피소드에서 이야기했던 내용을 잠시 상기해 드리자면, 성공적인 셰이핑 세션을 진행하는 방법이었죠. 그 부분을 조금 더 자세히 보충 설명해 드릴게요. 가장 중요한 것은 보상의 위계를 갖추는 거예요. 즉, 강아지가 아주 환장할 정도로 좋아하는 간식을 말하는 거죠. 여러분의 강아지가 보상으로 생각하는 것을 가지고 있나요? 아, 세상에, 맞아요. 그리고 강화물(reinforcement)은 있나요? 제가 위계(hierarchy)라고 말할 때, 만약 반려견이, 예를 들어 테이터 샐러드(Tater Salad) 같은 경우, 제가 그의 최고 강화물을 사용하면 너무 흥분할 수도 있어요. 그게 좀 웃긴 건, 사실 15개월 된 구조견으로 여기 처음 왔을 때는 정말 느긋하고 일하기 싫어하던 개였거든요. 그런데 지금은 너무 들떠버려요. 그래서 저는 더 정교한 행동을 형성(shaping)할 때는 그에게 더 낮은 가치의 강화물을 사용하곤 하죠. 이것이 바로 강화의 위계입니다.
they may not give up, but they may get a little bit frustrated. And with that frustration with the higher drive dogs, you might get vocalization. You might get them showing anxiety in other ways, like you might get a spin and then a behavior, or you might get a little nip and then a behavior. Now, not all dogs are like that, are they? So, with dogs that aren't driven like that, they're going to go, oh, I don't get it. And they might just go and lay down, or they might like scratch their ear or stare at you. And so, many people think of shaping as successive approximations. And that is something I want you to release. Send that out into the interwebs and let it be free to roam. Because what I want you to think about it, what I described in the last episode is outside the box shaping. It's shaping by using behavioral blocks that your dog knows. Now, every single one of you listening to this, I know there are behaviors that you have reinforced in your dog that you have either given them a cookie, you've given them permission to do something, you've given them a toy. So, I'd like you, if you're driving in a car, just think about this. But when you get a chance, write down what are your dog's behavioral blocks? Things that they have earned reinforcement for. It could be something like a sit, a down, a stand. I'm sure many dogs have reinforcement for those three things. Those are behavioral blocks that could come in handy. For example, if you were trying to teach a dog to crawl on their belly, yet they'd never ever received a cookie for lying down. Can you see how difficult it would be to get that dog to offer a down position when it's never been a previously reinforced position? So, list all of the behavioral blocks in a journal somewhere. So, you'll know what you have to work with when you're trying to shape a behavior. Okay. I'm going to get your questions. I'm going to first remind you of some of the things we talked about in the last episode, how to have a successful shaving session. And I'm going to flesh that out a little bit more. So, super important that you have a hierarchy of reinforcement, meaning food that your dog goes cuckoo for Cocoa Puffs about. Do you have reinforcement that your dog goes, oh my gosh, yes. And do you have reinforcements? When I say a hierarchy, if you have a dog, like for example, Tater Salad, if I use his number one reinforcement, he might get too crazy, which is kind of funny because he was a pretty laid back dog who really didn't want to work when he arrived here as a 15 month old rescue dog, but he gets so jacked up. So, I would use like lesser
7:48
강화의 위계입니다. 그리고 '우리 개는 정말 일하기 싫어해요'라고 말하는 분들이라면, 바로 거기서부터 시작해야 합니다. 반려견이 음식을 받아먹을 수 있도록 강화의 위계를 만드는 것이죠. 그 내용은 259화에서 이야기했습니다. 두 번째는 위치 특화 강화 마커를 갖는 것입니다. 두 번째는 위치 특화 강화 마커를 사용하는 것입니다. '쿡(Cook)', '서치(Search)', 그리고 '차우(Chow)'도 사용하시길 적극 권장합니다. '차우'는 밥그릇에 간식이 있다는 뜻이죠. 밥그릇에서 간식을 꺼내 먹어도 된다는 겁니다. 아니면 생식을 먹이는 분들의 경우, 저는 밥그릇에 한 숟가락 정도 사료를 담아둡니다. 만약 제가 밥그릇 옆에 간식을 던져주고 '서치'라고 하면, 개들은 밥그릇에 있는 음식을 먹으러 가면 안 됩니다. '차우'나 여러분이 정한 위치 특화 강화 마커를 들었을 때만 밥그릇에 있는 음식을 먹어야 합니다. 즉, '쿡'은 입으로 바로 받아먹는 것, '서치'는 바닥에서 찾아 먹는 것을 의미합니다. '서치'는 타겟팅을 연습할 때 아주 좋은 신호입니다. 259화를 기억하시나요? 제가 담요 위로 개를 유도해서 행동을 형성하게 했었죠. 우리는 그 담요를 점점 더 작게 만들었고, 이제는 앞발 타겟팅이 되었습니다. 음, 개가 타겟 위에 앞발을 올리는 것에서 더 큰 가치를 느끼도록 강화물을 더 많이 제공하는 방법은, '리셋 쿠키(reset cookie)'라는 것을 사용하는 것입니다. 그때 바로 '서치'라고 신호를 주는 것이죠. 그러면 개에게 타겟에서 내려와 바닥에 있는 쿠키를 찾아도 된다는 뜻이 됩니다. 저는 개 뒤쪽으로 쿠키를 살짝 던져서 다시 돌아오게 만듭니다. 그러면 개들은 그 타겟을 아주 쉽게 찾습니다. 이제 숙련된 개라면 저는 정말로 그 개들에게 도전 과제를 주고 싶습니다. 간식을 제 뒤로 던져줄 수도 있죠. 그러면 개들은 다시 돌아와서 타겟을 보고 저를 마주할 방법을 찾아야 합니다. 중간 정도 수준의 개라면 그 두 지점 사이 어딘가로 간식을 던져주면 됩니다. 따라서 리셋 간식은 '위치 기반 보상 마커 탐색'을 이해하는 개에게만 가능합니다. 이것은 매우 중요합니다. '잇츠 유어 초이스(It's your choice)' 훈련이요. 만약 여러분이 잇츠 유어 초이스를 이해하지 못하는 개에게 행동을 형성하려고 하면, 개는 손에 든 음식에 너무 집착하게 되어 스스로 어떤 행동을 제안해야 할지 생각하지 못하게 될 것입니다. 또한 만약 여러분의 개가 잇츠 유어 초이스가 제대로 되어 있지 않다면, '차우(chow)'라는 위치 기반 보상 마커를 사용할 수 없습니다. 왜냐하면, 예를 들어 개가 일직선으로 후진하게 만들고 싶을 때, 제가 할 수 있는 방법 중 하나는 뒤쪽에 간식이 든 그릇을 두는 것입니다. 그래서 개가 몇 걸음 움직였을 때 보상하는 방법으로 앞다리 사이로 간식을 굴려줄 수도 있지만, 그냥 '차우'라고 말해서 개가 뒤로 돌아
value reinforcements for him when I'm shaping a more precision behavior. So, hierarchy of reinforcement. And for those of you who say my dog really doesn't want to work, that is where you're starting. You're creating a hierarchy of reinforcement so that your dog will take the food. And I spoke about that in episode number 259. Second, you're going to have those location-specific reinforcement markers. Cook, search, and I really encourage you to use chow as well, which is there's a cookie in the bowl. You can take the cookie out of the bowl. Or for those of us who are raw feeders, I put a spoonful of the food in the bowl. Now, if I throw a cookie beside the bowl and I say search, they aren't to take the food from the bowl. Only take the food from the bowl if they hear chow or whatever location specific reinforcement marker you're going to use. So, cook, come into your mouth, search, look for it on the floor. And search is a great cue when we're working to create a targeting. Remember in episode number 259, I had you shape the dog onto a blanket. We made that blanket smaller and smaller. Now we have a paw target. Well, the way we're going to build in more and more reinforcement for that dog finding value in putting their paws on a target is we're going to use what's called a reset cookie. So, that's where you're cute. You would say search, which tells the dog you can get off and look for a cookie on the floor. And I'll throw that cookie a little bit behind my dog so that they come back up and they find that target super easy. Now for a dog that's more experienced, I really want to challenge them. I might throw the cookie behind me. So, then they have to figure out how to come around and face me again on that target. Now your dog in between, you're going to throw somewhere in between those two spots. So, the reset cookie is only possible if you have a dog who understands that location specific reinforcement marker search. Super important. It's your choice. If you are trying to shape a dog who doesn't understand it's your choice, the dog is going to be so obsessed with the food in your hand, they're not going to be able to think about what to offer. Also, if your dog doesn't have really good, it's your choice, you can't use the location specific reinforcement marker chow. Because if I say want my dog to back up in a straight line, one of the things I might do is put a bowl with a cookie in it somewhere behind so that as I get a couple steps, one place I might reinforce that dog is by rolling a cookie between their front legs. But I also might just say chow so they turn
10:21
바로 뒤에 있는 간식을 먹게 할 수도 있습니다. 그러면 후진할 때 개가 비뚤게 가는 것을 방지하는 데 도움이 될 수 있죠. 그래서 저는 요즘 강아지를 훈련할 때 매일같이 '차우'를 사용합니다. 훈련 중에 '차우'를 어딘가에 꼭 사용하죠. 그리고 이건 잇츠 유어 초이스 없이는 불가능한 일입니다. 자, 이제 '크레이트 게임'입니다. 크레이트 게임과 '핫 존(Hot Zone)' 말이죠. 둘 중 하나만 해도 된다고 말씀드렸지만, 사실 크레이트 게임의 모든 단계를 진행하면, 더 많은 행동 블록을 얻을 수 있습니다. 즉, 개가 크레이트를 보거나 여러분이 케이지로 들어가라는 신호를 주면, 곧바로 크레이트로 달려가는 개를 만들 수 있다는 뜻입니다. 그리고 그것은 여러분의 개가 여러분을 떠나 먼 거리까지 이동해서 무언가를 수행하도록 보상받아 왔다는 행동 블록이며, 이는 나중에 매우 유용하게 쓰일 것입니다. 우리가 거리를 두고 다른 것들을 하고 싶을 때요. 그래서, 크레이트 게임과 핫 존이 있습니다. 저는 정말로 집중할 겁니다. 크레이트 게임에요, 하지만 둘 다 확실히 할 수 있죠. 왜냐하면 핫 존은 한 번에 한 마리 이상의 개를 훈련할 수 있게 해주니까요. 짧은 세션들이요. 좋아요. 오늘 이후에 여러분이 해주셨으면 하는 게 있습니다. 여러분 모두에게 1분 이하의 세션 5번을 하겠다고 약속받고 싶습니다. 그 세션들을 영상으로 찍어서 돌아와 제게 알려주세요 무엇을 배웠는지요. 그리고 기억하세요, 그게 제가 드린 숙제의 전부는 아닙니다. 제가 또 여러분께 부탁드린 것은 지금 여러분의 개가 가지고 있다고 알고 있는 모든 행동적 장애물을 적어보는 것입니다. 그리고 6번은 관련이 있습니다. 5번과요. 부디 셰이핑을 영상으로 찍고 그 셰이핑을 검토해주세요. 그리고 두 가지를 확인하길 바랍니다. 첫째, 여러분이 무엇을 기대했고 개는 무엇을 했는지 살펴보는 것입니다. 이제 여러분은 다시 돌아가서 다음에 영상을 볼 때는 개를 무시할 겁니다. 여러분은 무엇을 했나요? 왜냐하면 여러분은 개가 했던 행동과 여러분이 어떻게 보상을 주었는지, 혹은 어떻게 서 있었는지, 어디를 보고 있었는지 사이의 연관성을 보게 될 것이기 때문입니다. 그래서, 여러분의 역학에 대해 정말 비판적이 되세요. 왜냐하면 제 친구여, 그것이 개가 성공할지 아닐지에 가장 큰 영향을 미치기 때문입니다. 선행 조건을 어떻게 정리했나요? 지난 에피소드에서 이것에 대해 이야기했지만, 정말 중요합니다. 만약 제가 개에게 뒷걸음질을 가르친다면, 아마도 저는 제가 먼저 시작할 겁니다. 바닥에 무릎을 꿇고요. 만약 여러분이 개에게 뒷걸음질을 시키는데 여러분이 서 있다면, 그러면 개는 여기 위에 있는 것들을 생각하고 있을지도 모릅니다. 여러분이 바닥에 있으면, 개는 조금 더 낮은 곳을 생각하게 될 겁니다. 게다가 보상 배치를 더 정확하게, 바로 개 다리 사이로 할 수 있죠. 그래서, 그것이 그들이 뒷걸음질을 더 하도록 장려하는 것입니다. 그러니, 선행 조건 정리라는 것은, 단지
around and get the cookie right behind them. That might help them to not go in a crooked line when they're backing up. So, the use of chow, I use it pretty much every day when I'm training my puppy right now. I will use chow somewhere in our training. And that's just not possible without it's your choice. All right. Crate games. Crate games and hot zone. Now I said you could do one or the other, but I got to tell you, if you work through all the stages of crate game, you have so many more of those behavioral blocks. Meaning you have a dog who will run to their crate when they see the crate or when you give them the cue to go in their kennel. And that is a behavioral block that means your dog has been reinforced for leaving you and traveling a great distance to do something that will come in handy when we want to do other things at a distance. So, crate games and hot zone. I would really focus on crate games, but you can do them both for sure because hot zone allows you to train more than one dog at one time. Short sessions. Okay. Here's what I'd like you to do after today. I want you all to commit to doing five sessions that are one minute or less. Video those sessions and come back and tell me what you learned. And remember, that's not the only piece of homework I gave you. I also asked you to write down all the behavioral blocks you know your dog has right now. And number six is related to number five. Please video your shaping and review the shaping. And there's two things I want you to look for. Number one, you're going to look at what did you expect and what did the dog do? Now you're going to go back and you're going to ignore the dog when you look at the video the next time. What did you do? Because you're going to see a connection in between what the dog did and how you delivered the reinforcement or how you were standing or where you were looking. So, really be critical of your mechanics because that my friend has the biggest impact on whether the dog has success or not. How did you arrange your antecedents? I spoke about this in the last episode, but it's just so important. Now, if I was teaching the dog to back up, I would probably start with myself kneeling on the ground. If you're getting a dog to back up and you're standing up, then the dog might be thinking of things up here. When you're on the ground, the dog is going to be thinking a little bit lower. Plus your placement of reinforcement can be more exact right between your dog's legs. So, that's encouraging them to back up more. So, the antecedent arrangements, it's not just
12:46
주변의 다른 방해 요소가 무엇인지가 아닙니다. 그것은 여러분의 보상 배치나 여러분이 어떻게 잡고 있는지를 말하는 겁니다 당신의 도구나 당신이 앉아 있거나, 서 있거나, 무릎을 꿇고 있는 방식이 그 개의 행동에 어떤 영향을 미치고 있을까요? 당신이 현재 개에게 요구하는 행동의 기본 요소를 개가 얼마나 잘 파악하고 있을까요? 제대로 한다면, 올바른 반응은 개에게 너무나 명확해야 합니다. 개가 무언가를 스스로 시도하기를 기다리며 마치 우주에서 무언가를 잡아내려는 듯이 애쓰게 하는 대신, 점진적 접근법(successive approximations)에서 종종 보이듯이, 당신과 개 모두가 빠르게 진행되는 과정에 참여하게 됩니다. 훈련 중에 잠시 정체되는 순간이 있을 수도 있지만, 그런 경우는 극히 드뭅니다. 또한, 훈련이나 셰이핑(shaping) 과정이 혼란스럽고 정신없어 보이는데 갑자기 클릭과 간식을 준다면, 이는 올바른 방향으로 나아가는 것이 아닙니다. 왜냐하면 당신이 가르치려는 행동 속에 그 모든 정신없고 불안한 에너지를 함께 심어주게 될 것이기 때문입니다. 누가 그런 것을 원할까요? 우리 중 누구도 개가 정신없거나 불안해하는 것을 원치 않습니다. 여덟 번째, 개에게 말을 걸거나 도와주고 싶은 충동을 참으세요. 개가 멈춰 서 있다면, 스스로 생각할 시간을 주세요. 만약 30초나 1분 정도 지났는데도 개가 아무런 진전을 보이지 않는다면, 그냥 중단하고, 핫 존(hot zone)에서 개를 뛰어오르게 한 뒤 선행 조건을 재설정하세요. 개가 멈췄을 때 언어적 신호로 돕거나, 몸으로 신호를 보내거나, 손가락으로 가리키거나, 공식적인 명령을 내리는 방식으로 도와주면, 그것은 셰이핑이 아니라 지시를 하는 것입니다. 또한 그와 더불어, 자신의 감정을 매우 의식해야 합니다. 한숨을 쉬거나 신음 소리를 내지 마세요. 개에게 마커를 주고 보상할 때는 기뻐해도 됩니다. 물론 그런 감정은 표현해도 좋지만, 실망감은 드러내지 마세요. 개들은 그것을 느낍니다. 중립적인 태도를 유지하세요. 네, 그렇습니다. 제 개들이 제가 기대했던 대로 행동하는 모습을 보면 저도 모르게 신이 납니다. 그러니 축하해 주는 것은 괜찮습니다. 강아지에게 간식을 줄 때 함께 기뻐해 주되, 다시 중립적인 태도로 돌아와서 무엇을 마킹하고 강화할지 살피는 행동 관찰자의 자리로 돌아오세요. 아홉 번째, 평가의 시간을 기억하세요. 모든 세션은 1분 이내의 평가 세션으로 시작해야 합니다. 마지막으로 한 번 더 강조하고 싶은데, 셰이핑은 연속적인 근사치만을 다루는 것이 아닙니다. 개가 이전에 강화받았던 경험을 바탕으로 성공적으로 다음 단계로 나아갈 수 있도록 행동 블록을 배치하는 과정이어야 합니다. 예를 들어, 반려견에게 코로
what are the other distractions around you? It's your placement of reinforcement or how you're holding your tools or how you are sitting, standing, or kneeling yourself. How is that impacting that dog's ability to grasp the concept of what behavioral building block you are looking for them to offer you right now? Because if you do this right, the correct response should be so obvious to the dog. And rather than waiting for the dog to offer something and the dog like trying to grasp something from outer space, like successive approximations sometimes look, both you and the dog are engaged in the process that you're a part of because it's happening fast. Now there may be some lulls in your training, but they're very, very few and far between. Also, if your training, if your shaping looks chaotic and frantic and all of a sudden the dog gets a click and a cookie, then that also is not taking you in the path in the right direction because you will be building in all that frantic anxiety into the behavior that you're building. And who wants that? None of us wants frantic or anxious in our dogs. Number eight, resist the urge to talk to your dog or help the dog. If the dog stalls out, give them a moment to process. And if it doesn't, maybe after like 30 seconds or a minute, the dog doesn't look like they're moving towards anything, then just break it off, have them hop it up in the hot zone and rearrange your antecedents. Because if the dog stalls out and you help them by giving them verbal prompts or prompting them with your body or giving them a finger point or giving them a formal command, then you're not really shaping, you're telling. Now also along with that, I want you to be really conscious of your emotions. Don't sigh like you're ever going to get, don't groan. You can be happy when you are marking and reinforcing the dog. Sure, show that kind of emotion, but don't show disappointment because your dogs are going to feel that. You're going to be neutral. And yes, I can't help but get excited when I see my dogs doing what I expected. So, it's okay to celebrate with the dog as you're giving them the cookie, but go back into the place of neutrality as you are just an observer of behavior looking for something to mark and reinforce. Number nine, remember that evaluation minute. Every single session has to start with an evaluation session that's one minute or less. And finally, I'm going to remind you one more time, shaping shouldn't be about successive approximations. It should be about arranging behavioral blocks so that the dog can successfully
15:25
찬장 문을 닫도록 가르치고 싶은데, '수잔, 저는 강아지가 코로 찬장 문을 닫게 강화한 적이 없는데요'라고 생각할 수 있습니다. 그럼 셰이핑을 할 수 없는 걸까요? 아니요. 강아지에게 코 타겟(nose target)을 가르친 적이 있나요? 그게 바로 첫 번째 행동 블록입니다. 이제 손에 테이프를 붙여보는 건 어떨까요? 아, 그게 두 번째 행동 블록입니다. 강아지는 테이프를 타겟팅할 수 있죠. 이제 그 테이프를 파리채 같은 곳에 붙여서 당신의 손에서 멀어지게 해보세요. 강아지가 그때도 타겟팅할 수 있을까요? 네, 할 수 있습니다. 이제 그것을 찬장 문에 붙이고 파리채를 길게 해서 당신의 손을 그곳에서 빼보세요. 강아지가 거기에서도 타겟팅을 할 수 있을까요? 짠! 아주 빠르게 우리가 원하는 행동을 하도록 유도하는 행동 블록들을 갖추게 된 것입니다. 좋아요. 마지막으로 여러분의 질문에 답해 드리겠습니다. 첫 번째 질문입니다. 셰이핑할 때 보상하지 않는 마커(non-reward marker)를 사용해야 할까요? 아니요. 선행 조건들을 잘 배치해서 여러분이 원하는 반응이 당연하게 나오도록 만드세요. 그러니 여러분의 도움은 필요 없습니다. 모든 것은 여러분의 계획, 강아지를 위한 환경 설정, 그리고 개가 이전에 강화받았던 반응들에 달려 있습니다. 개가 자발적인 행동을 하는 것을 편안하게 느끼도록 돕는 좋은 연습은 무엇일까요? 강아지가 스스로 무언가를 제안하는 것을 괜찮게 여기도록 돕는 좋은 연습은 무엇일까요? 반응이요? 글쎄요, '찾아(search)'라는 위치 특정 강화 마커처럼 간단한 것과 담요를 예로 들 수 있죠. 259화에서 말씀드렸듯이 그 연습은 모든 개가 발 타겟팅을 하도록 동기를 부여해 줄 겁니다. 만약 그래도 안 된다면, 강화제의 가치가 반려견에게 충분히 높지 않을 가능성이 큽니다. 행동에는 클릭을 하고, 위치에는 보상을 하라는 말을 들어본 적이 있습니다. 동의하시나요? 음, 그건 정지된 행동을 셰이핑(shaping)하고 있는지, 아니면 움직이는 행동을 강화하고 있는지를 구분하는 데 도움이 되는 아주 훌륭한 경험 법칙이죠. 사람들은 개가 자기에게서 멀어지기를 바라는 경우가 너무 많은데, 그러고 나서 다시 자기 쪽으로 불러 클릭하고 보상하곤 합니다. 하지만 저라면 무언가를 던져줌으로써 개에게 클릭하고 보상할 거예요. 솔직히 말해서 저는 클릭조차 하지 않고, 그냥 '좋아' 같은 언어적 마커를 사용해서 강화물을 개가 있는 쪽으로 던져줄 것 같습니다. '행동에는 클릭, 위치에는 보상'이라는 말이 항상 100% 우리가 하는 방식은 아닙니다. 이미 리셋 쿠키(reset cookies)에 대해 언급했으니까요, 그렇죠? '찾아'라고 말함으로써 우리는 사실 위치에 대해 보상하는 것이 아니라, 의도적으로 리셋을 유도하기 위해 강화하는 것입니다. 그 리셋은 개가 우리가 원하는 것을 할 수 있게 만들어 주죠. 저는 셰이핑을 시도하지만 제 메커니즘이 엉망이라 저와 반려견 모두 너무 좌절합니다. 그래서 셰이핑은 초보자를 위한 것이 아니라는 말을 들었는데, 사실인가요? 제 생각에 셰이핑은 누구에게나 필요합니다. 왜냐하면 많이 해볼수록 더 잘하게 되기 때문이죠. 만약 당신과 반려견이 좌절하고 있다면,
move through things that they've previously been reinforced before. Now, you might want to teach your dog to close the cupboard door with her nose, but, oh, Susan, I've never reinforced my dog for closing the cupboard with her nose before. So, that means I can't shape it. No. Have you taught your dog to nose target? Well, that's behavior block one. What if we put a piece of tape on your hand now? Ah, behavior block two, he can target tape. Now let's put that tape, say, on a fly swatter or something that you can get away from your hand. Can they target then? Yeah, they can. Now let's put that on a cupboard door and extend the fly swatter to get your hand out of there. Can they target it in there? Boom. We've got behavioral blocks that in a very speedy way has led our dog to offer the behavior we were wanting. Okay. And finally, yes, I'm getting to your questions. So, question number one, should I be using a non-reward marker when I'm shaping? No. Your antecedent arrangements are arranged in a way that the response you're looking for is the obvious response. So, no help from you. It is all on your planning, your setting up of the dog and the dog's offered previously reinforced responses. What is a good exercise to help a dog to learn to be okay with offering responses? Well, something as simple as the location-specific reinforcement marker of search and the blanket. As I spoke about in episode number 259, that exercise will get every dog motivated to offer paw targeting. And if it doesn't, chances are your reinforcement isn't high enough value to the dog. I've heard the comment, click for action and reward for position. Do you agree? You know, that is a great little rule of thumb to help differentiate between are you shaping a stationary behavior or are you reinforcing a behavior of motion? So, so often people want their dogs to run away from them and then they click and reward them back at them. But I would click and reward the dog by throwing something. And I wouldn't even click, honestly, I would just use a verbal marker like good and throw their reinforcement out there to them. Now click for action, reward for position isn't always 100% what we do because I've already mentioned reset cookies, right? By saying search, we aren't really reinforcing for position, but we're intentionally reinforcing to create a reset that allows a dog to do what we're looking for. I try to shape, but my mechanics suck and my dog and I get very frustrated. So, I've heard shaping isn't for novices. Is this true? So, shaping is for everybody in my opinion,
18:00
그건 선행 조건(antecedent arrangements) 설정 문제로 돌아가야 합니다. 저는 다시 발 타겟팅으로 돌아갈 거예요. 간단한 것부터 시작해서 거기서부터 확장해 나가세요. 다시 말하지만, 점진적 근사치(successive approximations)는 행동 블록(behavioral blocks)을 사용하는 셰이핑보다 당신과 반려견을 더 좌절시킬 가능성이 큽니다. 앞서 말씀드린 것처럼, 개가 좌절하면 짖거나, 낑낑대거나, 당신을 앞발로 긁는 것과 같은 '치팅 행동(cheat behaviors)'이 나타나게 됩니다. 원치 않는 행동들이 그 동작 안에 섞이게 될 것이고, 그러면 더욱 힘들어질 겁니다. 당신에게 좌절감을 줄 수 있겠네요. 수잔, 셰이핑 세션을 시작할 때 신호를 주나요? 아니요, 저에게는 그냥 개 훈련 세션일 뿐이니까요. 그래서 제가 선행 조건을 어떻게 설정했는지, 또는 환경을 어떻게 구성했는지가 제 개들에게는 우리가 곧 새로운 것을 배우거나 과거에 작업했던 내용을 연습할 거라는 아주 큰 신호가 됩니다. 개가 계속 틀리면 어떻게 하시나요? 저는 세션을 끝낼 거예요. 개들을 핫 존(hot zone)으로 이동하게 하고 그에 대해 보상을 줄 겁니다. 그리고 나서 제 영상을 보면서 선행 조건의 어떤 부분이 제가 정말로 원하는 개의 행동과 상충했는지 평가할 거예요. 만약 개가 특정 트릭을 셰이핑 받았는데 그 트릭을 계속해서 반복적으로 제시한다면, 목걸이를 살짝 잡거나 개를 이동시켜서 흐름을 끊을 수도 있습니다. 하지만 저는 개가 스스로 문제를 해결하는 것을 정말 좋아하지만, 그 누구도 좌절하지 않는 방식으로 그렇게 하길 원해요. 그래서 선행 조건을 재조정해서 올바른 행동을 하는 것이 매우 명확한 환경을 만들 수 있다면, 그것이 저의 첫 번째 선택이 될 것입니다. 셰이핑은 가르칠 수 있는 모든 행동과 트릭에 효과가 있나요, 아니면 특정 상황에서만 사용하나요? 모든 것을 셰이핑으로 가르치시나요? 셰이핑으로 가르칠 수 없는 것을 생각할 수가 없네요. 어떤 것들은 도저히 어떻게 셰이핑을 해야 할지 모르겠는 것들도 있죠. 좋아요. 댓글로 무엇인지 알려주세요. 모든 견종에게 효과가 있나요, 지능이 낮은 견종에게도요? 으악, 저는 개인적으로 지능이 낮은 견종은 없다고 생각해요. 저는 특정 견종이 다른 견종보다 어떤 기술에 더 적합한 경우가 있다고 믿습니다. 네, 셰이핑은 개뿐만 아니라 앵무새, 햄스터, 쥐, 뒷마당의 까마귀와 다람쥐까지 모든 동물에게 효과가 있어요. 정말 많은 것들이 있죠. 야생 동물에게 먹이를 주라고 권장하고 싶지는 않지만, 모든 동물은 셰이핑을 통해 배웁니다. 심지어 우리 인간들도요. 좋아요. 여기서 받아들여야 할 점이 참 많네요. 유튜브로 이동해서 댓글을 남겨주세요. 여러분 반려견의 행동 빌딩 블록이 무엇인지 알려주세요. 이전에 보상을 받아왔던 행동 단위들이 무엇인지, 그리고 다른 행동을 셰이핑할 때 어떤 것을 활용할 수 있는지 알려주세요.
because the more you do it, the better you get at it. If you and your dog are getting frustrated, that comes back to your antecedent arrangements. And I would go back to the paw targeting. Start with something simple and then grow from that. Again, successive approximations are probably going to frustrate you and your dog more than shaping with behavioral blocks. And as I mentioned earlier that when the dog gets frustrated, you'll get cheat behaviors like barking, whining, pawing at you, things you don't want are going to get built into that behavior. And that's going to be even more frustrating for you. Susan, do you cue shaping sessions? No, because to me, they're just dog training sessions. And so, you know, how I've arranged the antecedents or my environment is a pretty big cue to my dogs that we're about to learn something new or work on something that we've been working on in the past. What do you do if the dog keeps getting it wrong? I would end the session, have them jump in the hot zone, give them a reinforcement for that. And then I would go to my video and evaluate what part of the antecedent arrangements were in opposition to what I really wanted my dog to do. Now, if you have a dog that's been shaped a certain trick and they just keep offering that trick over and over and over again, you can interrupt it by maybe doing a collar grab and moving them. But again, I really like the dog to figure things out for themselves, but in a way that doesn't frustrate anybody. And so, if I can rearrange the antecedents and create an environment where the correct is super obvious, that would be my first choice. Does shaping work with all behaviors and tricks that can be taught or shaping only for specific things? Are you shaping for everything? So, I can't think of something that it can't be taught with. Some things I just can't fathom how to shape. Okay. Leave me a comment. Let me know what it was. Does this work with all breeds, even unintelligent breeds? Yikes. I personally don't think there are unintelligent breeds. I believe that there are breeds that are better suited for some skills than others. And yes, shaping works for all, not just dogs, but parrots and hamsters and rats and your backyard crows and squirrel. I mean, there's so many things. I don't want to encourage you to feed wildlife, but all animals learn by shaping even, yes, us people. Okay. A lot of things to take on board here. I want you to jump over to YouTube and leave me a comment. Let me know what are your dog's behavioral building blocks, the behavioral units that your dog has been previously reinforced for
20:28
만약 제가 구체적으로 어떤 행동에 대해 단계별로 설명해 주길 원하신다면 기꺼이 그렇게 하겠습니다. 다음 에피소드에서는 너무 흥분해서 셰이핑하기 어려운 반려견들을 위해 무엇을 할 수 있는지 다룰 예정입니다. 하지만 이 팟캐스트를 다시 듣고, 제가 진 도널드슨의 푸시 스틱 드롭에 대해 이야기했던 184번째 에피소드를 다시 듣는다면, 셰이핑 세션에서 겪고 있는 문제를 확실히 해결할 수 있는 모든 도구를 갖추게 될 것이라 확신합니다. 제 유튜브 채널로 와주세요. 그리고 그곳에 가시는 김에 이 에피소드에 좋아요를 눌러주시고 댓글도 남겨주세요. 셰이핑에 대해 더 알고 싶은 것이 있다면 알려주세요. 여러분을 위해 두 개의 에피소드를 더 계획해 두었거든요. 다음번에 이곳 'Shaped by Dog'에서 다시 뵙겠습니다. 셰이핑에 대해 더 알고 싶은 것이 있다면 알려주세요. 왜냐하면 여러분을 위해 두 개의 에피소드를 더 계획해 두었기 때문입니다. 다음번에 이곳 'Shaped by Dog'에서 다시 뵙겠습니다.
that you can use in shaping another behavior. And if there's a specific behavior you'd like me to walk you through what that looks like, then I'm happy to do that. And in an upcoming episode, I am going to talk about what we can do for those dogs that are just so frantic, it's hard to shape them. But I'm pretty sure that if you re-listen to this podcast and go back and listen to podcast episode number 184, where I talk about Gene Donaldson's push stick drop, that you will have all the tools to absolutely fix what's been going on in your shaping session. But jump over to my YouTube channel. And while you're over there, please give this episode a thumbs up, but leave me a comment. Let me know what more you want to know about shaping because I have two more episodes planned for you. I'll see you next time right here on Shaped by Dog.