5 Popular Ways To Train Your Dog With Food #170
Susan GarrettShaped by Dog with Susan Garrett · 팟캐스트 · 18분
강화 기반 반려견 훈련은 루어링, 관찰 학습, 행동 포착, 그리고 다양한 형태의 셰이핑을 포함하는 여러 방법론으로 구성됩니다. 루어링은 간식이나 장난감으로 개의 동작을 유도하고 보상하는 방식으로 접근하며, 이를 위해서는 보상물에 대한 의존도를 낮추는 제거 계획이 필수적입니다. 관찰 학습은 다른 개나 사람의 행동을 모방하는 방식과 버튼 조작 등을 통한 사회적 학습으로 분류됩니다. 행동 포착은 개가 본능적으로 수행하는 동작을 기다렸다가 타이밍에 맞춰 강화하는 기법이며, 셰이핑은 최종 목표 행동을 여러 단계로 세분화하여 개가 스스로 환경을 탐색하며 완성하도록 이끄는 과정입니다. 훈련사는 특히 셰이핑 과정을 여러 층위로 나누어 진행함으로써 개가 불안감을 느끼지 않고 최종 동작을 수행할 수 있도록 자신감을 북돋아 줄 수 있습니다. 체계적인 훈련 설계는 개가 원치 않는 부수적 행동을 학습하는 것을 방지하고 효율적인 학습 성과를 보장합니다.
음식으로 반려견을 훈련하는 인기 있는 5가지 방법 #170
5 Popular Ways To Train Your Dog With Food #170
0:00
누군가 당신에게 자신은 강화 기반 반려견 훈련사라고 말한다면, 그것은 실제로는 여러 가지 의미일 수 있습니다. 오늘 저는 강화 기반 반려견 훈련의 주요 차이점들에 대해 이야기해 보려고 합니다. 안녕하세요, 저는 수잔 개럿입니다. 'Shaped by Dog'에 오신 것을 환영합니다. 강화 기반 반려견 훈련사라고 해서 반려견을 훈련하는 방식이 모두 똑같은 것은 아니라는 점을 모르실 수도 있습니다. 자, 대부분의 사람이 가장 먼저 떠올리는 것은 루어링(luring)입니다. 손가락 사이에 간식을 끼우고 개를 유도하는 방식이죠. 하지만 이것은 누군가가 반려견을 훈련할 수 있는 네, 다섯, 혹은 여섯 가지 주요 방법 중 하나일 뿐입니다. 그리고 그 다섯, 여섯 가지 주요 방법이 모두 극적으로 다른 것은 아닙니다. 한 사람이 다섯, 여섯 가지를 모두 사용하는 경우는 드뭅니다. 사실, 가능한지도 잘 모르겠네요. 알아보도록 하죠. 첫 번째 방법은 방금 말했듯이, 루어링입니다. 여기에는 루어(유도)-행동-보상이라는 공식이 있습니다. 즉, 개가 좋아하는 것으로 유도한다는 뜻입니다. 그게 정말 중요해요. 그냥, 주머니에서 아무 간식이나 꺼내서 우리 개가 따라오겠지, 라고 생각해서는 안 됩니다. 개는 그것을 좋아해야 합니다. 좋아요. 자, 행동을 끌어내기 위한 루어입니다. 개가 앉기를 원한다면, 간식을 개의 머리 위로 가져가야 합니다. 그러면 중력의 영향으로 엉덩이가 바닥으로 내려가려고 하게 됩니다. 그렇게 개를 앉은 자세로 만드는 것이죠. 그것이 바로 루어-행동-보상에서 행동 부분입니다. 그런 다음 개에게 간식을 줍니다. 알겠죠? 그러니 루어링에 사용할 수 있는 것이 단순히 음식뿐만이 아니라는 점을 아는 것이 중요합니다. 전 세계의 많은 사람들이 수많은 기초 반려견 훈련을 일종의 루어를 사용해서 수행합니다. 그러니 테니스 공을 사용할 수도 있습니다. 만약 당신의 개가 테니스 공을 좋아한다면, 테니스 공을 들어 올려서 공중에 던지거나 테니스 공을 엉덩이 옆에 두세요. 반려견이 여러분 옆에서 목줄을 느슨하게 하고 걷게 하고 싶다면 말이죠. 사용할 수 있는 것들은 아주 많습니다. 도그 어질리티에서는 많은 사람들이 훈련 중에 도그 워크, 시소, 또는 A-프레임의 끝에 반려견이 가장 좋아하는 장난감을 두어 그것을 향해 달려가게 하거나 위브 폴 훈련을 할 때 사용합니다. 모두 같은 원리이지만 단지 다른 도구를 사용하는 것뿐이죠. 저는 개인적으로 훈련할 때 유도법(루어링)을 거의 전혀 사용하지 않습니다. 그리고 솔직히 이 경우를 제외하고는 딱히 생각나는 게 없네요. 저는 이럴 때만 유도법을 사용합니다. 아마도 강아지 스트레칭이겠네요. 강아지 코에 간식을 대고 그 간식을 엉덩이 쪽으로 가져가서 예쁜 스트레칭 동작을 유도하는 것인데, 그런 경우라면 저도
If someone were to say to you, I'm a reinforcement-based dog trainer, that actually could be a number of different things. And today I'm going to go over the main differences between reinforcement-based dog training. Hi, I'm Susan Garrett. Welcome to Shaped by Dog. And you may not realize that there's a lot of different ways somebody may train their dog if they are a reinforcement-based dog trainer. Now, the first one that most people would think about is luring, where you put a food lure between your fingers and you lure the dog. But that is one of four or five, or maybe even six main ways that somebody could train their dog. And those five or six main ways aren't all very dramatically different. It's rare for somebody to use all five or six of them. Actually, I don't even know if it would be possible. Let's find out. Methodology number one, I just said, it's luring. And so, there's a formula which is lure, behavior, reward. Meaning you get the dog luring with something they love. That's really important. You can't just say, oh, I'm just going to pick a piece of food out of my pocket. My dog's going to follow it. The dog has to love it. All right. So, lure to get a behavior. If you wanted your dog to sit, you would put the cookie over their head. And that would cause like gravity to help the butt to want to go to the ground. And so, you're going to get the dog in a sitting position. And that is the behavior part of the lure behavior reward. And then you deliver the food to the dog. Okay. So, it's important to know it's not just food that you can use for luring. A lot of basic dog training is done by many, many people all over the world done with a lure of some sort. So, you could use a tennis ball. If your dog loves tennis balls, you could put the tennis ball up in the air or put the tennis ball on your hip. If you wanted your dog to walk on a nice loose leash beside you. There's a lot of different things you could use. In dog agility, a lot of people will put like the dog's favorite toy at the bottom of a dog walk or a seesaw or an A-frame in training to get the dog to drive to it or in weave pulls. All of those are the same thing, but just using something different that the dog loves. Now, I personally virtually never use a lure in my training. And I honestly can't think of something that I can say, except for this. I would use a lure for this. Maybe puppy stretches. That would be a good one where I would put a cookie on the puppy's nose and I would put the cookie and touch it to their hip in order to get a nice stretch. That would be something I would use a
3:03
음식 유도법을 사용할 것입니다. 행동을 만드는 게 아니니까요. 아, 사실 만든 적은 있네요. 하지만 저는 일반적으로 그런 식으로 행동을 만들지는 않습니다. 제 반려견 모멘텀은 코를 엉덩이에 대고 있을 수 있어요. 그건 그 아이가 아는 행동이죠. 좋아요. 하지만 제 훈련의 99.9%는 어떤 종류의 유도도 없이 이루어집니다. 자, 수잔, 당신은 반려견을 어떻게 훈련하나요? 과거에 언급했듯이, 저는 음식과 장난감, 그리고 우리 개가 좋아하는 모든 것, 활동 등을 사용하지만 유도 수단으로는 사용하지 않습니다. 이번 팟캐스트 에피소드에서 제가 왜 유도법을 쓰지 않는지에 대한 자세한 내용을 다룰 수 있을지는 모르겠지만, 약속드릴게요. 조만간 제가 왜 행동 유도를 하지 않는지에 대한 자세한 이유를 공유할 에피소드가 나올 예정입니다. 부분적으로는 30년 동안 사람들을 가르친 전문가로서의 관찰 결과이지만, 큰 부분은 개가 어떻게 학습하는지에 대한 과학적 근거 때문입니다. 훈련 시 음식 유도법을 사용하지 않는 것이 왜 실제로 좋은 아이디어인지 뒷받침하는 수많은 과학 논문이 있습니다. 그 모든 말을 차치하고라도, 음식 유도법을 강화 기반 반려견 훈련의 유일한 형태로 사용하여 매우 성공적으로 훈련하는 많은 반려견 훈련사들이 있다는 점을 말씀드리고 싶습니다. 제가 왜 그렇게 하지 않는지에 대한 이유는 다음 에피소드에서 다룰 예정입니다. 자, 그럼 유도(luring)에 대해 이야기해보죠. 음식으로 유도하는 방법은 이제 세상에 널리 알려져 있다고 생각합니다. 본격적으로 넘어가기 전에, 제가 80년대에 훈련을 처음 시작했을 때 실제로 음식 유도를 사용했다는 점을 꼭 말씀드리고 싶네요. 그게 제가 강아지 훈련을 처음 접했던 방식이었거든요. 만약 여러분이 음식 유도를 사용하는 사람이라 하더라도 나쁜 훈련사라는 뜻은 아닙니다. 하지만 성공을 원한다면, 어떻게 유도물을 제거(fade)할지에 대한 계획을 반드시 세워야 합니다. 음식 유도나 다른 어떤 종류의 유도물이라도 사람들이 음식 유도를 멈추지 않으면 훈련이 제대로 되지 않기 때문입니다. 행동이 유도물에만 의존하게 되거든요. 여기에는 딱 한 가지 예외가 있습니다. 일에 대한 의욕이 너무 강해서 유도를 받으면서도 본능적으로 일을 수행하는 견종들이죠. 예시는 아주 많습니다. 많은 도그 어질리티 스포츠에 참여하는 셸티들은 음식 유도로 배우지만, 주인과 달리고 쫓는 것을 너무 좋아해서, 정말 아주 빠르게 배우곤 합니다. 많은 슈츠훈트(Schutzhund) 훈련사들도 유도물을 사용하지만, 그들은 말리노이즈나 저먼 셰퍼드처럼 일에 대한 의욕이 너무 강한 견종들과 함께하기 때문에, 훈련 초기에 유도물을 사용하는 것이 별로 중요하지 않을 정도입니다. 강화 기반 강아지 훈련의 두 번째 방법은 여러 가지 이름으로 불립니다. 관찰 학습 혹은 사회적 학습이라고 하죠. 가끔 모델링이라고 부르기도 하지만, 기술적으로 모델링이라고 부르는 게 맞는지에 대해서는 의문의 여지가 있습니다. 이는 강아지가
food lure for because I'm not going to create a behavior. Oh, I actually have. But I generally don't create a behavior like that. My dog momentum will hold her nose to her hip. And that's a behavior that she knows. Okay. But 99.9% of my training is done without any kind of a lure. All right. So, how do you train your dog, Susan? As I've mentioned in the past, I use food. I use toys. I use everything that my dog loves, activities, but I don't use it as a lure. And I don't know if we're going to be able to get into the details of why I don't in this podcast episode, but I promise you I've got one coming up, which will share the details of why I do not lure behaviors. Part of it is my observation as a 30-year professional teaching people, but a big part of it is the science behind how dogs learn. And there's tons and tons of scientific papers that support why it's actually a great idea not to use food lures in your training. Having said all that, let me just say there are a lot of very successful dog trainers using food lures exclusively as their form of reinforcement-based dog training. And I will get into the reasons why I don't in an upcoming episode. Okay. So, luring, I think the world knows about food luring. Now, before I leave food luring, I've just got to mention, I actually use food luring in my training way back in the 80s. It was the first way I was introduced to dog training. And I just want to share, if you are somebody who uses food lures, doesn't make you a bad person, but if you want success, you have got to have a plan of how you're going to fade the lure. Because that's when food lures or any kind of lures in training don't work, is when people don't get rid of the food lure because the behavior gets attached to the lure. There is only one exception to that. And that is with some breeds that are so driven to do the work, they work in spite of being lured. And I could give you many examples. A lot of Shelties in the sport of dog agility, they learn with food lures, but they learn to be really, really fast because they love to run and chase their owner. A lot of Schutzen trainers use lures, but they're working with breeds like Malinois or German Shepherds that are so driven by the work that it's just insignificant that they use lures to start them. Number two on the reinforcement-based dog training comes by a number of different names. Observational learning or social learning. Occasionally it's called modeling, but there's some question whether that's technically correct to call it modeling. And that is when a dog learns
5:56
사람이나 다른 강아지를 관찰하며 배우는 것을 말합니다. 그래서 다른 개가 하는 행동을 보면, 제 강아지는 수박을 던져줄 때 핫존(hot zone)으로 올라가는 법을 배웁니다. 왜냐하면 다른 모든 개들이 수박을 너무 좋아해서 다들 자기 자리에 가 있기 때문이죠. 강아지들은 다른 개들에게서 아주 좋은 점도, 그렇지 못한 점도 모두 배웁니다. 이제 공식적인 강아지 훈련 방식 중 하나가 있는데... 두 애즈 아이 두(do as I do)라고 불리는 훈련 방식이 있는데, 사람이 실제로 행동을 시범으로 보여주기 때문에 일부 사람들은 이를 모델링이라고 부르기도 합니다. 그들은 개에게 행동을 모델링합니다. 예를 들어 개가 회전하기를 원한다면, 그들은 직접 회전한 다음 개가 스스로 회전 행동을 할 때까지 기다리거나, 발 타겟을 터치하고 개가 발 타겟을 터치하는 행동을 스스로 할 때까지 기다립니다. 이제 개는 이전에 행동을 스스로 제안하는 것에 대해 어느 정도 이해하게 되었을 텐데, 그렇게 훈련하는 사람들도 있습니다. 다시 말하지만, 저는 제 훈련 과정에서 한 번도 해본 적 없는 방식입니다. 다시 말하지만, 저는 제 훈련 과정에서 한 번도 해본 적 없는 방식입니다. 점점 인기를 얻고 있는 또 다른 형태의 훈련이 있습니다. 그리고 이것은 저에게 흥미로운 주제입니다. 제가 주된 훈련 방식으로 사용하지는 않겠지만요. 저는 그것을 관찰 학습이나 사회적 학습 범주에 넣을 것 같지만, 솔직히 그것이 완전히 독자적인 범주에 속하는지는 잘 모르겠습니다. 소셜 미디어를 보신다면 켈피 종인 스텔라와 쉽어두들 종인 버니라는 유명한 두 마리의 개에 대해 잘 아실지도 모르겠습니다. 이제 이 두 마리의 개는 버튼을 통해 주인과 의사소통하는 법을 배운 것으로 보입니다. 그들은 서로 다른 단어들이 적힌 버튼이 있는 보드를 가지고 있습니다. 그리고 실제로 개들이 의사소통을 위해 완전한 문장을 만드는 것처럼 보이는데, 저는 이것이 꽤 놀랍다고 생각합니다. 아시다시피 전제는 개들이 영어를 이해하기 시작한다는 것인데, 만약 여러분이 개에게 산책 갈까라고 말해봤거나, 제가 '개에게 야한 말을 한다'고 부르는 것처럼 수영이나 시소같이 모멘텀(Momentum)이 알아듣는 단어, 혹은 어떤 어질리티 큐(agility cues)라도 제가 이해하는 단어를 사용할 때 저희 개들은 엄청나게 흥분하곤 합니다. 여러분도 사람들이 쿠키나 할머니 댁 등과 같은 단어를 말하는 소셜 미디어 게시물을 아마 보셨을 겁니다. 좋습니다. 그래서 개가 영어를 이해한다는 전제하에, 그들은 단어들을 활동이나 그 활동에 대한 묘사라는 범주로 분류합니다. 그래서 그들은 개가 밖으로 나가려 할 때 강아지에게 밖으로 나갈 거라고 말하며 버튼을 누르게 하는 식이죠. 그리고 이렇게 말합니다. 밖으로 나갈래? 하면서 밖으로 나가는 과정에서 '밖(outside)'이라는 단어를 다양한 형태로 사용하죠. 그러면 결국 강아지나 개가 스스로 버튼을 누르겠다고 의사를 표현하게 됩니다. 저는 제 개들에게
by watching either a person or another dog. So, if they see a dog do something like, my puppy will learn to get up in a hot zone when watermelon's being thrown around because all the other dogs seem to be in the dog beds because they love watermelon. So, puppies do learn both great things from other dogs and not so great things from other dogs. Now there is a form of formal dog training called do as I do, where the person actually models the behavior, which is why some people call this modeling. They model the behavior for the dog. Like if they wanted the dog to turn, they would turn and then wait for the dog to offer the turn, or they would touch a foot target and wait for the dog to offer the foot target. Now the dog will have some understanding of behavior offering before this, but there are people who do that. Again, nothing that I have ever done in my training. There is another form of training that's growing in popularity. And this one is interesting to me, although I would never use it as my main form of training. I would put it under observational or social learning, but I honestly really don't know if it belongs in a category all of its own. You may, if you watch social media, you may be familiar with two popular dogs, Stella the Kelpie and Bunny the Sheep-a-Doodle. Now it appears that these two dogs have learned to communicate with their owners through buttons. They have a board of buttons, all with different words on them. And the dogs actually appear to be making full sentences to communicate, which I think is pretty mind-blowing. And you know, the premise is that dogs start to understand English, which if you've ever said, do you want to go for a walk to a dog or what I call talking dirty to my dog when I use words that they understand like swimming or seesaw is one that momentum knows or any agility cues, my dogs get super, super excited. You've probably seen the social media posts where people say things like cookies and grandma's house and et cetera. Okay. So, the premise that dogs understand English, they group the words in categories of activities or descriptors of that activities. And so, they teach the dog through social learning in that when they are going to go outside with a puppy, they'll say, we're going to go outside and the person will touch the button. And they'll say, do you want to go outside? And they'll use the word outside in different forms as they're going outside. And eventually the puppy or the dog will offer to hit a button. Now, I have taught my dogs to hit a
8:38
밖으로 나가기 위해 버튼을 누르도록 가르쳤지만, 사회적 학습을 통해 가르친 것은 아닙니다. 저는 그 행동을 셰이핑(shaping)했죠. 현재 키우는 강아지에게 실험 삼아 사회적 학습으로 가르쳐 보았습니다. 믿기 어렵겠지만, 그 녀석은 우리가 밖으로 나갈래?라고 물어볼 때만 버튼을 누릅니다. 그래서, 그 점이 꽤 흥미롭습니다. 전에는 사회적 학습을 해본 적이 없지만, 실험으로서 빌리프(Belief)에게 적용해보고 있으며, 평소 강아지들에게 하는 다른 훈련들도 병행하고 있습니다. 즉, 관찰 학습은 두 가지 다른 범주로 나눌 수 있습니다. 내가 하는 대로 따라 하는 것(do as I do)과 버튼 누르기를 통한 사회적 학습이 바로 그것입니다. 다음 범주는 행동 포착(capturing behavior)입니다. 유튜브로 보고 계시다면, 제 뒤에 있는 사진에서 제 개 스웨거(Swagger)가 자기 엄마 머리에 앞발을 올리고 있는 모습을 보실 수 있을 겁니다. 그건 포착된 행동입니다. 제가 예전에 다른 개에게 팔을 두르는 귀여운 포즈를 가르쳤을 때처럼 말이죠. 그냥 그렇게 포착된 겁니다. 행동 포착이란 행동이 나타날 때까지 기다렸다가 보상을 주는 것을 의미합니다. 제가 가장 좋아하는 포착 행동 중 하나는, 솔직히 말해서, 개에게 다른 개를 안는 법을 가르치고 싶다면, 개가 마법처럼 저절로 그 행동을 보여주길 기다리며 포착하는 것보다 셰이핑을 하는 것이 훨씬 빠릅니다. 그래서 행동 포착에서 제가 좋아하는 것은 개들이 아침에 일어나자마자 기지개를 크게 켜는 행동입니다. 저 말고도 개에게 '기지개 크게 켜자(big stretch)'라고 말하는 분들이 분명 계실 겁니다. 그런 식으로 말하는 거죠, 맞나요? 그건 아주 큰 기지개예요. 오, 정말 크게 기지개 켜네. 그러고 나서 강화해 주는 거예요. 자, 그냥 "큰 기지개네"라고 말하는 것 자체가 사실 강화를 해주는 거예요. 하지만 저는 개들이 그런 행동을 한 뒤에 종종 간식을 줘요. 큰 기지개요. 앞다리를 쭉 펴고 뒷다리 발끝까지 쭉 뻗는 걸 보셨을 거예요. 그게 바로 큰 기지개를 포착(capturing)하는 거예요. 이건 제가 개들에게 뒷발을 뻗고 버티도록 셰이핑(shaping)하는 것과는 달라요. 완전히 다른 거죠. 저는 아침 기지개가 훨씬 더 본능적이라고 생각해요. 그래서 제가 정말, 정말, 정말 포착하고 싶어 하는 행동이에요. 왜 행동을 포착하냐고요? 음, 한 가지 이유는 셰이핑을 했다면 그렇게 본능적으로 나오지 않았을 수도 있기 때문이죠. 또 다른 이유는, 만약 개가 어떤 행동을 할 걸 안다면 포착하기 쉽기 때문이에요. 어떤 개들은 말하듯이 소리 내는 걸 좋아하죠. 다시 말하지만, 제가 소셜 미디어를 너무 많이 봐서 그런지는 몰라도요. 엄마라고 말하는 것 같은 개들 영상 보신 적 있죠? 맞아요. 프렌치 불독이나 퍼그들요. 감자 샐러드라고 말하는 것처럼요. 제 개 피처스는 예전에 우리가 "우우" 소리를 낸다고 말하곤 했는데, 마치 말을 하는 것처럼 들렸어요. 그건 포착된 행동이었어요. 개가 직접 그렇게 했고, 우리는 개에게 대꾸하거나 웃어주거나,
button to go outside, but I haven't taught it through social learning. I've shaped the behavior. With my current puppy, just as an experiment, I did teach it through social learning. And believe it or not, she is hitting that button. Only when we ask her, do you want to go outside? Okay. So, that is kind of interesting. And it's something that I've never done social learning, but I am as an experiment doing it with belief, but I am doing all of the other training that I normally do with puppies as well. So, that observational learning could be divided into the two different categories. The do as I do, and then the social learning through hitting buttons. The next category is capturing behavior. Now, if you're watching this on YouTube, you can see a picture behind me of my dog Swagger putting his paw on his mother's head. Now, that is a captured behavior in that I actually taught one of my dogs, the cute pals, where she put her arm around another dog and it was just captured like that. Capturing a behavior means you wait until it happens and then you just reinforce it. One of my favorite captured behaviors, and honestly, if I do want to shape my dog to put their arm around another dog, shaping is far faster than waiting for them to just magically offer it, capturing it. So, capturing, one of the behaviors that I love to capture is when my dogs wake up first thing in the morning and they give that great big morning stretch. And I bet I am not alone to say to my dog, big stretch. And we all probably say it in that way, right? That's a big stretch. Oh, a big stretch. And then you reinforce it. Now, just you saying big stretch actually reinforces it. But I often give my dogs cookies after they do that big stretch. You know that they walk out the front end and then they point their toes with the back end. That's capturing a big stretch. Now that's different than what I do when I shape my dogs to stretch their rear paw and hold it out there. Very, very different. I think a morning stretch is really more authentic, which is why it's a behavior that I love, love, love to capture. Why would you capture behaviors? Well, for the one reason that they might not offer them as authentically, if you were to shape them. The other reason it's easy, if you know your dog is going to offer something like some dogs like to talk. Again, I spend too much time on social media, but if you heard those dogs that sound like they're saying mama. Right? French bulldogs or pugs, tater salad. My dog features, she used to, we'd say she was giving a woo-woo and she would sound like
11:18
때로는 갈비뼈 쪽을 쓰다듬어 주거나 간식을 주는 방식으로 강화해 줬어요. 그런 것들이 바로 포착하는 행동들이죠. 이제 행동 포착의 단점은 개가 그 행동을 하고 싶어 하지 않으면, 우리가 할 수 있는 게 아무것도 없다는 거예요. 반면에 행동을 셰이핑했다면, 거기에는 많은 기초 단계가 깔려 있어요. 그래서 개가 그 행동을 하기 싫어하면, 저는 그냥 기초 단계로 돌아가서 다시 차근차근 쌓아 올려요. 그러면 개는 이 다른 단계들을 통해 많은 보상을 받았기 때문에 "그래, 할게"라고 반응하게 되죠. 기지개 같은 행동은 만약 개가 기지개를 켜고 싶어 하지 않으면, 그냥 끝이에요. 그걸로 끝이죠. 그래서 포착된 행동이란 건, 마치 사진을 찍는 것과 같아요. 다시 말하지만, 그건 개가 스스로 보여주는 행동을 우리가 타이밍 좋게 보상으로 포착하는 거니까요. 당신은 단지 최종 행동만 얻게 됩니다. 되돌아가서 수정하고 싶을 때 도움이 될 만한 단계가 없습니다. 다음 범주는 셰이핑입니다. 하지만 관찰 학습처럼, 저는 셰이핑을 두 가지 범주로 나눕니다. 그래서 자유 셰이핑(free shaping)이라는 것이 있는데, 이것은 제가 90년대에 행동을 셰이핑하기 시작했을 때의 방식입니다. 그리고 우리가 하는 방식이 있는데, 그것은 곧 설명하겠습니다. 자유 셰이핑은 클리커와 간식 한 그릇을 들고 방에 서서 강아지가 무언가를 하기를 기다리는 것입니다. 개에게 고개를 돌리는 것 같은 행동을 유도하고 클릭하고 보상하는 식이죠. 그리고 개들이 문 쪽으로 움직이면 클릭하고 보상을 줍니다. 물론, 염두에 둔 결과가 있겠지만, 당신은 개가 무언가를 하기를 기다리는 것입니다. 그것은 자유롭습니다. 개들은 그냥 행동하는 것이죠. 자유 셰이핑을 할 때 염두에 둔 결과가 없을 수도 있습니다. 그냥 '어디 한번 개가 무엇을 보여주려나 보자'라고 생각할 수도 있죠. 그래서 개가 하는 모든 행동을 클릭하고 보상하면서 어떤 일이 일어나는지 지켜보는 것입니다. 소품을 놓아두고 강아지나 개가 그것과 상호작용할 때 클릭하고 보상함으로써 자유 셰이핑을 할 수도 있습니다. 가끔 저는 개가 셰이핑에 흥미를 느끼게 하려고 구멍이 뚫린 플라스틱 용기 같은 것에 고기가 붙은 큰 뼈를 넣고, 개가 그 물체와 상호작용하기만 해도 클릭하고 보상을 줍니다. 그렇죠? 그것이 개가 자발적으로 행동을 제시하도록 하는 방법입니다. 그래서 자유 셰이핑은 자유롭습니다. 만약 누군가 개에게 예를 들어 물건을 물어오는 법을 가르치고 싶다면, 바닥에 공을 놓고 개가 그것을 쳐다보거나 다가갈 때 클릭할 것입니다. 제 블로그에는 90년대에 제가 자유 셰이핑으로 행동을 가르쳤던 훌륭한 예시와 지금 제가 하는 방식으로 셰이핑하는 영상이 있습니다. 그 차이는
she was talking. That was captured. She did it. We reinforced it by talking back to her or laughing or giving her a big pat on the ribs, sometimes giving her a cookie for it. And so, those are behaviors that capture. Now, the downside of capturing behavior is if the dog doesn't feel like giving it to you, there's nothing you can do. So, if I've shaped a behavior, there's a lot of foundational layers to it. So, if my dog doesn't feel like giving it, then I just go back to some of the foundational layers, build that up again. And so, that it's more, yes, I will do it because I've been given all this value for these other layers. Something like stretching. If my dog doesn't feel like stretching, we're done. We're done. So, captured behaviors, it's like, you know, capturing a picture. You're just getting the end behavior. There's no stages to get you to a place if you want to go back and fix it. The next category is shaping. But just like observational learning, I'm breaking shaping into two categories. So, there's something called free shaping, which is how I started way back in the nineties shaping behavior. And there's what we do, which I'll get to in a second. Free shaping is where you get yourself a clicker and a bowl of cookies, and you stand in a room and you wait for your dog to do something. And maybe they like, you know, turn their head and you click and you feed that. And maybe they move towards the door and you click and you feed that. Now, obviously you're going to have an outcome in mind, but you're waiting for the dog to do something. It's free. They're just doing things. Now, you might not have an outcome in mind when you're free shaping. You might just go, hey, let's just see what they want to offer. And so, you just click and reward anything the dog does and you see what happens. You can also free shape by putting a prop down and click and reward the puppy or the dog for interacting with that. Sometimes I'll get a dog interested in shaping by putting a container, like a plastic container with holes in it and a big meaty bone in there and click and rewarding the dog for just interacting with that thing, right? It's a way of getting them to freely offer behaviors. So, free shaping is free. So, if somebody wanted to teach a dog, let's say to retrieve something, they might put like a ball down on the floor and they'd click the dog for looking at it and walking towards it. There's a video on my blog that shows a great example of me free shaping a behavior way back in the nineties and then me shaping it the way I do now. And the difference
13:57
제 생각에 제가 90년대나 2000년대 초반에 제 보더 콜리 스토니(Stoney)와 했던 방식인 것 같습니다. 잘 기억나진 않지만, 제 보더콜리 앙코르에게 그 행동을 가르치는 데 3분 30초가 걸렸습니다. 서로 아주 가까운 혈연관계인 개들인데도 똑같은 행동을 배우는 데는 3초밖에 걸리지 않았죠. 그렇습니다. 그래서 저는 자유 셰이핑(free shaping)은 하지 않습니다. 만약 제가 개가 하고 싶은 대로 하도록 놔두고 클릭하고 보상해주고 싶었다면 했을지도 모르겠네요. 솔직히 90년대에는 그렇게 하곤 했지만, 지금은 정말 하지 않습니다. 제가 자유 셰이핑을 하지 않는 이유는 첫째, 원치 않는 쓸데없는 행동들이 많이 형성되어 나중에 없애기 어려울 수 있기 때문입니다. 예를 들어, 제가 개에게 무언가를 집어 올리라고 시켰는데 개가 좌절감을 느끼고 불안해하면서 짖기 시작합니다. 그런 다음 물건을 집어 올리면 그 짖는 행동까지 행동의 일부가 되어버립니다. 행동에 불안함이 섞이게 되는 거죠. 그래서 저는 그런 쓸데없는 행동을 원치 않습니다. 짖거나 아니면 발로 땅을 파는 등의 행동이 나타날 수 있는데, 왜냐하면 이전에 그런 행동들로 보상을 받았기 때문이죠. 제 블로그에서 언급했던 영상 클립을 보면, 제 개 스토니가 쿨러 위에 올라가게 하려고 할 때 전에 강화되었던 기억 때문에 갑자기 점프해서 로프를 물어버리는 장면이 나옵니다. 그래서 자유 셰이핑은 자유롭긴 하지만 체계가 훨씬 부족합니다. 우리가 하는 방식은, 딱히 정해진 이름이 있는지 모르겠네요. 솔직히 말해서 저는 레이어드 셰이핑(layered shaping)이나 효율적인 셰이핑, 혹은 의도적인 셰이핑이라고 부르고 싶습니다. 우리가 하는 방식은 환경을 조작해서 개가 올바른 선택을 하도록 명확하게 만드는 것입니다. 제가 두 마리의 똑똑한 보더콜리, 서로 매우 가까운 관계인 두 보더콜리에게 같은 행동을 가르쳤던 이유입니다. 지금의 훈련 방식으로는 백 배는 더 빠르게 해낼 수 있었죠. 가르치려 했던 행동은 개가 쿨러 위에 올라가게 하는 것이었습니다. 스토니를 훈련할 때는 방 건너편에 서 있었죠. 오늘 다시 한다면 아마도 아이스박스 뚜껑을 열고 그 위에 덮개를 씌우곤 했죠. 강아지가 그 위로 올라가게 하는 겁니다. 저는 덮개를 치우고 강아지가 그 덮개 위로 올라가게 한 다음, 다시 아이스박스에 덮개를 씌웁니다. 그러니까 저는 더 많은 단계로 나누기도 해요. 이게 무슨 장점이 있냐고요? 행동을 잘게 나눌수록 행동이 성장하면서 자신감도 더 커집니다. 어질리티를 하려고 시소 장비를 새로 샀는데, 강아지를 바로 시소 위에 올려놓는 사람들을 본 적이 있어요. 강아지는 발톱을 세우고 조금 불안해 보이죠. 맞아요. 단계가 없기 때문이에요. 저는 뭐랄까,
was I believe the one I did with my Border Collie Stoney way back when, it might've been in early 2000, I can't really recall, but it took three and a half minutes for her to get the behavior with my Border Collie Encore, dogs very closely related. It took like three seconds to get the same behavior. Okay. So, free shaping to me, I don't do it. The only time I would do it is if I just wanted to click and reward a dog for doing whatever they felt like doing, which honestly I used to do back in the nineties, but I really don't anymore. The reason why I don't free shape is because I believe, number one, you build in a lot of cheap behaviors you might not want, and that might be difficult to get rid of. For example, if I wanted my dog to pick something up and they're getting frustrated and anxious and they'll start vocalizing, and then they pick something up and you get that vocalizing into the behavior. You get the anxiety built into the behavior. So, I don't want the cheap behaviors like barking or they might start, you know, digging their feet or whatever they're doing because they got rewarded for that before. In the video clip that I spoke about on my blog, my dog Stoney jumps up and grabs a rope because it was reinforced before when I was trying to get her to jump on a cooler. So, free shaping, it's free. It's a lot less structured. What we do, I don't know that there's a name for it. I would call it layered shaping or efficient shaping, honestly, more intentional shaping. And what we do is we manipulate the environment so that the correct choice for the dog is the obvious choice. And that's why I trained two border collies, two brilliant border collies, two border collies very, very closely related to each other, the same behavior. And it took like a hundred times less time to do it the way I train now. The behavior was get your dog to jump on a cooler. And I was standing across the room when I did it with Stoney. If I was doing it today, I probably would take the lid off the cooler and put a cover on it. The dog would get up on that. I would take the cover off, the dog would get on that cover and then I put the cover on the cooler. So, I'd even break it into more layers. What's the advantage of this? The more layers you put into a behavior, the more confidence that grows as the behavior grows. I've seen people who say, Hey, I got a new seesaw. And so, I want to do agility with my dog. I've put my dog over it. They look like their nails wrote and they're a little bit worried in there. Yeah. Because there's no layers involved. I've got like, I don't know,
16:32
시소를 가르칠 때 20~30단계는 거칩니다. 그래요. '레이어드 셰이핑(Layered Shaping)'이라는 용어, 괜찮은 것 같네요. 이번 팟캐스트에서 새로운 용어를 만든 건가요? 레이어드 셰이핑, 정말 많은 단계가 있어서 최종 행동으로 이어질 때 자신감이 생깁니다. 그래서 아주 많은 자신감을 가진 채 최종 행동에 도달하는 거죠. 차이점은 만약 조금 흐트러진다면, 그 단계 중 하나로 되돌아가서 보상을 더 많이 주면 됩니다. 그러면 훌륭한 행동이 완성되죠. 네, 지금까지 보상 기반 훈련 프로그램에서 강아지를 훈련하는 5가지 범주, 아마 6가지 정도의 방법론을 다뤄봤습니다. 보시다시피 강아지를 훈련하는 7가지의 아주 다른 방법들이 있죠. 다음 에피소드에서는 제가 왜 전통적인 유도 방식의 보상 기반 훈련보다 레이어드 셰이핑을 더 선호하는지 더 깊이 파고들어 보려고 합니다. 궁금한 점이 있으시면, 유튜브에 오셔서 여러분은 어떤 방식의 보상 기반 훈련을 하는지 알려주세요. 혹시 제가 빠뜨린 방법론이 있을까요? 없기를 바라지만, 그럴 수도 있겠죠. 다음 시간에 'Shaped by Dog'에서 다시 만나요.
20 or 30 layers to how I teach a seesaw. Okay. So, with layered shaping, I think that's a good name. Did we just point a phrase here on this podcast episode? Layered shaping, there's so many layers that grow confidence leading to the final behavior. So, you get to the final behavior with a lot of confidence. And the difference is if something gets a little bit sloppy, go back to any one of those layers and put more reinforcement in it. And then you've got brilliant behavior. Okay. So, there you have it. Five categories, probably six different methodologies of training a dog in a reinforcement-based program. But as you can see, seven vastly different ways of training a dog. In our next episode, I'm going to dig a little deeper into why I prefer layered shaping over the more traditional lured approach to reinforcement-based dog training. If you have any questions, jump on over to YouTube and let me know, what is your approach to reinforcement-based dog training? And am I forgetting a methodology? I hope not, but it's possible. See you next time right here on Shaped by Dog.