5 Simple Hacks to Help Your Dog Learn Faster (Reinforcement Process) #31
Susan GarrettDogs That · 영상 · 20분
반려견이 더 빨리 배우도록 돕는 간단한 팁 5가지 (강화 과정) #31
5 Simple Hacks to Help Your Dog Learn Faster (Reinforcement Process) #31
0:00
여러분 안녕하세요, Shaped by Dog에 오신 것을 환영합니다. 저는 수잔 개럿입니다. 오늘 저는 여러분의 반려견 훈련을 훨씬 더 쉽고, 훨씬 더 효과적이며, 훨씬 더 효율적으로 만드는 방법을 알려드리려고 합니다. 반려견 훈련의 효과를 극대화하는 방법을 가르쳐 드리고 싶습니다. 이 방법은 여러분들의 반려견 훈련 여정의 어느 단계에 있든 동일하게 적용됩니다. 저처럼 전문가이든, 독 스포츠 선수이든, 특히 복종 훈련 스포츠 대회에 참가하시는 분들을 위해 준비한 특별한 내용이 있습니다. 초보자이거나, 아직 반려견이 없어서 기다리는 중이라도 상관없습니다. 지금 이 내용을 잘 기억해 두시면, 새로운 반려견이나 강아지를 맞이했을 때 훈련 방식에 엄청난 차이를 가져올 것입니다. 우선 기본 개념부터 시작해보죠. 모든 훈련은 가치의 전달입니다. 무엇을 배우든 어떤 행동을 하든 상관없이, 강아지든 사람이든 그 과정에서는 가치의 전달이 일어납니다. 예를 들어, 여러분은 지금 이 팟캐스트를 듣고 계십니다. '개'라는 단어를 구글링하다가 우연히 처음 발견했을 수도 있지만, 이상적으로는 다시 찾아오게 되겠죠. 여러분이 다시 듣는 이유는 이 방송이 여러분에게 가치를 제공했기 때문입니다. 제 목표는 정보, 교육, 그리고 재미라는 가치를 전달하는 것입니다. '더블 E'를 달성했죠. 제가 추구하는 바입니다. 여러분을 즐겁게 해드리는 한편, 훌륭한 반려견 훈련 교육을 제공하고 싶습니다. 여러분은 직장에 나갑니다. 거기에도 가치의 전달이 있죠. 여러분이 업무를 수행하면 급여를 받습니다. 만약 급여를 받지 못한다면, 계속 직장에 나가시겠습니까? 다른 사람들을 위해 일하면서 기쁨을 느끼는 가치를 얻는다면 그럴 수도 있겠죠. 모든 행동은 가치의 전달입니다. 우리 반려견들이 우리와 함께 어질리티를 하거나, 옆에서 나란히 걷거나, 혹은 우리를 위해 장난감을 가져오는 이유는 우리가 훈련할 때 그들에게 가치를 제공했기 때문입니다. 모든 행동은 가치의 전달입니다. 우리 반려견들이 우리와 함께 어질리티를 하거나, 옆에서 나란히 걷거나, 혹은 우리를 위해 장난감을 가져오는 이유는 우리가 훈련할 때 그들에게 가치를 제공했기 때문입니다. 그런 행동들을 하기 위해서요. 좋습니다. 모든 훈련은 가치의 전달입니다. 어떻게 하면 여러분의 훈련을 강화할 수 있을까요? 여러분의 훈련을 어떻게 더 효과적이고 효율적으로 만들 수 있을까요? 여러분의 훈련을 어떻게 더 효과적이고 효율적으로 만들 수 있을까요? 어떻게 하면 우리는 여러분의 훈련을 더 효과적이고 효율적으로 만들 수 있을까요? 어떻게 하면 여러분의 훈련을 더 효과적으로 만들 수 있을까요? 로스트 비프 조각으로 시작한다고 가정해 봅시다. 이것은 여러분의 반려견에게 어떤 가치를 지닙니다. 그러니 뭔가를 고르세요. 반려견에게 엄청난 가치가 있는 것이요. 반려견은 '아, 너무 좋아'라고 하겠죠. 자 이제, 우리는 그 가치를 행동으로 전달해야 합니다. 하지만 반려견 훈련 과정에서 그 가치의 일부를 잃게 됩니다. 그 가치는 어떻게 손실될까요? 바로 반려견 훈련사로서 우리의 기법 때문입니다.
Hey everybody, welcome to Shaped by Dog. I am Susan Garrett and today I'm going to share with you how you can make your dog training a lot easier, a lot more effective, a lot more efficient. I want to teach you how to supercharge your dog training. And it's going to be the same regardless of where you are in your dog training journey. If you're a professional like me, if you're a competitor in sports, something special for those of you who compete in the sport of obedience, it doesn't matter if you're brand new or if you don't even have a dog and you're just waiting. This is something you can take note of and that when you get your new dog or puppy, it will make a massive difference when you train. To start, let's get a base level. All training is a transfer of value. It doesn't matter what you're learning or what behavior you are doing, whether you are a dog or a human being, there's a transfer of value that occurs. For example, you're listening to this podcast. You might happen on it the first time maybe just because you Googled the word dog, but you would return to it ideally. You're returning to it because it brought you some value. My goal is to bring the value of information, education, and entertainment. I got the double E. That's what I'm looking for. I'm hoping to entertain you, but to give you some great dog training, education. You go to work. There's a transfer of value. You do a behavior, you get a paycheck. If you didn't get a paycheck, would you still be going to work? You might if you got the value of doing something for others that filled you with joy. Every behavior is a transfer of value. Our dogs will do agility with us, or they will walk by our side, or they will retrieve a toy for us because of the value we brought to them when we were training them to do those behaviors. All right. All training is a transfer of value. How are we going to supercharge your training? How are we going to make your training more effective and efficient? How are we going to make your training more effective and efficient? How are we going to make your training more effective and efficient? How are we going to make your training more effective? Let's say we start with a piece of roast beef and it has a value to your dog. So, pick something that's got a massive value to your dog. The dog says, this is my, oh, I love that. Now, we need to take that value and transfer it into the behavior. But the process of dog training loses some of that value. And how is the value lost? By our mechanics as dog trainers.
3:12
그래서 우리가 해야 할 일은 가치를 유지하기 위해 우리의 기법을 어떻게 개선할지에 대한 이해를 높이는 것뿐입니다. 만약 로스트 비프가 반려견에게 10의 가치를 가진다면, 우리는 그 10을 전달하여 우리가 훈련 중인 행동의 끝에서 10의 가치를 얻고 싶어 합니다. 앉기, 엎드리기, 장난감 쫓기, 손님에게 앞발 올리지 않기, 식탁에서 구걸하지 않기 등 무엇이든 상관없이, 혼자 있을 때 짖지 않기 등 어떤 행동이든 우리는 그 전달을 원합니다. 훈련할 때 가치를 잃고 싶지 않은 거죠. 그 가치 전달 말입니다. 우리가 로스트 비프의 10이라는 가치를 끝까지 유지할 수 있다면 우리의 훈련은 가장 효율적이고 효과적일 것입니다. 그것이 바로 저와 같은 전문가, 즉 최고 수준의 도그 스포츠에서 탁월한 성과를 거두고 성공을 경험한 사람과 그렇지 않은 사람의 차이입니다. 제 가치 전달 방식이 더 효과적이고 효율적이기 때문이죠. 지금 바로 여러분도 그렇게 할 수 있도록 도와드리겠습니다. 네 가지가 있습니다. 우선, 제 멘토인 밥 베일리가 즐겨 하던, 정말 훌륭한 만트라를 공유할게요. 강화는 하나의 과정이지, 이벤트가 아니라는 것입니다. 즉, 여러분의 반려견이 앉았을 때 '착하지'라고 말하며 간식을 줍니다. 강화는 간식이 개의 입으로 들어가는 것만이 아닙니다. 강화에는 과정이 있습니다. 그 과정을 이해하고 존중하기 시작하면, 그 과정이 바로 로스트비프가 가진 가치를 '기다려'로 전달하는 방법이며, 그 과정에서 가치를 잃지 않고 행동을 수행하는 도중에 가치가 떨어지지 않게 하는 방법입니다. 왜 우리는 그렇게 유지하고 싶어 할까요? 개에게 로스트비프를 사랑하는 마음이 있다면, 우리가 '엎드려'라고 말할 때 그 로스트비프에 대한 사랑을 그 행동으로 옮길 수 있기 때문입니다. 그러면 개는 빠르게, 그리고 신나게 엎드릴 것입니다. 개는 매번 그렇게 할 것입니다. 그것이 바로 가치 전달이 제대로 이루어졌음을 알 수 있는 방법입니다. 이제 훈련 자체를 살펴보겠습니다. 먼저 우리는 개가 보여주는 반응들을 평가합니다. 개가 하고 있는 행동들 중에서 우리가 원하는 것을 고르는 거죠. 방금 이야기했던 '엎드려'를 예로 들어 봅시다. 엎드려의 기준이 무엇인지 알아야 합니다. 무엇이 엎드려인가요? 개가 더 이상 서 있거나 앉아 있지 않을 때라고 할 수 있겠죠. 좋아요. 그런데, 배는 바닥에 닿았지만 신이 나서 팔꿈치가 떠 있다면, 그것도 엎드려일까요? 그 부분에 대해 고민해 보세요. 사실 고민할 필요조차 없어야 합니다. 여러분만의 기준을 알고 있어야 하니까요. 왜냐하면 기준을 모르면 훈련할 수 없기 때문입니다. 훈련에 들어가기 전에 여러분의 기준이 무엇인지 확인하세요. 이제 여러분의 훈련은 보상하고 싶은 행동이 나올 때까지 개가 하는 행동을 평가하는 것입니다.
And so, all that we have to do is improve our understanding of how to improve our mechanics in order to maintain the value. If roast beef is a value of 10 to your dog, we want to transfer that 10 and get a value of 10 at the end of the behavior that we're training. Regardless if it is sit down, chase a toy, keep your feet off my guests, don't beg at the table, don't bark when you're alone. Whatever the behavior is, we want the transfer. We don't want to lose the value by when we're training. That transfer value. Our training is most efficient and effective if we can hold that 10 value for the roast beef all the way through. That's the difference between a professional like me, somebody who has excelled at the highest level of dog sport and had success because my transfer of value is more effective and efficient. I'm going to help you get yours there right now. There's four things. First of all, I'm going to share with you. My mentor, Bob Bailey, he has this great mantra and it is that reinforcement is a process. It's not an event. Meaning when your dog sits and you say, good boy, here's a cookie. The reinforcement isn't the cookie getting to the dog's mouth. There is a reinforcement process. And if you understand it and you start respecting it, that process is how we get the transfer of value from the roast beef of 10 to stay and not lose any value, not go down on the way to getting into the behavior. Why do we want to keep it that way? Because if your dog loves roast beef and we can transfer that love for roast beef into something like lying down when I say, your dog's going to lie down fast and they're going to lie down excited. They're going to do it every single time. That's how you know you've got a great transfer of value. And so, let's look at training itself. First, we're just evaluating the responses that our dogs get, our dogs offering, our dogs are doing. And then we see the one we like. Let's pick the down since we were talking about that. You have to know what is the criteria of a down. What makes it a down? Is it when the dogs, you can say, well, the dog's no longer in a stand or a sit. All right. So, their stomach's on the ground, but their elbows are off because they're excited. Is that a down? Let me think about that. You shouldn't have to think about it. You know your criteria. Because you can't train it if you don't know it. Know what your criteria is before you get into that. Now your training is evaluating what your dog is doing until you see what it is that you want to reward.
6:11
자, 개가 '엎드려'를 이해했고 여러분은 더 빠른 엎드리기 등에 보상하고 싶다고 가정해 봅시다. 첫 번째 단계는 모든 반응을 살펴보고 그중에서 마음에 드는 것을 포착하는 것입니다. 그러니까, 그게 강화 과정의 1단계입니다. 기억하시나요? 강화는 이벤트가 아니라 과정이라는 것을요. 첫 번째는 마음에 드는 행동을 골라내는 것이고, 그다음은 표시(마킹)를 하는 것입니다. 그런데 항상 그렇게 하는 것은 아닙니다. 모든 행동에 마킹을 하되, 대부분의 행동에 마킹을 한다고 가정해 봅시다. 클릭커를 사용해 마킹할 수도 있습니다. 저는 그 방법을 좋아합니다. 단어로 마킹할 수도 있죠. '좋아' 같은 말이요. '예스'처럼 다른 단어로 마킹할 수도 있습니다. 제 말은 단어가 짧을수록 좋다는 것입니다. 그것이 강화 과정의 일부이니까요. 자, 그래서 우리는, 첫 번째 단계는 여러분이 무엇을 좋아하는지 인지하거나 파악하는 것입니다. 쾅. 두 번째는 마킹하는 것입니다. 그리고 사이에는 간격이 있습니다. 좋아하는 행동을 보고, 좋아하는 부분을 골라내고, 그다음에 마킹하는 과정 사이의 간격이죠. 그것이 강화 과정의 일부입니다. 세 번째는 보상을 전달하는 것입니다. 이것 역시 강화 과정의 일부입니다. 여러분들이 좋아하는 행동에 마킹을 한 후 보상을 전달하기까지의 간격입니다. 그리고 마지막은 보상을 배치하는 것입니다. 마지막이라고 말하면 안 되겠네요. 한 가지가 더 있는데, 그것은 바로 행동으로부터의 해제, 즉 반려견에게 다음 단계로 넘어가도 좋다는 허락을 주는 것입니다. 그러니 다섯 가지가 있다고 해보죠. 좋습니다. 좋아하는 것을 보는 것. 1단계는 기준을 알고 그것을 평가하는 것입니다. 그리고 좋아하는 것을 보는 것과 그것을 인지하는 것 사이의 간격입니다. 이때 감정을 배제해야 합니다. 알겠죠? 왜냐하면 팟캐스트 16편 '그것 이전의 그것'을 기억하시나요? 음, 강화 과정에서 반려견은 '그것 이전의 그것'을 매우 빠르게 알아차립니다. 그래서, 훈련 중에 지나치게 몰입해서 머리를 한쪽으로 기울이고 있다면요. 그러다가 개념을 잡기 시작할 때, 여러분이 움직이기 시작합니다. 간식을 주러 가야 하니까요. 반려견은 여러분이 몰입한 상태에서 보상을 주러 이동한다는 것을 알아차리고, 행동을 멈추거나 서둘러 끝내버릴 것입니다. 세기말에 있었던 말에 관한 행동 연구가 있는데 그 말은 '영리한 한스(Clever Hans)'라고 불렸습니다. 독일에서 있었던 일이죠. 이 남자는 말에게 어떤 수학 문제든 낼 수 있었는데, 나눗셈, 산수, 곱셈, 덧셈, 뺄셈 등 온갖 종류의 수학 문제를 풀게 할 수 있었습니다. 심지어 문장제 문제도 있었던 것 같아요. 그리고 그 말은 80% 이상의 확률로 정답을 맞혔죠. 제 생각엔 87% 확률로 맞혔던 것 같아요. 그래서 행동주의자들은 매우 놀랐죠. 그리고 말 주인이 아닌 다른 사람들이 질문을 해도, 다른 사람들이
Now let's say your dog understands a down and you want to reward, you know, faster downs or whatever. The first part is that you are able to look at all the responses and see the one you like. So, that's step number one in the reinforcement process. Remember? Reinforcement is a process, not an event. The first is picking out what you like. The next is marking it. Now you don't always mark all behaviors, but let's assume most of them you do. It could be marking with a clicker. I like it. It could be marking with a word. Good. Could be marking with a different word like, yes. I mean the shorter the word, the better. That's part of the reinforcement process. All right. So, we've, first step is acknowledging or recognizing what you like. Boom. Second is marking. And there's a gap between seeing what you like, picking out what you like, and then marking it. That's part of the reinforcement process. The third is delivering the reward. That's part of the reinforcement process. The gap between you marking what you like and you delivering the reward. And the final thing is the placement of the reward. And I shouldn't say the final thing. There is one other thing and that is the release from the behavior, giving the dog permission to move on. So, let's say there's five things. All right. Seeing what you like. Step number one, know your criteria and then evaluate it. And the gap between seeing what you like and acknowledging it. And you need to be non-emotional. All right. Because remember podcast episode number 16, the thing before the thing. Well, the reinforcement process, the dog picks up the thing before the thing very quickly. So, if you are training and you're like all intense and your head's all cocked to the side. And then when the dog's starting to get it, you start moving. Cause you're going to move towards your cookie. And the dog recognizes you went from intense to moving towards reinforcement and they will stop the behavior or they will cut it short. There is an old turn of the century behavior study with a horse called Clever Hans. And it was in Germany. This fellow could give the horse any kind of math problem, division, arithmetic, multiplication or addition or subtraction, anything, all kinds of math problems. I think there was actually even word, word problems. And the horse like over 80% of the time got it right. I think it was 87% of the time got it right. And so, that was astonishing to these behaviorists. And other people like could ask the question, not only the fellow who owned the horse, other people could
9:09
질문을 해도 말이 정답을 맞혔어요. 결론부터 말씀드리면, 그들이 발견한 것은 말이 포커에서처럼 신호를 알아챘다는 거예요. 제 추측으로는 처음에 한스의 주인은 말의 발을 아주 유심히 보고 있었을지도 몰라요. 말이 정답을 맞혔을 때 주인이 고개를 들면 말은 '아, 저 사람이 고개를 움직이니까 보상을 받겠구나'라는 걸 알게 된 거죠. 그게 주인이 보상을 주기 전에 하는 첫 번째 행동이었으니까요. 결국 말은 사람들의 눈을 보기 시작했어요. 그래서 계속 발로 땅을 긁는 거예요. '5 더하기 5는 뭐야?' 하면 발로 땅을 긁기 시작하는 거죠. 그러다가 사람들이 고개를 들기 시작하는 걸 보면 발로 땅을 긁는 속도를 늦췄어요. 그러다가 사람들이 고개를 특정 방향으로 돌리면 땅 긁기를 멈췄죠. 그게 바로 '행동 이전의 행동'이에요. 그것이 여러분의 훈련에 영향을 미치고 효율성을 떨어뜨릴 거예요. 그래서 1단계를 평가할 때, 여러분이 보고 있는 기준을 평가할 때는 감정을 배제하고 마커를 찍을 때까지 움직이지 않아야 합니다. 이제 행동을 본 순간과 마커를 찍는 순간 사이의 간격이 중요합니다. 예를 들어 반려견이 무언가를 하고 있는데 '아, 그거, 그거 잘했어. 내가 좋아하는 게 바로 그거야.'라고 생각하고 클리커를 가져와서 클릭하려는데, 어, 클리커가 거꾸로 있네. 바르게 고쳐 잡고, 그러고 나서 클릭을 한다고 칩시다. 자, 그 간격 동안 여러분이 보상하고 있는 건 원래의 행동이 아니라, 그 행동을 유지하는 시간입니다. 그러니 가정을 해보자면 개가 엎드린 자세를 취하고 그대로 유지했다고 칩시다. 여러분은 지금 엎드리는 행동이 아니라 그 자세를 유지하는 행동을 마킹하고 있는 거예요. 실제로, 아 좋아요, 당신의 팔꿈치가 바닥에 닿았네요. 그거 마음에 들어요. 아니면 이런 일이 일어났을 수도 있죠, 당신의 개가 엎드리기(down)를 했을지도 모르는데, 당신이 클리커를 만지작거리는 동안 개가 그 상태에서 일어나 앉기(sit)를 했을 수도 있습니다. 그래서 당신은 잘못된 행동을 클릭하게 되어, 훈련을 비효율적으로 만들게 되는 거죠. 그래서 첫 번째 단계는, 침착하게 당신이 원하는 것을 골라내는 것입니다. 두 번째 단계는 마킹하는 것이고, 세 번째 단계는 강화물을 제공하는 것입니다. 자, 마킹을 했습니다. 좋아요. 마음에 듭니다. 그러고 나서 주머니에 손을 넣어 쿠키를 꺼냅니다. 그런데 '아, 이건 너무 크네' 싶죠. 좋아요. 하나를 반으로 쪼갭니다. 그리고 나머지는 다시 주머니에 넣고 쿠키를 주려는데, 아마 손에 몇 개 더 쥐고 있다가 그걸 다시 빼겠죠. 그럼 그 전달 과정은 엉망인 겁니다, 그렇죠? 주머니에 손을 넣어 무엇이 있는지 뒤적거리고, 쿠키가 너무 커서 쪼개는 데 걸린 시간 때문에요.
ask the question, the horse got it right. What they, long story short, they found was the horse picked up tells like poker tells, right? So, my hallucination is at first Hans's owner might've been looking at his feet really intently. And when the horse got it right, he put his head up and then the horse knew, oh, I'll get a reward because his head's moving. And that's the first thing he does before he rewards me. Eventually the horse was just looking at people's eyes. So, he would just keep pawing the ground. Like what's five plus five? He'd start pawing the ground. And when he'd see that people were starting to lift their head, he'd slow the pawing the ground down. And then when they turned their head a certain way, he stopped pawing the ground. That is the thing before the thing. That is going to influence your training and it's going to make it less efficient. So, when you are evaluating step one, when you're evaluating the criteria you're seeing, you need to be unemotional and not be moving until you've marked it. Now the gap between seeing it and marking it is important. Because let's say your dog is doing something and you go, oh, that, that was good. I think that's the one I like. Yeah. I got to get my clicker and then I'm going to, I'm going to click, oh, the clicker's upside down. I'm going to turn it up right side up, and then I'm going to click it. Well, that gap, what you're rewarding now is not the behavior, you're rewarding the duration. So, let's assume the dog went into a down and stayed into a down. You are now marking the stay of the down, not the actual, oh good, your elbows hit the ground. I like that. Or what might also have happened, your dog might've gone into a down and while you were fumbling with your clicker, he might've got up out of the down and then gone into a sit. And then you clicked the wrong thing, making your training inefficient. So, step number one, going through and unemotionally picking out what you want. Step number two, marking it. Step number three is the delivery of the reinforcement. So, you've marked it. Good. I like that. And then you go into your pocket and you pick out a cookie. And then it's like, oh, that one's too big. Okay. I'll break that one in half. And then I'll put these ones back in my pocket and then I'm going to give you the cookie, but maybe I got a couple other in my hand and I'm going to pull that away. So, that delivery was crap, right? The time it took to reach into your pocket and to sift through what you have and to break them up because they were too big.
11:26
그래서 그만큼의 시간이 지연된 거죠. 기억하세요, 당신은 매우 비효율적으로 하고 있는 겁니다. 게다가 다른 쿠키들을 손에 쥔 채로, 즉 손 안의 주머니에 담긴 상태로 쿠키를 전달합니다. 그러면 개는 한 개를 받으면서 열 개가 사라지는 것을 봅니다. 잠깐, 나 계산할 줄 알거든. 나한텐 별로였어. 그러니 보상을 줄 때는 빠르고 효율적으로 전달하세요. 쿠키 한 개만 꺼내서 그 한 개만 주는 겁니다. 왜냐하면 많은 개들이 보상의 가치가 사라지는 것을 목격하기 때문입니다. 그리고 당신이 준 그 작은 한 개는 보상이 아니라 오히려 처벌이 될 수도 있습니다. 왜냐하면, '어, 나머지는 다 어디 갔지?' 훈련 중에 불안감을 유발할 수도 있죠. 물론 어떤 개들은 그것을 감수하는 법을 배우겠지만, 제가 알기로 테리어 품종을 훈련할 때, 그들은 가치를 빼앗길 때 절대 좋아하지 않습니다. 그래서 전달 방식이 중요합니다. 강화물을 미리 준비하세요. 손에 쥐고 있거나... 저는 항상 쿠키 한 개를 가지고 있습니다. 훈련할 때 제 손에 들고 있는 거죠. 그러니까 쾅, 쾅, 쾅, 쾅 이렇게 할 수 있는 겁니다. 그 간격을 줄이고 싶으신 거죠. 원하는 것을 보고, 마킹하고, 보상을 전달하세요. 쾅, 쾅, 쾅, 쾅. 아주 빠르게 말이죠. 조금 더 쾅, 쾅 하고요. 쾅, 쾅. 여러분의 반려견 훈련에 쾅, 쾅을 좀 더 더하세요. 이게 이번 에피소드 제목이어야겠네요. 어쩌면요. 이야기가 샜네요. 네 번째 요소는 위치 선정입니다. 그런데 수잔, 방금 보상 전달이라고 하셨잖아요. 그게 그거 아닌가요? 아니요, 아니죠. 위치 선정은 비법입니다, 여러분. 정말 핵심 비법이죠. 예를 들어 반려견이 엎드려 자세를 취하게 하고 싶다고 해보죠. 마킹을 하고 '좋아'라고 말한 뒤 보상을 전달하러 갈 때 정말 빠르게 움직여야 합니다. 보상을 줄 때 반려견이 간식을 받으려고 일어서려고 할 수 있거든요. '그냥 내가 너에게 갈게' 하는 거죠. 그러니까 제가 중간에서 미리 가서 반려견이 간식을 쉽게 먹을 수 있도록 도와주는 겁니다. 그리고 반려견이 간식을 받고 바로 엎드려 자세로 돌아가게 하는 거죠. 만약 반려견이 일어났다면 사실 기준을 어긴 것에 대해 보상한 셈이 됩니다. 그러니 보상의 위치 선정은 여러분의 목표가 무엇이든 그 목표에 기여해야 합니다. 복종 훈련을 하시는 분들을 위해 말씀드리자면, 복종 훈련에서는 물어오기 훈련을 많이 합니다. 반려견은 아주 구체적인 물건을 물어와야 하죠. 나무나 플라스틱 막대 양 끝에 종이 달린 것 같은 물건 말입니다. 이걸 덤벨이라고 부릅니다. 반려견이 이걸 물어오면 입에서 놓을 때 보상을 주죠. 그래서 많은 반려견들이 물고 있기를 싫어합니다. 제 반려견들에게 이 훈련을 할 때는 덤벨을 물고 있을 때 보상을 줍니다. 반려견이 입에 덤벨을 문 상태에서 간식을 먹게 하고 그다음에 제가 덤벨을 가져갑니다. 그건 그 자체로 하나의 긴 과정이라서 이번 팟캐스트 범위에는 포함되지 않습니다. 하지만 제 요점은 보상의 위치가 궁극적인 목표에 기여하느냐는 것입니다. 제 궁극적인
So, that took time away from the gap. Remember you're getting very inefficient. And then you deliver the cookie with other cookies in the smaller of your hand, in the pocket of your hand. So, the dog gets one and sees 10 going away. Wait a minute. I can do math. That sucked for me. So, when you're delivering the reward, deliver it fast and efficient. One cookie goes in and you give the one cookie. Because a lot of dogs are going to see the value being taken away. And that little one you give may become more of a punishment than a reward because, oh, hey, what about the other ones? It could cause some angst in your training. Now some dogs will learn to live with it, but I know training terriers, they are not happy when you take away the value. So, the delivery is important. Have your reinforcement ready. Maybe have it in your hand or have it in a… I always have one cookie in my hand when I'm training. So, I could be boom, boom, boom, boom. You want to cut down those gaps. See what you want, mark it, deliver it. Boom, boom, boom, boom. Super quick. A little more boom, boom. Boom, boom. Get a little more boom, boom into your dog training. Should be the name of this episode. Maybe. I digress. The fourth element is the placement. Well, Susan, you just said the delivery. Isn't that the same thing? Oh, nay, nay. The placement is secret sauce, guys. It's secret sauce. So, let's say you want your dog to go into a down. You mark it, you say good, you go to deliver it really quickly. And as you deliver it, your dog kind of comes up to reach you. I'm just going to meet you. So, I'm going to help facilitate you giving me that cookie by meeting you part way. And then they take the cookie and they go right back into the down. You actually have rewarded them for breaking criteria. So, the placement of the reinforcement needs to contribute to your goal, whatever your goal is. So, for you obedience people out there, a lot of people in obedience, we have a retrieve and the dog has to retrieve something very specific. It's like a piece of wood or plastic dowel with bells on the end. It's called a dumbbell. And so, the dog retrieves it. They get rewarded when it comes out of their mouth. So, a lot of dogs don't want to hold it. Now, when I train my dogs to do this, I actually deliver the cookies to them while they're holding it. They actually get cookies in their mouth and then I take it out. That's a process all on its own. Not really in the scope of this podcast. But my point is the placement of the reinforcement, does it contribute to your ultimate goal? My ultimate
14:08
목표는 반려견이 덤벨을 물고 있게 하는 것이니, 덤벨을 물고 있는 상태에서 보상을 줄 것입니다. 도그 어질리티 훈련을 할 때, 반려견이 시소 위에 머물기를 원하실 겁니다. 반려견이 시소 위에 있을 때 보상을 주고 계신가요? 아니면 '아, 착하다'라고 말하고 반려견이 당신에게 걸어와서 보상을 받게 하시나요? 보상 위치가 최종 목표를 달성하는 데 도움이 될까요? 매우 중요합니다. 보상 위치와 관련된 다른 부분으로 첫째, 목표 달성에 도움이 되는가입니다. 둘째, 당신과의 거리가 동일한가입니다. 예를 들어, 식사 중에 반려견이 식탁에서 구걸하지 않도록 자기 침대에 머물게 하고 싶다고 해보죠. 그래서 가끔 일어나서 침대에 있는 반려견에게 간식을 줍니다. 매일 그렇게 반복하면 당신은 반려견이 자기 침대에 머무는 법을 배우고 있다고 믿게 됩니다. 행동에 관해서는 이것이 핵심입니다. 실제로 우리가 무엇을 훈련하는지 확실히 아는 것은 반려견뿐입니다. 우리는 우리가 무엇을 훈련하고 있다고 생각할지 모르지만, 사실 알 수 없습니다. 항상 마지막 결정권은 반려견에게 있으니까요. 어쩌면 반려견의 입장에서는 간식을 받기 위해 당신 곁에 머무는 것이 전부일지도 모릅니다. 그래서 식탁이 너무 멀어지면 반려견은 더 가까이 다가오기 시작할 겁니다. 어떻게 하면 좋을까요? 예를 들어, 반려견에게 나로부터 멀어지는 직선 방향으로 달리게 가르치고 싶다면, 보상을 주기 위해 반려견을 다시 내게로 부르지는 않을 겁니다. 그 대신 그곳에 원격 급식기를 사용하거나, 직선으로 달려 나간 것에 대한 보상으로 장난감을 던져 줄 것입니다. 보상의 위치와 당신과의 근접성이 중요합니다. 그게 어떤 모습일까요? 반려견이 방을 나갈 때 짖지 않도록 가르치고 싶다면, 계속 방으로 돌아가서 간식을 줘서는 안 됩니다. 다른 사람이 방에 들어가 간식을 주게 하거나 원격 급식기에 투자하세요. 좋습니다. 이런 점들을 알고 있어야 합니다. 무엇을 평가해야 할지 빨리 파악하고 결정을 내릴 수 있어야 합니다. 원하는 행동을 본 순간 바로 결정을 내리고, 즉시 보상하거나 표시를 해줘야 합니다. 만약 여러분이 그렇게 하고 있다면 행동을 마킹하고 보상을 빠르고 효과적으로 전달하세요. 보상을 전달할 때는 보상의 위치가 어디인지, 그리고 그것이 내 최종 목표와 나로부터의 거리, 그 최종적인 근접도에 어떻게 기여하는지 고려해 보세요. 마지막으로, 개가 보상받은 자세를 여전히 유지하고 있을 때 릴리스 신호를 주어야 합니다. 그것은 강화의 효과를 극대화하는 것입니다. 제가 이곳 'Shaped by Dog'에서 허락의 힘에 대해 이야기했던 것을 기억하시나요? 그 에피소드에서 저는 우리의 말, 즉 릴리스 신호가 어떤 개들에게는
goal is to have my dog hold the dumbbell. I will reward him with his mouth still holding the dumbbell. When you're training in dog agility, you want your dog to stay on the seesaw. Are you rewarding the dog when they are on the seesaw? Or are you saying, Oh, good boy. And he walks towards you and gets his cookie walking towards you. Does the placement of the reinforcement contribute to your ultimate goal? Super important. The other part of placement of reinforcement, number one, does it contribute to your goal? Number two, is it the same proximity to you? So, for example, you want your dog to just hang out in his bed while you eat dinner so he's not begging at the table. And every now and again, you get up and you feed him in his bed. And you keep doing that day after day. And you believe the dog's learning to stay in his bed. Here's the thing about behavior. Only the dog really knows for sure what we're training. We may think we know what we're training, but we don't know. The dog always has the last word. So, maybe in the dog's mind, it's all about staying close to you to get cookies. So, if the table gets too far away, the dog's going to start moving closer. What could you do? If for example, I wanted to teach my dog to run in a straight line away from me, I wouldn't call him back to me to reward him. I'd either use a remote feeder out there, or I would throw a toy to reward him for running out in a straight line. The placement of the reinforcement is important and the proximity to you. What does that look like? If you want a dog to learn to not bark when you leave the room, then you can't keep going back to the room to feed them. Either have somebody else go back to the room to feed them or invest in a remote feeder. All right. So, you've got to know those things. Be able to see what you want to evaluate quickly, make the decision. And the moment you see what you want and you've made that decision, you've got to be able to get in there and reinforce or mark the behavior if that's what you're doing. Mark the behavior and then deliver the reinforcement fast and effectively. And when you deliver the reinforcement, then consider what is the placement of the reinforcement and how does that contribute to my end goal and the end proximity away from me. And finally, you're going to give your release while the dog is still holding the position you rewarded. That is like supercharging the supercharge. Remember I talked about the power of permissions here right on Shaped by Dog. In that episode, I said our words, our releases are as reinforcing
17:02
우리가 반려견 훈련에서 사용하는 간식보다 더 큰 강화물이 될 수 있다고 말했습니다. 따라서 가치 전달의 마지막 단계이자 음식이나 장난감을 사용할 경우 그 높은 가치를 유지하는 방법의 마지막 조각은 개가 릴리스 신호를 들었을 때 여러분이 설정한 기준을 유지하고 있는지 확인하는 것입니다. 물론, 여러분은 항상 릴리스 신호를 주어야 합니다. 그러니 개에게 '엎드려'라고 한 뒤 그냥 일을 하러 가버리면 안 됩니다. 그러면 개가 혼란스러워할 테니까요. 항상 그래야 합니다. 개에게 앉아, 엎드려, 서 있어 같은 자세 신호를 주거나, 어질리티에서 타겟에 멈춰 서 있거나 출발선에서 기다리게 하는 등 무엇을 하든, 개가 통제된 자세를 유지해야 하는 모든 상황에서는 반드시 릴리스 신호가 뒤따라야 합니다. 제가 제안하는 것은 릴리스 신호를 주기 전에 개가 기준을 유지하고 있는지 확인하라는 것입니다. 예를 들어, 개가 앉아 있기를 원하는데 개가 앞으로 기울기 시작하고, 엉덩이가 바닥에서 조금씩 들리기 시작한다면, 개가 그런 행동을 하고 있을 때 릴리스 신호를 준다면, 여러분은 무엇에 대해 보상하고 있는 걸까요? 맞습니다. 차를 운전하면서 대답하시는 소리가 들리네요. 여러분은 바로 그 행동에 대해 보상하는 것입니다. 엉덩이를 바닥에 대고 있어야 한다는 기준을 어겼을 때 개를 풀어주는 것입니다. 자, 이제 막 새로운 반려견을 맞이할 준비를 하시는 분들께 이번 내용이 너무 어렵게 느껴지지 않았기를 바랍니다. 이런 기초적인 행동 훈련에서 우리 같은 전문 반려견 훈련사들만큼 성공을 거두는 것은 매우 쉽습니다. 여러분은 그저 사용하고 있는 간식이나 장난감의 가치를 인식하기만 하면 됩니다. 그리고 그 가치를 유지하고, 이상적으로는 다음 다섯 가지 사항을 염두에 두어 훈련의 효과를 극대화하고 계신가요? 무엇을 보상할지 확인하고, 마킹하고, 보상을 전달하고, 보상 위치를 선정하고, 릴리즈 신호를 주는 것입니다. 이 팟캐스트가 도움이 되셨다면, 부탁 하나만 드려도 될까요? 강아지를 사랑하는 다른 한 분, 혹은 그 이상에게 이 팟캐스트를 공유해 주시겠어요? 여러분의 소셜 미디어 페이지에 공유해 주신다면 정말 감사하겠습니다. 제 목표는 전 세계의 반려인들이 자신의 반려견을 더 잘 이해하도록 돕고, 반려견들이 전 세계 어디서나 최상의 삶을 누릴 수 있도록 돕는 것이기 때문입니다. 그 목표에 기여해 주시는 점 미리 감사드립니다. 그럼 다음 'Shaped by Dog'에서 뵙겠습니다. Shaped by Dog.
into some dogs more reinforcing than the cookies we are using in our dog training. And so, the final piece to this transfer of value in maintaining that high value of your food or your toys, if you're using toys, maintaining that value, the final piece is making sure that the dog is maintaining the criteria you've established when you deliver the release word. And of course, you are always, always giving a release word. So, you're not going to tell your dog down and then go to work because then you're going to confuse your dog. Always. If we give our dogs a positional cue, sit down, stand, or in agility doing a target and stopping at the bottom of something or waiting at a start line. Whatever we do, anything that requires a dog to hold a control position must always be followed up with a release word. And in what I'm suggesting is make sure the dog is maintaining criteria before you give that release word. So, if you want your dog to sit and he starts leaning forward and it turns into like his butt just starts lifting off the ground a little bit and a little bit, you actually, when you, if you gave a release word when the dog was doing that, you would be releasing them for what? That's right. I heard you say that while you're driving in your car. You would be releasing the dog for breaking your criteria of keeping your butt on the ground. Now, I hope this wasn't too overwhelming for those of you who are just getting ready to get a new dog. It's super easy to be as successful in these foundational behaviors as any one of us professional dog trainers. All that you need to do is recognize the value of the food or the toy that you're using. And are you maintaining that value and ideally supercharging your training by being mindful of those five points? See what you're going to reward, mark it, deliver it, placement of reinforcement and give a release. Hey, if you're finding value in this podcast, would you do me a favor? Would you share it with one other dog loving person or maybe more? Share it on your social media pages. I would forever be indebted to you because my goal is to help dog owners worldwide better understand their dogs and to help dogs worldwide to have the best life possible. And I thank you in advance for contributing to that goal. I'll see you next time on Shaped by Dog.