In Reward Based Training Is Telling A Dog "No" An Aversive?
Michael EllisLeerburg Dog Training Podcast · 팟캐스트 · 8분
반려견 훈련은 마커를 활용한 소통 체계를 기반으로 진행되며, 보상 기반 훈련의 핵심은 개가 보호자와의 상호작용을 놀이처럼 인식하게 하는 것입니다. 훈련 초기에는 준비, 예, 아니오, 굿, 다 했어와 같은 다섯 가지 기본 단어를 사용하여 개에게 행동의 성공 여부와 지속 시간을 정확히 전달합니다. 보상 없음 마커는 개가 올바른 행동을 수행하지 못했을 때 사용하는 피드백 도구로서, 개가 좌절하지 않고 목표 행동을 찾도록 돕는 역할을 합니다. 프리 셰이핑 방식은 개가 스스로 행동을 제안하도록 유도하지만, 훈련사는 적절한 마커를 통해 개가 수행해야 할 행동의 기준을 명확히 제시함으로써 훈련의 효율성을 높입니다. 특정 행동을 수행하는 동안 개가 다른 행동을 시도하면, 훈련사는 아니오 마커를 통해 그 행동이 목표가 아님을 알려주고 다시 올바른 자세를 유지하도록 유도합니다. 성공적인 보상 기반 훈련은 개와 보호자 간의 신뢰 관계를 형성하며, 개 스스로 보호자가 요구하는 행동을 하고 싶어 하는 의욕을 고취합니다.
보상 기반 훈련에서 개에게 "안 돼"라고 말하는 것은 혐오 자극인가요?
In Reward Based Training Is Telling A Dog "No" An Aversive?
0:00
이 내용은 마커 트레이닝과 개에게 혐오 자극으로 '안 돼'를 사용하는 것에 대한 것입니다. 먼저 읽어본 뒤, 그 기본 원리와 언제 사용해야 하는지에 대해 이야기하겠습니다. '안 돼'. 최근 마커 트레이닝을 주제로 한 기사를 읽었습니다. Leerberg 기사는 아닙니다. 구체적으로 '보상 없음 마커(no reward marker)', '부정적 마커(negative marker)', 또는 '조건적 처벌자(condition punisher)'의 사용에 관한 것이었습니다. 그 기사는 보상 없음 마커의 사용이 겉보기엔 그렇지 않아 보여도, 사실 개에게는 혐오 자극이 되며, 보상 없음 마커를 사용하는 것이 개가 성공하지 못했거나 잘못된 행동을 했다는 것을 알려주는 데 도움이 되지 않는다고 주장합니다. 그들은 올바른 행동을 했을 때 격려만 해주는 것이 개가 더 빨리 배운다고 말했습니다. 그런 생각은 처음 들어봅니다. 당신은 경험이 많으시니 이 문제에 대한 당신의 생각을 듣고 싶습니다. 좋은 질문입니다. 그 기사는 프리 셰이핑(free shaping)을 하는 사람들이 쓴 것입니다. 프리 셰이핑을 하는 사람들은 훈련의 학습 단계에서 개를 전혀 교정하지 않습니다. 그들은 개가 행동을 수행하거나 점진적 근사(successive approximation)를 보일 때만 보상합니다. 이것은 최종 목표한 방식대로 완벽하게 수행하지는 않았더라도 시도했기 때문에 그 시도에 보상하겠다는 의미입니다. 그렇게 그 시도에 보상을 주는 것이죠. 그리고 개가 조금 더 노력하고, 조금 더 노력하게 유도하여 점차 목표 행동에 도달하도록 합니다. 그것이 점진적 근사입니다. 그리고 그것이 프리 셰이핑이 작동하는 핵심입니다. 우리는 그것에 동의하지 않습니다. 프리 셰이핑은 효과가 있습니다. 다만 시간이 오래 걸립니다. 하지만 우리가 생각하기에 보상 기반 훈련은 개와 소통하는 방법입니다. 그것은 하나의 소통 체계입니다. 그리고 우리는 훈련 과정에서 언제 여러 마커를 사용해야 하는지 이해해야 합니다. 새로운 반려견이나 새로운 훈련사를 교육하기 시작할 때, 마커 트레이닝이 무엇인지 이해하려면 우리는 다섯 가지 기본 단어로 시작합니다. 그 단어들은 훈련이 진행됨에 따라 점점 더 많아질 수 있으며, 숙련된 훈련사들은 매우 많은 마커를 사용할 수 있습니다. 하지만 아주 처음에는 다섯 가지 단어로 시작합니다. 그 단어들이 무엇인지, 그리고 어떻게 서로 맞물리는지 여기서 알려드리겠습니다. 시작하는 단어, 음, 우선 그 단어들은, 준비되셨나요? 예, 아니오, 굿, 그리고 다 했어(all done)입니다. 이것이 바로 다섯 가지 단어입니다. 이제 시작할 때, 우리는 반려견에게 훈련 세션을 시작할 것임을 알려주어야 합니다. 그리고 우리는 '준비됐니?(Are you ready?)'라고 말함으로써 그것을 합니다. 준비됐니? 마커를 충전하고 나면, 즉 마커가 고가치의 보상으로 이어진다는 것을 반려견이 이해하게 되면 그들이 깨닫는 데는 오래 걸리지 않습니다. 우리가 '준비됐니?'라고 말할 때 그들이 흥분하는 데는 그리 오랜 시간이 걸리지 않습니다. 두 번째 단어는 '예(Yes)'입니다. '예'는 특정 시점에 반려견이 하고 있던 행동을 표시합니다.
This is about marker training and using no as an aversive to the dog. First I'll read it and then I'm going to talk about the foundation for it and when to use no. I saw an article, not a Leerberg article, recently on the subject of marker training, specifically the use of a no reward marker or a negative marker or a condition punisher. The article argues that the use of the no reward marker, while it may not seem like it, is actually an aversive to the dog and that using a no reward marker does not help to tell the dog that they did not succeed or that they did something wrong. They said that the dogs learn faster if just encouraged through what they do correctly. This is the first I've ever heard of that idea. I know you have a lot of experience, so I'd appreciate your thoughts on the matter. A good question. The article was written by people that free shape. People that free shape do not correct their dogs at all during the learning phase of training. They only reward when the dog either does the behavior or does successive approximation, which means he didn't do it totally the way we're going to end up with it, but he tried and we're going to reward the try. And we're going to reward the try. And then we're going to reward that he has to try a little bit harder and a little bit harder and then gradually the dog's going to get to the behavior. That's successive approximation. And that's the guts of how free shaping works. We don't agree with it. Free shaping does work. It takes a long time. But as far as we're concerned, reward-based training is a method of communicating with your dog. It's a communication system. And we need to understand when to use various marks within our training. When we start a new dog, when we start a new dog trainer, understanding what marker training is, we start with five basic words. Those words, as training progresses, can get more and more and more and experienced trainers can have many, many marks. But the very beginning starts with five words. I'll tell you what they are and how they fit together here. The beginning word, well, first of all, the words are, are you ready? Yes, no, good, and all done. Those are the five words. Now when you start, we have to tell our dogs that we're going to start a training session. And we do it by saying, are you ready? Are you ready? It doesn't take long for them to realize once we've charged the mark, once they understand that a mark results in a high value food reward. It doesn't take them long to get excited when we say, are you ready? The second word is yes. Yes marks a specific thing a dog was doing at a point in time.
3:10
저는 마이클 엘리스, 우리의 친구 마이클 엘리스가 처음에 우리에게 설명했던 방식이 마음에 드는데, 행동을 마킹할 때, 그것은 마치 반려견이 그 보상을 받기 위해 방금 했던 행동을 마음속으로 사진을 찍는 것과 같습니다. 개들은 사람처럼 생각하지 않습니다. 많은 사람들이 그렇게 생각하지만, 실제로는 그렇지 않습니다. 완벽한 예가 여기 있습니다. 우리가 반려견에게 앉아서 5초 동안 앉아 있기를 기대한다면, 개에게 그것은 하나의 행동이 아닙니다. 그것은 두 가지 행동입니다. 개에게 '앉아'는 엉덩이를 바닥에 붙이는 것을 의미합니다. 엉덩이가 바닥에 닿는 순간, 그것으로 그 행동은 끝납니다. 왜냐하면 그것이 전부이기 때문입니다. 그가 저에게 앉으라고 했고 저는 앉았습니다. 만약 그가 제가 그곳에서 5초를 더 앉아 있기를 바란다면, 그건 다른 문제입니다. 그러니 한 번 더 해보죠. 우리는 개에게 앉으라고 합니다. 개는 앉습니다. 우리는 5초 동안 앉아 있기를 기대하기 때문에 이 시점에서는 마크(mark)하지 않을 것입니다. 따라서 개가 앉을 때 우리는 '예스(yes)' 명령을 하지 않습니다. 우리는 '좋아, 좋아, 좋아'라고 말합니다. 이제 충분히 오래 앉아 있었으니, 그때 마크를 할 것입니다. 우리는 지속 시간을 마크하는 것입니다. 우리는 '예스'라고 말하며 주변을 뛰어다니고, 개는 일어나서 음식 보상이나 장난감 보상을 받을 수 있습니다. 하지만 개의 입장에서는 그저 두 가지 행동을 수행한 것뿐입니다. 개는 앉았고, 5초 동안 앉아 있었습니다. 보통 '노 리워드 마커(no reward marker)'는 지속 시간 중에 나옵니다. 여기에는 주의 사항이 하나 있습니다. 그것은 만약 개가 '앉아'를 알고, '엎드려'를 알고 있으며, 당신이 마음속으로 100% 확신할 때, 개에게 '앉아'를 시켰는데 하지 않는다면, '아니, 앉아'라고 말하는 것입니다. 이에 대한 주의 사항은 '앉아'나 '엎드려'를 가르치고 있고 개가 당신과 완전히 교감 중일 때, 개에게 '엎드려'라고 했는데 앉는다면, 그냥 '아니, 엎드려'라고 말하는 것입니다. 그것이 훈련 초기에 노 리워드 마커를 사용하는 한 가지 방법입니다. 훈련이 진행됨에 따라 개가 더 오래 앉아 있기를 원할 때, 우리가 할 일은 개에게 앉으라고 하는 것입니다. 개는 앉습니다. 개는 '앉아'를 알고 있습니다. 우리는 그것을 마크하지 않을 것입니다. 우리는 개가 5초 동안 앉아서 기다리길 원합니다. 그러니 '앉아'. '앉아' 잘했어. '앉아' 잘했어. '앉아' 잘했어. 이제 우리는 '예스'로 마크할 것입니다. 그래서 우리는 '예스'라고 말하고, 개는 튀어 올라 보상을 받거나, 장난감을 얻거나, 무엇이든 하게 합니다. 보상을 받으려면 제대로 해야 한다는 것을 '안 돼'라는 말로 개에게 알려주세요. 왜냐하면 개가 튀어 오르더라도 '앉아'는 할 줄 알거든요. 하지만 튀어 오르며 방방 뛰기 시작하면 '아니'라고 말하세요. 개에게 통제권을 유지할 수 있도록 리드줄을 잡고 움직이게 하거나 다시 데려오세요. 앉으라고 명령하세요. 잘했어. 잘했어. 옳지. 앉아. 앉아. 앉아. 앉아. 앉아. 앉아. 앉아. '앉아, 잘 앉았어, 좋아, 잘했어, 앉아'라고 말하고 '옳지'라고 한 뒤 보상을 주세요. 그건 개에게 혐오 자극이 아닙니다. 만약 개가 튀어 오르며 '내 보상은 어디 있어?'라는 듯 행동하면 '아니, 앉아야지, 앉아'라고 하고 개가 앉으면, 잘 앉았어, 좋아, 잘했어, 옳지. '옳지' 마커를 사용할 때는 생동감 있게 반응하세요.
And I like what Michael Ellis, our friend Michael Ellis, the way he originally explained it to us was, when you mark a behavior, it's like that dog takes a photograph in his mind of what he just did to earn that food reward. Dogs don't think like people. A lot of people think they do, but they don't. Here's a perfect example. If we expect a dog to sit and sit for five seconds, that's not one behavior to a dog. That's two behaviors. To a dog, sit means drop my butt on the ground. As soon as my butt is on the ground, that's the end of that behavior, because that's all he asked me to do was sit and I sat. If he expects me to sit there for five more seconds, that's a different thing. So one more time. We ask the dog to sit. He sits. We're not going to mark it at this point because we expect him to sit for five seconds. So when he sits, we don't give a yes command. We say, good, good, good. Now he sat long enough, then we will mark that. We are marking the direction. We'll say, yes, and jump around and he can jump up and he can get his food reward or his toy reward. But in the dog's mind, he's just done two behaviors. He sat and he sat for five seconds. Now, usually, usually the no reward marker comes in during duration. There's a caveat to that. And that is, if the dog knows sit, if the dog knows down, and you're 100% sure in your mind that he knows sit, you ask him to sit and he doesn't do it, you say, no, sit. The caveat to this is if you're teaching the sit or the down, and the dog's all engaged with you, and you ask him to down and he sits, just say, nope, down. That's one way of using the no remarker in the beginning of training. As we progress, as we progress and we want him to sit longer, what we'll do is we'll ask him to sit. The dog sits. He knows sit. We're not going to mark that. We want him to sit and stay for five seconds. So sit. Good sit. Good sit. Good sit. Now we're going to mark it with a yes. So we say yes, and he can bounce up, get his reward, get his toy, whatever. Using no to tell a dog, no, if you want your reward, you're going to have to do it right. Because if he would bounce up, he knows sit. But if he would bounce up and start jumping around, say, nope. Move him around, bring him back, keep him on leash so you have control of him. Ask him to sit. Good. Good. Yes. Sit. Sit. Sit. Sit. Sit. Sit. Sit. You're saying sit, good sit, good, good sit, yes, and get the reward. That is not an aversive to the dog. If the dog bounces up and acts like where's my reward, nope, got to sit, sit, he sits, good sit, good, good, yes. Get animated when you use your yes marker.
6:27
생동감을 가지되, '안 돼' 마커를 사용할 때 화난 것처럼 들리지 않게 하세요. '안 돼, 안 돼'라고 말할 필요는 없습니다. 그건 개와의 관계를 쌓는 데 도움이 되지 않으니까요. 보상 기반 훈련의 목표는 개와 관계를 쌓고 당신이 요구하는 행동을 개 스스로 하고 싶게 만드는 것입니다. 개는 당신과 나와서 놀고 싶어 합니다. 제대로 훈련하면 개는 훈련을 놀이로 인식합니다. 개들은 방방 뛰며 오히려 당신을 훈련시키려 할 겁니다. 알아서 여러 행동을 시도할 테니까요. 개들은 다양한 행동을 스스로 제안할 것입니다. 자유 형성(free shaping) 훈련을 하는 사람들의 경우, 훈련이 그 단계에 이르고 개가 방방 뛰기 시작하면, '이봐, 진정해'라고 말해줄 의사소통 수단이 없어서 개들이 좌절하게 될 겁니다. 앉거나 엎드리거나, 내가 요구하는 행동을 해야 한다는 것을 알려줘야 합니다. 개들이 이리저리 뛰기 시작할 텐데, 아무런 피드백도 받지 못하는 상태가 됩니다. 아무런 보상도 받지 못하는 거죠. 그러면 짖기 시작하고 위아래로 점프하기 시작할 겁니다. 그건 모두 역효과를 낳을 뿐입니다. 처음 단계에서의 '노(no)' 마커는 혐오 자극이 될 수 있지만, 이는 생산적인, 생산적인 혐오 자극으로서 여러분의 반려견 훈련에 긍정적인 효과를 가져다줍니다. 보상 기반 훈련을 처음 접하시는 분들을 위해 저희는 여러 온라인 강좌와 이해를 도울 수 있는 스트리밍 강좌들을 제공하고 있습니다. 저는 마커를 활용한 반려견 훈련의 힘에 관한 강좌를 진행한 적이 있습니다. 저희는 캘리포니아에 있는 친구 마이클 엘리스와 함께 반려견과 줄다리기 놀이를 하는 힘에 관한 온라인 강좌와 영상을 제작했습니다. 또한 '음식을 활용한 반려견 훈련의 힘'이라는 제목의 또 다른 강좌도 진행했습니다. 그래서 저희 웹사이트에는 마커 훈련을 이해하는 데 도움이 될 아주 좋은 정보가 많습니다. 그리고 보상 기반 훈련에 대해서도 마찬가지고요. 어떤 사람들은 이를 마커 훈련이라고 부르기도 합니다. 저는 워낙 오랫동안 해와서 그렇게 부르는 경향이 있지만, 사실 이는 보상 기반 훈련입니다. 저는 이 일을 정말 오랫동안 해왔으니까요.
Be animated, don't sound mad when you're using a no marker. You don't have to say no, no. That's not going to build a relationship with your dog. In reward based training the whole goal is to build a relationship with a dog and make him want to do the things you ask him to do. He wants to come out and play with you. When you do this right, they look at our training as play. They jump around and they try and train you. They offer behaviors, okay. They'll offer different behaviors. And what happens with the free shaping people is once they get to that stage of training and their dog is jumping around, the dogs are going to get frustrated because there's no line of communication there to tell them, hey, settle down. You got to sit or you got to down or you got to do what I'm asking you to do. The dogs will start to jump around and they're not getting any feedback. They're not getting any rewards. They'll start to bark, they'll start to jump up and down. That's all counterproductive. The no marker to the beginning can be an aversive, but it's a productive, it's a productive aversive that has a positive effect on your dog training. And for those people that are new to reward-based training, we have several online courses and streaming courses that'll help you understand. I did one on the power of training dogs with markers. We did an online course and a video with Michael Ellis, our friend, in California on the power of playing tug with your dog. We did another one titled, The Power of Training Your Dog with Food. So we have very good information on our website to use to help you understand marker training and reward-based training. Some people call it marker training. I've been doing it so long I tend to do that, but really it's reward-based training. I've been doing it so long.