Michael Ellis on Common Terminology Used in Dog Training
Michael EllisLeerburg Dog Training Podcast · 팟캐스트 · 18분
반려견 훈련은 고전적 조건화와 조작적 조건화라는 두 가지 주요 학습 원리에 기반을 둔다. 고전적 조건화는 중립적인 자극을 일차 강화물과 연합하여 반응을 이끌어내는 과정으로, 음식과 같은 보상을 예측하게 하는 신호를 형성하는 데 활용된다. 반면, 조작적 조건화는 행동과 결과 사이의 인과관계를 조절하여 특정 행동의 발생 빈도를 결정한다. 정적 및 부적 강화는 행동을 높이고, 정적 및 부적 처벌은 행동을 낮추는 역할을 수행하며, 이 과정에서 긍정과 부정은 보상의 추가나 제거라는 과학적 정의를 따른다. 훈련 효율을 높이기 위해 핸들러는 마커를 사용하여 정확한 의사소통 체계를 구축하고, 루어링, 공간 압박, 셰이핑 등의 기법을 통해 목표 행동을 세분화하여 가르친다. 또한 개의 동기 부여 수준을 파악하여 헝거 드라이브와 프레이 드라이브를 적절히 활용하고, 가림 현상을 방지하기 위해 신호의 우선순위를 조절하는 것이 중요하다. 지속적인 집중 상태인 인게이지먼트를 유지하면서 보상 이벤트의 상호작용성을 높이는 방식은 훈련의 전반적인 생산성을 향상시킨다.
마이클 엘리스(Michael Ellis)의 개 훈련 전문 용어 해설
Michael Ellis on Common Terminology Used in Dog Training
0:00
네, 반려견 훈련에는 반려견 훈련만의 전문 용어나 언어가 있습니다. 우리가 반려견 훈련 원칙에 대해 서로 소통하려면 같은 언어를 사용하고 있는지 확인하는 것이 중요합니다. 용어를 사용할 때마다 자주 들을 수 있는 몇 가지 일반적인 용어들을 정의해 보겠습니다. 좋습니다, 첫 번째 용어는 고전적 조건화입니다. 고전적 조건화는 파블로프 조건화 또는 연합 학습이라고도 합니다. 이것은 이반 파블로프라는 러시아 심리학자가 발견했는데, 그는 실험실에서 개의 타액 반응에 대한 실험을 하던 중 방 안에 있는 연구원들의 존재가 자신의 실험에 영향을 미치고 있다는 것을 알아차렸고, 그래서 그는 연구원들이 없는 상태에서 개들에게 먹이를 주기 위해 자동 급식 시스템을 설치했습니다. 종소리가 먹이 공급을 알리고 자동 급식기를 통해 먹이가 제공되었는데, 이전에는 개에게 아무런 의미가 없던 중립적인 신호였던 종소리가 먹이를 예견하게 됨으로써 개들에게 먹이와 동일한 반응을 이끌어내게 되었습니다. 그래서 종소리 후 먹이, 종소리 후 먹이를 반복하여 개들이 종소리에 반응해 실제 먹이에 반응하는 것과 똑같이 침을 흘리게 되었고, 그는 이를 고전적 조건화라고 불렀습니다. 고전적 조건화는 반려견 훈련에서 항상 작용하며, 우리는 훈련의 특정 부분에서 의도적으로 이를 사용하여 곧 이야기할 우리 의사소통 시스템의 일부인 마커를 조건화하고, 환경 내의 특정 자극에 개들을 조건화하기도 합니다. 또한 우리가 원하든 원하지 않든 언제든지 발생하는데, 개에게 의미 있는 자극을 예견하는 중립적인 자극을 반복적으로 노출할 때마다 고전적 조건화가 일어날 수 있습니다. 우리가 알아야 할 중요한 사항 중 일부는 반려견 훈련에 적용되는 고전적 조건화에 대해 기억해야 할 점은 파블로프는 중립 자극이 제시되는 시점이 보상, 즉 이 경우 일차 강화물이라 불리는 음식과 동시에 이루어지면 고전적 조건화가 일어나지 않는다는 것을 발견했습니다. 또는 중립 자극이 개가 음식을 먹는 도중에 제시되어도 고전적 조건화는 일어나지 않았습니다. 조건화가 일어나려면 반드시 음식의 등장을 예측해야 하며, 바로 직전에 발생해야 합니다. 이 점을 명심하세요. 이것은 이 DVD 전반에 걸쳐 아주 많이 다루게 될 개념입니다. 소리나 중립 자극은 반드시 의미 있는 자극인 일차 강화물보다 앞서야 합니다. 고전적 조건화에 대해 우리가 기억해야 할 또 다른 점은 고전적 조건화가 일어난 후, 즉 조건화가 완성된 후에 나타나는 개 내부의 반응은
All right, dog training has its own set of jargon or language specific to dog training and it's important if we are going to communicate with each other about principles of dog training that we make sure we're speaking the same language. I'll try to define as we use the terms some of the common terms you'll hear. Okay, our first term is classical conditioning. Classical conditioning is also known as Pavlovian conditioning or associative learning. It was discovered by a Russian psychologist named Ivan Pavlov who was doing experiments in laboratories on salivary responses in dogs and he noticed that the presence of the researchers in the room was affecting his experiments and so he set up an automated feeding system in order to feed the dogs with the researchers not present and so a bell would signal the delivery of food and food would be dispensed by automatic feeders and the bell which was previously a neutral signal to a dog meant nothing to the dogs through pet predicting food came to prompt the same response in the dogs that food did. So bell food, bell food, repeatedly until the dogs were responding to the bell by salivating in the same way that we would actually respond to food and he called this classical conditioning. Classical conditioning is at work all the time in dog training we use it deliberately in certain aspects of our training to condition our markers which are part of our communication system that we'll talk about shortly and to condition dogs to certain stimuli in the environment as well and then it's happening whether we want to or not anytime we're repeatedly exposing the dog to a neutral stimuli that predicts a meaningful stimuli classical conditioning can occur. Some of the important things we need to remember about classical conditioning as it applies to dog training is that Pavlov discovered that if the neutral stimuli was presented at the same time as the reward or the food what we call the primary reinforcer in this case then no classical conditioning happened or if the neutral stimuli was presented while the dog was eating the food there was no classical conditioning. In order for it to happen it had to predict the production it had to happen right before and so keep this in mind this is a this is a concept that we're going to come back to extensively in this DVD. The sound or the neutral stimuli must precede the meaningful stimuli the premier primary reinforcer. Another thing we want to keep in mind about classical conditioning is that after classical conditioning occurs after it has been achieved the response the internal response in the dog is
2:36
불수의적이라는 것입니다. 이것은 매우 중요하며 나중에 다시 다룰 예정이지만, 사람들이 고전적 조건화에 대한 개의 반응은 선택이 아니라 불수의적인 반응임을 이해했으면 합니다. 다음 용어는 조작적 조건화로, 도구적 학습이라고도 합니다. 조작적 조건화는 반려견 훈련의 모든 인과관계를 다룹니다. 이것은 반려견 훈련의 일환이자 모든 동물이 학습하는 방식입니다. 사실 인간도 고전적 조건화와 조작적 조건화를 통해 학습하지만, 조작적 조건화는 개가 자신의 행동과 결과 사이에서 만드는 연관성을 지배하며, 개는 결과에 대한 이전 경험을 바탕으로 특정 행동을 하기로 선택합니다. 저는 이것을 반려견 훈련의 인과관계 부분이라 부르며, 조작적 조건화에는 네 가지 사분면이 있습니다. 네 가지 사분면은 정적, 부적, 강화, 그리고 처벌이라는 단어들의 모든 조합입니다. 처벌입니다. 여기에는 어떠한 판단도 포함되어 있지 않으므로, '긍정(positive)'이 '좋다'는 뜻은 아니며 '부정(negative)'이 '나쁘다'는 뜻은 아닙니다. 긍정이란 우리가 방정식에 무언가를 더한다는 의미입니다. 부정은 우리가 방정식에서 무언가를 제거한다는 의미입니다. 강화는 개가 앞으로 그 행동을 할 가능성이 더 높아진다는 것을 의미하고, 처벌은 앞으로 그 행동을 할 가능성이 더 낮아진다는 것을 의미합니다. 이것이 조작적 조건 형성의 4가지 사분면입니다. 가능성이 더 낮아진다는 것을 의미합니다. 이것이 조작적 조건 형성의 4가지 사분면입니다. 이것이 조작적 조건 형성의 4가지 사분면입니다. 첫 번째는 정적 강화입니다. 이는 개에게 보상을 주는 고전적인 방식입니다. 개가 앉으면 저는 간식 한 조각을 건넵니다. 제가 방정식에 더한 음식, 즉 일차적 강화물인 보상입니다. 그 결과로 제가 개에게 간식을 주게 만든 행동인 '앉기'가 앞으로 일어날 가능성이 더 높아집니다. 정적 강화입니다. 부적 강화. 우리는 보통 이것을 교정이나 혐오적인 결과라고 잘못 생각합니다. 그렇지 않습니다. 부적 강화에서 개는 자신의 행동을 통해 불쾌한 일이 일어나는 것을, 즉 어떤 혐오적인 경험이 일어나는 것을 멈추게 합니다. 개는 자신의 행동을 통해 불쾌한 일이 일어나는 것을, 즉 어떤 혐오적인 경험이 일어나는 것을 멈추게 합니다. 그래서 예전 방식대로 개가 앉을 때까지 목줄을 위로 당겨서 앉게 가르치는 것은 부적 강화의 일종입니다. 저는 목줄을 당깁니다. 개가 앉습니다. 저는 당기는 것을 멈춥니다. 부적 부분은 제가 목줄을 불쾌하게 당기는 것을 제거했다는 것입니다. 그 행동을 멈추기 위해 개가 수행한 행동은, 이 경우에는 앉는 것인데, 이제 앞으로 일어날 가능성이 더 높아집니다. 부적 강화입니다. 불편함을 멈추게 하는 개의 행동은 이제 앞으로 일어날 가능성이 더 높아집니다. 조작적 조건 형성의 다음 용어는 정적 처벌입니다. 정적 처벌은 전통적인 교정 방식입니다. 이는 혐오적인 결과입니다. 예를 들어, 만약 제 개가 저에게 뛰어오르고 제가 개의 가슴을 무릎으로 쳤다면, 또는
involuntary and this is going to be very important we'll come back to this later but I want people to understand that the dog's response to classical conditioning is not a choice it's an involuntary response. Our next term is operant conditioning also called instrumental learning. Operant conditioning covers all of the cause and effect of dog training. This is the part of dog training where or how any animal learns actually humans as well through both classical and operant conditioning but operant conditioning governs the associations dog make with their behaviors and outcomes and they choose to do certain behaviors based on their previous experience with the outcomes. I call this the cause and effect part of dog training and there are four quadrants to operant conditioning. The four quadrants are all combinations of the words positive, negative, reinforcement, and punishment. There's no judgment implied in this so positive does not mean good and negative does not mean bad. Positive means we add something to the equation. Negative means we remove something from the equation. Reinforcement means that the dog is more likely to perform that behavior in the future and punishment means they're less likely to perform it in the future. So the four quadrants of operant conditioning. Number one positive reinforcement. This is the classical give your dog a reward. My dog sits I hand him a piece of food. The food I add to the equation the primary reinforcer the reward. The behavior that prompted me giving the dog a piece of food, sitting, is now more likely to occur in the future. Positive reinforcement. Negative reinforcement. We generally erroneously think of this as a correction or an aversive consequence. It's not. In negative reinforcement the dog stops something unpleasant from happening, some aversive experience from happening with their behavior. So the old way we used to teach sit by pulling up on a dog's collar until they sat is a form of negative reinforcement. I'm pulling on the collar. The dog sits. I stop pulling. The negative portion is I remove the unpleasant pulling on the leash. Whatever behavior the dog performed to stop that, in this case sitting, is now more likely to occur in the future. Negative reinforcement. The dog's behavior that stops the discomfort is now more likely to occur in the future. The next term in operant conditioning is positive punishment. Positive punishment is the traditional correction. It's an aversive consequence. So for instance if my dog jumped up on me and I were to knee it in the chest. Or if
5:08
만약 제 개가 카운터로 뛰어오르고 제가 개에게 '안 돼'라고 소리를 질렀다면. 그 부정적인 결과, 즉 불쾌한 결과입니다. 결과가 적용됩니다. 정적 부분입니다. 개가 카운터 위로 뛰어오르면 저는 소리를 지릅니다. 부적 부분, 정적 부분은 제가 상황에 고함을 추가하는 것입니다. 그리고 카운터 위로 뛰어오르는 행동은 이제 미래에 발생할 가능성이 낮아집니다. 처벌을 받은 것입니다. 조작적 조건 형성 패러다임의 마지막 부분은 부적 처벌입니다. 부적 처벌은 단순히 개가 원하는 것을 보류하거나, 개가 원하는 것에 접근하는 능력을 처벌의 형태로 제거하는 것입니다. 예를 들어 어린 개를 음식으로 훈련하고 있다고 해보죠. 개를 유도하다가 음식을 위로 들어 올리면 개가 저에게 뛰어오릅니다. 만약 개가 뛰어올라 발로 저를 건드리면 저는 음식을 치우고 예를 들어 등 뒤로 숨깁니다. 저는 개가 원하는 것(부적 부분)을 제거하는 것이고, 그 원하는 것을 제거하게 만든 행동인 뛰어오르기는 이제 처벌을 받아 미래에 발생할 가능성이 낮아집니다. 정말 간단합니다. 다음 용어는 조건 강화물입니다. 조건 강화물이란 단순히 개에게 아무런 의미가 없는 소리나 시각적 신호와 같은 중립 자극을 취하는 것을 의미합니다. 그리고 개에게 본질적인 가치가 있는 무언가, 여기서는 음식이라고 합시다, 그런 의미 있는 무언가를 예측하게 함으로써 중립 자극은 개에게 일차 강화물이나 보상 물질, 여기서는 음식과 같은 의미를 갖게 됩니다. 우리는 이것을 사용할 수 있으며, 소리나 시각적 신호 등 우리가 선택하는 무엇이든, 여기서는 주로 소리를 사용할 것입니다. 그 소리는 이제 고전적 조건 형성을 통해 강화 작용을 하도록 조건화되었습니다. 따라서 언어적 마커를 사용하거나 클리커를 사용하는 것은 고전적인 조건 강화물입니다. 예를 들어 제가 클릭 소리를 내고 개에게 음식을 줍니다. 클릭 소리를 내고 개에게 음식을 줍니다. 제가 '예'라고 말하고 개에게 음식을 줍니다. '예'라고 말하고 개에게 음식을 줍니다. 이것을 반복하면 제 개는 그 클릭 소리나 '예'라는 말에 반응하게 됩니다. 그들이 음식에 반응하는 것과 같은 방식입니다. 이제 '예스'나 클릭커는 강화제로서 조건화되었습니다. 이 용어는 우리의 경우 마커와 상호 교환적으로 사용되며, 흔히 브릿지라고 표현하는 것도 들으실 수 있을 겁니다. 마지막 용어는 조건 강화제였습니다. 다음 용어는 마커입니다. 우리는 개가 맞았을 때와 틀렸을 때를 소통하기 위해 언어적 마커를 사용합니다. 그래서 우리는 개에게 그들이 언제 맞았고 틀렸는지 매우 정확하게 알려줄 수 있는 의사소통 시스템을 갖추고 있습니다.
my dog were to jump on a counter and I were to yell no at them. The negative consequence, the unpleasant consequence, is applied. The positive portion. My dog jumps on the counter and I yell. The negative portion, the positive portion is me adding the yelling to the equation. And the behavior, jumping on the counter, is now less likely to occur in the future. It's been punished. The final part of our operant conditioning paradigm is negative punishment. And negative punishment is simply the withholding of something the dog wants or removing the dog's ability to access something they want as a form of punishment. So let's say I'm training with a young dog with food. I'm luring the dog around and as I lift the food up the dog jumps up on me. If the dog jumps up and hits me with their feet and I take the food away, put it behind my back for instance, I'm removing something they want, the negative portion, and the behavior that prompted the removal of what they want, jumping up, has now been punished and is less likely to occur in the future. It's really that simple. Our next term is conditioned reinforcer. Conditioned reinforcer simply means we take a neutral stimuli again, some sound or visual prompt that means nothing to the dog. And by predicting something that does mean something to the dog, something of intrinsic value to the dog, in this case let's say food, the neutral stimuli comes to have the same meaning for the dog as the primary reinforcer or the rewarding substance, in this case food. We can use this and the sound or visual prompt or whatever we're choosing, in this case we're primarily going to use a sound. The sound is now been conditioned to be reinforcing through the use of classical conditioning. So our use of verbal markers or someone's use of a clicker is a classic conditioned reinforcer. So for instance I make a click, I give my dog a piece of food. I make a click, I give my dog a piece of food. I say yes, I give my dog a piece of food. I say yes, I give my dog a piece of food. If I do this repeatedly, my dog responds to the click or the yes, the same way that they would respond to food. And now yes or click has been conditioned to be reinforcing. This term is interchangeable with marker in our case or frequently you'll hear it described as a bridge as well. Our last term was conditioned reinforcer. The next term is marker. And for us, we say use verbal markers to communicate to our dog when they're right and wrong. So we have a communication system that is geared around being able to tell the dog very precisely when they're right and wrong.
7:39
그리고 우리는 소리가 개에게 의미를 갖도록 조건화합니다. 따라서 마커는 조건 강화제와 상호 교환적으로 쓰입니다. 이 경우 우리가 사용하는 소리는 우리가 내는 언어적 소리, 즉 단어입니다. 우리는 '예스'를 많이 사용하지만, 그게 무엇인지는 사실 중요하지 않습니다. 즉, 마커는 단순히 조건 강화제나 브릿지와 같은 의미입니다. 우리는 목소리를 사용합니다. '인게이지먼트(집중)'는 지속적인 집중력과 동기 부여를 뜻하는 우리의 용어입니다. 우리가 개에게 무언가를 가르치기 시작하기 전에 보고 싶은 것 중 하나는 그들이 우리에게 집중하는 것입니다. 우리는 개가 우리에게 집중하는 것이 보상이라는 것을 가르치는 과정을 인게이지먼트라고 부릅니다. 지속적인 집중력을 보이고 우리로부터 무언가를 원하며, 우리로부터 일종의 보상을 얻고자 하는 상태를 우리는 '인게이지드(집중된)'라고 합니다. 보상과 보상 사이에 우리에게 부분적으로만 주의를 기울이거나 다른 것을 보기 위해 시선을 돌리는 개는 인게이지드 된 상태가 아닙니다. 그리고 인게이지먼트는 모든 훈련의 필수 조건입니다. 다음 용어는 '루어링(유도)'입니다. 루어링은 새로운 행동을 가르칠 때 개를 조종하기 위해 사용하는 물리적 도구 중 하나입니다. 우리는 개를 움직이게 하고 행동을 조종하여 무언가를 가르치기 위한 다양한 방법을 가지고 있습니다. 루어링은 우리가 이를 행하는 간단하고 직관적인 방법 중 하나입니다. 한마디로 루어링이란 단순히 개가 내 손을 따라오게 하는 것입니다. 그래서 저는 손에 간식을 쥐고 개에게 제 손을 머리로 따라오는 것이 유용하다는 것을 가르칩니다. 그런 다음 손을 움직여서 개의 몸이 뒤따라오게 함으로써 개를 유도할 수 있습니다. 이 DVD에서 유도(luring)에 관한 몇 가지 기본 규칙에 대해 이야기하겠습니다. 하지만 유도는 정말 간단합니다. 개가 머리로 여러분의 손을 따라오게 하는 것입니다. 공간 압박(spatial pressure)은 우리가 개의 행동을 유도하기 위해 사용하는 또 다른 물리적 도구 중 하나입니다. 우리는 유도와 마찬가지로 공간 압박을 사용하여 개에게 특정 사항을 가르치기도 합니다. 공간 압박은 단순히 제가 개 쪽으로 다가가면 개가 제 공간에서 벗어나는 것을 말합니다. 그래서 저는 제 몸을 사용하여 개를 밀거나 당길 수 있습니다. 개를 특정 위치로 이동시키고 새로운 행동을 가르치려 할 때 말이죠. 다음 용어는 셰이핑(shaping)입니다. 그리고 우리는 여기에 점진적 접근(successive approximation)이라는 용어를 포함했습니다. 저희 훈련사들은 이를 점진적 접근이라고 부릅니다. 셰이핑은 단순히 행동을 만들어내는 것입니다. 우리는 개가 특정 유형의 행동을 하도록 셰이핑하는 작업을 하고 있습니다. 어떤 행동은 아주 빠르게 포착(capture)합니다. 그리고 어떤 행동은 더 어렵습니다. 포착하는 데 여러 단계가 필요합니다. 개에게 행동을 가르치거나 행동을 하도록 유도하는 셰이핑 과정은
And we condition sounds to have meaning to the dog. So marker is interchangeable with conditioned reinforcer. In this case, the sound we're using happens to be a verbal sound that we make, a word. We use yes a lot, but it doesn't really matter what it is. So a marker is simply interchangeable with a conditioned reinforcer or a bridge. We're using our voice. Engagement is our term for sustained focus and motivation. So one of the things that we want to see in our dogs before we start trying to teach them things is that they're paying attention to us. And we call the process of teaching dogs that paying attention to us is rewarding engagement. So a dog that has sustained focus and wants something from us, wants a reward of some type from us, is what we call engaged. And a dog that is giving us a portion of its attention or checking out to look around and do other things in between rewards is not engaged. And engagement is a prerequisite for all of our training. Our next term is luring. And luring is one of the physical tools we use to manipulate dogs when we're teaching them new behaviors. So we have a variety of different ways to move a dog around and manipulate their behavior in order to teach them to do things. And luring is one of the simple and straightforward ways we do this. And luring in a nutshell is simply a dog following my hand. So I put a piece of food in my hand and I teach the dog that following my hand around with their head is useful. And then I can manipulate them by moving my hand around and having their body follow behind. We'll talk about some of the basic rules around luring in this DVD. But luring is really as simple as that. A dog that will follow your hand around with their head. Spatial pressure is another one of the physical tools we use to manipulate dogs behavior. We use spatial pressure to teach the dog certain things as well, just like we did with luring. And spatial pressure is simply if I move into my dog, my dog moves out of my space. So that I can use my body to push or pull the dog. And when I'm trying to move them into certain positions and teach them new behaviors. Our next term is shaping. And we've thrown in what we call a successive approximation. Our trainers call successive approximation. Shaping is simply the creation of a behavior. We're working on shaping the dog into doing a certain type of behavior. Some behaviors we capture really quickly. And some behaviors are more difficult. They take multiple steps to capture. The act of shaping or teaching the dog or manipulating the dog into doing a behavior.
10:05
흔히 점진적 접근이라 불리는 단계들로 나눌 수 있습니다. 완성된 행동으로 나아가는 과정에서 개에게 가르치거나 단계를 보상하는 것입니다. 그 과정을 점진적 접근이라고 합니다. 그래서 우리는 셰이핑과 점진적 접근을 사용하여 우리가 개에게 가르치고자 하는 복종 행동 중 일부를 포착할 것입니다. 페이딩(fading)은 단순히 훈련에서 도움을 없애는 것입니다. 우리는 개가 행동을 하도록 유도하는 데 사용하는 특정 도구들을 가지고 있습니다. 우리는 유도(lure)를 합니다. 우리는 바디 랭귀지를 사용합니다. 우리는 개에게 무엇이 옳고 그른지 알려주기 위해 우리의 의사소통 체계를 사용합니다. 그리고 시간이 지나면서, 우리는 개가 앉기와 같은 특정 행동을 할 수 있도록 돕습니다. 개가 이러한 행동에 능숙해지면, 우리는 도움을 제거해야 합니다. 도움을 서서히 제거해 나가는 과정을 페이딩(fading)이라고 합니다. 행동을 유도하는 신호를 제거하거나 도움을 없애고 싶을 때마다, 우리는 이것을 페이딩이라고 부릅니다. 다음 용어는 헝거 드라이브(hunger drive), 즉 궁극적으로 음식에 대한 동기 부여입니다. 따라서 훈련할 때, 특히 음식으로 훈련할 때는, 우리 반려견의 성공 여부는 보상에 대한 욕구, 즉 동기 부여에 의해 크게 좌우됩니다. 헝거 드라이브는 단순히 먹이에 대한 개의 본능적인 동기 부여를 의미합니다. 그리고 개들은 선천적으로 다양한 범위의 헝거 드라이브를 가지고 있습니다. 이것은 우리가 개에게 얼마나 많은 사료를 주는지, 그리고 훈련할 때 어떤 종류의 음식을 사용하는지에 따라 어느 정도 조절될 수 있습니다. 하지만 사실, 어떤 개들은 다른 개들보다 음식에 대한 동기가 더 강합니다. 그리고 어떤 개들은 다른 것들에 더 큰 동기를 느끼기도 합니다. 우리는 이것을 조절할 수 있지만, 음식으로 훈련하는 초기 단계에서 이에 주의를 기울여야 합니다. 다음 용어는 프레이 드라이브(prey drive, 포획 본능)입니다. 프레이 드라이브는 도그 트레이너들이 많이 사용하는 용어이지만, 행동학자나 다른 사람들에게서는 좀처럼 듣기 힘든 용어입니다. 간단히 말해서, 프레이 드라이브는 입으로 무언가를 쫓고 붙잡으려는 반려견의 본능적인 욕구입니다. 우리는 놀이 방법을 가르치거나 동기를 부여하기 위해 개의 프레이 드라이브를 활용합니다. 어떤 개들은 무언가를 쫓고, 물고, 입으로 붙잡는 데 매우 높은 동기를 보입니다. 따라서 우리는 이것을 훈련 시 강화 체계나 보상 체계로 사용할 수 있습니다. 다음 용어는 습득 기반 행동(acquisition-based behaviors)이라고 부르는 것입니다. 특정 행동들, 즉 개가 원하는 것을 얻게 되는 과정에서의 행동들은, 탐색이나 추격 같은 행동들은 개에게 본질적으로 강화가 됩니다. 우리는 이것을 자기 강화적(self-reinforcing)이라고 부릅니다. 그래서 우리는 행동 사슬을 특정 유형의 행동들로 나눕니다. 보상을 받기 위해 앉는 것과 같은 일부 행동은 개에게 본질적으로 강화가 되지 않습니다. 앉기라는 행동은 개에게 실제적인 의미가 없습니다. 하지만 추격이나 탐색은 빈번하게 의미가 있습니다. 그리고 개가 보상을 얻으려고 노력할 때 강화가 된다고 느끼는 행동 유형들을,
Can be broken down into steps that are frequently called successive approximation. Where I teach the dog or reward steps on the way to the finished behavior. That process is called successive approximation. So we'll use shaping and successive approximation to capture some of the obedience behaviors we're looking to teach the dog. Fading is simply the elimination of help in our training. So we have certain tools we use to manipulate a dog into doing behaviors. We lure them. We use body language. We use our communication system to tell them when they're right and wrong. And over time, we help the dog do certain behaviors like sit. As the dog becomes fluent in these behaviors, we need to remove the help. And the process of slowly removing the help is called fading. Anytime we want to remove a signal that's prompting a behavior and eliminate the help, we call this fading. Our next term is hunger drive or motivation for food ultimately. So in training, especially when we're training with food, a lot of our dog's success is dictated by their desire for the reward, their motivation. And hunger drive is simply the dog's natural motivation for eating. And dogs have a wide range of hunger drives naturally. This can be manipulated to some degree by how much we feed the dogs and what types of food we're using when we train. But the truth is, some dogs are more hunger motivated than others. And some dogs are more motivated by other things. We can manipulate this, but we need to pay attention to it in our early stages of training with food. Our next term is prey drive. And prey drive is a term that you'll hear dog trainers use a lot, but seldom hear behaviorists and other people use. In a nutshell, prey drive is simply your dog's desire to chase and grab things with its mouth. And we manipulate a dog's prey drive to teach it how to play or to motivate it. So some dogs are highly motivated to chase and bite and grab things with their mouth. And so we can use this as a reinforcement system or a reward system in our training. Our next term is what we call acquisition-based behaviors. So certain behaviors, the acquiring of something the dog wants, whether it's searching, chasing, those sorts of things, are intrinsically reinforcing to dogs. We call them self-reinforcing. So we break our behavior chains down into certain types of behaviors. Some behaviors, like sitting to get a reward, aren't inherently reinforcing to the dog. Sitting has no real meaning to a dog. But chasing or searching frequently does. And the types of behaviors that dogs find reinforcing in trying to acquire a reward,
12:45
우리는 획득 행동(acquisition behaviors)이라고 부릅니다. 다음 개념은 강화제(reinforcer)와 보상(reward)이라고 부르는 것입니다. 엄밀히 말하면 이 둘은 약간 다릅니다. 하지만 우리 시스템에서는 이 용어들을 다소 혼용해서 사용하는 경향이 있습니다. 보통 우리가 강화제나 보상이라고 말할 때, 우리는 같은 것을 의미합니다. 하지만 조작적 조건화에 대한 우리의 논의를 기억하세요. 엄밀히 말하면 그것들은 약간 다릅니다. 만약 제가 반려견의 리드줄을 당기고 있어서 개가 불편함을 느끼고 무언가를 했을 때, 제가 당기는 것을 멈추면, 당김을 멈추는 행위 자체가 본질적으로 강화제가 됩니다. 그것은 강화 작용을 합니다. 개는 그 멈춤을 유발한 행동을 할 가능성이 더 높아집니다. 그러니 이것이 반드시 보상인 것은 아닙니다. 하지만 우리가 강화제라고 말할 때 대부분은 보상을 의미합니다. 우리는 개가 어떤 행동을 한 것에 대해 개가 원하는 것을 주는 것을 의미합니다. 다음 개념은 보상 이벤트(reward event)를 만드는 것이라고 부르는 것입니다. 훈련 초기 단계에서 우리는 물건이 개를 보상하는 데 있어 가장 큰 원동력이라고 생각하곤 했습니다. 우리가 개에게 무엇을 주는지가 가장 중요한 부분이었습니다. 시간이 흐르면서 우리는 상호작용적 보상 이벤트(interactive reward event)라고 부르는 것을 만드는 방향으로 발전해 왔습니다. 즉, 개를 보상하는 과정에서 저와 개가 함께 무언가를 한다는 의미입니다. 반려견에게 가치 있고 동기 부여가 될 뿐만 아니라, 반려견과 저 사이의 유대감을 쌓는 경험이기도 합니다. 그래서 제가 상호작용하는 방식으로 보상을 주고 반려견이 그 상황을 좋아하게 된다면, 반려견은 자신이 수행한 행동에 대해 강화될 뿐만 아니라, 저에게 더 집중하기 시작하고 저와 더 협력하게 됩니다. 따라서 보상에 대한 반려견의 태도가 전반적으로 개선됩니다. 우리는 이러한 짧은 보상 시퀀스를 보상 이벤트라고 부릅니다. 다음 용어는 구두 신호입니다. 구두 신호 또는 구두 프롬프트는 사실상 명령을 의미합니다. 즉, 우리가 내는 모든 소리, 사용하는 단어, 행동을 유도하기 위해 내는 구두 소리를 말합니다. 프롬프트나 신호는 단순히 반려견에게 학습한 특정 행동을 하라는 신호를 보내는 것입니다. 신체적 신호는 말 그대로 행동을 유도하기 위해 우리가 하는 신체적 움직임이나 동작입니다. 명령이 구두 신호나 구두 프롬프트인 것과 마찬가지로, 신체적 프롬프트는 반려견을 유도하여 앉게 하거나, 보디랭귀지, 몸을 숙이는 등 그러한 모든 행동일 수 있습니다. 반려견이 특정 행동을 하도록 유도하기 위해 우리가 신체적으로 하는 모든 것을 말합니다. 강화 스케줄은 특정 행동에 대해 얼마나 자주 반려견에게 보상하는지를 의미합니다. 따라서 반려견의 강화 스케줄을 조절함으로써 반려견의 행동을 근본적으로 변화시킬 수 있습니다. 이는 시간상 보상의 빈도일 수도 있고, 특정 행동에 대해 얼마나 자주 보상하는지를 의미할 수도 있습니다.
we call acquisition behaviors. Our next concept is what we call reinforcer versus reward. Now technically, they're slightly different. But in our system, we tend to use them kind of interchangeably. Normally when we say reinforcer or reward, we mean the same thing. But remember our discussion of operant conditioning. And they are technically slightly different. So if I'm pulling on my dog's leash and it's uncomfortable and he does something, and I stop pulling, the stopping of pulling is inherently a reinforcer. It's reinforcing. The dog's more likely to do the behavior that made it stop. So it's not necessarily a reward. But most of the time when we say reinforcer, we mean reward. We mean giving the dog something they want for having done a behavior. Our next concept is what we call creating a reward event. So in our early stages of training, we used to think that the item was the main driving force in rewarding a dog. What we gave them was the most important part. Over time, we've evolved into trying to create what we call an interactive reward event. Meaning the dog and I do something together during the process of rewarding the dog that the dog finds both valuable and motivating, but also sort of a bonding experience between the dog and I. So if I'm rewarding the dog in an interactive way and the dog likes what's happening, not only are they reinforced for the behavior that they're performing, but also they start to pay better attention to me, like working with me more. So there's an overall improvement in their attitude about rewards. And we call these little sequences of rewards, reward events. Our next term is verbal cue. Verbal cue or verbal prompt really just means a command. So any kind of sound we make, a word we use, a verbal sound that we make that prompts behavior. So a prompt or a cue is something that just simply signals the dog to do a certain behavior that they've learned. A physical cue is simply that, a physical move or motion that we make that prompts behavior. And so in the same way that a command is a verbal cue or verbal prompt, a physical prompt might be luring the dog into a sit, body language, leaning over, any of those things. Anything physically we do to prompt the dog to do a certain behavior. Reinforcement schedules are frequently how often we reward a dog for a certain behavior. And so by manipulating a dog's reinforcement schedule, we can radically manipulate their behavior. So it might be the frequency of rewards in terms, in time, or it might be how often we reward a given behavior.
15:21
훈련의 일반적인 패러다임은 우리가 연속 강화 스케줄이라고 부르는 것으로 시작하는 것입니다. 특정 행동을 반복할 때마다 반려견에게 보상을 주는 방식이며, 이를 높은 강화율이라고 합니다. 우리는 반려견의 주의를 유지하기 위해 매우 자주 보상을 제공합니다. 시간이 지나면서 우리는 강화 스케줄을 다양하고 무작위로 변경하는데, 이는 모든 행동 반복마다 보상을 주지는 않는다는 의미입니다. 그리고 우리는 시간이 지남에 따라 성과를 향상시키기 위해 보상 간의 간격을 더 길게 합니다. 또한 모든 행동에 대해 지속적으로 개에게 보상해야 하는 필요성을 없애기 위해서죠. 우리가 '흥분도(arousal)'라는 용어에 대해 많이 이야기하는 것을 들으실 겁니다. 흥분도란 단순히 개의 흥분 상태나 자극 수준을 의미합니다. 그래서 개가 어느 정도 흥분하거나, 들뜨거나, 동기 부여가 되거나, 고조되는 등 무엇이라고 부르든 간에 결국 신이 난 상태를 말합니다. 그럴 때 그 흥분은 우리가 의도하는 특정 행동을 돕기도 하고 방해하기도 합니다. 따라서 우리가 개의 흥분 수준을 조절한다고 말할 때는 그들의 들뜬 정도를 조절하는 것을 의미합니다. 우리가 하는 훈련을 더욱 생산적으로 만들기 위해서죠. 때로는 높은 흥분도가 더 좋을 때도 있고, 낮은 흥분도가 더 좋을 때도 있습니다. 또 다른 핵심 개념은 '가림 현상(overshadowing)'이라고 불리는 개념입니다. 가림 현상이란 단순히 우리가 개에게 두 가지 프롬프트, 신호, 또는 입력을 동시에 줄 때, 개가 더 관련성이 높은 것에 집중하고 다른 것은 무시하는 현상을 말합니다. 예를 들어, 제가 개에게 언어적 프롬프트와 신체적 프롬프트를 동시에 준다면, 개에게는 신체적 입력이나 신체적 프롬프트가 언제나 언어적 프롬프트보다 더 관련성이 높습니다. 개는 언어를 사용하는 생명체가 아닙니다. 우리는 평소 개 주변에서 말을 많이 하지만, 그 대부분은 개에게 아무런 의미가 없습니다. 그래서 개들은 신체적 단서에 더 많은 주의를 기울입니다. 만약 두 가지를 동시에 수행하면, 신체적 프롬프트가 언어적 프롬프트를 가려버립니다. 우리는 개가 언어적 프롬프트에 집중하고 있다고 생각하지만, 사실 개는 신체적 프롬프트에만 집중하고 있는 것입니다. 그래서 제가 언어적 프롬프트만 준다면, 개는 반응하지 않습니다. 하지만 신체적 프롬프트를 주면 반응하죠. 이 개념은 앞으로의 모든 훈련에서 필수적인 부분이 될 것입니다. 그리고 이 개념은 사람들의 훈련 과정에서 많은 문제를 일으키기도 합니다. 다음 용어는 동기입니다. 동기는 간단히 말해 특정 활동이나 물건에 대한 반려견의 욕구를 의미합니다. 제 반려견의 경우, 음식이나 놀이, 또는 추격에 대한 반려견의 욕구를 동기 수준이라고 말할 수 있습니다. 즉, 우리는 특정 활동에 대한 반려견의 욕구 수준과 강도를 설명하고 있는 것입니다.
The general paradigm in training is we start out on what we call a continuous reinforcement schedule, where we reward the dog for every repetition of a certain behavior, and what we call a high rate of reinforcement. We reinforce the dog very frequently to hold their attention. Over time, we vary and randomize the reinforcement schedule, meaning we don't reinforce the dog for every single repetition. And we go longer in between rewards to improve performance over time, and to eliminate the necessity of constantly rewarding the dog for every behavior. You'll hear us talk a lot about the term arousal. And arousal just means the dog's excitement or stimulation level. So when a dog gets aroused to varying degrees, excited, motivated, worked up, whatever you want to call it, they're excited. Then it can either help a certain behavior or hinder a certain behavior we're trying to capture. So when we talk about manipulating a dog's arousal level, we're talking about manipulating their excitement level to make the training that we're doing more productive. So sometimes higher arousal rates are better, and sometimes lower arousal rates are better. Another key concept is a concept called overshadowing. And overshadowing is simply if we give a dog two prompts or two signals or two inputs at the same time, the dog will pay attention to the more relevant one and ignore the other. So for instance, if I give my dog a verbal prompt and a physical prompt at the same time, for dogs, physical inputs or physical prompts are always more relevant than verbal prompts. Dogs aren't verbal creatures. We talk around them frequently. Most of it means nothing to them. And so they pay more attention to physical cues. So if I do them at the same time, the physical prompt overshadows the verbal prompt. And we think the dog's paying attention to the verbal prompt, but they're really only paying attention to the physical prompt. So if I give them only the verbal prompt, they don't respond. But if I give them the physical prompt, they do. This concept is going to be an integral part of all of our training going forward. And this gives people lots of problems in their training. Our next term is motivation. And motivation simply means the dog's desire for some activity or some item. So my dog, we can talk about my dog's desire for food or play or chasing as their level of motivation. And we're simply describing the level and intensity of their desire for a certain activity.
18:05
우리에게 동기란 무엇일까요?
What is motivation for those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those of us as those