<![CDATA[Special Episode: The Seat Belt Analogy and Negative Reinforcement]]>
Ivan Balabanov<![CDATA[Training Without Conflict® | Dog Training Podcast]]> · 팟캐스트 · 35분
부적 강화 학습 모델에서 흔히 사용되는 안전벨트 비유는 강화의 본질을 오해하게 만드는 부적절한 예시입니다. 훈련사나 보호자는 이 비유가 자극의 강도를 고정된 상수로 전제하고 있다는 점을 인지해야 합니다. 진정한 부적 강화는 개가 수행해야 할 행동을 명확히 이해하고 있는 상태에서, 개별 동기 수준에 맞춰 혐오 자극의 강도를 유연하게 조절하는 과정을 포함합니다. 자극은 행동을 멈추거나 발생시키는 통제 수단으로서의 목적을 가지며, 단순히 불편함을 유발하는 상기 신호나 경고음과는 엄격히 구분됩니다. 학습자는 개가 무엇을 해야 할지 모르는 상황에서 불필요하게 혐오 자극을 가해서는 안 되며, 적절한 교육을 통해 반사적이거나 의도적인 반응을 이끌어낼 수 있는 환경을 조성해야 합니다. 효율적인 부적 강화는 학습자가 좌절감에 치우치지 않고 안전한 긴급 상태 내에서 체계적으로 강도를 조절하여 원하는 행동을 유도할 때 성공적으로 작동합니다.
스페셜 에피소드: 안전벨트 비유와 부적 강화
<![CDATA[Special Episode: The Seat Belt Analogy and Negative Reinforcement]]>
0:00
부적 강화와 유명한 안전벨트 비유에 대해 이야기해보겠습니다. 어떤 분들, 특히 유럽에 계신 몇몇 분들은 이 비유를 완전히 이해하지 못하시는 것 같습니다. 그 비유 말입니다. 그 비유가 도대체 무엇이냐는 메시지와 이메일을 받곤 했습니다. 사실 이 비유는 부적 강화가 어떻게 작동하는지를 아주 간단한 방식으로 이해하기 쉽게 설명하기 위한 것입니다. 그리고 제가 생각하기에 부적 강화를 친근한 방식으로 제시하는 것이기도 합니다. 안전벨트 비유는 기본적으로 그것이 너무 강압적으로 느껴지지 않게 해줍니다. 무슨 말씀인지 아시겠죠. 본격적으로 시작하기 전에, 처벌과 강화에 대해 이야기해 봅시다. 이것들은 무엇을 위한 것일까요? 행동을 통제하기 위한 것입니다. 그것이 유일한 목적이죠. 처벌과 강화는 우리가 행동을 통제할 수 있게 해줍니다. 물론 강화는 행동을 발생시킵니다. 강화는 행동을 멈추게 합니다. 장려와 억제죠. 기본적으로 그게 전부입니다. 정말 간단하죠. 하지만 이 대화를 계속 이어갈 때 우리가 명심해야 할 점이 있습니다. 자, 이제 시작해 보죠. 이상하게도, 왜 그런지는 모르겠지만 부적 강화는 잘못 이해되는 경우가 많습니다. 그리고 이에 대한 연구도 많지 않습니다. 알다시피 그만큼 많이 연구된 분야가 아닙니다. 정적 강화나 처벌과 비교하면, 그쪽에는 훨씬 더 방대한 문헌들이 존재하죠. 그렇다고 부적 강화에 관한 연구나 문헌이 전혀 없다는 뜻은 아닙니다. 단지 그만큼 방대하지 않을 뿐입니다. 안전벨트 비유가 적절하지 않다고 말하는 것이, 그것이 틀렸다는 뜻은 아닙니다. 보시다시피 부적 강화라는 요소는 전혀 없습니다. 하지만 그것은 단지 매우 부적절한 비유일 뿐입니다. 그래서 우선, 어떻게든 무의식적인 수준에서조차 안전벨트 비유는 우리가 강화를 위해서는 낮은 수준의 혐오 자극을, 처벌을 위해서는 더 높은 수준의 혐오 자극을 사용할 것임을 암시합니다. 그래서 그것 자체가 문제입니다. 이것은 정말로, 그러니까, 강화와 처벌, 부적 강화와 정적 처벌을 잘못 구분한 것입니다. 네. 여기 흥미로운 점이 있습니다. 우리가 항상 역사 속으로 되돌아가서 무엇이 어디서 유래했고 왜 그런지 알아보려 하는 방식에 관한 것인데요. 어디서 어떻게 유래했는지에 대해서 말이죠.
Talking about negative reinforcement and the famous seat belt analogy. There are some people, especially I think some people in Europe don't quite understand the analogy. I've gotten messages and emails like, well, what is that analogy? And it's really a way to explain how negative reinforcement works in a very supposed to be very simple way that will make sense. And also what I believe is present the negative reinforcement as something friendly. Not so like the seat belt analogy basically doesn't allow it to be too intense. You know what I'm saying. Now before we get going there, punishment and reinforcements. They're for what? Control behavior. That's their sole purpose. Punishment and reinforcement is so we can control behavior. Reinforcement of course makes behaviors happen. Reinforcement makes behavior stop. Encouragement and suppression. And basically that's it. It's really that simple. But this is something that we need to keep in mind as we go on with that conversation right now. One of the kind of strange, I don't know why, but negative reinforcement is poorly understood. And there is not as many studies on it. There is not as much, you know, it's not explored as much. If you compare it to positive reinforcement, to punishments, you know, there is way more extensive literature. Now this is not to say that there is no studies in literature on negative reinforcement. It's just not as extensive. When I say that the seat belt analogy is a poor example, I am not suggesting that it's not, that there is no element of negative reinforcement as you will see. But it's just a very poor one. And so to begin with, somehow the, even on a subconscious level, the seat belt analogy implies that we're going to use low level of aversion, low level of aversive for reinforcement and higher level for punishment. So that already is a problem. This is, this is really not, not, you know, it's incorrect distinction between the reinforcement and punishment, negative reinforcement and positive punishment. Yeah. Something, something of interest here. It's, you know, how we, we always go back in history and we try to see where, what's coming from where and why.
3:53
흥미롭게도 스키너 역시 이런 종류의 혼란에 대해 부분적으로 책임이 있습니다. 그가 처음 부적 강화라는 용어를 도입했을 때, 그는 사실상 처벌도 같은 의미로 사용했습니다. 부적 강화와 정적 처벌 사이에 구분이 없었습니다. 그 점은 주목할 만한 중요한 부분입니다. 그런데 재미있는 것은, 그가 나중에 50년대 후반쯤에 그것을 수정했다는 점입니다. 재정의를 내린 것이죠. 그는 두 개념을 분리하는 방향으로 정의를 다시 내렸습니다. 그가 자신의 논문이나 강연 어디에서도 원래 두 개념을 하나로 묶어 다루었다는 사실을 인정한 적은 없었던 것으로 압니다. 그래서 다른 모든 것과 마찬가지로, 그것도 관성을 갖게 됩니다. 하지만 다시 강조하자면, 그것들을 같은 것이라고 생각하는 것은 정말 좋은 생각이 아니며, 부적 강화는 항상 낮은 수준의 혐오 자극을 사용하고 정적 처벌은 항상 높은 수준의 혐오 자극을 사용한다고 단정 짓는 것 역시 분명히 좋지 않은 생각입니다. 혐오 자극을 말이죠. 이것은 매우 큰 오해입니다. 말씀드렸듯이 강화, 처벌, 행동 통제는 제가 무언가를 하라고 요청하고 그것을 강화하거나, 아니면 무언가를 하지 말라고 요청하고 그것을 제지하는 것입니다. 이제 누군가, 예를 들어 훈련 캠프에서 전기 목줄을 많이 사용하는 사람들이 있다면 분명 낮은 단계의 자극만 사용한다고 말하는 훈련사들이 있을 겁니다. 이것은 오해의 소지가 있으며 옳지 않은 말입니다. 이는 안전벨트 비유가 왜 틀렸는지와 같은 이유 때문입니다. 왜 틀렸는지 말씀해 주실 수 있나요? 그럼 제가 말씀드리죠. 부담 갖지 마세요, 말씀하셔도 됩니다. 기본적으로 그것을 제시하는 한 가지 이유는 '그렇게 나쁘지 않을 거야'라는 점 때문이죠. 그렇게 심하지 않을 거예요. 겨우 낮은 단계일 뿐이니까요. 약간의 불편함이고, 조금 불쾌한 정도죠. 별일 아니라는 겁니다. 그럴 수도 있겠지만, 항상 그렇지는 않습니다. 안전벨트를 생각해보세요. 안전벨트는 어떤 역할을 하나요? 안전벨트는 일정한 경고음을 냅니다. 차량마다 조금씩 다른 변형이 있긴 합니다. 하지만 모두 신호음이라는 신호를 주죠, 맞나요? 어떤 차에서는 안전벨트를 매지 않으면 잠시 후에 안전벨트 경고음이 조금 더 높아집니다. 그다음엔 무슨 일이 일어나나요? 대부분의 차에서는 경고음이 멈춥니다. 왜 멈출까요? 어느 시점에 이르면, 정말 고집스럽게 안전벨트를 매지 않는다면 분명 더 이상 매지 않을 것이기 때문입니다. 그렇죠? 자, 여기에는 이런 두 가지 특징, 즉 신호음이라는 특징이 있습니다. 하지만 이건 대부분 신호에 가깝습니다. 대부분 경고 신호인 셈이죠. 마치 '이봐, 그거 하는 거 잊지 마'라고 말하는 것과 같습니다. 그렇죠? 또한 우리는 이것을 논할 수 있는데, 이를테면 일종의 이정표라고 말할 수 있습니다.
And interestingly, like Skinner is partially to be blamed for, for this kind of confusions. When he first introduced it as negative reinforcement, he pretty much meant punishment as well. There was no, a distinction between negative reinforcement and positive punishment. That's a, important thing to, to note. Then interestingly, he, later in like late 50s or so, he changed it. He redefined. He made a different definition to where he separated the two. He never really admitted anywhere that I know in his papers and, and lectures that he originally put them in a, in one and the same. So as with everything else, that takes a momentum. But again, it's a, it's really not a, not a good idea to think that they are, um, one, one and the same, and it's definitely not a good idea to always con-sting that negative reinforcement uses low level of aversive and positive punishment always use high level of aversive. This is, this is big misconception. As we said, reinforcement, punishment, controlling behavior, either I ask you to do something and I reinforce it or ask you not to do something and I stop you from it. Now when somebody is, somebody like let's say in the training camps that use electric collars a lot and you will have trainers that will say, well, I only use low level stimulation. And this is misleading and it's incorrect. This is for the same reason why the seat belt analogy is incorrect. Can you tell me why is it incorrect? I'll tell you then. Don't, don't feel, you can say. But basically, you know, the, the, the one reason is to present it as, well, it's not going to be so bad. It's only low level. It's a little bit of discomfort, a little bit unpleasant. No big deal. And that might be the case, but it doesn't have to be always the case. When you think of the seat belt. What does it do? It has a constant tone. Different seat belts have different, a little bit different variations. But they all give a tone, a signal, right? After a little bit, if you don't put your seat belt on, in some cars, the seat belt tone goes a little bit higher. What happens next? The tone in most cars stops. And why does it stop? Because at some point, if you really were that stubborn not to put the seat belt on, you're obviously not going to anymore. Right? Now, it has, it has these two features, the, the tone. But it's mostly a signal. It's mostly a warning signal. It's like, hey, remember to do that. Right? It's also, we can argue it, like we can say that it's a, some sort of a signpost.
7:52
이정표에 대해 이야기할 때 우리가 어떻게 말하는지 아시죠? 그러니까, 길가에 있는 표지판을 생각하면 됩니다. 표지판은 당신을 안내합니다. 무엇을 해야 할지 알려주죠. 무엇을 하라고 제안하는 겁니다. 왜냐하면 아마도 그에 따른 결과가 따를 것이기 때문입니다. 어쩌면 사고가 날지도 모르죠. 어쩌면 다칠지도 모릅니다. 어쩌면 경찰에게 큰 범칙금을 부과받을 수도 있겠죠. 하지만 그것은 여전히 제안이자 상기시켜 주는 것, 신호일 뿐입니다. 그 신호가 반드시 당신을 통제하는 것은 아니죠. 그렇지 않나요? 우리가 행동을 통제하는 것에 대해 이야기하는 방식처럼요. 자, 그게 성가신가요? 그렇죠. 그 끊임없이 들리는 소리는 많은 사람에게 확실히 혐오 자극이 될 수 있습니다. 하지만 저는 대부분의 사람이 그 혐오감 때문이 아니라 상기시키는 역할 때문에 반응한다고 생각합니다. 안전벨트를 매야겠구나 하는 생각이죠. 범칙금을 받고 싶지 않거든요. 사고가 났을 때 안전벨트를 매지 않은 상태이고 싶지는 않으니까요. 그 둘 중 하나이거나 둘 다 결합된 것이겠죠. 그건 혐오 자극으로서 충분하지 않습니다. 누군가 하기 싫다고 한다면, 하지 않을 테니까요. 그렇죠? 그러니 정적 강화에 대해 생각해 보세요. 배불리 잘 먹은 래브라도가 있다고 칩시다. 방금 맛있는 식사를 마쳤죠. 그런데 당신에게는 오래된 개 간식이 하나 있습니다. 그리고 강아지에게 어떤 행동을 하라고 요구하고 있죠. 그 행동을 강화하려고 합니다. 그런데 강아지가 말하기를, 저는 그 개 간식은 싫어요. 사실 지금은 아무것도 먹고 싶지 않아요라고 한다면 어떨까요. 그 강화가 효과를 거두길 원한다면, 그런 상황에서 우리는 어떻게 해야 할까요? 그렇죠? 우리는 어떻게든 강도를 높여야 합니다. 우리는 변화를 주어야 합니다. 우리는 개가 간식을 원하게 만들어야 합니다. 아침 식사를 주지 않는 편이 좋을지도 모릅니다. 아니면 더 흥미로운 간식을 줄 수도 있겠죠. 개가 실제로 반응하게 만들 그런 것이 필요합니다. 그렇죠? 부적 강화도 이와 매우 똑같습니다. 그런 말투로는 효과가 없습니다. 우리가 진정으로 부적 강화를 하려면, 또 다른 수준의 혐오 자극 강도가 필요합니다. 그렇죠. 그러면 '자, 이제는 어때?'라고 말하는 셈이 되죠. 이제 안전벨트를 맬 건가요? 여전히 그것과 씨름하고 있는 상황이죠. 이제 제가 무슨 말을 하려는지 아실 겁니다. 그 강도는 당신이 안전벨트를 매도록 설득할 때까지 계속 높아질 것입니다. 정적 강화를 상상해 보세요. 당신은 간식을 가지고 있죠. 그 래브라도 리트리버를 앞에 두고요. 그리고 앉으라고 말하고 있죠. 그런데 개는 그냥 '대단히 감사합니다'라고 하는 것 같아요. 앉을 필요가 없다는 거죠. 안전벨트를 뒤로 채우는 것과 별반 다를 게 없죠. 안전벨트를 매겠다는 당신의 결정은 동기 부여의 수준에 달려 있습니다. 그렇죠? 그리고 그 동기 부여의 수준은 변할 것입니다. 그것이 부적 강화의 핵심입니다. 모든 강화와 모든 처벌에서 핵심이 되는 부분이죠. 고정된 것은 아무것도 없습니다. 예를 들어, 시각적인 이해를 돕기 위해 쉽게 말해볼게요.
You know how we talk when we talk about signposts? I mean, it's really like you think of a sign on the, on the street. It guides you. It tells you what to do. It suggests you what to do. Because probably there is consequences. Maybe, maybe you'll crash. Maybe you'll hurt yourself. Maybe the police writes you a big ticket. But it's still a suggestion, a reminder, a signal. It's not, that signal not necessarily controlling you. Right? How we talk controlling behavior. Now, is it annoying? Yeah. That, that persistent sound, you know, it sure can be aversive for many people. But I believe that most people respond to it, not because of the aversiveness, but the reminder of, I should put my seatbelt on. I don't want to get a ticket. I don't want to get in a crash and not have my seatbelt on. One or the two are both in combination. It's not aversive enough. If any person says, I don't want to do it, you're not going to do it. Right? So, think of positive reinforcement. You have some Labrador that's really well fed. He just had a good meal. And you have some kind of old dog biscuit. And you're asking him to do behavior. And you're going to reinforce it. And he says, I don't like that dog biscuit. I actually don't feel eating at all right now. What would we do in that situation if we want that reinforcement to work? Right? We have to escalate somehow. We have to change. We need to make the dog want the cookie. We need to maybe not feed them for the morning feeding. Maybe give a more interesting treat. Something that actually is going to make the dog respond to. Right? So it's very much the same with the negative reinforcement. And that tone doesn't do the job. If we really want to have negative reinforcement, then we have to have another level of aversive intensity. That's going to say, well, how about now? Would you put your seatbelt now? And you're still fighting with it. And I think you know now where I'm going. That intensity will increase until it convinces you to put the seatbelt on. Imagine positive reinforcement. You have that cookie. You have that lab. And you're telling him sitting. He just goes, thank you very much. I don't need to sit. Not that different than just clipping the seatbelt behind. Your decision to put the seatbelt on, it depends on the level of motivation. Right? And that level of motivation is going to change. And that's the thing with negative reinforcement. That is the thing with any reinforcement and any punishment. Nothing is constant. Like if you, just for the sake of visual, you know, is easy to understand.
11:52
전기 충격 칼라의 강도가 1에서 10까지 있다고 가정해 봅시다. 내가 당신에게 무언가를 시키고 싶거나, 무언가를 하지 않게 만들고 싶을 때 말이죠. 그리고 레벨 2로 살짝 누를 수 있는데, 그건 아마 멈춰서 집중해야 느낄 수 있는 정도일 거예요. 어떤 것을요. 하지만 당신은 '와, 알겠어'라고 말하죠. 안 할 거야. 아니면 기꺼이 할 거예요. 내일, 아니 내일도 아니고 한 시간 뒤에 똑같은 행동을 할 때 더 높거나 낮은 강도가 필요할 수 있어요. 무엇에 따라 다르냐고요? 무엇에 따라서요? 그 순간 그 일을 어떻게 느끼는지, 얼마나 동기부여가 되었는지에 따라 다르죠. 동기부여 수준이 행동을 강화하거나 멈추게 하도록 설득하기 위해 필요한 강도 수준을 결정합니다. 행동을 강화하거나 멈추게 하는 것이죠. 강화와 처벌이요. 그렇죠? 하지만 그걸 결정하는 건 당신, 혹은 개, 아니면 안전벨트를 매는 사람의 동기예요. 그게 결정하게 될 겁니다. 거기에 문제가 있습니다. 안전벨트 비유가 왜 나쁜 비유인가요? 상수이기 때문이죠. 그 작은 경고음보다 더 강해질 수 없고, 그렇게 되지도 않으니까요. 그리고 끝납니다. 게다가 또 무슨 일이 일어날 수 있죠? 지름길을 찾을 수 있죠. 맞아요. 안전벨트를 뒤로 해서 끼워버리면 소리를 무력화할 수 있죠. 하지만 실제로는 요구된 행동을 하지 않게 되죠. 해야 할 행동을 보이지 않는 거예요. 그러니 그건 나쁜 특징 중 하나인데, 강화하려는 거라면 실제로 그 행동을 강화할 수 있어야 하거든요. 행동을 강화할 수 있어야 해요. 저도 원래 이 생각을 하고 있었거든요. 지금도 가끔 머릿속에서 떠나지 않아서 계속 생각해보곤 해요. 차에 타고 있을 때 나탈리아가 '좋아, 당신 무슨 짓 하는지 다 알아'라고 말하곤 하죠. 처음에는 그녀가 '안전벨트 매요'라고 했으니까요. 그래서 저는 그냥 안 매요. 그러고 나서 기다리고, 또 기다리고, 또 기다리죠. 경고음이 좀 더 커지는데 뭐 어쩌겠어요. 그러다 결국 삐 소리가 나요. 그제야 그냥 매는 거죠. 제 말은, 그냥 운전하다 심심해서 장난치는 거예요. 그렇죠? 하지만 이게 왜 좋은 본보기가 되지 않는지 확실히 보여주죠. 그래서 두 가지 문제가 있어요. 사실 두 가지보다 더 많지만, 큰 문제는 두 가지예요. 하나는 자극이 일정하다는 거예요. 운전자의 동기나 개의 동기를 반영하지 못하죠. 우리는 자극의 강도를 조절할 수 없어요. 요령을 피우기가 더 쉽죠. 제대로 된 강화, 즉 부적 강화를 활용하려면 혐오 자극의 강도를 조절할 수 있어야 해요. 이건 필수 조건입니다. 왜냐하면 때로는 아주 약한 자극이나 자극 없이도 신호만으로 바로 반응할 수 있으니까요. 하지만 상대의 고집이나 몰입도, 거부하려는 동기에 따라서는 상대를 납득시킬 수 있는 수준까지 강도를 높여야 합니다. 그리고 그 수준이 곧 처벌인 것은 아니에요. 낮은 수준은 부적 강화를 위한 것이고 높은 수준은 처벌을 위한 것이라고 생각하는
Let's say we have from scale 1 to 10 electric color. And I want you to do something or I don't want you to do something. And I can tap level 2, which is you probably need to kind of stop and concentrate to feel something. But you say, wow, okay. I'm not going to do it. Or I will eagerly do. That same behavior tomorrow, or not even tomorrow, maybe in an hour, may require more or less level of intensity. Depending on what? Depending how you feel about it at the moment, how motivated you are. Level of motivation dictates the level of intensity to convince you either to be able to reinforce the behavior or to stop a behavior. Reinforcement and punishment. Right? But it's your motivation or the dogs or the person that's putting the seatbelt on that's going to dictate that. And so there is your problem. Why is it poor analogy, the seatbelt? Because it's a constant. It does not, it cannot escalate more than that little louder tone. And then it ends. On top of that, what else can happen? You can make a shortcut. Exactly. You can get that seatbelt, clip it behind you, and you can neutralize the sound. But actually not perform what's required. Not show the behavior that you should. So that's another poor feature of, you know, if you're reinforcing, you have to be able to actually reinforce the behavior. Because I was originally thinking about this. And even now sometimes it's almost stuck with me and I play with it. And when we're in the car and Natalia will be like, okay, I know what you're doing. Because in the beginning she will like, put your seatbelt on. And I'm like, just not putting it on. And then I wait, I wait, I wait. It goes a little louder, whatever. Then it ends up beeping. And then I just put it on. I mean, I'm just playing games because I'm bored on the road. Right? But it really shows you how it's not the best example. So two problems. Actually more than two problems, but two big problems. One is it stays constant. It doesn't reflect the motivation of the driver or the motivation of the dog. We cannot play with the levels. It's easier to cheat. So when you want to have a good reinforcement, negative reinforcement, you have to be able to control the intensity of the aversive. This is a must. Because sometimes you may need very little or nothing and respond straight to the signal. And sometimes depending on your stubbornness and commitment and motivation not to, it has to reach a level that's going to convince you. And that level is not punishment. That's the big kind of bubble that somebody can get into again with thinking that the low
15:48
일종의 잘못된 통념에 다시 빠지기 쉽다는 게 큰 함정이죠. 처벌은 할 수 있어요. 수업 시간에 우리가 나눈 이야기를 생각해 보세요. 경찰이 차를 세우고 경고를 주는 상황 말이에요. 그게 처벌일까요? 당연히 처벌이죠. 어느 정도 강도가 있는 자극이니까요. 네, 꽤 가벼운 수준이죠. 정말 가벼운 수준이에요. 그냥 다시는 그러지 말라고 주의를 준 것뿐이니까요. 그래서 당신은 알겠다고 대답한 거고요. 하지만 그 경고를 무시하기로 결정한다면, 벌금은 더 높아지고 또 더 높아질 것입니다. 그리고 결국 면허가 정지될 때까지 이런 상황은 계속될 것입니다. 그럼 보험료도 훨씬 비싸지겠죠, 그렇지 않나요? 만약 경찰관이 매번 '이봐요, 다음부터는 그러지 마세요'라고만 말한다고 상상해 보세요. 당신은 '네, 경관님'이라고 대답하겠죠. 감사합니다. 감사합니다. 그리고 당신은 떠납니다. 그리고 이틀 뒤에 그가 다시 멈춰 세웁니다. 그는 '그래, 기억나지? 그러지 말라고 했잖아'라고 말합니다. 그런 일이 계속 반복되는 거죠. 안전벨트 경고음이 작동하는 방식이 바로 그런 식입니다. 그러니 동기 부여의 수준은 다시 한번 말하지만 필수적입니다. 그리고 강화에 대해 이야기할 때 반드시 논의되어야 합니다. 우리가 강화를 올바르게 설명하고 싶다면, 그 안전벨트 비유는 아주 좋지 못한 예시입니다. 다시 말하지만, 그것은 수준이 낮거나 높거나 그 중간 어디에 있다는 것을 의미하지 않습니다. 그저 자신의 목적을 다해야 할 뿐입니다. 자신의 목적을 달성하고 당신의 행동을 제어해야 하죠. 그 예시에서, 그것은 강화 작용을 하기 때문에 기본적으로 당신에게 확신을 주어야 합니다. 당신이 무엇을 해야 하는지 아는 일을 하도록 말이죠. 추측하는 게 아니잖아요, 그렇죠? 우리가 부적 강화에 대해 이야기할 때, 적어도 TWC에서 가장 중요한 것 중 하나는 무엇일까요? 그게 무엇일까요? 쉬운 탈출구, 맞죠? 너무 조용하게 있지 마세요. 우리는 개가 고생하는 것을 원치 않습니다. 처음 차에 탔는데 안전벨트를 어떻게 매는지 전혀 모르는 상황을 우리는 좋아하지 않습니다. 경고음은 계속 울리고 있고요. 결국 일주일 뒤에야 방법을 알아낸다면요. 그럼 '와, 알겠다' 하는 식이죠. 누군가 방법을 알려주길 원할 겁니다. 그게 훨씬 쉽고 훨씬 기분 좋잖아요, 그렇죠? 그래서 그게 알아두어야 할 또 다른 중요한 점입니다. 하지만 네, 말씀드린 대로 스키너까지 거슬러 올라가는 이야기죠. 그는 그것을 한 가지로 요약했습니다. 아, 그것은 혐오적 강화, 즉 처벌을 사용하는 것이라는 거죠. 그러다가 시간이 좀 지나면서 상황이 조금 변합니다. 하지만 다시 말하지만, 오늘날에도 여전히 널리 믿어지고 있는데, 단순히 반려견 훈련사들 사이에서만 그런 게 아닙니다. 저는 과학자, 행동학자, 진화심리학자들과 같은 사람들과 대화를 나누는데, 심리학자들도 그렇죠. 그래서 정말 널리 오해받고 있는 개념입니다. 안전벨트 비유가 워낙 널리 쓰이고 있으니, 방금 제가 말씀드린 대로 그냥, 그게 잘못된 비유라고 설명하는 게 낫겠네요. 그렇다면 좋은 비유는 무엇일까요? 중요한 요소들이 포함된다면 어떤 비유든 좋을 것입니다.
levels are enforced for reinforcement and the high levels are for punishment. You can punish. Think of how we talk in class. The, you know, a police pulls you over, gives you a warning. Is it a punishment? Of course it is. Some level of intensity. Yeah, it's pretty mild. It's pretty mild. They just told you don't do that again. And you said, okay. But if you decide to disrespect that warning, you will get higher and then you will get higher. And then this will keep going until eventually probably your license gets suspended. And you pay very different price for your insurance, right? Imagine if the police officer always says, hey, don't do that next time. And you're like, yes, officer. Thank you. Thank you. And you go. And then two days later, he stops. He's like, yeah, remember, don't do that. And that goes on. That's kind of how the negative, how the seatbelt thing works. So the level of motivation, again, essential. And it must be talked about when we talk about reinforcement. If we want to explain reinforcement correctly, that seatbelt analogy is very, very poor. Again it doesn't imply any level, being low, being high, being anything in between. It just, it has to do its purpose. It has to serve its purpose and control your behavior. And in that instance, because it is to reinforce, basically it's going to have to convince you to do something that you know what to do. It's not a guess, right? We talk about what negative reinforcement, like one of the most important things, at least in TWC, is what? The easy way out, right? Don't get so quiet. We don't like the dog to struggle. We don't like if you go first time in the car and you have no idea how to put a seatbelt, and that tone is beeping. And eventually, a week later, you figure it out. It's like, wow, okay. You want somebody to show it. Much easier, much nicer, right? So that's another important thing to know about. But yeah, it goes all the way back to Skinner, as I said. He kind of summed it up all in one thing. Oh, it's using aversive reinforcement, punishment. Then, after some time, things change a little bit. But again, even today, it's very widely believed that, not just among dog trainers. This is like I talk with people that, you know, scientists, behaviorists, evolutionary psychologists, psychologists. So it is really widely misunderstood concept. So since the seatbelt analogy is so widely used, you might as well just do it just what I just told you and explain that that's a bad analogy. Now, what will be a good analogy? Any analogy will be good if what are the important pieces.
19:44
탈출할 수 있어야 하고 나중에는 회피할 수 있어야 하죠. 그건 강화 범주에 속하는 거니까요. 또 다른 건요? 동기 부여 수준을 극복하는 것이죠. 동기 부여 수준을 극복하는 것, 맞죠? 즉, 당신을 통제하고 설득하려면 당신의 고집 수준에 맞추거나, 실제로는 그 이상으로 올라가서 '아니, 넌 이걸 해야 해'라고 말해야 한다는 뜻입니다. 일단 설득하고 나면 그다음 일어나는 일도 매우 흥미롭습니다. 예를 들어, 혐오 자극이 무엇이든 간에 다시 1에서 10까지의 척도로 올라가야 한다고 가정해 봅시다. 우리는 3에서 시작해서 5로, 그리고 7로 올라갑니다. 그러면 '알겠어,' '안전벨트를 맬게'가 되는 거죠. 그렇다고 계속 7에 머물러야 한다는 뜻은 아닙니다. 다음번에 경고 신호가 오면 '그래, 준비됐어, 가자'라고 하는 거죠. 그게 바로 부적 강화의 핵심입니다. 조절을 해야 한다는 것이죠. 고정된 것이 아니니까요. 그래서 반려동물 보호자나 스포츠 트레이너 등 누구와 대화하든 부적 강화에 대해 논의할 때, 그것은 동기 부여 수준에 맞춰져야 합니다. 그리고 그것이 행동을 수행하게 만드는 거죠. 하지만 절대로 특정 수준을 의미해서는 안 됩니다. 소셜 미디어에서 가끔 대화를 하다 보면 항상 이런 일이 생기는데, 제가 뭔가를 언급하면 말이죠. 어째서인지 사람들은 다들 낮은 수준을 쓰지 않으면 개에게 가혹하게 대하는 것이라고 생각합니다. 아니요, 전혀 그런 뜻이 아닙니다. 제가 강하게 하는지 결정하는 사람은 제가 아닙니다. 만약 개가 간식을 원하지 않으면, 다른 간식을 줘볼 겁니다. 그 간식도 원하지 않으면, 공을 줘볼 수도 있겠죠. 아니면 식사를 한 끼 거르게 할 수도 있고요. 하지만 저는 개가 그것을 원하게 만들고 행동을 하게끔 설득할 것입니다. 그리고 만약 부적 강화를 사용하는 상황에서 제가 "이거 해"라고 말하며 리드줄을 살짝 당겼는데, 개가 "아니, 지금은 이게 훨씬 더 중요해"라고 반응한다면요. "네가 뭘 원하는지 알지만, 안 할 거야"라는 거죠. 그럴 때 우리는 개에게 다가가 "아니, 정말 해야 해"라고 말합니다. 그러고 나서 개를 설득하는 거죠. 알았어, 할게. 그리고 바로 다음에 우리가 다시 요구할 때는 혐오적인 자극을 전혀 줄 필요 없이 그저 신호만으로도 반응을 얻어낼 수 있습니다. 우리는 알 수 없죠. 그럴 수 있다는 게 핵심입니다. 상황마다 다르니까요. 항상 일정한 것은 아닙니다. 그렇죠? 혹시 떠오르는 비유가 있다면 말씀해 보시겠어요? 글쎄요, 자동차 안전벨트 경고음이 울리지 않을 때까지 차를 주행하는 것은 부적 벌에 해당할까요? 맞습니다. 맞습니다. 그러니 확실히 그렇게 할 수도 있죠. 안전벨트를 매기 전까지는 차가 출발하지 않으니까요. 그건, 이제는 그게 부적 처벌의 나쁜 예라고 반박할 수 없겠네요. 왜냐하면 완전히 목적에 부합하니까요. 어디론가 가고 싶어 하고, 그냥 안전벨트를 가지고 장난치는 게 아니라고 가정한다면 말이죠, 그렇죠?
It has to be able to escape and later avoid. It puts it in a reinforcement category. What else? Overcome the level of motivation. Overcome the level of motivation, right? So that means to control, to convince you, we have to match and actually go above your level of stubbornness and say, no, you have to do that. And once we convince that what happens next is very interesting too, because let's say we need to go from 1 to 10 again on scale, regardless of what the aversive is. But we start with 3 and then we go to 5 and then we go to 7 and it's like, well, okay, I'm going to put the seatbelt on. That doesn't mean that we have to stay on 7. Next time the warning signal comes and it's like, yep, I'm ready, let's go. So that's what negative reinforcement is about. You have to adjust. It's not a constant. So that's really when you, anytime you talk to your pet client or a sport trainer or your whatever, whatever, and it's negative reinforcement discussion, it matches the level of motivation. And it makes you perform, do something. But it doesn't, it never ever should imply a certain level. Like whenever I talk like on social media sometimes and we have some conversation, there is always, I would mention something. And for some reason everybody thinks that, oh, if you're not using low level, that means that you're hammering dogs. Like, no, it really does not mean that. I'm not the one to decide if I am. If the dog doesn't want the cookie, I will offer some different treat. If he doesn't want that treat, maybe I will offer a ball. Maybe he will skip a meal. But I will convince him to want it and make the behavior. And if we have negative reinforcement, and I say, hey, do that, and I make a little pop on the leash, and he's like, no, this is way more important right now. I know what you want, but I'm not going to. Then we meet him and say, no, you really have to. And then we convince him. Okay, I'll do it. And the very next time we ask him, we may not even need to do anything aversive, just the signal, and you have the response. We don't know. It can, this is the thing. It varies. It's not a constant. Yeah? So if you think of any analogy, go ahead. Well, I was just wondering if the car, say, put into the car and drive unless your self is on, would that be negative punishment? Correct. Correct. So you can certainly do that too. Until you put the seatbelt on, the car doesn't go. It's a, now I cannot argue that that's a bad example of negative punishment, because it totally is going to serve the purpose. Assuming that you want to go somewhere, and you're not just playing with the seatbelt, right?
23:27
그러니 그 예시에서는, 네, 맞습니다. 또 다른 멋진 예시를 생각해 볼 사람 있나요? 아예 예시를 다 바꿔보는 건 어떨까요? 그 예시가 얼마나 대중적이고 또 얼마나 오해의 소지가 있는지 정말 말도 안 되는 수준이죠. 제 말은, 다시 말하지만, 저는 정말 다양한 교육 수준을 가진 다양한 사람들과 이야기해 왔고 각기 다른 일을 하는 사람들을 만나봤는데, 그들은 알 거라고 생각했거든요. 그런데 가서 들어보면, 뭐랄까요, 유튜브에 가면 모든 걸 찾을 수 있지만, 강화나 조작적 조건 형성에 관한 강의를 찾아가더라도 십중팔구, 대학 강사나 교수가 수업 시간에 부적 강화가 무엇인지 설명할 때, 99%의 확률로 안전벨트 비유를 듣게 될 겁니다. 아니면 다른 예로, 두통이 있어서 아스피린을 먹는다거나 하는 식의 것이거나, 비가 와서 우산을 쓰는 것 같은 예시들이죠. 네, 안전벨트 말고도 부적 강화의 예는 더 많지만, 여전히 강아지 훈련에서 말이죠, 부적 강화와 정적 강화에 있어서 우리는 단서를 만들고 싶어 합니다. 신호를 주고 싶은 거죠. 우리는 개가 반응하기를 원합니다. 그리고 어떻게 강화할지, 어떻게 개가 단서 없이도 특정 방식으로 반응하도록 설득할 수 있을지에 대한 선택지를 활용하는 겁니다. 그렇죠? 그러니까 밖에 나가려는데 비가 오고 우산을 들고 있는 상황인 거죠. 그게 '앉아'나 '엎드려' 혹은 '이리와'와 꼭 일치하는 건 아닙니다. 그 안전벨트 알림이 얼마나 느린지 이해하시나요? 유튜브에 들어가서 조작적 조건 형성 강의를 찾아보면 예일 대학교 강의든 누구의 강의든 간에 이 비유를 듣게 될 가능성이 매우 높습니다. 그만큼 이게 혼란스럽다는 겁니다. 그리고 아주 오랫동안 이런 상태였죠. 조금만 생각해보면 반박하기가 매우 어렵습니다. 안전벨트를 어떻게 하면 부적 강화물로서 개선할 수 있을까요? 우리가 무엇을 할 수 있어야 할까요? 안전벨트가 단순히 경고 신호나 이정표, 혹은 성가신 알림이 아니라, 정말로 제 역할을 하게 하려면 어떻게 해야 할까요? 부적 강화는 원하면 아주 쉽게 무시할 수 있습니다. 어떻게 해야 이것이 정말 진정한 의미의 부적 강화가 되게 만들 수 있을까요? 좌석에 전기를 흐르게 하는 거죠. 좌석에 전기를 흐르게 하는 겁니다. 어떻게 해야 '레드 점프'처럼 좌석에 (효과가) 나타나게 할 수 있을까요? 그렇죠, 맞습니다. 80년대에 캐런 프라이어의 '개는 훈련시키지 마라(Don't Shoot the Dog)'에 나온 방법이 하나 있었습니다. 그녀는 주차된 차의 주차 미터기 시간이 만료되었는데 경찰이 차를 폭파해버린다는 아주 기발한 비유를 제시했죠. 좀 불합리하죠, 그렇죠? 우린 그 정도까지는 필요 없습니다. 하지만 그럴 수도 있는 거잖아요. 그 사이 어딘가에 단계가 필요하다는 거죠, 맞나요? 이렇게 갑자기 '펑!' 하고 터지면 안 되니까요.
So in that instance, yes, that's correct. Anybody can think of cool example? Maybe we can change altogether the examples. You know, like it's really insane how popular that example is, and how misleading it is. I mean, again, I've talked to so many different people from different, different education and different, you know, things that they do, and you would think that they would know. But you go, you listen to, I don't know, go to, I mean, on YouTube you can find everything, but even if you go to some lectures on reinforcement or operant conditioning most likely, and some college teacher or professor explaining to the class what negative reinforcement is, chances are 99% that you will hear the seatbelt analogy. Or another one will be, well, you have a headache and you take aspirin, or something of that nature, or it's raining and you put an umbrella. Yes, there are more negative reinforcement than the seatbelt, but it's still, like in dog training, the thing with negative reinforcement and positive reinforcement is like we want to have a cue. We want to have a signal. We want the dog to respond. And then we are playing with options how we can reinforce, how we can convince the dog to respond in a certain way without cue. Yeah? So you're going outside and it's raining and you have an umbrella. It doesn't necessarily match as your seat or down or cum. Do you understand how slow that seatbelt thing is? And it really, like if you go to, on YouTube and go operant conditioning lectures, it can be Yale University, it can be anybody, really the chances are that you will hear this analogy. This is how confusing it is. And it's been like this for a very long time. And it's very hard to argue it when you actually think a little bit. How we can improve the seatbelt as a negative reinforcer? What do we need to be able to do? What needs to happen if we're going to improve the seatbelt to really serve not as a warning signal and signpost and half us annoyance, negative reinforcement, that is very easy to override if you want to. How can we make it so it actually really truly is negative reinforcement? Electrify the seat. Electrify the seat. How can we make it so it's on the seat like the red jump? Sure, sure. There was a way back in the 80s when Don't Shoot the Dog, Karen Pryor came with it. She came up with a very cool analogy where you have the car parked and the meter is already expired and the cop comes and blows up the car. It's kind of unreasonable, right? We don't need that. But we can get there. It's just there's got to be levels, something in between, right? It cannot be this and ta-pah!
27:13
그러니 소리를 점진적으로 높일 수 있겠죠. 제 말은, 소리로 할 거라면 그 소리를 키우자는 겁니다. 경적을 울리는 축구 광팬들처럼 거의 그런 정도로요. 테슬라나 일론 머스크라면 안전벨트를 맬 때 이상한 소리가 나게 만들 수도 있을 거예요. 하지만 단계가 있어야 합니다. 이것이 중요한 점입니다. 이런 이야기를 들을 때마다 굳이 세상 사람들을 설득할 필요는 없습니다. 하지만 고객과 함께 작업할 때는 사물이 어떻게 작동하는지 설명하는 것이 여러분의 의무라고 생각합니다. 개가 자신에게 요구되는 것을 이해하고 있는 한, 조건부 혐오 자극을 적용하는 것과 비조건부 혐오 자극을 적용하는 것은 매우 다릅니다. 이는 거의 모든 연구에서 다루는 내용인데, 전기 칼라나 프롱 칼라 등을 금지하는 논의에는 항상 '개가 상황을 이해하지 못하면 어떻게 되는가?'라는 의문이 따릅니다. 그건 나쁜 생각이며, 굳이 연구까지 할 필요도 없는 문제입니다. 나쁜 생각이죠. 하지만 세상에는 나쁜 생각이 많습니다. 어떤 바보가 포크로 자신을 찔렀다고 해서 우리가 포크 사용을 금지하지는 않잖아요. 그렇다면 차가 폭발하기 전에 단계적으로 강도가 높아지는 안전벨트를 어떻게 만들 수 있을까요? 바로 레벨을 두는 것입니다. 저도 생각해 봤습니다. 소리부터 시작할 수 있습니다. 소리를 점점 크게 만들 수도 있고, 낮은 단계의 전기 자극부터 시작할 수도 있죠. 전기 자극의 단계를 높일 수도 있고요. 무엇이든 간에 특정 대상이 반응하도록 설득하기 위해서는 그 대상에게 충분히 불쾌한 자극이어야 합니다. 그러니 자극의 강도가 훈련사로서의 기분이나 좌절감을 절대 반영해서는 안 됩니다. 그건 잘못된 훈련 방식이죠, 우리는 그걸 압니다. 강도를 높이기 전에 대상이 무엇을 해야 하는지 확실히 이해하고 있는지 확인해야 합니다. 아까 말했듯이, 만약 당신이 차에 앉아 있는데 경고음이 울리기 시작하고 계속해서 무언가 요란한 상황이 이어지는데, 정작 무엇을 해야 할지 모른다면 그것은 잘못된 일입니다. 당신에게 매우 불공평한 처사가 되겠죠? 따라서 우리는 대상이 무엇을 해야 할지 알고 있는지, 교육을 받았고 수행 능력이 있는지 확실히 해야 하며, 아니면 반사적인 행동이어야 합니다. 예를 들어 뭔가를 만졌는데 작은 전기 충격을 느꼈다고 해봅시다. 그러면 어떻게 하시겠어요? 더 세게 밀어붙이지는 않겠죠. 손을 뒤로 뺄 겁니다. 따라서 그런 상황에서는 굳이 아주 자세한 지시가 필요하지 않습니다. 왜냐하면 반사적으로 본능에 따라 반응하게 될 테니까요. 하지만 안전벨트의 경우, 네, 이렇게 하는 겁니다. 안전벨트를 착용하고 그 이유를 설명해 주는 거죠. 그게 장점이고, 그게 단점입니다.
So we can increase the sound. I mean if it's going to be a sound, let's increase it. Let's go to almost like the soccer maniacs with the horns. Like I'm sure Tesla, Elon can come up with some weird sounds that you are putting that seatbelt on. But there's got to be levels. This is what is important. Like anytime you hear this, you don't need to convince the world otherwise. But when you are working with a client, I think it's your duty to explain how things work. And that as long as the dog understands what is asked from them, it's very different if we're applying contingent and non-contingent aversive. This is almost in all the studies, you know, and the banning electric collars and prong collars and whatever, there is always this, well what happens if the dog cannot make sense of it? It's a bad idea, which there is no need to make a study off of that. It's a bad idea. But there is many bad ideas in the world and we don't, you know, we don't forbid forks because some dumbass just popped himself and, you know. So, how do we make a seatbelt that can escalate but before it blows up the car, we have levels. I've thought of it. You can start with the sound. You can start to make, like the sound can increase this, you can start some low level of electric. You can increase the level of electric. Whatever it is, it has to be unpleasant enough, whatever that is, to the, that particular subject to convince them to respond. So, that intensity never really reflects your mood as a trainer and your frustration. That's bad training, right? We know that. Before the intensity increases, we need to make sure that you understand what you need to do. So, as I said, if you're going to sit in the car and the tone starts and then it keeps going and then it keeps doing all sorts of things, but you have no idea what to do, that will be wrong. That will be very unfair to you, right? So, we have to be certain that either the subject knows what to do, it's instructed and we know that it can, or it's a ref, I cannot say the word now, by reflex. You know, like let's say you're touching something and you feel that little electric impulse, what are you going to do? You're not going to push back harder. You're going to go back. So, in that situation, you do not necessarily need very deep instruction because by reflex you will respond in the way that you, you should be responding. But with the seatbelt, we, yeah, this is what you do. You put your seatbelt on and you explain your why and, and that's the pros, that's the cons.
30:56
어떻게 할지는 스스로 결정하세요. 하지만 강화를 주고 싶다면, 이런 방식으로 진행될 겁니다. 자, 우리가 말했듯이 네, 우리는 정말 강력하게 강도를 높일 수 있지만, 혐오 자극의 강도는 적절하게 적용되어야 합니다. 지능적인 수준이어야 하죠. 무엇을 고려해야 할까요? 방금 말씀하셨듯이, 혐오 자극의 강도입니다. 안전한 긴급 상태요. 안전한 긴급 상태, 맞습니다. 제가 한번 읽어드릴게요. 학습을 위한 이상적인 감정 상태는 '안전한 긴급 상태'입니다. 다시 말해, 불안이라는 부정적인 영향 없이 높은 수준의 집중력을 유지하는 상태를 의미합니다. 그래서 우리는 아주 의도적이고 체계적으로, 좌절감 없이 교육을 진행합니다. 이것이 바로 가르침이며, 강화가 어떤 모습일 수 있는지 보여주는 것입니다. 제가 생각한 것은, 만약 자동차에 안전벨트를 맸는지 감지하는 센서가 있어서, 안전벨트를 맬 때까지 보험 회사가 보험료를 계속 올린다면 어떨까 하는 점입니다. 음, 그렇게 하면 안전벨트를 매도록 강제하는 효과가 있겠죠. 하지만 정말 좋은 비유네요. 그건 정말 값을 매길 수 없을 정도로 훌륭합니다. 네, 기술적인 측면에서 그렇죠. 마치, 오, 팝, 팝, 팝, 팝, 팝, 팝, 팝. 오, 그거 아세요? 오, 네. 그에 대해 이야기하는 게 편하신가요? 왜냐하면 고객들이 '이건 사용하고 싶지 않아요'라고 말할 수 있으니까요. '이건 비윤리적이에요. 이건, 뭐랄까, 내 개를 망치게 될 거예요. 난 내 개를 사랑한다고요'라고 말이죠. 그리고 사실, 그거 아시죠, 때로는 그것이 매우 합리적인 용도로 쓰일 때가 있습니다. 요약하고 이번 주제를 마무리하자면, 우리가 무엇을 배웠을까요? 혐오 자극의 강도와 동기 부여의 수준이 서로 일치해야 합니다. 계속해서 일정하게 유지되어서는 안 됩니다. 그렇지 않으면 부적 강화의 아주 나쁜 예시가 되기 때문입니다. 궁극적으로 안전벨트의 경고음은 경고 신호로서의 목적이 더 크다는 점을 깨달으셨으면 합니다. 상기시키는 것, 이정표, 그게 전부입니다. 부적 강화의 과정인 거죠. 그렇죠? 현시점에서는요. 감사합니다.
You decide what you do. But if you want to make a reinforcement, then that's how it's going to go. Now, as, as we, we said, you know, yes, we can really, really escalate, but the level of aversive has to be applied. It has to be a smart level. What are we looking on? You, you mentioned that already. The level of aversive. Safe state of emergency. Safe state of emergency, right? So, the, I, I'll read it to you. The, the ideal emotional state for learning is a state of safe emergency. In other words, there is a high level of attention without the negative impact of anxiety. So, we, you know, we're very purposefully, methodically, without frustration. We're, this is, this is what teaching is, what reinforcement can look like. So, the, the thing I thought of is, if the car had a sensor to detect whether you put in the seatbelt, your insurance company raises up your premium, it raises the price, until you start putting on your seatbelt. Um, and that would, you know, compel you to put on your seatbelt. But it's really good one. That's, that's just priceless. Yeah, the technology is. It's like, oh, pop, pop, pop, pop, pop, pop, pop. Oh, guess what? Oh, yeah. You feel comfortable talking about it? Because they, you know, like your client will say, hey, I don't want to use this. This is unethical. This is, you know, we're gonna blow up my dog and I love my dog. And you can really, you know, there is a very reasonable place for it sometimes. To, to summarize and close this one, what did we learn? It has to match the intensity of the aversive, has to match the level of motivation. It cannot be constant. Otherwise, it's very poor example of negative reinforcement. And ultimately, I hope you realize the, the beeping tone in the seatbelt serves more the purpose of a warning signal. A reminder, a signpost, done. Negative reinforcement event. Yeah? At this point. Thank you.