I wouldn't even think to justify such a thing. The llm gives a better accuracy to a negative weighted token input, I don't understand how this is so upsetting to people?
I'm actually very shocked to see the responses - as everyone I know uses these tactics to get more accuracy, and there's nothing remotely abusive or meaningful to us.
Maybe there are more 'ai is sentient' type people on hackernews than I realized.
Being an asshole to a machine is still being an asshole.
So boxing is violent. And I have chosen to box in my past. Does that mean I'm a violent person now? Even though I go out of my way to deescalate real fights?
I play games as the villain and and mass murder people in the game. Does that mean I'm a violent extremist?
If you are an asshole to a computer, then you have created similarly biases in my expectations about your potential behavior toward humans.
Observations create expectations. Be an asshole in any context, and people will assume you can be an asshole in other contexts.
Simple solution: never be an asshole.
If you are scared and mad about it - I guess we won't be friends, which is likely the best for both our mental sanity.
And the "me" that lives in a tiny southern town just to help my 95 year old grandma in her last years at the expense of my economic prospects is a facade.
The "me" that helps my aging neighbor when she's sick for no reason is a facade.
The "me" that hugs and loves my wife when I get home is a facade.
The "me" that brushes my aging dogs teeth every night because she has dental issues is a facade.
The "me" that flies to my friend I haven't seen for years and takes care of them after extreme health issues is a facade.
But,the "me" that puts tokens in a token machine in a way that gets better accuracy is the "real" me.
Oh. I also play violent video games where I murder people sometimes as well. Do you think that makes me secretly a murderer too?
This is not a game of having done X good things in life and therefore being afforded the right to do Y bad things. You are making a choice to say, "I am allowing myself to treat this thing I believe is lesser than me in a way I willingly acknowledge is bad." That's your thesis. I wholeheartedly disagree with it.
So yeah, I whole heartedly with 100% of my being think llms are just an input/output/processing computer, I don't think they are aware, feeling, sentient beings.
So yeah, putting negative sentences in a processing machine that forces it to return higher accuracy results is something I don't have any feelings about.
I'd never yell at a cat or a dog. I'd never be mean to another person. As those aren't just hardware/software. I'd be fine smashing a rock violently. Or entering a negative text in a language model.
Putting negative tokens in a machine is no different than playing a violent video game to me. It's not about, oh I'm a good person - so I can do bad things. It's just a neutral thing.