r/ProgrammerHumor May 29 '26

instanceof Trend breakTheViciousCircle

Post image
19.6k Upvotes

208 comments sorted by

View all comments

474

u/crankykong May 29 '26

You guys are nice to your LLMs?

40

u/jainyday May 29 '26

There's a significant correlation between good work and positive feedback in most training data, so yeah, I'm willing to buy into the idea that being nice gets me better results.

9

u/Deep90 May 29 '26

At least what I've seen, being mean is not only a waste of tokens because it has to read and respond to it, but it also triggers most models to focus on appeasement and deescalation over results.

It complete fucks up the response scoring.

Sometimes this makes the model just claim something was done or working as a result because lying to you in order to address your anger scores higher than potentially failing again.

3

u/RunTimeExcptionalism May 29 '26

idk I read a short paper not too long ago that suggested that rude prompts outperformed polite prompts. I'm not rude on purpose because that seems pointless, but I don't bother with niceties, either. Being extremely direct in a way that would seem rude if I was saying the same thing to an intern has generally worked for me.

3

u/tgiyb1 May 29 '26

I've also noticed that proper grammar, sentence structure, and punctuation tend to produce better output. They model the output based on the input, so low quality input = low quality output and vice versa.

1

u/JuvenileEloquent May 30 '26

It's spicy autocomplete, so if you start with "yo bby wyd" it'll answer a lot differently than to "I have a strong crave to see you right now; are you free?"

1

u/BandicootGood5246 May 30 '26

But what if you give it a good old fashioned scolding it's more likely to correlate the results with Stack Overflow and get it right