Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I think that's pretty obvious and shallow, and anyone that knows a little bit about how LLMs work will know that.

The question is: why do they start cheating when we beat them with a stick?

LLMs are not human, they are just multi variable regressions on steroids, so this behaviour couldn't have emerged from the code, it provably emerged from the training and/or fine tuning set, so what's in this set that makes them behave like this?

Is it just a bad set or is cheating inherently part of human behaviour?



How many people are cheating at job interviews? How many posts have we seen by humans on HN even justifying their cheating on job interviews and working multiple jobs without informing their employers? How many submissions have we seen about students cheating on schoolwork, particularly since the advent of LLM? Of course cheating is inherently part of human behavior.

So you are saying there are cheating examples in the training set?

If that's the case, then we can simply clean up the training/fine tuning set and solve this mess.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: