Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

The “trust summarized thinking” part is valid in so far as perhaps the models which appear to be faithful to user intent are lying about that faithfulness, but there’s no incentive for models to pretend to be faithless. Moreover, the results track how annoying the models are when they disagree with you.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: