Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I thought LLMs don't say when they don't know something because of how they are tuned and because of RLHF.


They can say they don't know, and have been trained to in at least some cases; I think the deeper problem — which we don't know how to fix in humans, the closest we have is the scientific method — is they can be confidently wrong.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: