Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

There are two different things:

- was item X in the training data

- did the inclusion of X in the training data lead to Y

I understand why the second is hard, but why is the first one hard?

 help



Yep, the last part in my post was suggesting some ways we could determine if "item X was in the training data" (as well as some potential blockers for that from OpenAI's perspective.)



Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: