Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

It's very important on translations at least. Watermarking will result in poorer results.

Do they also do it with code? Do you think deliberately picking tokens that are not the highest probability in code is acceptable for the consumer?



What's your evidence that it will result in worse translations?

I'm skeptical that such a thing as a universally optimal translation exists in cases beyond the trivial. But if it does, I see no reason to think LLMs are anywhere close to it, so I think nobody will be able to tell the difference with watermarking.

That's certainly true for code. LLM code is at best mediocre. There is oceans of room to subtly watermark generated code without practical impact.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: