Damus
myrmepropagandist · 5w
I said I *didn't* think Taylor Lorenz was an industry shill but... wow... now I'm starting to rethink that. You know I might be wrong, but what is wrong with watermarks? https://cdn.masto.host/saur...
myrmepropagandist profile picture
Let me outline my understanding of "LLM Watermarking"

It is possible to embed special characters and patterns in the output of LLMs that would make text generated by these system easier to reliably detect. This is mainly being done so that when LLMs scrape the web for new information they can avoid ingesting machine generated content.

When you train an LLM on machine generated content it leads to "model collapse" since, like a compressed jpeg, the content is already compressed.
2
myrmepropagandist · 5w
The companies that run LLMs can also use this for PR to calm concerns from the public about the proliferation of such content. From a CS perspective I can't think of any way to have a watermark that couldn't be easily defeated through additional processing. LLM dependent people currently take pain...
Andy L. · 5w
nostr:nprofile1qy2hwumn8ghj7un9d3shjtnyd968gmewwp6kyqpqzdp33shl69xr0uq3x8n5gsjykq9upycwh6nqm02c3f6x0frrn0dqlrpfuj My (weak) understanding of AI Text watermarking is that they tweak the probabilities of words in certain positions so they're more likely to line up with complex watermark patterns. I w...