Understanding the Impact of LLM Watermarking on AI Agent Behavior
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Understanding the Impact of LLM Watermarking on AI Agent Behavior
Unofficial Hacker News client; not affiliated with Y Combinator.
skybrian · · focus · HN ↗
<a href="https://chatgpt.com/s/t_6ab7d694885481918083b8cbf0ba9040" rel="nofollow">https://chatgpt.com/s/t_6ab7d694885481918083b8cbf0ba9040
In particular: sometimes they measure “churn”, which doesn’t show whether the results are better or worse on average. They sometimes only test with one random seed. There are multiple-comparison issues. And they’re not testing Anthropic’s algorithm.
possibilistic · · focus · HN ↗
This is a spy tool.
It's not "watermarking", it's "spymarking": <a href="https://news.ycombinator.com/item?id=49794615">https://news.ycombinator.com/item?id=49794615
cubefox · · focus · HN ↗
sodapopcan · · focus · HN ↗
cubefox · · focus · HN ↗