Microsoft exec called AI scraping 'the largest theft of labor in human history'
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Microsoft exec called AI scraping 'the largest theft of labor in human history'
Unofficial Hacker News client; not affiliated with Y Combinator.
jacquesm · · focus · HN ↗
It is said that at the heart of every great fortune there is a great crime, so it should be no surprise that the most valuable companies on the planet will most likely result from this crime. And given that justice can be bought by those with the most money you can forget about anything coming of this.
user43928 · · focus · HN ↗
The 'sell it back to us' argument falls short in my view.
Free versions are abundant, and in some time useful models will ship preinstalled on all mobile phones.
The comment here seems incredibly pessimistic and quite dramatical.
SecretDreams · · focus · HN ↗
It's asinine that you think the sell it back to us argument falls short.
Not only does it distill our history to try to sound like some average version of us, it sounds like the blandest versions of us... And then sells this back to us.
From a coding standpoint, the tech is good and gets the job done. The pillaging of all other aspects of human history is just sad. With the only solace I'm seeing is that future training has to train on the dogshit versions of the internet that are now infected with LLM content.
user43928 · · focus · HN ↗
Making it accessible, understandable, and usable is another matter.
How LLMs sound is not a fundamental limitation of the technology.
The current model's poor writing style and tone are currently a main focus of research and I would expect improvements there soon.
You do not have to train on anything you do not deem up to standard. This supposed poisoning of training data remains a common fantasy.
SecretDreams · · focus · HN ↗
I see no evidence that this is what the use of LLMs is accomplishing for most users. Rather, they see to get distilled answers without the depth required to fully understand the response. Partially because that's what they like, and that's what the LLMs serve. Deeper understanding is not being given by LLMs. Instead, it's the SEMBLANCE of depth and laypeople don't know the difference. It's effectively eroding comprehension for some cool knowledge dopamine hit.
frozenseven · · focus · HN ↗
SecretDreams · · focus · HN ↗
frozenseven · · focus · HN ↗
SecretDreams · · focus · HN ↗
Anamon · · focus · HN ↗
1) LLM content can at best be as good as the source material it was trained on. That's the upper bound. "Out of distribution" output of LLMs is mostly unusable.
2) LLM-generated content increasingly drowns out original content, online and elsewhere.
This spells "monotonically decreasing content quality" to me.