> Present the original and new writing to the model and ask it which is better.
LLMs have absolutely terrible taste when it comes to writing. I don't find their feedback useful at all, beyond trivial spelling/grammar mistakes, which you don't really need an LLM for in the first place.
Proofreading is all you need.
Edit: I do sometimes ask an LLM for a fact-check, though.
As a "proof" of their poor judgement, take a paragraph you like. Ask the LLM to rewrite it to make it better (which I think we both agree will not make it better), and then in a fresh session ask it which it thinks is best. It'll almost always pick its own writing, even when it sucks.
You are specifically recommending asking the model which of two versions is better (the quote in my top-level comment).
We both agree that they are poor "make it better" machines, but I also believe they are bad A/B testers and I'm using the former to demonstrate the latter.
It doesn't matter who writes what, what matters is that LLMs have a preference for LLM-shaped writing. By A/B testing against an LLMs opinion, you are optimizing in the direction of LLM prose even if the LLM never writes any of the prose itself.
LLM style is not literally anticorrelated with quality. There are some things that they tend to do poorly, but you're not going to do those because you're writing the text yourself. If you have it judge your writing, and are careful to avoid the failure modes that the post goes into, it can be helpful by serving as a competent editor that doesn't share your blind spots.
Retr0id · · focus · HN ↗
LLMs have absolutely terrible taste when it comes to writing. I don't find their feedback useful at all, beyond trivial spelling/grammar mistakes, which you don't really need an LLM for in the first place.
Proofreading is all you need.
Edit: I do sometimes ask an LLM for a fact-check, though.
tptacek · · focus · HN ↗
Retr0id · · focus · HN ↗
tptacek · · focus · HN ↗
Retr0id · · focus · HN ↗
We both agree that they are poor "make it better" machines, but I also believe they are bad A/B testers and I'm using the former to demonstrate the latter.
tptacek · · focus · HN ↗
Retr0id · · focus · HN ↗
ameliaquining · · focus · HN ↗