I think I get your point but I think it's reductionist to the point of being incorrect. LLMs must be better at some semantics than others. Programming languages don't have random semantics, they have what matches the world and what matches our languages and so on. And the current frontier LLMs aren't so generic that they can predict any phrase no matter the quality of the content and grammar. Concretely I mean some languages are harder to reason about (predict) than others.
The semantics they are trained on include latent representations of the world that are superior to any pre-existing PL corpus. That is my guess and it's not any less rigorous than your guess.
jmull · · focus · HN ↗
Implement whatever abstractions you think LLMs should work in terms of in whatever language is handy, and have your LLM use those abstractions.
hollars · · focus · HN ↗
jmull · · focus · HN ↗
hollars · · focus · HN ↗