So predicting the next word given all humanity’s knowledge is surely going to max out at slightly less good (we probably can’t get perfect data) than the best human in any specific field. What test does the AI do to be able to understand it is improving? At some point it becomes impossible to know that the output is actually better right?
LLMs went from predicting the next output to performing search to find the underlying rule that produces the output. It is like going from memorising the Fibonacci sequence to uncovering the rule that produces it. The second type generalises much better to unseen data.
andy_ppp · · focus · HN ↗
password54321 · · focus · HN ↗