It seems like the docs[0] are a better explanation? The comparison to llm tokens is kinda confusing.
It looks like the model takes as input a state (structured text? not sure if multi-modal) and a question (as a "Choice", "Score", or "Noul") with some additional augmentations possible. Then outputs the question's answers as appropriate (e.g. a choice, accompanying probabilities, confidence).
Edit: On the AI primer page, it looks like they do the RLCD on a pre-trained base model?
I do agree that the comparison to LLM tokens is hard to understand (also because output tokens are not comparable).
But yes, text or structured state (like a JSON with multiple pieces of text in) -> decisions out (e.g. choice maps to "match" statement, "score" maps to sorting, "noul" short for bernoulli maps to if-statements)
I see this super interestingly as the "subconscious" to the llms "conscious" for lack of better terms. I'm super interested in this for broad and rapid decision making in the context of consumer agents so will be signing up for sure.
1. I am extremely on the same page
2. I do think that subconscious is not only much smarter than we give it credit for, but also much more robust than the "jagged frontier" of current LLMs
(shilling my blog post on that jaggedness: <a href="https://www.completeskeptic.com/p/lies-damned-lies-and-benchmarks" rel="nofollow">https://www.completeskeptic.com/p/lies-damned-lies-and-bench...)
big_toast · · focus · HN ↗
It looks like the model takes as input a state (structured text? not sure if multi-modal) and a question (as a "Choice", "Score", or "Noul") with some additional augmentations possible. Then outputs the question's answers as appropriate (e.g. a choice, accompanying probabilities, confidence).
Edit: On the AI primer page, it looks like they do the RLCD on a pre-trained base model?
[0]:<a href="https://docs.typesafe.ai/concepts/system-one" rel="nofollow">https://docs.typesafe.ai/concepts/system-one
CompleteSkeptic · · focus · HN ↗
I do agree that the comparison to LLM tokens is hard to understand (also because output tokens are not comparable).
But yes, text or structured state (like a JSON with multiple pieces of text in) -> decisions out (e.g. choice maps to "match" statement, "score" maps to sorting, "noul" short for bernoulli maps to if-statements)
ianbutler · · focus · HN ↗
CompleteSkeptic · · focus · HN ↗
(shilling my blog post on that jaggedness: <a href="https://www.completeskeptic.com/p/lies-damned-lies-and-benchmarks" rel="nofollow">https://www.completeskeptic.com/p/lies-damned-lies-and-bench...)
cheesecakegood · · focus · HN ↗