Kev: Tiny Jev-like family of decision models built on top of Qwen3.5
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
Kev: Tiny Jev-like family of decision models built on top of Qwen3.5
Unofficial Hacker News client; not affiliated with Y Combinator.
hbarka · · focus · HN ↗
mohsen1 · · focus · HN ↗
also tried myself: <a href="https://console.typesafe.ai/playground?share=shr_1690a3160f19c2e4851b6a2883bcf17a790" rel="nofollow">https://console.typesafe.ai/playground?share=shr_1690a3160f1...
bityard · · focus · HN ↗
Unless specifically told in a system prompt, the pile of weights has absolutely no knowledge of itself. You could hypothetically train it to answer such questions, but nobody bothers to do this, and ALL "knowledge" embedded in the weights is probabalistic anyway.
(I feel like this should be common knowledge in LLM discussions on HN by now.)
akx · · focus · HN ↗
bityard · · focus · HN ↗
Historically, many do not and there are lots of counter-examples proving this. They merely hallucinate an answer just like anything else. The SAME model may even give different answers to the same prompt when asked multiple times... sometimes they claim to be ChatGPT, sometimes Gemma, etc. The fact that the answer is delivered confidently fools people who don't understand this, and these people then run straight to social media with "proof" of their conspiracy theory that one AI lab "stole" another AI lab's model.
My point stands that unless specifically trained or told, big bags of weights do not possess any inherent introspection. LLMs have many fascinating emergent properties, but this is not one of them.
mohsen1 · · focus · HN ↗
temperature?