A single function Jev-like wrapper for LLMs, including vision models
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
A single function Jev-like wrapper for LLMs, including vision models
Unofficial Hacker News client; not affiliated with Y Combinator.
TeMPOraL · · focus · HN ↗
Most interactive tech on Star Trek is like this - from phasers to consoles to communicators to voice interactions with the ship's computer. The computer seems to be aware of the user and surrounding, and actively infers intent from context, to DWIM ("do what I mean") and when they mean it, instead of doing dumb things[1] on simple triggers.
--
[0] - The direction, not final implementation - surely we can work out how to do it more efficiently than wrapping around final stage of LLM. But the point is, multimodal.
[1] - Obviously it's a fictional show, but in this, both Watsonian and Doylist explanations align near-perfectly: this is/portrays advanced technology, that Just Works and doesn't do stupid shit. Same intent recognition algorithm is there - fictionally in the computer, in reality in the minds of on-set technicians.
[deleted] · · focus · HN ↗
[deleted]