I’m thinking about implementing a Jev like model into an agentic harness I’m building. Still it woildnt be enough since Jev like model woild only judge single actions, the case is that agent can build a rogue strategy step by step where each one in isolation is totally safe but as a whole they make up danger behaviour.
We come down to the question - who observes the agent and how its implemented
piterrro · · focus · HN ↗
We come down to the question - who observes the agent and how its implemented
ramkumar2606 · · focus · HN ↗
[dead]