How AI tool calling works (40 lines of vanilla JavaScript)
Thread
Unofficial Hacker News client; not affiliated with Y Combinator.
How AI tool calling works (40 lines of vanilla JavaScript)
Unofficial Hacker News client; not affiliated with Y Combinator.
stratos123 · · focus · HN ↗
It also spends an entire section trying to convey that large tools waste tokens:
but surely that's wrong - it's part of the same conversation, it only gets processed once and then cached. Otherwise doing long conversations would always cost an amount quadratic with length.dvt · · focus · HN ↗
This isn't necessarily true. I'm working on a local harness that doesn't do this and instead coerces everything to YAML (including tool calls) for better bucketing. Some models are indeed trained on the `<|tool_call>...<tool_call|>` token schema (or something similar—e.g. jinja), but it's vendor-specific and often times inconsistent (so you're constantly fixing calls or going back to the LLM).
stephenblum · · focus · HN ↗