I believe this is the popular "batteries included" kit: <a href="https://github.com/can1357/oh-my-pi" rel="nofollow">https://github.com/can1357/oh-my-pi
Yeah I just gave it a run and it starts off burning 6% of my tokens on start.
Pi sits at 0%.
I'm so use to the command keys in Pi that having to type something out in OMP slows me down. I'll stick with Pi, I've built it to my needs and having WSL finally working. I still have to run 2 CLIs, one for LLama.ccp and the other for Pi, not sure if this is normal.
OMP provides a significant amount of structure to the work. That costs some tokens in terms of the system prompt. When I use these systems, I balance the return of the extra guidance in the system prompt with the cost of the context. You've talked about how much context it takes, but you haven't talked about any difference in behavior.
omp offers some nice tools, like /shake, to manage context efficiently and offset some of that consumption.
Not sure about your point with command keys. OMP has all kinds of shortcuts and is highly configurable. It seems very unlikely that you cannot achieve what you're looking for in OMP in terms of shortcuts. Specifics would be immensely helpful here.
And yeah, you'll need to run your inference engine separately from your harness, for the same reason the web server and the web browser are different applications.
wasting_time · · focus · HN ↗
personjerry · · focus · HN ↗
mcast · · focus · HN ↗
razster · · focus · HN ↗
I'm so use to the command keys in Pi that having to type something out in OMP slows me down. I'll stick with Pi, I've built it to my needs and having WSL finally working. I still have to run 2 CLIs, one for LLama.ccp and the other for Pi, not sure if this is normal.
rpdillon · · focus · HN ↗
omp offers some nice tools, like /shake, to manage context efficiently and offset some of that consumption.
Not sure about your point with command keys. OMP has all kinds of shortcuts and is highly configurable. It seems very unlikely that you cannot achieve what you're looking for in OMP in terms of shortcuts. Specifics would be immensely helpful here.
And yeah, you'll need to run your inference engine separately from your harness, for the same reason the web server and the web browser are different applications.