I've been so far using pi like this:
> pi --offline --no-extensions --no-skills --no-prompt-templates -nc -nt --thinking low --system-prompt "$(cat <custom purpose path for system/role prompt>)
Purpose was to reduce token usage to an absolute minimum (zero extra token) as often I use this per use API keys (and not my GLM key, which let's say, is a bit "different" when it comes to conversations).
Are there any other tools specifically designed for such tasks? Even though hax looks like the absolute bare minimum one can go while still being usable.
I wish there was also tools that would refuse and reject (or prevent) models from inserting "thinking" messages into responses. They keep appending those and that very soon that part starts snowballing on steroids.
it grew out of annoyance of dependencies on js runtimes, probably similar to you. mine additionally works on solaris and esp32.
could be interesting to collaborate!
If anyone's aware of a smaller agent than this, hit me up!
[1] https://gist.github.com/fourlexboehm/a60e4ef9306744483731cd1...
So the points about it being a well behaved unix tool, installing it via brew and it not being react are points I - love - to see.
Thank you for making it, I will definitely try it out.
It's a command you copy/paste, but you are free to edit it to suit your whims.
A lot of people have multiple files for their injected system prompts (a.la openclaw or hermes), I think it would be a good idea to either add or modify this point to be able to handle multiple file injections (system_prompt_append_folder or the like). Fitting that shape would make it easy to compare to those systems and make it easier for people to transition from those systems to yours.