15 karma · joined January 19, 2026
I'd prioritize audit logs + correlation IDs, and short-lived creds per tool call. Do you expose tool capabilities to the planner without exposing creds?
Whenever I need to transform JSON, I spend 20 minutes guessing filters until something works.
So I built a CLI tool: give it input JSON and desired output, it generates the jq filter.
Example:
Input:
[{"name": "Alice", "email": "alice@example.com"},
{"name": "Bob"},
{"name": "Charlie", "email": "charlie@example.com"}]
Wanted:
["alice@example.com", "charlie@example.com"]
Generated:
[.[] | select(.email != null) | .email]
How it works:1. Takes your input/output examples
2. Generates a filter, runs jq, verifies the output matches
3. If wrong, retries automatically
Works with local models (Ollama) or cloud (OpenAI/Anthropic).
~450 tests, MIT licensed.
Curious what edge cases break it.
1. Crash edge case: If an agent executes a side-effect and dies before signing the receipt, is that action orphaned? Any WAL-style intent/completion model?
2. Multi-step workflows: Do receipts chain natively (parent pointers/Merkle) or via external linking? (I see storage/ledgers are out of scope, but curious about the linkage design.)
The negative proof angle (proving AI didn't touch prod) is compelling for compliance.