For comparison the whole git log body on that branch is 3KB; the pack holds 33x the reasoning.
33 karma · joined March 10, 2021
For comparison the whole git log body on that branch is 3KB; the pack holds 33x the reasoning.
I agree agents don't always self-talk decisions, that's why we distill the whole transcript after the fact instead of asking them to log anything. Your baselining idea is good!
- eval is public here: https://github.com/evansjp/grepathy/blob/main/docs/REPORT.md
- fair point on the last one too. For the original Clerk incident the transcript is deleted, so I can't prove nobody approved it. Maybe I even waved it through. That's kind of the point though. Right now there's no record either way. I'd rather have some sort of receipt.
I'd be flattered for sure. I think there's a lot of uncaptured value in AI to API requests, in general. Imagine an app that schedules "Flight at LAX at 2pm", but also offering to order the Uber to the Airport in time to get there 2 hours before departure. Sounds pretty cool to me and until OpenAI's GPT Plugins have broader support, I'll build that :)
That's what a "scheduling assistant" does. :)
AI is good at reasoning what other events should be scheduled along w/ an event.
Thanks for your question!