Wrong, actual code and tests provided than show 10000x speed up. Users can run it themselves and have been seeing better results.
Appreciate if you didn’t make up stuff.
Appreciate if you didn’t make up stuff.
Instantiating an agent is not the bottleneck for LLMs. Two hundredths of a second is a rounding error compared to what the model costs in time.