Show HN: LLM 100k portfolio management benchmark
github.com
PoC for something some the potential to yield some interesting results eventually.
github.com
That's the hypothesis this experiment is trying to validate but so far I have no reasons to believe they will behave much worse than human portfolio managers.
* Update portfolios based on model decisions
In the project overview at the top of the readme
This task may be a good proxy to measure how well LLMs are able to coordinate the aforementioned efforts.