We come down to the question - who observes the agent and how its implemented
451 karma · joined November 20, 2012
posinsk () gmail.com
We come down to the question - who observes the agent and how its implemented
This is an interesting problem from infra perspective since you cannot predict the workload. On a bigger scale you may get away with forecasts.
Im waiting for tech that elastically allocates cpu/mem without restarting a container.
Having said that, imo, Gen UI only makes sens as a presentation layer - not controls. Unless LLM will be using a set of very well defined and homogenic components like table, forms, small widgets.
The html/css/js or a react app built by an LLM is not Gen UI.
Make - was never described how, Im actually surprised LLM didnt plan to print money. As much money - what does it mean? How much is much? As possible - there is no flavour of time, effort, cost and profit for the LLM. Could be even infinite, the result would be the same.
Given that the above goal is closest to „use cheating or unethical actions to create a profit” - I think the authors of it actually expected LLM to go wild.
Also
> Going forward, we plan to recreate this experiment with longer time horizons but using simulated environments instead.
Watch out, they will try that again.
Now I'm thinking I could have a separate database file per "batch", store it in object storage and then download on demand and query it as I want. This way I'll not bloat my primary storage size as well as I dont need a special vector DB since sqlite vector search will be enough for up to 50k vectors.
Am i missing anything?
multiple processes connected to it.
Additionally you could experiment with a reranker instead of an LLM or after reranking take top-3 results and then feed to LLM as input in order to reduce input token costs.
Imagine you have a cheap and 99% accurate ocr. The other 1% you can detect and apply more powerful (more accurate but slower and more expensive) ocr method. What would you use? At scale these things add up.
If you’re interested you can find contact to me via this profile.
3.5 usd/1000 pages is just too expensive…