1,793 karma · joined January 19, 2023
Well if you do the math, the number of agent-compute time in total, given the insane number of agents thrown at the problem, might end up being comparable in time, if not for the budget.
https://github.com/harbor-framework/terminal-bench-science/t...
Then it's not a valid benchmark. I agree though they're not reliable enough to just put results in a paper.
> “Not just those who pose an immediate threat, there are people there who are not worthy of life. They shouldn’t live. They’re not even people,” added Ben-Gvir.
Not sure how the full quote is supposed to help.
I remember HN comments used to be of higher quality than this.
Being snarky doesn't really help your position.
You are doing that implicitly by fitting a Gaussian curve.
That's a very big word you're using there for what is basically making shapes out of clouds. A bell-curve is the amortised function of a random variable with a mean and standar deviation. What does that have to do with a timeseries dataset?