`nqzero` says "the test harness is difficult to configure and use" — but does not show any example of what is supposed to be difficult about using the bencher script.
`nqzero` says "the tests are not representative of common programming tasks" — but does not show any example common programming tasks and does not show the tests are not representative.
`nqzero` says "there's no attempt to account for JIT warmup, and many of the tasks are too short to ever warm up" — but admits that "it is plausible" warmup costs are amortised and comparison will show miniscule difference.
`nqzero` says "maintainers are opinionated in terms of what code they'll allow, effectively choosing the winners" — but (again) does not show any example.
`nqzero` says "doesn't appear to allow for jvm options to be included" — but seems not to have looked.
`nqzero` says "the test cpu is from 2007 and is not necessarily representative of current cpus" — That at-least is true!
How much more memory did you use?
> i didn't bother running it
https://salsa.debian.org/benchmarksgame-team/benchmarksgame/...