1,164 karma · joined October 26, 2018
It depends of the town/location and region. In our town, police had to be involved several times due to different issues with creeps targeting 12yo or less.
Edit:And I forgot, as it's Switzerland, police moved only because parents were accompanying their kids and almost done justice by themselves. Protecting creeps is a national sport here too.
Basically 90% are just crap done by computers (pre AI), targeting preschoolers.
And I don't remember to have been able to have pushed to 200k context Qwen 3.6. 3.8 is running on my RTX 5090.
/on The prose is load-bearing unbearable — every sentence feels like it was engineered to sound profound rather than to be read.
The cost of the prebuilt is now lower than today 5090 price.
https://github.com/Neroued/ninfer/blob/master/docs/performan...
Category MTP3 stochastic sampler DFlash stochastic sampler DFlash greedy Code 1/15 natural stops; 0/15 prompt-complete 2/15 natural stops; 0/15 prompt-complete 0/15 natural stops Story 9/15 natural stops; the nine Chinese outputs pass requested division and minimum length 8/15 natural stops; the eight Chinese outputs pass requested division and minimum length 10/15 natural stops; five Chinese dialogue outputs are under length Translation 15/15 natural stops; 15/15 pass structural checks 15/15 natural stops; 15/15 pass structural checks 15/15 natural stops; 15/15 pass structural checks Structured 0/15 satisfy the requested complete record/script contract 0/15 satisfy the requested complete record/script contract 0/15 satisfy the requested complete record/script contract
And on my "own" "quick" benchmark, it's slower than vllm.
DeepSeek is okay for random API-based stuff, as it's cheap.
Local open models running on a 5090 are hit or miss. I feel that most GGUFs/quants are awful...
Suddenly, everyone are expert writing the most elegant, clean and without bug code.
Both sucks in French but at least chatgpt prose is readable while Claude is awful.