That would be the litmus test.
"Does not hallucinate" is not the same as "is never wrong".
So the ATC test could be the benchmark.
4,455 karma · joined May 21, 2020
That would be the litmus test.
"Does not hallucinate" is not the same as "is never wrong".
So the ATC test could be the benchmark.
Now with LLMs, they won't gain attention and all knowledge would be stuck in LLMs but there would be no knowledge base to train the LLMs on so... would this all collapse in a decade or two?
PS: Not anti AI rant, just a genuine open question.
Hypocrites hyping for the IPO.
PS: True story.
During off-peak hours, the unit price is reduced from $0.007 for input cache hits to $0.003, $0.22 for input cache misses to $0.15, and $0.12 for output to $0.6
I also find the DeepSeek models to be more precise than Claude models (last I used 4.7) in that I yet had not the occasion where model did something unintentional that I did not direct it to.
EDIT: Updated percentage reduction.
"We are in favour of 3D printers but there should be a body that tests and certifies that a 3D printer cannot print anything that can be used as weapon. Anything pointy or with a spring and recoil or... or..."
Someone is calling corrupt as corrupt. Surprising.
I don't know about that but based on my own experience with Deepseek v4 Lite alone (with high effort) I have no doubt in my mind that anyone claiming such great things about GLM 5.2 must be true because Deepseek v4 already is really awesome.
EU is looking and charting its course already. Yeah, we can joke about it, we can mock it but it is in momentum already, one step at a time.
But I vote for these heroes with my wallet. Just yesterday did again.
Change my mind.
Going forward would be such open source, open data and open recipe models possibly someday even with the training being crowd sourced if not inference like the BitTorrent model.
Lastly, even Chinese models (GLM, Deepseek, MiMax) work really really good and any user would testify that they do not miss OpenAI/Anthropic/Gemini at all if they're using those Chinese models which is argument enough that with such models, no one is going to miss Chinese models as well.
Hence, closed source is what's next probably. Unfortunately.
As if OpenAI and Anthropic are giving us ball to ball commentary on how their training runs go. Deepseek did train it on domestic hardware, model might be out in public soon (open weights or not) and then anyone can see what is it about.
The only problem is - when US services are available, there's no incentive to bring anything to the market.