4,609 karma · joined July 6, 2015
That's surprising considering how good their documentation is. A tool using LLM should have no problem with that. WolframLanguage is almost ideal for an LLM actually.
I also played around with the canvas and couldn't draw anything, I tried taking the dot product of a "raven" and a "crow", piping it to a "sun", and didn't see anything. I would have expected something since a raven and crow are similarish so should have a non-zero dot product. But for that matter a dot product of a raven and a raven piped to a sun doesn't show anything either so I'm just completely lost.
The deepseek v4 paper talks about one variant of this (related to failures) and how they mitigate it.
>During preemption, we pause the inference engine and save the KV cache of unfinished requests. Upon resumption, we use the persisted WALs and saved KV cache to continue decoding. Even when a fatal hardware error occurs, we can re-run the prefill phase using the persisted tokens in WAL to reconstruct the KV cache.
>Importantly, it is mathematically incorrect to regenerate unfinished requests from scratch, as this introduces length bias. Because shorter responses are more likely to survive interruption, regenerating from scratch makes the model more prone to producing shorter sequences whenever an interruption occurs. If the inference stack is batch-invariant and deterministic, this correctness issue could also be addressed by regenerating with a consistent seed for the pseudorandom number generator used in the sampler. However, this approach still incurs the extra cost of re-running the decoding phase, making it far less efficient than our token-granular WAL method.
(I noted the same thing a few weeks back, https://news.ycombinator.com/item?id=48341224 but his recent blogposts should make it crystal clear if there was any lingering doubt).
See CLIP https://github.com/openai/CLIP
For some reason most of the uses of "agents" are to build yet other AI products, it's turtles all the way down. Maybe that says more about the field of harnesses than it does about the power of "agents".
No, you miss that "lying flat" is only possible when cost of food/living is low and housing is abundant.
https://www.youtube.com/watch?v=5MdSE-N0bxs is remarkably prescient given that it was written before LLMs
Because people believe it will go up, it's a prediction market for attention. The stock market seems to have decoupled from underlying "real value" of a company and is closer to cryptocoins now. Of course the flipside is that you also have the same volatility of cryptocoins.
Why does a remote control require a RTOS?
Also this is basically a replacement for the zombie TotalSpaces 3
The point of "X is real" in a therapeutic context is to make the person feel seen and acknowledged, that his struggles are real to him and really do weigh on his mind, even if it is technically "all in his head".
Here it doesn't even make sense, of course the VRAM is real. Is it going to tell me that my keyboard is real next?
I wonder if this was generated with the local model, this seems to be a case where it memorized the style but not the meaning and intent.