2,841 karma · joined March 22, 2023
I'm on the other side, I hate webp because if I download such an image a lot of apps don't support it. So typically I need to screenshot it instead to get a usable jpg/png.
Now one more format to hate.
I would say webp/jxl/avif is anti-general compute, since they require very recent software.
On Linux is free for all, a regular app running as a user can read/exfiltrate anything owned by that user.
And sandboxing like snap/flatpak is not popular.
Do you think they also are mostly deficient of scruples?
Which can be fine, people go to therapists, but not without knowing it !!!
Now it seems they want to go head to head with DuckDB.
Last week I've have about 4 resets in 48 hours:
- reset done by OpenAI around Friday
- banked reset expiring on Saturday
- banked reset expiring on Sunday
- regular weekly reset expiring on Sunday
So basically I was unable to use them, given weekend and all at once, unless you have the software factory ready to spin up...
I'm not complaining, they are free after all, but it's clear they are not randomly distributed.
This reminds me of the early days of Covid: "it's only 10 cases, worry about the flu"
most of this stuff will be "unconscious" the model will have a pull in a particular direction without being aware
> stray thoughts from previous context floating around. We normally dismiss them
It's not that simple, see https://en.wikipedia.org/wiki/Priming_(psychology).
> Priming is a concept in psychology and psycholinguistics to describe how exposure to one stimulus may influence a response to a subsequent stimulus, without conscious guidance or intention
Also, we are AGI, the model isn't, it can't (re)organize its thoughts as easily as we can.
But nobody knows, what will happen is people will experiment with this, you can run the benchmarks, if it works it will be used, if not, well...
However given it's a super-obvious thing to do, drop middle tool calls from cache and keep the rest without re-prefilling, I would guess it degrades performance quite a lot, otherwise the labs would have been doing this already.
Example:
Trajectory: >>What is the capital of France? Let me think<<
the KV cache tokens for "let me think" will have baked in them among other things "France", "Paris", "major city", "question"
Now you compact, and reuse the later tokens (suffix)
Compacted trajectory: >>Let me think<<all right, let's see<<
Now the compacted KV cache for "Let me think" will have baked in them "France", "Paris", and the model can be "that is weird, why am I thinking about France and Paris? there is nothing related in the previous context"
But this is the kind of thing you could ask your agent to test locally.
Also, FT is too big to not be aware of AI writing, and to not have internal checks (Pangram, ...)
AI speak infects human language now, people start to talk like LLMs.
Having it also on the desktop makes things, less compatibility issues.
Yes, it's the manufacturers fault, not the unbelievable market demand.
We are all Capitalists, until the Market comes after the stuff we love.
I see that Slack/Discord/... are on the roadmap, but I also see that Slack can be added as an Integration, so I guess what's missing is inviting Agenta to Slack or messaging it directly?
Also, you might want to update the changelog (or remove it), I thought initially that development slowed down, last release listed there 3 weeks ago, but on github I see frequent recent releases.
One thing you could try is use it as an Oracle "is P = NP", YES or NO.
Or it can output a Lean proof, which gets checked on another air-gapped computer, the computer shows a single bit - proof valid or not and then the computer is destroyed (together with the proof that might contain a trojan).
I scrolled this website in 20 seconds and I have no idea what it does. Seems like some sort of dynamic summarize widgets.
You can stay on your horse. It's perfectly usable. Don't fall for the hype.