Still impressive that it did so much just to answer my simple question.
824 karma · joined April 9, 2026
Still impressive that it did so much just to answer my simple question.
Asked pi agent it to identify the main hero sprite size of game I was running. It had a ton of shader effects so it was hard to determine.
It used some cli tools to identify that it was a game made with Godot, decompiled the executable but data was encrypted, broke the encryption after writing a brute force tool to test keys extracted from the exe, then proceeded to extract the game gd scripts and assets, only to answer the question of the sprite size.
And I wonder if Google's main monorepo is already in Anthropic/OpenAI training data because of some stubborn dev.
Even Anthropic moved to auto-approve by default and they are kings of fearmongering.
timeline suggests not.
Had western labs figured that out before, they would have used it to make kv caching cheaper before and not only now.
The burden of proof here is on western labs. But I doubt they'll try to lie that much.
But slug wins in perspective in my eyes.
I use MSDF to render crisp text in my webgl hobby game. Hope to publish it with source code when I get the time.
thanks for sharing the article. I'll take a deeper look at it later.
For CRUD shoveling, models like DS4.1 are enough.
And the intelligence gap between cheap and premium is closing, as can be seen from the title of this post.
I'm told that their mods tolerate left leaning hate speech much more than right leaning counterpart so unaware visitors think current left leaning government is beloved by a great majority and that it has been for years.
But they told me that last election ended very tight, something like 49/51% and I'm told current imminent election is at similar 50/50 split.
Same probably happened with German election.
If most devs were used to Lisp, it would be the other way around.
It's a chicken and egg problem, at this point.
Siri is still subpar despite leveraging third party AI. What a genius move.
I often use Study Together which is a large discord server (50k users online) for folks to work and study in groups/body doubles, but someone told me a lot of Brazilians have to use VPN to circumvent the ban.
> prompt: /tmp/iterReverse.png what could this image be?
Thinking:
> Identifying the visual dataflow language from the node layout and iterReverse label, likely Prograph.
> Confirming Prograph identification and mapping node roles to list deconstruction and reconstruction.
Answer:
> This is a Prograph method window, for a method called iterReverse.
> Prograph is a visual dataflow programming language - you program by wiring boxes together instead of writing text. What you're looking at is the implementation of iterative list-reverse:
- big box = method editor, top pins = inputs, bottom pins = outputs
- black circle top-left = method icon
- yellow boxes = controls: start*+, when*+, unless*+, repeat*+
- blue boxes = list primitives: null test, decons split list into head/tail, cons prepend
- red lines = data links, green = the small control / boolean wiring
> In other words: if input is null, when succeeds, else unless loops with decons / cons + repeat to build the reversed list.> The window chrome with Save / Close and that style of drawing is the Marten clone of Prograph running on Linux.
> You have 14 calendar days to return an item from the date you received it.
That throws a NullReferenceException
https://store.steampowered.com/hwsurvey/Steam-Hardware-Softw...
I use many languages. Some more than PHP. But that's besides the point.
Signals their ease to spit out incorrect information with confidence. That has no place in engineering.
In my experience models just don't take forever to mark tasks as done.
For my coding usage I don't set time limit so benchmarks that do so are providing me less interesting use cases.
With that said, I can understand having budget limits for expensive LLMs for those of us that don't have infinite VC money.
As for no internet access, the issue is that for most problems we do want Llms to be able to search docs, API specs, GitHub issues, etc. So not allowing that usually just favours larger LLMs that were able to memorize more data, not necessarily smarter ones when both have internet access.
- capped per-task budget and time limit
- No internet access
- different harnesses mixed
A $10/mo subscription to OpenCode Go would have done the job for you.
They have models like Kimi K3, Grok 4.6 , GLM-5.3, Mimo 2.6 Pro (launched today, already available) which are happy to follow your orders without accusing you of being a terrorist.
It's a bold strategy cotton, lets see if it pays off for em.
I ask because my wife has the 15T and the camera is better than my iPhone 17 Pro. And while toying around with it I didn't notice any bloat.
Plus hers support native split screen which I kinda need to multitask on the go.
I'm so pissed at how bad Siri is compared to her android phone that I'm thinking about selling the iPhone to get a Huawei Pura Ultra.