HNHacker News
TopNewBestAskShowJobs

pdntspa

3,539 karma · joined August 8, 2022

submissionscomments
pdntspa··on Singapore govt dating app uses Gale-Shapley stable marriage algorithm
I don't know... mid 2010's pre-match.com OkCupid and PlentyOfFish seem to get a lot of love. Not saying they didn't have problems but they did actually match a lot of happy couples and online dating didn't feel like nearly as much of a wasteland as it does today
pdntspa··on The AI Race Just Got Awkward
It isn't, that is whole point. A normie hears that word and they think vodka.
pdntspa··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
I just ran a huge text/image extraction grudgematch against all the current inexpensive models except gpt-5.5/5.6/6 (due to some issues with openrouter and bugs in my code) and DS4 ranked very poorly. Accuracy winner was Gemini 3.8 flash with minimax M3 and qwen 3.8 placing, and the chinese models beat the incumbent (Gemini 2.5 Flash) on cost whilst keeping like 95% of the accuracy.

I haven't used deepseek for anything else but the above results make me question its overall capability. Meanwhile qwen3.8 has continued to impress.

pdntspa··on Microsoft Abandons Personal AI Chatbot Race with Copilot Reboot
Again?
pdntspa··on Claude Status – Elevated errors for multiple models
Agreed, I believe strongly that human-in-the-loop is the way to go. LLMs are fantastic thought partners but ask it to critique something and it will go absolutely nuts. For Claude Code's /code-review command the highest I will set it is 'medium', otherwise it will produce so much feedback that you'd never get anything done.
pdntspa··on Frontier AI on Your Own Hardware
Anecdotally, the volume of recruiter spam for senior-ish positions I've been getting through email has picked up. Still nowhere near pre-COVID levels.
pdntspa··on Claude Status – Elevated errors for multiple models
I recently had Astra review a fairly detailed design doc I have for an audio VST fork, that I originally wrote with Opus and/or Fable a few months ago. The doc reaches deep into signal flow and module topology while lifting most of the DSP code from other open-source projects. I had it review for feasibility and architectural soundness.

Astra found a number of flaws that would have come up during implementation and we worked through them. But then I had Fable 5.1 review that document and it found a number of issues with Astra's changes, the least of which had was that Astra duplicated a lot of technical notions that it added rather than using references to an authoritative section. It also flagged some of Astra's designs as technically impossible, pointing out why and I'm actually in the process of digesting its feedback and updating the design spec. (I hand-review each point and we work through a solution together -- I don't trust either model to come up with something that follows my vision on their own)

I'm not promoting one or the other, I just found it interesting how this sort of adversarial review found pretty significant flaws in the other model's work. I am curious as to whether this process will eventually converge on a document that both agree on or if the models are going to perpetually nitpick each other.

I haven't actually started implementation yet, so maybe one or the other is full of shit. Just trying to come up with an architecturally sound design for something I want to write, when I lack the DSP knowledge to be able to write it myself. But the intent is to pass an agent the design doc and list of milestones and let it handle implementation.

pdntspa··on GPT-5.6 Luna vs. GPT-6 Astra: Is a $1.20 Model Good Enough for Code Review?
With the way memory systems work, I can see the value in having a different person's AI conduct the review as that AI's 'memory' is going to have a slightly different perspective aligned with the developer piloting it
pdntspa··on iOS 27, iPadOS 27, and macOS 27
Windows 10 start menu is the same way. How is it these things are so bad??
pdntspa··on YuE2 · Frontier Music with Symbolic Planning
Is there some way to get this to output just the vocals? Without running a stem splitter?

It would be useful to be able to generate backing vocals/hums/melismas and whatnot, that could be plugged into a DAW.

pdntspa··on YuE2 · Frontier Music with Symbolic Planning
Have you observed the sheer monotony of this avalanche of AI-generated "creative" expression at all?

You have people pumping out an entire album in a few days simply because they can, repeatedly, and then they flood streaming services with their garbage. I forget which network it was but as of a few months ago, AI-generated submissions were at least half of all submissions.

pdntspa··on DeepSeek v4.1 Flash
If computers degraded the way flesh does once you stop fueling it I feel like that would defeat a lot of your argument. And yes the brain has inherent statefulness (you're referring to memories, I'm guessing?), we have also jerry-rigged some degree of statefulness into LLMs. Mechanically it is very different and inferior, and there is a notion of separation that probably doesn't map to brains, but I would argue that LLMs, when you look at how inference is used in situ, are not necessarily stateless.
pdntspa··on DeepSeek v4.1 Flash
At first I was inclined to agree with you, but then I realized that the brain requires this whole complicated contraption (the body) to run and, really, do anything at all. And while I'm not familiar with the notion of 'quantum mind' I do think that biological processes aren't deterministic (at a cellular level).

And I think this does mirror the situation with LLMs -- you need this whole computer contraption and GPU, also running on electricity, to support the LLM's "thought" processes. And that if we model the brain's neurology sufficiently (which it seems we've done) we can achieve results that appear to be like thinking, even if it is an emergent behavior from "relatively" simple math.

Which actually makes me wonder the opposite -- are we, as humans, not much better than these LLMs? Suppose the body is just that super complicated computer contraption, honed by thousands/millions of years of evolution to achieve some semblance of homeostasis? If you reject the idea that we have a soul, we start to look very similar to the machines we build. "You are a brain inside a skull cockpit, piloting a bone mech covered in meat armor and skin" feels more and more relevant. That I'm just a meat circuit running brain chips and once you pull the plug on the source of electricity it all just... stops

pdntspa··on No Man's Sky Cosmos
Idunno man I thought Far Cry 5 was brilliant. Sure there was a fair amount of copy pasting but it was such a blast that I didn't care and the game was over before I was ready for it. The whole psychedelic religious cult critique was a really interesting premise and I really liked the main antagonist and his sister. I just wish they had some way of concluding the story that allowed you to continue to get into all the ridiculous chaos it offers. Once everything is settled its just too peaceful.
pdntspa··on No Man's Sky Cosmos
> what else do people want?

A little bit of depth to go with all that breadth?

I'm not saying there aren't deep systems in NMS. Its just that they are sparse. You can tool around all day having only the most cursory interactions with most of the game's systems. They've added so many interesting things, yet last time I played it felt there wasn't much of anything to do with them. For example, the Pokemon clone they added a few updates ago -- fun for five minutes. Super simple battles. However many months of development for something that got old after 5 minutes.

pdntspa··on Claude Fable 5.1 and Claude Mythos 5.1
It learned from the best, no? There is a lot of sarcasm and implication on the internet.
pdntspa··on Claude Fable 5.1 and Claude Mythos 5.1
You should try setting claude code to opus 4.6. With the style instructions I set in my user CLAUDE.md it does exactly that. It's like night and day: Opus 5 gave me a page and a half of word-vomit, yet the exact same task and prompt with 4.6 and I got maybe 100-150 words total, entirely readable.
pdntspa··on Playa Phone
They're the fringes, along with the increasingly large contingent of tourists/festies/sparkle ponies. The majority are actual burners and they are a wild and unorthodox bunch. Look up the 10 Principles if you want to understand the actual culture

https://burningman.org/about-us/10-principles/ -- "Decommodification" is the most important one to keep in mind, particularly with this audience, as I imagine some here would have trouble not trying to productify/commodify everything

pdntspa··on StemDeck, a free, open-source and local AI stem separator
My only interest is if there are any newer, cleaner algorithms/models. htdemucs (as well as everything else in UVR) are good but not great and leave a lot of artifacts. The acapellas it stamps are 'good enough' til you get to final mixdown and then the flaws stand out real heavy.
pdntspa··on StemDeck, a free, open-source and local AI stem separator
Damn, was hoping this would be an upgrade over (the seemingly abandoned) UVR5
pdntspa··on GLM-5.3 is now open-weight
I was running one of the older llamas (3.1 I think?) at slow-ish (10-20 tok/sec at Q4?) but OK speeds on 12 year old DDR3 ECC Xeon machine
pdntspa··on Show HN: OpenTIE and OpenXWA, Modern Ports of Tie Fighter and X-Wing Alliance
Cool! I played the original with a joystick back in the day. I tried the previous port to XWA with Tie Fighter more recently, and it had been so long that I couldnt even complete the tutorial with the joystick! Proper gamepad support would be clutch
pdntspa··on MS Paint and Photos inivisibly watermark even locally generated output with GUID
You can incentivize cooperation without having to compel action with a warrant. This sort of corrupt quid pro quo is quite common, in other countries at least...
pdntspa··on DFlash 2: Keep Drafting Parallel
I am working this out with Fable right now, for getting this running on my DGX spark homelab; it mentioned that there might be issues with the 'optimized LM-head restrictions' that unsloth NVFP4 ships with. Have you had any issues here?

Are you trying this with vLLM? Or a different engine? I am getting about 15 tok/s on my spark on my current setup using the 0.26 nvidia vLLM image and MTP.

pdntspa··on Beware Management Consultants
Dr Bronners is local to me and they are very active in the burner scene (I believe the founder is an OG burner). I have a few friends who work there and Bronners often has a presence at local burner events (and Burning Man) -- notable because these are environments where commercial expression is generally forbidden. Yet they sometimes get a pass, I think because they embody the burner spirit and often supply a needed service (showers) totally free and in a way that fits in with the culture. It is the most guerilla marketing you ever saw and I would not be suprised if that wasn't their intent -- rather its just the company gifting itself and embodying the spirit of the ten principles.

They remind me a lot of Valve in that they are able to chart their own path and find success.

Bronners certainly isn't without its problems as a company. But I find it to be very inspirational that companies can find success without having to devolve into the bland milquetoast mediocrity that feels like 95% of US business culture. We need more people being visibly, authentically, and unashamedly weird.

pdntspa··on Oxide Computer raises $445M (SEC Form D)
Oh man I cannot tell you how many times I've gotten into arguments on here and other technical forums years ago about self-hosting and how 'stupid' I was to not want to rent cloud servers
pdntspa··on Why DNA damage from smoking and UV rays cause cancer in some but not others
The actual infection rates for some STIs are also similarly low, per-encounter. But they are high enough that if you are promiscuous and have a large number of at-risk encounters then you will eventually get a bad roll of the dice.
pdntspa··on 98.css
XP? What with its Fischer-Price primary color scheme? You have got to be kidding me.
pdntspa··on 98.css
Finally, a throwback lib that seems to be mostly pixel-accurate! Some of the element spacing seems to be off but it does actually look like Windows 98.

The lack of accuracy and attention to detail in a lot of these efforts is incredibly irksome and annoying, just about as obnoxious as pixel-art games with inconsistent pixel sizes (like how hard is it to render to a smaller buffer and do a simple linear upscale?!?)

pdntspa··on Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
I'm running price-sensitive data extraction workloads on flash 2.5 and its still the king when it comes to accuracy + cost, all the gemini 3 variants perform a bit worse and cost a lot more. Low-key freaking out, ngl
Page 1 of 34Next →