HNHacker News
TopNewBestAskShowJobs

creamyhorror

2,886 karma · joined March 5, 2012

Fintech leader. Node/Laravel/Rails, Typescript/Javascript/React, React Native/Capacitor for mobile, C#, modular monoliths, and a penchant for problem-solving and applying math for fun and profit.
submissionscomments
creamyhorror··on I vibed a proof of Conway's conjecture
Absolute bad-metaphor-laden slop, in a dramatic writerly voice. These LLMs are trained on too much pretentious writing.
creamyhorror··on New stealth model: Union Alpha
Yeah, people are guessing it's Sakana AI's Fugu or a similar approach.
creamyhorror··on Ask HN: What model is new stealth Union Alpha model?
Quite possibly Sakana AI's Fugu (i.e. an orchestrator-router), since it seems to produce results of extremely varying quality, and doesn't seem to be politically censored.
creamyhorror··on Open-sourced jev architecture last year with model,paper and dataset
The world's heavily about marketing, resources, connections, and signaling, unfortunately. You probably needed to market it in a bigger forum with shinier claims to attract attention (I don't think a paper on Arxiv is enough).
creamyhorror··on A Selection of Los Alamos Rolodex Business Cards
If anything, I think General ___ is now a means of sounding established and technically advanced (implying "we're already covering everything in this area"). "General Assembly" got some visibility in the 2010s, and I've seen at least one recent YC startup use this naming pattern.
creamyhorror··on I used AWS cognito for a startup. I wouldn't do it again
Yep, if someone thinks AWS docs are rough, wait till they try Azure docs!
creamyhorror··on No leap second will be introduced at the end of December 2026
Master of Time. One of the Masters of the Universe.
creamyhorror··on GLM-5.2 is the new leading open weights model on Artificial Analysis
It's a real step forward, getting closer to SOTA. It seems to be very epistemically cautious in its reasoning. I hope Deepseek and the other open-weights labs stay in the game and catch up too.
creamyhorror··on I'm Eric Ries, author of "The Lean Startup" and new book "Incorruptible" – AMA
Touche. Honestly, if there's going to be speculative fever in society that you can't suppress, it should be captured for better purposes, such as through your LTSE. Bring access to it to Asia sometime.
creamyhorror··on I'm Eric Ries, author of "The Lean Startup" and new book "Incorruptible" – AMA
> that incredible sense of idealism

The idealism that has been sucked out of the tech industry. It was so (naively) hopeful at one point, and now the arms race and profit-maximization has eroded it all. Your observations really resonate with me.

I'm surprised I hadn't heard of the Long-Term Stock Exchange, it seems like a much healthier direction for the market.

creamyhorror··on VoidZero Is Joining Cloudflare
So long, and thanks for all the fish.
creamyhorror··on It is an amazing time for programmers
This is amazing, a true gem. I need to get it set up.
creamyhorror··on What is a Demand Coop
I'm not sure it's so black and white. Directing capital is powerful, and directing spending is powerful (but probably harder; this is marketing or government). I think it's more that directing spending requires influencing a lot more people than directing capital.
creamyhorror··on UnDUNE II
Tunes as captivating and evocative as the day I first heard them.
creamyhorror··on RSS feeds send me more traffic than Google
I'm building Subweb.net (not ready yet, it's just a few test feeds without the LLM pipeline turned on yet) to LLM-tag RSS feed items with topic, relevance/interest, location, and translations, and present them as feeds. I'm thinking I could maybe let users specify their preferred custom prompts and ranking params or similar, though the standard prompt is already fine.

I think the open web needs to come back, but in a fair way for everyone, giving readers control over their feeds while also sending traffic and comments back to the original sources. Not quite sure how to do that yet.

creamyhorror··on SubQ: a sub-quadratic LLM with 12M-token context
Whether this is real or not, multiple commenters here look like astroturfers - created in the past year (or hours) with very low karma
creamyhorror··on Where the goblins came from
"Seam" has been stretched by AI from its original legacy-code context to any point in code where something can be plugged in. I actually asked an AI about this a few weeks ago because I was surprised by the consistent, frequent use of "seam".

Frequent words I see from GPT: "shape", "seam", "lane", "gate" (especially as verb), "clean", "honest", "land", "wire", "handoff", "surface" (noun), "(un)bounded", "semantics" (but this one is fair enough), and sometimes "unlock"

It feels like AI really likes to pick the shortest ways to express ideas even if they aren't the most common, which I suppose would make sense if that's actually what's happening.

creamyhorror··on Where the goblins came from
Goblinmaxxing. Clean.
creamyhorror··on DeepSeek v4
No, the Deepseek V4 paper itself says that DS-V4-Pro-Max is close to Opus 4.5 in their staff evaluations, not better than 4.6:

> In our internal evaluation, DeepSeek-V4-Pro-Max outperforms Claude Sonnet 4.5 and approaches the level of Opus 4.5.

creamyhorror··on DeepSeek-V4 Technical Report [pdf]
Two key quotes:

• Reasoning: Through the expansion of reasoning tokens, DeepSeek-V4-Pro-Max demonstrates superior performance relative to GPT-5.2 and Gemini-3.0-Pro on standard reasoning benchmarks. Nevertheless, its performance falls marginally short of GPT-5.4 and Gemini3.1-Pro, suggesting a developmental trajectory that trails state-of-the-art frontier models by approximately 3 to 6 months. Furthermore, DeepSeek-V4-Flash-Max achieves comparable performance to GPT-5.2 and Gemini-3.0-Pro, establishing itself as a highly cost-effective architecture for complex reasoning tasks.

• Agent: On public benchmarks, DeepSeek-V4-Pro-Max is on par with leading open-source models, such as Kimi-K2.6 and GLM-5.1, but slightly worse than frontier closed models. In our internal evaluation, DeepSeek-V4-Pro-Max outperforms Claude Sonnet 4.5 and approaches the level of Opus 4.5.

While they're some months behind closed SOTA (though benchmarks put them close), I wonder if Deepseek 4's longer context capabilities and kv-cache advantage will make up for this

creamyhorror··on Website streamed live directly from a model
Are TYPE-MOON relationship diagrams the new pelican benchmark?
creamyhorror··on Migrating from DigitalOcean to Hetzner
The accounts are worth something later (e.g. for spreading opinions or promoting something) and can be sold.
creamyhorror··on Darkbloom – Private inference on idle Macs
oh boooy, it's a benchmarking script, but still...
creamyhorror··on The Gombe Chimpanzee War
Reposting this for comparison in light of the Ugandan chimpanzee war. Another multi-year war between members who were originally part of the same tribe.
creamyhorror··on ChatGPT Pro now starts at $100/month
r/codex is reporting that $20 (Plus) seems to have had its usage limit reduced (some people are saying it feels like 1/3 the previous limit now). The theory[1] is that reducing $20's limit lets them claim $200 has 20x $20's limit (and $100 has 10x).

If that's true, then the value comparison is not so positive for Codex any more

[1] https://old.reddit.com/r/codex/comments/1sgxy71/so_did_they_...

creamyhorror··on New laws to make it easier to cancel subscriptions and get refunds
Nope, unlike in the US, there's no easy way to create virtual credit cards freely in Singapore (afaik). Might be a result of Singapore law, monopoly power of the banks, or just a lack of awareness that such a thing is possible.
creamyhorror··on New laws to make it easier to cancel subscriptions and get refunds
It seems to me like it ought to be possible for the consumer to cancel a payment arrangement via their card provider.

Yet my banking app (here in Singapore) doesn't let me block any prior authorizations. It feels like the payment networks don't want to make it too easy to cancel periodic payments? Which isn't surprising, of course, but it feels like something I'd change banks for.

creamyhorror··on Universal Claude.md – cut Claude output tokens
I already do this manually each time I finish some work/investigation (I literally just say

"write a summary handoff md in ./planning for a fresh convo"

and it's generally good enough), but maybe a skill like you've done would save some typing, hmm

My ./planning directory is getting pretty big, though!

creamyhorror··on Universal Claude.md – cut Claude output tokens
I've started saying "gate" and "bound(ed)" and "handoff" a lot (and even "seam" and "key off" sometimes) since Codex keeps using the terms. They're useful, no doubt, but AI definitely seems to prefer using them.
creamyhorror··on What young workers are doing to AI-proof themselves
The end of ZIRP (cheap money) is precisely what ended the new-ventures/new-projects drive among big companies and turned them all to cost-cutting and maintenance mode.
Page 1 of 23Next →