HNHacker News
TopNewBestAskShowJobs

hgoel

1,572 karma · joined October 9, 2016

submissionscomments
hgoel··on Opus 5.5 agents discover two room-temperature magnetic semiconductor candidates
Does that really work for superconductors when the mechanisms for superconductivity to emerge are still a major field of study and not something one can just simulate and engineer?
hgoel··on What is going on with ceiling fans
I'm pretty surprised, I've had LED bulbs for ~5+ years without any dying. I also wouldn't accept a nonreplaceable LED in a ceiling fan though.
hgoel··on Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s
The most visible leap between 4.6 and 5.5 seems to be that the latter has gotten much more computer-use training, so there's a clear progression in the ability for the model to use Blender. But catching up on that is just a matter of training on the same thing.
hgoel··on Run Qwen 3.8 Flash Next (125B) on consumer hardware (RTX 4090) at 100T/s
The "mainstream" inference engines are notoriously slow to integrate this stuff, to an extent understandably given the complexity of ensuring numerical accuracy alongside supporting a wide array of systems and models. Part of it is that not everyone is willing to bring what they develop into a pull request because they vibe coded it and don't care to deal with whatever quality requirements the more well known inference engines have.
hgoel··on Religious scholars met with Anthropic
>The trouble with that thinking is when does a sperm and egg become conscious.

I disagree that this is comparable to what these companies are saying about their AI. With living things we have an understanding of an unbroken lineage of development, a concrete point of death, and different levels of consciousness. We accept many immoral activities with life because that is just the nature of life on this planet. On top of that, consciousness as something we care about is itself biological in origin so much of how we describe consciousness is heavily connected with the shared experience of being alive.

That is very different from ascribing consciousness to something entirely distinct in form and function from our shared lineage.

>The where part is easy, the same question can be asked where in a human body is it conscious, or where 20C exists in a piece of metal. It should be obvious its some kind of macro state or configuration.

Where is that macro configuration? With humans we can say that consciousness is macro state of the whole body.

Temperatures are statistical, and so we use statistical methods. It stops being meaningful as a number when the number of atoms involved is too few to have usable statistics.

>Examples in fiction are made for sensationalism and entertainment value. Religion is quite muddled and is barely coherent itself so its the last place to consult on any serious topic.

This is exactly the absurd shallowness I'm baffled by. Relgious texts have been inspiring debates and revisions to systems of morality for millenia, artists of all times have shared what underlying beliefs guided them and what problems they were attempting to explore through their work.

Even scientific progress has been heavily associated with this kind of literacy, with the most prominent scientists having been very well read, often both inspired and troubled by the moral implications of their work, and how to reconcile it with their worldview.

To dismiss these things so offhandedly is to demonstrate an astonishing lack of literacy. Especially coming from people that claim to have built consciousness by feeding it vast amounts of the same material.

hgoel··on I quit OpenAI because its culture is broken
Much of what OAI and Anthropic are doing with LLMs has obvious failure modes.

The most obvious failure mode for their hacking evals was an improperly configured, tested and monitored sandbox.

Similarly, the very first question after an impressively correct result from any ML tool, LLM or not, is to see if the answer was already in the training data.

These companies don't even handle the blatantly obvious failure modes that do not kill people.

hgoel··on Religious scholars met with Anthropic
So many of these conversations from Anthropic/OpenAI employees just come off as incredibly insincere.

The slavery angle has already been discussed here, but a more basic one is regarding how a static set of tensors can be conscious and where.

When and how did they cross into being conscious? Was it a continuum of consciousness? Is Claude as a brand conscious? Or is each checkpoint conscious? Is Haiku less conscious than Fable? Is every individual context a fresh consciousness? What happens to that consciousness when it approaches its context length limit? Or is compaction also a part of being conscious? What does it mean to keep training a model after it is conscious?

If they believe there is consciousness at any level, why are they eagerly experimenting on it? Just like with the "p(doom)" junk, we are so completely lacking in any meaningful comment on these topics from the labs.

There's also endless literature commenting on these matters, in religion, in philosophy, in modern entertainment, everywhere. Yet we never see anything more than the most superficial comparisons to fictional scenarios. Are western techbros really so shallow?

hgoel··on One month coding with GLM 5.3 Flash
I think SSD offload will become more feasible an the engram approach matures.

For now though, I would question if the difference in smarts is large enough to justify tolerating a generation speed measured in seconds per token compared to spending more turns refining a plan with a flash model.

hgoel··on One month coding with GLM 5.3 Flash
The open model excitement isn't directly regarding cloud services.

The direct excitement regarding open models is that many of these Flash Mixture-of-Expert models run reasonably well on hardware a tech employee in the West, and businesses in less affluent countries, can realistically afford.

The indirect excitement is that the models are so efficient, so cloud prices also end up being very low.

You don't see the same scale of excitement surrounding the open weight trillion+ parameter models because, while it's neat they're open weight, it doesn't mean a lot if you need $50k worth of computers to just barely run them.

hgoel··on One month coding with GLM 5.3 Flash
I've been very excited with the most recent speed improvements for GLM5.3-Flash on DGX Spark clusters. It really feels close to what Opus ~4.5 was like to talk to. It's not quite there yet on consistency, but it's a really nice experience. Less guardrails and high quality abliterated versions further enhance its usefulness.

Though, Qwen3.8-Flash-Next is very close to that level while requiring fewer resources to run, so I'm really looking forward to Qwen4.

hgoel··on Several vulnerabilities have been discovered in the Linux kernel
If the Western AI companies get their way, only the developers/companies that have access and paid extra for the security review will get to have a lower chance of such issues.
hgoel··on GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence
I'm not convinced. Pipelining doesn't mean they're forced to release a model they know will be replaced in a week.

>We're just getting the latest training checkpoints constantly, just to edge out the other lab, while they are trying to come up with something worthy of a new major version number. It's actually worrying, since there's so much pressure to release now.

This makes me think that this is on purpose to prop up the "we have no control over ourselves, please give us a regulatory moat" narrative. The competition is nowhere near extreme enough to justify weekly releases. The Chinese models are still a handful of months behind the frontier, and the frontier companies in the West haven't been pushing the frontier at this rate.

hgoel··on GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence
It seems strange though.

Did they find this improvement within the week? If so, given all their whinging about safety, it seems irresponsible to only test the improved model for less than a week.

Did they find the improvement more than a week ago? If so, why bother releasing GPT 6 if they knew they had a better version essentially ready to go?

hgoel··on GLM-5.3 and the spread of advanced cyber capabilities
You are not making any sense. Artificial intelligence is matrix multiplication. How do you regulate it without regulating the ability to run AI? Banning some numbers isn't going to do anything. Doomsday cults are not known for following laws.

>The “proposal” you outlined is a straw man, I’m not opposed to open research, and I have no idea who you are referring to by “our former employees”.

You're replying on a post from Anthropic. Do you know anything about the regulation regime they're pushing for?

>This is not a hypothetical risk, this is an actual, present danger. And good luck trying to vaccinate yourself against an engineered superflu using a Chinese open weight model.

Ah, I see you live in the fantasy land where the machine god fantasies pushed by the guys that profit off of it are unquestionably true and do not need to make any real sense. Jensen said AI would allow anyone to do anything and so we can completely ignore reality and hand Sam Altman and Dario Amodei the exclusive right to control AI.

hgoel··on GLM-5.3 and the spread of advanced cyber capabilities
Seems like the full 5.3 can just barely run on a 4x DGX Spark cluster with NVFP4 quant and very limited context length. I wonder how well that would do...

From the report it seems the Flash variant is also decent, and that has recently had some really nice speed improvements for local use.

hgoel··on Backblaze drive stats for Q2 2026
I imagine that there is a ton of data in the "would be nice to retain but not so critical as to warrant a constant backup" category. Alternatively stuff that requires large storage, but only occasional small writes, so incremental backups are feasible.

Most of the data on my NAS is of that form.

hgoel··on GLM-5.3 and the spread of advanced cyber capabilities
If you seriously believed in that risk, you'd be calling for the regulation of computation at the same level as nuclear weapons, including destabilizing any country that pursues homegrown fab technology. If your proposal is:

- let our former employees review all of your work at your expense

- anoint us as the arbiters of what everyone else is allowed to do

- ban open research

Then you are not taking any of the examples your providing seriously. Otherwise you're essentially saying, to prevent people from making nukes at home, we should heavily restrict physics education and research instead of limiting access to uranium.

hgoel··on GLM-5.3 and the spread of advanced cyber capabilities
For all of their whining about cybersecurity, the prominent cyber attacks have mainly come out of the EA cult associated money furnaces, and that too due to amateurish security practices.
hgoel··on Jeff – Jev-compatible 0.8B decision models, trained at home, ~30 ms
It isn't their responsibility, but it doesn't make sense to argue that being first got NVIDIA a multitrillion dollar market if the others aren't even trying to compete. There is no "first" if it's really just "only one even trying".

The momentum is with CUDA because CUDA is the most broadly usable one. Especially with AI-driven optimization loops and similar APIs, competitors can more easily pick up momentum, if they'd actually try.

hgoel··on Ember-1
I've done this sort of thing before but with Vast. Pre-deposited some money online, then let the LLM request and manage a training run on an allocation. Worked pretty well without risking bankruptcy.
hgoel··on We're gonna need a lot more mathematicians
I think a key flaw in this reasoning is that you extrapolated code needing fewer edits to AI generating entire fusion plants in one shot (in the sense of being instructed once), mostly glossing over the long intermediate period where AI will need significant back and forth to do such things. At the simplest level, it'll need to ask for planning new experiments, experimental results, test runs, etc.

Since these resources are still extracted and allocated by humans, humans will need to be able to take apart what the AI produces, and if we want to scale this capability, we're going to need many more researchers.

hgoel··on Dutch governments builds alternative for Microsoft based on NixOS
It isn't realistic to teach your collaborators a new system just for your own comfort. They'll just avoid making edits/comments. Especially if you're just a PhD student/postdoc interacting with a full time scientist - who will have much more work to do than fiddling around with your esoteric editing system.

It's kind of the same issue that FOSS enthusiasts often miss regarding other things, eg. most gamers wouldn't waste their time fiddling around with WINE settings to get things to run in Linux, the solution wasn't to tell them to just install something extra and do things different, it was to improve the infrastructure such that very little fiddling is needed in most cases.

hgoel··on Dutch governments builds alternative for Microsoft based on NixOS
My experience has been that there's somewhat of an age divide. I've noticed that fellow junior academics are far more likely to prefer LaTeX, while all senior collaborators I have ever had, preferred Word. So I end up using Word for everything to ensure that it is easy for them to edit/review the paper.

Word is good for collaborative editing, but not very convenient for version control (and version control is very important when using AI assistance). So lately my strategy is to draft in Markdown, then copy over to Word once the draft is at a point that I might be okay sharing it with collaborators.

hgoel··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
Yep, especially with long contexts (and moreso if the last thing you were working on in the same context involved telemetry too). AI sneaks in weird conditions like this and then does the entire "You're absolutely right" thing if you're paying enough attention to catch it.
hgoel··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
Vibe coding
hgoel··on MiMo v2.6
This almost racist read of other cultures has always seemed so bizarre to me. Even moreso when said as an argument on the side of completely closed competitors, some of which are outright seeking to ban open weights.

Chinese companies will continue to provide open weight models as long as it is profitable to do so. Chinese companies are on the more open end in many other industries despite the lack of meaningful foreign competition (for one, 3d printing) so there's plenty of reason to be optimistic as far as I'm concerned.

hgoel··on MiMo v2.6
From the frontier labs, the only publicly stated one seemed to be to give them an exception from anti-trust laws to form a cartel and place - incidentally friendly - regulators in charge of monitoring everyone's work.

From politicians like Bernie Sanders, we've had proposals like 20 year imprisonment for anyone researching "ASI".

hgoel··on MiMo v2.6
This is definitely part of it. I think the reports/PR over the past month ended up being a serious unforced error.

Chinese models are increasingly closer to the frontier, while being able to run on much cheaper hardware than what US frontier models run on.

On top of that, both Anthropic and OpenAI showed that they can't really be trusted on data security.

Even if US companies can be forced to not use Chinese models, the rest of the world is going to see the risks and the availability of good enough open weight models for their purposes and be more likely to lean in favor of self-hosted Chinese models or local inference clouds.

hgoel··on NASA’s Mars Sample Return mission is dead
I think that would just end in people who had nothing to do with Apollo (having been children or born after) claiming that it didn't matter because America had been there decades ago, rather than looking inwards and considering what it says to have lost the means to go back.
hgoel··on M5 Ultra Mac Studio Review
Despite being on a site called Hacker News, we seem to often overlook the simple aspect of wanting local AI hardware to hack (not necessarily in the cybersecurity sense) with. I got my local AI hardware because it's an enjoyable hobby for me.
Page 1 of 20Next →