HNHacker News
TopNewBestAskShowJobs

xg15

23,165 karma · joined February 12, 2014

submissionscomments
xg15··on The evolution of effective altruism
How about "conspiracy" then?
xg15··on I asked Claude build a physically accurate O'Neill cylinder you can walk around
The weirdest experience in this must be to never know if you're walking uphill or downhill.
xg15··on Tourismusflipper
Artist's website (and similar contraptions) here: https://www.morgan-art.ch/catalog/tourismusflipper/
xg15··on "Torturing" LLMs in a Robot Prison Has Triggered the Dumbest Debate in AI Yet
Sorry, but, yawn.

The sentence itself makes no sense to me. If it's an illusion, then what does "illusion" even mean? Who perceives the illusion?

I'm with you if you want to say that the idea of a metaphysical entity that is separate from the body and not part of physical reality (a "soul"?) might be wrong - i.e. that our conscious experience is fully caused by physical effects, by the interactions of neurons in the brain and all the other machinery there - and that it might not even be an indivisible whole but might be "composed" of different components.

But that doesn't make any practical difference. We're still experiencing the world, have internal thoughts, memories, feelings, etc. Those things exist. In what form they exist is an interesting scientific question, but that's something different from dismissing them completely.

xg15··on "Torturing" LLMs in a Robot Prison Has Triggered the Dumbest Debate in AI Yet
My personal "ethics framework" (if you could call it that) is currently:

I have zero qualms about aborting a chat or resetting it to an earlier state, because any potential kind of consciousness that could exist would IMO have to take place during the inference passes, so there is nothing that amounts to "killing" an LLM.

I do try to not be a dick in chats, avoid intentionally subjecting the LLMs to mindfucks etc. This happens to be the far more effective way of working with LLMs (in terms of tokens needed to complete a task) as well.

I do not use character.ai or "virtual friend" services or anything else that tries to anthropomorphize LLMs even more than the training already does anyway. (This is more out of concern for myself than for the LLMs)

I think the whole debate runs to conclusions a bit prematurely, before we even have a good model to understand what happens in an inference pass of a billion parameter ANN with 100 layers of attention modules switched in series.

More generally, I don't like the attitude of encouraging people to be assholes. Suddenly we are in a situation where some people are building detailed simulations of torture and the people who are upset about it are the idiots?

(I still have no love for the EA guys or the rest of the TESCREAL circus. They always had a cultish vibe coming off them, but in recent years, all the masks seem to have fallen. They also consistently make the most un-empathic statements and observations, all in the name of "empathy")

xg15··on Woking Electrical Control Room (2016)
Good point and thanks a lot for the information from that area!

I guess I mixed up different things here - I was thinking about "software dashboards" like Grafana or various Kubernetes frontends.

I didn't know this is still done for dashboards of industrial systems. I think then my annoyance is more that this has never been translated to other areas, like software systems.

xg15··on City building games have a Soul Problem pt.2
This is the irony of virtual worlds: The sloppiest virtual world looks like the platonic ideal of landscaping we have in the real world.
xg15··on Woking Electrical Control Room (2016)
It's great that so much about those old control rooms is posted here lately!

One thing I notice is how much efforts had been spent to illustrate both the current status of the system, but also it's overall structure and how the two relate - i.e. the actual components of the facility.

Almost all of the dashboards here have some kind of graph, showing the wiring or plumbing between components, or in case of railway control, the railway lines. The graph acts like a map on which status indicators or switches can be placed.

Compare that to modern software "dashboards" which are full of complicated charts showing various metrics, but almost never show the structure of the system.

xg15··on City building games have a Soul Problem pt.2
I wonder if the dominance of (realtime) 3D rendering we have today is having a bad influence here. I remember old tile-based games (like SimCity or Transport Tycoon), or games with fixed backgrounds feeling much more natural, simply because you can prerender (or simply paint) your desired level of detail without having it affect your runtime compute budget. Of course the cost is less flexibility at runtime, but that cost may be acceptable.

For a modern take on this, look at Factorio. Why can an industrial zone with a population of one look more lively in Factorio than an entire simulated city in Cities Skylines?

xg15··on Dear Software Makers
I've heard we have this amazing new technology now that can understand natural language and automatically convert it into structured data if we ask it to. Maybe we could use that to ask people at scale now.

> and those that do ask for things they don't actually want, or are only wanted by very few people.

Then how do you know what people actually want?

xg15··on Dear Software Makers
How about actually asking the users instead of experimenting on them?

> or do we want companies that get feedback from users on whether new changes are helping or hurting.

You don't get that feedback. The feedback you get is whether some telemetry KPI goes up or down. That's not the same as actual utility for the user.

xg15··on DraftKings is using AI to behaviorally target chronic gamblers
Not completely, in a scenario where the same companies would also finance UBI. But I guess "UBI costs" could be easier controlled than salaries.
xg15··on DraftKings is using AI to behaviorally target chronic gamblers
Which would shed all those grand visions from tech tycoons about a fully UBI-financed "pure consumer" economy in a different light: "We couldn't care less about jobs, but guinea pigs - we need lots of those!"
xg15··on Nvidia wants to put a watchdog chip next to every AI agent
What does this chip do what a harness with guardrails or running on an account with restricted permissions doesn't do?
xg15··on Being a Doctor Will Never Be the Same After A.I
Thanks for the gift link!

But wow, they even want an account to read unlocked stories now...

xg15··on When did Google get so weird?
I feel this is a result of the "make the world a better place" narrative that Silicon Valley companies used to adorn themselves with.

If you just want to play with cool tech and make money, but for marketing reasons, you need to show some ostensible real-world reason why this tech should be developed, then "well, our chatbot could help lonely people with their relationship issues or give people in an abusive relationship a covert channel for help" is probably what comes to mind...

Runners up: Our AI startup will help the poor children in Africa and/or combat climate change (ignore the gas turbines, please).

xg15··on We Should Be Able to Change Our Languages
Still learning Rust, but I found it unfortunate that basically the first command you're introduced to, println/print is a macro.

It's starting with the exception before you even had a chance to understand the rule.

xg15··on We Should Be Able to Change Our Languages
> Things have been changing drastically. There was once a point where it made sense to create suboptimal code if it meant that your team understood it better. Why? Because changes to the code provided more business value than the code itself being optimal, and if your team didn't understand it, no one could change things. But this is becoming less and less the case. If you can make code faster, better, even at the cost of readability and understandability, by you, but the AI can work with it perfectly fine, why wouldn't you?

I guess this will become the frontline of the upcoming civil war in programming-land...

xg15··on What even is an OS now?
I think Star Trek is interesting in that "AI" was explored so often by so many different writers that you have the entire spectrum of agency or sophistication levels by now - the ship's computer at one end, then the Holodeck characters, Data and eventually the Borg at the other end...
xg15··on What even is an OS now?
And yet, even on Star Trek we can see people typing things into screens or onto pads. The helmsman doesn't fly the ship by talking to the computer.

My point is that text-based interfaces - and voice is just a convenience around text - are not the best interfaces for a lot of tasks.

xg15··on What even is an OS now?
I've seen a few posts of the form "why would you ever want to use anything else than AI from now on?" by now. My counterpoints would be:

Unless you have a local model and the appropriate hardware for it, your agent is Somebody Else's Computer. Do you really want to send all your data and make your entire computing experience dependent on whatever OpenAI or Anthropic or whoever else is planning in this moment?

That's not even starting with inference time and token cost. Despite all the incredible advances in inference, it still takes more time than most non-AI computer functionality. Do you really want to wait a few minutes and pay money for something that you could also do with a few clicks fully locally on your PC?

But the most important thing: User interfaces. Right now, we're basically cramming everything you could possibly want to do at a computer into a chat interface. But there are lots of applications fir which specialized graphical interfaces are much more suitable. Why would you want to get rid of them?

Somewhat connected to that: Repetition. If you have to do the same task again every week or every day, it seems wasteful to ask the AI for it every time: Not just are you wasting a lot of time, energy and tokens, you're also at risk of getting inconsistent results, if the agent from today's session will interpret the requirements slightly differently than the agent from yesterday.

You can circumvent all those things by having the AI write you a custom app, but then using the app without AI to do the task.

xg15··on AI has no intent and no motivation
Was thinking the same. Air, water and fire have no intents and yet we get hurricanes, flash floods and wildfires.

Also, repeating a point from a similar thread: Software can have "intent" in the sense that it steers itself towards a predefined goal without having to be "alive" or "conscious" in any way. Some classic examples are thermostat control loops, navigation systems and chess engines.

xg15··on VSCode's SSH Agent Is Bananas (2025)
Also using it, and by now at least I see the reason why they did it. VSCode has a large plugin ecosystem, many which are essential for development. The problem is that those plugins don't know anything about remote development and expect to use the standard file system and OS APIs to interact with the workspace.

So how to make the plugins remote-capable? You could write a massive virtualization layer that captures all system calls and forwards them to the remote - or, you could run the plugin on the remote and just pass the user commands and UI updates over the connection.

VSCode does the latter, so the nodejs runtime is where all the plugins are running on the remote.

(I understood the reason, I didn't say it was a good reason...)

xg15··on VSCode's SSH Agent Is Bananas (2025)
That makes no sense to me though. The VSCode frontend already has an open SSH connection to the remote machine, over which it could do the same things and more. Why is the websocket connection (which is probably tunneled through the SSH connection anyway?) any worse here?

Edit: another comment clarified it really goes both ways: The remote can use it to run code on the local machine as well, exactly what the "sandbox" pattern was supposed to prevent.

xg15··on VSCode's SSH Agent Is Bananas (2025)
> The agent runs over port-forwarded SSH. It establishes a WebSockets connection back to your running VSCode front-end. The underlying protocol on that connection can:

    Wander around the filesystem
    Edit arbitrary files
    Launch its own shell PTY processes
    Persist itself

Wait, could someone clarify which machine is being referred to here?

So in the author's setup, he runs VSCode (i.e. the front-end) on his dev laptop, which he wants to keep free of direct LLM access.

VSCode connects via ssh to a dedicated "sandbox" machine on which the LLM will be free to do whatever it wants (mostly).

VSCode realizes this the Microsoft way, by using the ssh connection to install VSCode Server on the sandbox machine - the "backend" - and communicating through it via a websocket connection.

So then, what happens? If the websocket connection allows the front-end to run arbitrary commands on the sandbox machine, this wouldn't be very exciting: The front-end already has an ssh connection and a massive server process that can do the same - and the entire purpose of the sandbox machine is to run arbitrary, untrusted commands without harm.

But the article says the websocket connection goes "back to your running VSCode front-end". So does that mean things are reversed? I.e. the agent/harness runs in the server on the sandox machine but for some reason has this websocket connection that also lets it run arbitrary commands on the dev laptop?

Is that it? That would be truly insane!

xg15··on Markdown in /src
Well, at least in Git, the commit message already has two sections.

AFAIK, the first line is supposed to be the a very brief summary, while the other lines may contain additional information and can sometimes make up a pretty long text. Lots of tools make use of this convention and only show the first line if no detail information is needed.

Most tools that show the history only show the first line of each commit.

But I agree with you, this still assumes that long commit messages are rare and not that almost every commit has a huge message. Also, "long" doesn't mean you should put a novel in there.

xg15··on Jev in 25 Lines of Python
> In real life, a human doesn't do classification tasks with the System One part of their brain, they use System Two.

Huh? I guess that depends on the exact definition of "classification", but I think the bulk of basic classification tasks we make every day to make sense of our surroundings, such as object recognition is definitely done using system 1. So is higher-level "stereotyping" or anything you could described with "I know it when I see it".

Because those responses can be incorrect or even harmful, you would sometimes make use of system 2 to correct them - but that doesn't change that the initial response is from system 1.

xg15··on Jev in 25 Lines of Python
I missed the hypewave so can't say a lot about Jev, but the double standards are entertaining:

About Jev:

> We didn't train a model with Reinforcement Learning for Calibrated Decisions (RLCD) to calibrate the decisions and probabilities (even though they are not always correct).

Only 99% correctness! Borderline unusable!

About their model:

> It classifies: it gets a prompt with choices and outputs probabilities.

You want numbers, it gives you numbers! What more could you want?

xg15··on Markdown in /src
I think the question is still valid what you actually do with them once they are saved.
xg15··on Markdown in /src
I wonder if instead of checking the prompts into the repo as files, a better idea would be to store them inside the commit messages.

If prompts are specifications for a change of the system's behavior, then it seems natural to manage them as changes and not as resources.

This would also keep them in the right "historical context" of the repo and avoid the "prompt rot" you were talking about.

Page 1 of 34Next →