HNHacker News
TopNewBestAskShowJobs

ACCount36

380 karma · joined March 5, 2025

submissionscomments
ACCount36··on Experimental surgery performed by AI-driven surgical robot
[flagged]
ACCount36··on Experimental surgery performed by AI-driven surgical robot
"If?" This thing has a goddamn LLM at its core.

That's true for most advanced robotics projects those days. Every time you see an advanced robot designed to perform complex real world tasks, you bet your ass there's an LLM in it, used for high level decision-making.

ACCount36··on Quantitative AI progress needs accurate and transparent evaluation
And I am tired of "mentioning the ethical and social issues".

If the best you can do is bring up this garbage, then you have nothing of value to say.

ACCount36··on Quantitative AI progress needs accurate and transparent evaluation
Your options for evaluating AI performance are: benchmarks or vibes.

Benchmarks are a really good option to have.

ACCount36··on Quantitative AI progress needs accurate and transparent evaluation
No wonder. Bluesky is where insane Twitter people go when they get too insane for Twitter.
ACCount36··on Quantitative AI progress needs accurate and transparent evaluation
People tried VLMs on "closed set" GeoGuessr-type tasks - i.e. non-Street View photos in similar style, not published anywhere.

They still kicked ass.

It seems like those AIs just have an awful lot of location familiarity. They've seen enough tagged photos to be able to pick up on the patterns, and generalize that to kicking ass at GeoGuessr.

ACCount36··on Against the censorship of adult content by payment processors
No. That's an often-repeated bullshit excuse.

Payment processors have ways of passing some of the chargeback risks onto the stores, and it's not like Steam itself is chargeback central. If you just want free games, pirating them is extremely easy, and trying to abuse chargebacks gets you banned.

ACCount36··on Covers as a way of learning music and code
The same applies to learning languages.

You can learn a lot from textbooks... or you can use textbooks to give you the absolute bare minimum, and then just use the language itself repeatedly.

ACCount36··on CARA – High precision robot dog using rope
The "DIY Surgical Robot" vid is so good. And it's also posted over a year ago with no follow-up. Damn.
ACCount36··on Stop Building AI Tools Backwards
Not really a suggestion, but OpenAI has dropped some major hints that they're working on "AIs collaborating with more AIs" systems.

That might have been what they tested at IMO.

ACCount36··on AI groups spend to replace low-cost 'data labellers' with high-paid experts
If it was just "noisy", you could compensate with scale. It's worse than that.

"Human preference" is incredibly fucking entangled, and we have no way to disentangle it and get rid of all the unwanted confounders. A lot of the recent "extreme LLM sycophancy" cases is downstream from that.

ACCount36··on AI coding agents are removing programming language barriers
It's not too obscure. It's also about the point where some coding LLMs get weak.

Zig changes a lot. So LLMs reference outdated data, or no data at all, and resort to making a lot of 50% confidence guesses.

ACCount36··on Mistral reports on the environmental impact of LLMs
The press focus is a mix of the usual "new thing BAD", and the much more insidious PR work by fossil fuel megacorps.

Fossil fuel companies are damn good at PR, and they know well that they simply can't make themselves look good. The next best thing? Make someone else look worse.

If an Average Joe hears "a company that hurts the environment" and thinks OpenAI and not British Petroleum, that's a PR win.

ACCount36··on Subliminal learning: Models transmit behaviors via hidden signals in data
In this study, it required a substantial similarity between the two models.

I don't think it's easy to get that level of similarity between two humans. Twins? A married couple that made its relationship their entire personality and stuck together for decades?

ACCount36··on Subliminal learning: Models transmit behaviors via hidden signals in data
1. You train a model to exhibit a certain behavior

2. You use it to make synthetic data, data that's completely unrelated to that behavior, and then fine tune a second model on that data

3. The second model begins to exhibit the same behavior as the first one

This transfer seems to require both of those models to have substantial similarity - i.e. to be based on the same exact base model.

ACCount36··on What will become of the CIA?
What kind of tin foil hat conspiracy land did I stumble into?
ACCount36··on Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic
It's the way it should be.
ACCount36··on AI comes up with bizarre physics experiments, but they work
In practice, we'll just let that AI have a direct internet connection, and also give it enough access to push code straight to prod. For the good measure.
ACCount36··on What will become of the CIA?
The last major intelligence coup CIA had (that we know of) was when the agency called the Russia's invasion of Ukraine months in advance.

Going public with that was a bold call - CIA put its reputation on the line. But Ukraine was more prepared because of it - and so were its allies.

A lot of Ukrainian officials didn't believe that the war was about to start up until the moment it did. Imagine how much worse the situation could have been without US beating the drum.

ACCount36··on Agents built from alloys
Task length is increasing over time - and many AI labs are working on pushing it out further. Which necessitates better attention, better context management skills, better decomposition and compartmentalization and more.
ACCount36··on How Tesla is proving doubters right on why its robotaxi service cannot scale
What? An FSD Tesla has its very own "world model". It doesn't try to reconstruct a world "photons in", from scratch, 60 times per second. It continuously updates and refines the data it already has based on the sensor inputs, and then uses this internal representation to make driving decisions.

This "world model" is what you get to peek into through the car's screen. By now, it even has basic "object permanence". Nowhere near as good as a human yet. But AI is getting better, and an average driver isn't.

ACCount36··on How Tesla is proving doubters right on why its robotaxi service cannot scale
Your "cameras" have about 20 MP of active resolution between the two of them. With dead zones and pixels spread out unevenly. A modern smartphone has you beat.

There's a small, sharp, high resolution color-enabled area in each eye - but the bulk of your vision field is monochrome, and mostly sensitive to motion.

You don't notice that, because your image data is stacked and post-processed to shit to make it presentable. Your brain has been doing computational photography before it was cool - 90% of what you see at any moment in time is effectively AI-generated.

ACCount36··on How Tesla is proving doubters right on why its robotaxi service cannot scale
Statistically, SOTA self-driving cars are already superhuman. That holds for Waymo and Tesla both. They crash less, like for like, and the incidents they get into are less severe. But that's not because self-driving cars outperform a "top of the line" human driver. It's because they outperform the absolute worst bottom of the barrel human driver.

A big part of a self-driving car's "safety edge" is that it isn't going to go 80 in a 40, doesn't fall asleep at the wheel, and isn't capable of DUI.

Self-driving cars still struggle in some situations most human drivers wouldn't find challenging - AI issues - but they don't make the worst, the most unforced and avoidable "human factor" mistakes.

ACCount36··on How Tesla is proving doubters right on why its robotaxi service cannot scale
No amount of LIDAR wankery can solve self-driving.

Take any self-driving car crash where the self-driving car was found at fault. Dump the blackbox, extract the raw sensor data. What will you see?

You'll see that the car had all the sensory data it needed to make the right call, many times over. And it didn't make the right call. That's not a "sensors" problem. The sensors are good enough. The main bottleneck for self-driving is, and always was, in AI.

Which is why you get things like that Cruise car dragging a pedestrian despite being equipped with 360 cameras and a total of 5 overlapping LIDARs. It had the sensors. What it didn't have was object permanence.

ACCount36··on How Tesla is proving doubters right on why its robotaxi service cannot scale
Bullshit. And I am tired of having to call people out on it.

Autopilot shuts down when it can't handle the situation it's in. This doesn't help it "avoid blame" at all. Because Tesla considers Autopilot implicated in any crash that happened within 5 seconds from Autopilot being disengaged.

> To ensure our statistics are conservative, we count any crash in which Autopilot was deactivated within 5 seconds before impact, and we count all crashes in which the incident alert indicated an airbag or other active restraint deployed.

NHSTA's reporting requirements are even more conservative:

> Level 2 ADAS: Entities named in the General Order must report a crash if Level 2 ADAS was in use at any time within 30 seconds of the crash and the crash involved a vulnerable road user being struck or resulted in a fatality, an air bag deployment, or any individual being transported to a hospital for medical treatment.

ACCount36··on How Tesla is proving doubters right on why its robotaxi service cannot scale
You clearly have a pet issue. Why do you think that it's in any way relevant to the conversation at hand though?

Do you seriously think that the main challenge Tesla is going to face when trying to scale Robotaxi up is that there isn't enough room on the roads for all the Teslas? In a world where there's currently a dozen Robotaxi Teslas per city?

ACCount36··on Local LLMs versus offline Wikipedia
That's a very safe assumption. There are more smartphones on Earth than there are humans.
ACCount36··on Death by AI
One use of AI tech is that it can enable megacorps to take and process actual fucking feedback, for once.
ACCount36··on Local LLMs versus offline Wikipedia
Currently, there are billions of devices that are capable of storing and running a 4B LLM locally. Hundreds of millions for 32B LLMs. It would take an awful lot of effort to destroy all of that.

If you're doomsday prepping, there's no reason not to have both. They're complimentary. Wikipedia is more reliable, but also much more narrow in its knowledge, and can't talk back. Just the "point someone who doesn't know what he's dealing with in a somewhat sensible direction" is an absolute killer feature that LLMs happen to have.

ACCount36··on Open-Source BCI Platform with Mobile SDK for Rapid Neurotech Prototyping
I'm incredibly skeptical of any non-invasive BCIs. EEG was around forever, and has completely failed to become useful as an interface system.

Neuralink N1 is a fully invasive BCI, with over 1000 recording channels that go down to neuron level. In practice, that's barely enough to provide a useful, reliable control interface. It's still SOTA - anything else that exists is straight up worse.

The pathway to better BCIs seems to be in more invasive interfaces with greater channel counts - not the other way around.

← PreviousPage 3 of 7Next →