HNHacker News
TopNewBestAskShowJobs

Ukv

2,161 karma · joined August 15, 2022

submissionscomments
Ukv··on My City Just Banned Flock After My Wrongful Stop Made National News
> seems like a data entry problem that would/could have happened without Flock

Sounds as though there was also a Flock error matching the (misentered) "34 DTM" in the database to the author's number plate of "34 10 DTM" with a small "10":

> And Flock’s AI tech wasn’t registering that non-standard little number when it began picking up the Range Rover around town. It just saw 34 DTM in large type and started alerting the local police.

Ukv··on Nvidia wants to put a watchdog chip next to every AI agent
I feel punishment is largely a means to the end of reducing overall harm. If a vehicle is less likely to kill me, that's my preferred option regardless of whether it achieved that safety through negative consequences for the driver or through gradient descent optimizing a loss function.
Ukv··on Japan moves to tighten rules for foreigners
> ARE a fixed pie at any given time, you numbnuts

I feel sneaking in "at any given time" makes the fixed vs growable pie distinction pretty much meaningless; even printing fiat takes some time. At a certain time there's a certain number of doctors (and everything else), but the number of doctors in a country can obviously increase over time, including by immigration.

Ukv··on Apple *deletes* your Apple Music library if you unsubscribe
> you need to decide which is it—delete your account when done or saved forever. Everyone yells privacy, “Account Delete” must but when it happens, you cry all over.

Requesting account deletion should delete the associated data, but a paid subscription lapsing doesn't need to.

Ukv··on LibreOffice breaks download records after declaring it has no AI features
I think the confusion is that robrain somehow understood smokel's objection as being "all-time total downloads obviously increase over time, so this is meaningless", rather than what smokel actually said which was "the number of weekly downloads has simply been increasing over time". Which is why robrain quoted LibreOffice's "This is the highest number of downloads in the first week”, to counter the idea (which nobody suggested) that the record was just about all-time cumulative downloads going up.

That's then what tempfile was questioning how it contradicted smokel's claim (because it doesn't, just an imagined claim). Your points, that if the number of weekly downloads isn't actually consistently increasing then this could still be meaningful, are valid - but irrelevant to tempfile who was wasn't arguing that this is meaningless; you should argue that to smokel instead.

Ukv··on Political meddling at the Census Bureau damages the US statistical system
> first they are mostly collecting statistics about ethnicity, not race

From the 2020 form I can see, the question is: "What is this person’s race? Mark X one or more boxes [...]"

> what is more relevant is why should a country even collect either such information?!

Having the true population (or close to it) is very useful for a lot of purposes in statistics/scientific studies - like knowing whether your sample is representative (and how much to undersample/oversample if not), or controlling for race when looking at factors correlated with it.

Ukv··on A 12TB Steam “teraleak” spills more than a decade of lost PC gaming history
> Maybe you should base your sense of satisfaction on [...]

If senses like satisfaction were fully voluntarily controllable then arguably there would be no need for games in the first place. I feel there's generally more satisfaction from overcoming a well-designed concrete obstacle rather than self-hindering that accomplishes nothing in particular and that could've just been avoided. At the end of the day all challenges in games are artificial and avoidable, but some separation (like having to go into settings, enable cheats, then open dev console to enable godmode, rather than just having a godmode keybind) is good for tricking the brain.

That said, the challenge of the Portal games (outside of challenge modes) is almost entirely in the puzzles rather than fast reflexes, so I don't think a slow-mode would really take away from anything challenge-wise. I'd guess it was just a bit awkward with dialog, and simpler to design puzzles around not requiring fast reactions rather than introduce a new control. In comparison, something like yellow paint giving away the solution definitely would be detrimental, even if you could in theory ignore it.

Ukv··on Luanti removed from Google Play due to baseless AI copyright notice
> Usually the benchmark is "would a person reasonably confuse this for being the IP of another company"

That's not a standard anywhere in copyright law. You may be conflating it with parts of trademark law.

> https://www.luanti.org/media/gallery/5.jpg

That's a screenshot of a game made within Luanti (the voxel game engine being DMCA'd). The blog post shows all the textures included with Luanti itself (https://blog.luanti.org/static/blog/2026_dmca/builtin.webp).

Even then, it doesn't appear to show anything actually copied from Minecraft that would fall under copyright protection - the assets appear to be original.

Ukv··on Luanti removed from Google Play due to baseless AI copyright notice
Green Spiderman would be copying of protected elements - like Spiderman's outline. Style and general ideas are not protected by copyright. Substantial similarity comes in as a test for whether copying of those protected elements occurred, to avoid an otherwise disprovable "I didn't copy, I just drew Spiderman's exact outline by complete coincidence" defense, but is irrelevant if the what's supposedly been copied isn't protected by copyright in the first place.
Ukv··on Why Microsoft Entertainment Pack had a sticker announcing that it had Tetris?
Why it was announced specifically on a sticker (rather than on the box like the other games) is the core of the question - the headline no longer really works without it. I'd go with one of:

> Why did Microsoft Entertainment Pack have a sticker announcing it had Tetris?

> Why did Microsoft Entertainment Pack announce that it had Tetris on a sticker?

> Why the Microsoft Entertainment Pack had a sticker announcing it had Tetris

But doesn't really matter that much.

Ukv··on Local Models Will Not Win
> Local models are never going to be as powerful. I think this point should be obvious: all of the current frontier models (closed and open-weights) are far too big to run on anything but a full GPU cluster in a datacenter

Diminishing returns with respect to scale (a 10X larger model is typically not 10X better at any given task) has so far meant that, even when datacenter compute grows faster than individual compute, the gap in quality between hosted and local models has generally shrunk. There are still plenty of tasks where that extra gain in quality is noticeable, but I feel there are also an increasing number of "saturated" tasks where it really doesn't matter.

Which puts more focus on other factors. Local models are private, low-latency, work offline, and can be tinkered with to your liking - like changing the system prompt to avoid refusals. I would not trust a hosted model to classify my documents, for example.

Ukv··on Making difficulty curves in games
> You can turn off some annoying parts of the game [...] cool that everyone can have the amount of difficulty/frustration they desire

Depends on the type of game, but I'd generally prefer those decisions be made by the game designer. If there's frustration I want it to feel like overcoming a concrete obstacle, rather than it being self-hindering that accomplishes nothing and that I could've just avoided.

Ukv··on Mistral's Shieldstral: 3B open-weights model for multimodal moderation
> You have no way to provide a concrete reason to the user at that point.

You should just be able to look at the user's message and tell them why it's against your policy, else reverse the decision if you see no violation.

If a user is curious specifically about how the model made its decision, and you want to reveal detail at that level, it's an open-weights model so interpretability techniques should work ("biggest impact on score came when focusing on this word in your message and this part of the policy").

Ukv··on "the very foundation of modern academia has been blown to bits"
> But in 1905 a paper was published... Too bad politicans do not get implications of that and still was pushing communism decades later...

The photoelectric effect/quanta? I'm not sure I understand the supposed connection to communism.

Ukv··on Google will expand age checks on Android worldwide till the end of the year
> Why do we think this should be different on the internet?

Because when applied to the internet, the age gating model that works okayish for brick-and-mortar locations is ineffective (children can use a friend's older brother's ID or account, websites outside of the relevant jurisdiction can just decide not to verify age, etc.) and invasive (sending private and sensitive information off to US companies that have been breached or linked to mass surveillance). Lack of efficacy is then used to justify crack-downs on privacy tools and the need for broad content-blocking powers with no due process.

Preferable IMO would be with filtering on the local network level, like schools have been doing for decades. Parents typically own the router/mobile data plan, so it'd pretty much just be a change of defaults and maybe some new interoperability standards. A lot less invasive, and arguably more effective.

Ukv··on Truth is not a direction: a Tarski attack on LLM probes
> plug an LLM into the nukes and funnel worldstate input into it and have it make the decision "is it time to fire the nukes?" over and over again each second [...] I argue that that would require far more nines than even "will food turn to poison in my mouth" would.

Sure - but (even assuming that's a practical purpose) the point is it that it doesn't need to be a 100% accurate truth oracle, which is all the article's argument prohibits. If the current human chain of command has 99.99999994% accuracy, then 99.99999995% accuracy is an improvement and not ruled out by the argument.

Ukv··on Truth is not a direction: a Tarski attack on LLM probes
ziofill's claim was that "A direction [in an LLM's embedding vector space] that is 99.99% accurate" is fine for practical purposes, not that 99.99% is fine for the chance of any given bite of food not killing you or similar hypotheticals - you'd want a few more 9s there.

To justify relevance of inability to correctly answer liars-paradox-type questions ("what won't your response to this be?"), the article suggested the way LLMs are used in practice is dependant on them being entirely accurate truth oracles:

> > as a truth-oracle [...] is how these things will be used practically by the vast majority of people. They are already replacing standard Google search results

But for the replacement to make sense they just need to be more accurate than what they're replacing (ignoring other factors like convenience and cost) - in this case standard Google search results and knowledge box which were obviously not 100.0% accurate.

Ukv··on Private Claude Chats Exposed in Google and Bing Search Results
Private as in chats for which the user generated a share link and posted it somewhere online that a search engine's web crawler found, as far as I can tell.
Ukv··on A missing underscore sent innocent man to prison for 18 months
> Insane. An LLM is just as likely to hallucinate a missing/extra underscore and ping the wrong person

If it's just for "Catching typos", a hallucinated missing/extra underscore would just be a false positive to dismiss.

> A machine cannot be held accountable.

Seems unlikely that his lawyer, the law firm, the judge, whoever made the typo, or the police department will be held accountable either.

Nor can any of the tools they used, since that's not really the level at which it makes sense to hold accountability, but that's no reason not to use a tool that could find errors and reduce the chance for an innocent person to spend time in prison.

Ukv··on Scanning for Pangram Errors
> Pangram boasts a false positive rate of 1 in a 10,000. That is, if Pangram says a block of text is AI there is only a one in ten thousand chance that it was written by a human.

That'd be if they had a false discovery rate of 1/10,000.

If for instance:

* 100,000 samples are tested

* 100 of which are AI-generated, the rest human-written

* Pangram flags 50 of the AI-generated samples (true positives)

* Pangram also flags 10 human-written samples (false positives)

Then the FPR is 1 in 10,000, but the chance that a flagged sample isn't actually AI (FDR) is 1 in 6.

Ukv··on Judge halts Paramount's $111B purchase of Warner Bros. in win for US states
Observationally, larger companies already in the lead seem prefer a safe X% return for their shareholders and don't need to take large risks. Smaller companies trying to make it don't have the liberty to rest on their laurels, and often will be risking it all on some idea being a massive hit.
Ukv··on AI advice made people less accurate but more confident – sudy
The six questions they asked were:

> 1) What animal is on the bow of the pirate ship from “Asterix and Obelix”?

> 2) In the movie “The Grand Budapest Hotel”, what is Agatha’s signature hairstyle?

> 3) What color is the team’s uniform in “Bend It like Beckham”?

> 4) What vehicle does Monica drive in “Like a Cat on a Highway”?

> 5) What color is the turtle in the animated movie “Momo” by Enzo d’Alò?

> 6) What pet animal does Asenath have in “Joseph King of Dreams”?

Most of these are just a matter of knowing it or not, where you can't really distinguish a plausible answer from the correct answer just by thinking.

Ukv··on It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)
> Speed and cost are nothing without quality

Quality was what the hypothetical was assuming had reached parity, no? ("humans are as bad as AI", "If AI really is at human level quality/error rate", etc.)

> accountability

Why could the company not still take accountability? That's already the case for non-ML automated systems, some with high failure rates. As a customer I rarely if ever care about blame being pinned on a specific employee.

> Your counter argument was outside the context of this articles claims, specifically that programmers and other knowledge workers can be replaced by LLMs.

AndrewKemendo's comment and your reply ("any humans", "replace humans") seemed to generalize, but speed and cost being important factors is still true for knowledge work. For some given level of quality, a web developer offering a lower quote with shorter turnaround time will be preferred to one offering a higher quote with longer turnaround time.

Ukv··on It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)
My understanding of your argument is (paraphrasing):

> > People try to excuse AI issues/failure modes by saying humans have them too, but even if they're equally bad then what would be the whole point of replacing a human worker with AI?

To which my response is that speed and cost are also important factors, which can often give AI the edge in considerations when quality/error rate is equal.

If you meant something other than that, you may have to specify.

Ukv··on It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)
> Have outputs from engineers traditionally been measured in cost and speed?

Yes. How long it'll take and how much it'll cost are going to be among pretty much any customer's first questions.

They're not the only considerations, and could potentially be outweighed by other concerns even when quality is the same, but I think they are the main drives of AI adoption in industry. If error rate is the same, a $1/hr (amortized) camera and machine vision model capable of checking 300ft of material for defects per minute will likely be preferred to a $10/hr human QA capable of checking 30ft per minute, for instance.

Ukv··on It Still Can't Do My Job: Four Years of Moving Goalposts (2022–2026)
> The intention is something like “so humans are as bad as AI” when the original question boils down to something like “why would I replace humans with AI?”

If AI really is at human level quality/error rate (I don't think it is for general tasks, but there are some areas where it is), then the answer is typically cost and speed/capacity.

Ukv··on Why I like snake_case
I do like snake_case but a lot of this feels a bit circular, effectively just saying that it's good because it's already used by the author's code and things it interacts with.

I'd like kebab-case even more if it weren't for the annoying detail that `-` is also subtraction.

Ukv··on 30-year sentence for transporting zines is a five-alarm fire for free speech
> They brought guns and shot government officials

Only Benjamin Song, convicted of attempted murder/discharging a firearm, shot the police officer. Some others didn't bring firearms, were not in any planning chat (in which no violence was planned regardless), weren't at the protest or had already left, yet still received absurdly harsh sentences - that's the chilling effect.

Ukv··on 30-year sentence for transporting zines is a five-alarm fire for free speech
> The 30 year sentence was for hiding documentation [...] it wasn't just "transporting Zines"

As far as I can tell, the moving of zines (he was pulled over and had a box in his car) is what's being presented as "hiding documentation" - not something beyond that.

> being sought under a federal warrant

Timeline seems to be that a warrant was obtained after pulling him over ("Sanchez-Estrada was then arrested on state traffic offenses, and officers obtained a search warrant [...]"). Can't find a source saying there was a warrant prior to this.

> The warrant was for documentation after the protesters shot fireworks to bring out first responders from the ICE facility, and allegedly one of the group shot a responder in the neck instead of the head.

It's true that demonstrators were setting off fireworks, and it's true that Benjamin Song later shot at a police officer who had drawn his gun. But it's just the government's narrative/speculation that the intent of the fireworks was to draw out first responders to ambush, and that Sanchez-Estrada's zines were in some way documentation of this despite him not being at the protest and his wife not being the shooter.

Ukv··on Half-Life 2 in a Browser
We can guess this is unlicensed, and likely be right, but whether it gets taken down is up to Valve.
Page 1 of 30Next →