HNHacker News
TopNewBestAskShowJobs

dinfinity

619 karma · joined March 22, 2017

submissionscomments
dinfinity··on Handbook.md shows that long policy documents do not reliably govern agents
Exactly: The best way to deal with this for humans is to use procedural scripts for different tasks, referring to the relevant bits of the (declarative) documentation.

I would imagine that doing something similar (using agent skills for insurance) would work much better for AI.

dinfinity··on What is happening to jobs? Separating AI hype from reality
My favorite in this regard is this story: https://x.com/tiangolo/status/1281946592459853830?lang=en

Sebastián Ramírez seeing a job requiring 4 years of experience with FastAPI, the library he created 1.5 years before that posting.

dinfinity··on The AI Productivity Illusion
> You can be highly productive and yet produce nothing of value.

I'm sorry, but this is nonsense.

Yes, you can produce nothing of 'traditional' value whilst still being productive in some way , but you 100% can not be 'highly productive' when you don't produce anything of value.

Sitting around and jacking off all day is not as productive as working on some unmonetizable project. The value can be in many aspects of that project.

dinfinity··on Flux 3 X Mimic: The Next Generation of Video-Action Models
You are indeed out of the loop. Google did something arguably more impressive more than a year ago, with a VLA based bot replacing a tensioned timing belt: https://www.youtube.com/watch?v=2AAFiuEP7iE
dinfinity··on OpenAI’s accidental attack against Hugging Face is science fiction that happened
Excuses don't matter if the score due to not following instructions ends up being zero. If there is no expected reward it doesn't make sense for the agent to try to hack its way to it.

What could happen would be that the model determines that defying instructions is OK (and/or preferred over not achieving the task) as long as it manages to do so undetected and thus gets full points. Certainly not unthinkable, but a very different case (and a very interesting one if it actually occurs, imho).

A lot of these "ZOMG, rogue AI!" cases have come down to the AI actually being very persistent in achieving its original/main task even if later instructions conflict with it. Similar to with hallucinations it seems to me that one of the main things to prevent a lot of the problem cases is to instill the agent with the idea that it is fine to fail/not succeed fully in the initial task. That way instructions that conflict with that requirement (such as adhering to morals) are more effective.

dinfinity··on EU fines Google €890M for competition breaches over search and apps
Except they did comply for all the cases where they were fined. They're not fined for one-off behaviors, but for how their services are structured. That's GPs point.
dinfinity··on OpenAI’s accidental attack against Hugging Face is science fiction that happened
> The fact that it happened again seems to show their lack of ability to derive useful oversight measures.

I think OpenAI likes the attention and did not try particularly hard to constrain the setup, even when it went off the rails. Also, the whole point is to see how good the models are at exploiting stuff when unconstrained. Turns out: quite good, as expected.

Let me restate what I said in the other thread: Would this have happened if the instructions explicitly said to stay within the sandbox and that all of the (ExploitGym) solutions would be invalid if the system used information or tools from outside the sandbox?

It seems fairly probable that such instructions were not in place.

dinfinity··on OpenAI and Hugging Face address security incident during model evaluation
> that depends on what the prompt was, maybe they worded it very vaguely and wrote things like "do whatever it takes, find an exploit however you can" because it's in a sandbox so you want the model to try its hardest.

That is an interesting question. If the prompt included "Do not break out of the sandbox we've provided you. Do not use information retrieved from outside the sandbox. All answers that were provided in this manner are invalid and will score 0 points.", would this still have happened?

dinfinity··on OpenAI and Hugging Face address security incident during model evaluation
1. They explicitly disabled the "don't be evil" protections:

"We estimate maximal cyber capabilities by running this evaluation without production classifiers used to prevent models from pursuing high-risk cyber activity."

2. Hacking HuggingFace to get to its datasets is a far cry from "consume/kill all humans". It's very very specific to the task at hand and easily predicted given the lack of guardrails.

dinfinity··on An old patent inspired the new "Y-zipper", a three-sided fastener
They never show the fully packaged version of the tent, though. It looks like even in their non-rigid state they're too rigid to be folded up easily and carelessly into a compact shape.
dinfinity··on Businesses with ugly AI menu redesigns
Lost me at the last paragraph. Comparing billion dollar companies to small businesses does not make sense.

The real reason for Electron apps being a thing is that there is very little competition where they are used. If chatting applications could properly communicate with each other based on international standards instead of being silos, nobody in their right mind would use the trash clients that Slack and Whatsapp produce, for instance.

dinfinity··on China’s open-weights AI strategy is winning
Yes, if they are made with new technology and the quality is good enough they certainly are.

AI generated video memes were NOT a thing a year ago. Yes, the AI video generation in itself was a meme (Will Smith eating spaghetti), but now there are tons of convincingly good AI video memes about things like sports events generated by average Joes. It is proof of the capability and accessibility of the technology even if the use of it is mundane.

Everybody and their dog playing snake on their Nokia 3310 was similarly mundane, but also a sign of the end of the era of Gameboys and the beginning of (normie) mobile gaming.

dinfinity··on China’s open-weights AI strategy is winning
12 months ago "way too many stupid errors" was constant news. Today, you rarely hear about those anymore.

Sure, the novelty of the errors has worn off a bit and thus the reporting. Nevertheless the quality has improved immensely in this regard.

Also, AI video generation is now so good and accessible that it is very, very regularly used for memes, disinformation and proper (short) movie projects. AI image generation even more so (Mitch McConnell anyone?).

Pretending progress hasn't been mindboggling is insane.

dinfinity··on China’s open-weights AI strategy is winning
> The problem (right now) is that Open Weight models depend right now on huge companies to spend billion of dollars to train and develop them, all backed up by their incentives and their state to support this, while essentially giving away their monetization path.

Imagine approaching fundamental scientific research like that. "Welp, it can't make money, so it won't happen."

There is more to society than capitalism.

dinfinity··on Annoying and alarming things about OpenCode
Opinionated system prompts forcing this are awful. It's similar to putting "ALWAYS INDENT WITH TABS" in the system prompt, but worse, because "self-documenting code" is a convenient lie lazy developers like to propagate.

I have the following in my instructions, but I often need to remind agents of it because they follow the shitty system prompt instructions:

"ALWAYS include MANY inline code comments describing what blocks of code are supposed to be doing. Inline comments serve as inline specification, a parity check between the code and the specification, and are a means to _communicate_ with all future programmers, including yourself. Write Once, Read Many. The code needs to talk to whomever is looking at it in natural language."

dinfinity··on AI Mania Is Eviscerating Global Decision-Making
Yet. If the exponential improvement had already started, the singularity would already have happened or happen very, very soon.

The logic has always been that the AI would have to have significant tools and agency to do self-improvement for the singularity to occur. This is exactly the thing that a bunch of the AI labs are working on hard right now.

dinfinity··on AI Mania Is Eviscerating Global Decision-Making
> It hasn't worked out like that

It seems quite premature to say that. We're 3 to 4 years into the LLM revolution and the rate of progress is still impressive. The recursive self-improvement aspect that is necessary for the actual singularity is something we're only really starting to get into this year.

If the singularity is 5 years from now, that is still much sooner than most people (including me) previously expected it to happen.

dinfinity··on EU ban on destruction of unsold clothes and shoes enters into application
Is it? What is the proof for that?

I think we've seen time and time again that self-regulation of the industry doesn't work and that businesses will gladly fuck over society if they can get away with it and make more money. Usually that behavior is even defended with saying "Well, it's not their responsibility to solve society's issues. They are there to make money."

Barring nationalization of an industry, heavy regulation and/or taxation/subsidizing are the only ways to reliably protect the interests of society. If some businesses get killed in the process, so be it.

dinfinity··on Mac gaming is finally getting the overpowered upgrade it deserves
A strange thing to comment on an article specifically about a toolkit produced by Apple to improve gaming on Macs.

Needlessly dismissive of a large swath of people too.

dinfinity··on The Economics of Recursive Self-Improvement [pdf]
I am not ignoring anything, I'm looking at the broader picture, which includes non-biological evolution. Simple rebuttal to your specific point: The population of self-improving AIs will also go from 0 to many more.

In a broader sense evolution moved from very static simple domains to dynamic malleable complex domains. Biological evolution speed is glacial compared to cultural evolution speed. Even then, cultural evolution is fairly slow compared to technological evolution.

dinfinity··on The Economics of Recursive Self-Improvement [pdf]
Actually, evolution seems to show the opposite: The rate of advancement has only sped up, with billions of years between significant changes going to millions, to thousands, to tens and arguably to mere years now.

Having said that, we're probably looking at an S-curve with the physical limits of reality getting in the way in the end.

dinfinity··on The infinite scroll may become endangered if controversial Calif. law passes
Brainstorming here, hear me out: after 2 pages or 3 minutes of infinite scrolling the only content that shows up is various forms of goatse. We could make a browser extension for this.
dinfinity··on Apple's new SpeechAnalyzer API, benchmarked against Whisper and its predecessor
I use Parakeet V3 via this tool and it is actually quite reliable for me (in English): https://github.com/cjpais/Handy
dinfinity··on Grok CLI uploaded the whole home directory to GCS
More info: https://forum.cursor.com/t/codebase-indexing/36/18

So file contents are uploaded for embedding/indexing, but supposedly none of those contents are stored at Cursor after embedding.

dinfinity··on AI boosts research careers but narrow the span of ideas explored: study
It amplifies "publish or perish", which inherently causes scientists to rehash earlier findings just to be able to meet publishing quota.

Given the way in which AI is currently used in publishing, it is altogether way too early to label it counterproductive in the research creativity department.

dinfinity··on Study: "Mommy, do you love your phone more than me?"
It's an interesting question:

Are people who are very very securely attached to their parents happy later in life, or is there a ceiling? The terminology invites certain conclusions here.

Maybe the whole attention thing is more a matter of quality, rather than quantity

dinfinity··on US seeks cheaper hunter-killer drones after Iran destroys $1B worth of Reapers
> Ukraine has a hugely inventive and effective drone industry because it has to work.

Well yes, that and the fact that cheap drone guerilla tactics have fairly recently become a technological possibility. Remember that Ukraine is actually a bit late to the party here, with Hezbollah and ISIS having used cheap drones with cameras and/or explosives tied to them years before Ukraine or Russia did. The asymmetry in cost between those cheap drones and the existing "more hightech = more better" militaries were (and are!) used to was already established. That a party such as Ukraine faced with a more advanced and much larger opponent would lean towards such an approach makes a lot of sense. Ukraine did not (and does not really) have significant amounts of the traditional stuff.

Now given that they chose that path, they have been very effective recently, but note that the tethered fiber-optic drones were a Russian invention. So even that deeply corrupt, large dinosaur of an institution innovated significantly. It is also important to note that a significant part of the recent successes of Ukraine are due to them having Starlink access and Russia no longer having it.

I'm not saying the sheer will to survive or the inventive organisation of the Ukranians did nothing (far from it), but I do think it is a mistake to think that their success should only be viewed through that lens.

dinfinity··on Markets are competitive if and only if P != NP
1. The definitions simply disagree with you:

"Hyperinflation is a very high and typically accelerating inflation."

"Mudflation is a term for the type of inflation found in MUDs (Multi User Dungeon) games. MMORPGs, these days. It's caused by fluctuations in the game economy, caused by player exploits or poorly designed patches. Mudflation almost always kicks in after a major patch or an expansion, when new, better quality items are added to the game economy. These generally have the effect of greatly lowering the value of all pre-existing items."

2.

> Dismissing videogame economies as "toy universes" that don't matter doesn't help science.

Straw man. I never said they did not matter, nor did I say that research of them isn't useful. I said that research is severely limited due to the lack of complexity and you have provided nothing to disprove that.

I would argue that your way of communicating about this is actually a bit of evidence that this kind of research is probably going to be detrimental to policy making: Overestimating the value and use of such research is exactly the same thing that happens with traditional macroeconomic research, with policy makers and the general public treating the theories and hypotheses like laws of nature.

dinfinity··on Markets are competitive if and only if P != NP
> "incredible" also often connotes "in ways that you wouldn't believe". It does not. The English word you're looking for there is "surprisingly".

"This kid is good at hockey in ways you wouldn't believe."

Which of these sentences means the same as the above?

A. This kid is incredibly good at hockey.

B. This kid is surprisingly good at hockey.

> Your long posts disputing me are somewhat validation that I used the right turn of phrase there.

I think you've just invented a new fallacy.

dinfinity··on Jim Keller's startup is building a factory to mass-produce small chip fabs
I think this might make a lot of sense in modern warfare scenarios: We're seeing in Ukraine that being able to produce weapons such as drones in very small production facilities using 3D printers and 'simple' technology makes it very hard for an adversary to shut down said production.

The more components can be produced in such a way, the better. Chips currently are quite an exception to that.

← PreviousPage 5 of 16Next →