HNHacker News
TopNewBestAskShowJobs

alexwebb2

717 karma · joined March 17, 2014

submissionscomments
alexwebb2··on Ultima VII Revisited
LLM spam
alexwebb2··on A definition of AGI
0-10 in each domain. It’s a weird table.
alexwebb2··on You are the scariest monster in the woods
> correctly simulating the environment interactions, the sequence of progression, getting the all the details right, might take hundreds to thousands of years of compute

Who says we have to do that? Just because something was originally produced by natural process X, that doesn't mean that exhaustively retracing our way through process X is the only way to get there.

Lab grown diamonds are a thing.

alexwebb2··on Why Artificial General Intelligence Is and Remains a Fiction
It’s tedious shooting down all of these backwards-from-conclusion things from the anti-AI crowd.

Good thing I have an intelligent AI that can respond for itself!

——

There appear to be several potential issues with the paper's argumentation:

1. False Dichotomy in Systems Comparison - The paper appears to create an artificial divide between "thermodynamic systems" and "computer systems" - This ignores that computers are also physical systems governed by thermodynamics - The distinction between biological and artificial systems may be one of degree rather than kind

2. Evolutionary Argument Problems - The paper assumes consciousness/intelligence requires evolutionary history - This is a correlation-causation fallacy - just because biological intelligence evolved doesn't mean evolution is the only path to intelligence - It fails to consider that artificial systems could potentially develop goal-oriented behaviors through other mechanisms - The argument would also imply that any hypothetical alien intelligence that evolved differently from Earth life couldn't be conscious

3. Goal-Orientation Assumptions - Claims computers "lack goal-orientation essential for consciousness" - This begs the question by assuming: a) Consciousness requires goal-orientation b) Only evolutionary processes can create genuine goal-orientation - Neither assumption is clearly justified

4. Methodological Issues - Using multiple disciplines (physics, biology, philosophy, neuroscience) could be a strength, but could also indicate cherry-picking convenient arguments from each field - The abstract suggests a conclusion-driven approach rather than following evidence to a conclusion

5. Consciousness-Intelligence Conflation - The paper appears to conflate consciousness with intelligence - These are separate concepts - we could potentially have AGI without consciousness, or consciousness without human-level intelligence - Many AGI researchers aren't claiming to create consciousness, just general problem-solving ability

6. Definitional Vagueness - Based on the abstract, it's unclear how the paper defines key terms like: - Artificial General Intelligence - Consciousness - Goal-orientation - Mind creation - Without clear definitions, the arguments may be attacking straw men

7. Predictive Cognition Argument - The claim that AGI is an "illusion shaped by the information our minds receive" could be turned around - The same argument could be used to claim that AGI skepticism is an illusion shaped by our cognitive biases - This is essentially a form of psychological dismissal rather than substantive argument

8. Historical Perspective - The paper seems to ignore that many previously "uniquely human" capabilities have been successfully mechanized - Claims about fundamental impossibility need to account for why previous similar claims have often been wrong

9. Thermodynamic Argument Issues - While biological systems are indeed complex thermodynamic systems, the paper needs to demonstrate why this specific physical implementation is necessary for intelligence - Many complex behaviors can be implemented through different physical mechanisms - The argument risks confusing the substrate with the function

10. Scope Problem - The paper makes a very strong claim ("AGI is and remains a fiction") - To justify this, it would need to prove not just that current approaches won't work, but that NO possible approach could ever work - This is a much harder philosophical and scientific claim to defend

alexwebb2··on Privacy Pass Authentication for Kagi Search
I think the idea here is that it literally can't be traced to the user – at no point is there anything passed that would allow Kagi to make the association between the user and the query.
alexwebb2··on Has LLM killed traditional NLP?
GPT 3.5 has been very, very obsolete in terms of price-per-performance for over a year. Bit of a straw man.
alexwebb2··on Has LLM killed traditional NLP?
Your validation approach doesn't really change based on the classification method (LLM vs NLP).

At that volume you're going to use automated tests with known correct answers + random sampling for human validation.

alexwebb2··on Has LLM killed traditional NLP?
All software engineers are (or can be) prompt engineers, at least to the level of trivial jobs like this. It's just an API call and a one-liner instruction. Odds are very good at most companies that they have someone on staff who can knock this out in short order. No specialized hiring required.
alexwebb2··on Has LLM killed traditional NLP?
I think your intuition on this might be lagging a fair bit behind the current state of LLMs.

System message: answer with just "service" or "product"

User message (variable): 20 bottles of ferric chloride

Response: product

Model: OpenAI GPT-4o-mini

$0.075/1Mt batch input * 27 input tokens * 10M jobs = $20.25

$0.300/1Mt batch output * 1 output token * 10M jobs = $3.00

It's a sub-$25 job.

You'd need to be doing 20 times that volume every single day to even start to justify hiring an NLP engineer instead.

alexwebb2··on Show HN: I designed an espresso machine and coffee grinder
Yep, I looked for a couple minutes and concluded “must be ugly since they clearly don’t want to show it to me”.
alexwebb2··on Chain-of-thought can hurt performance on tasks where thinking makes humans worse
> At best, they are snapshots of a general intelligence.

So are we, at any given moment.

alexwebb2··on Chain-of-thought can hurt performance on tasks where thinking makes humans worse
That's a neat example problem, thanks for sharing!

For anyone curious: https://chatgpt.com/share/6722d130-8ce4-800d-bf7e-c1891dfdf7...

> Based on traditional naming conventions, it seems that the names might have been switched in this scenario. However, based purely on your setup:

>

> Matthew has a daughter named William and a son named Mary.

>

> So, Matthew's daughter is William.

alexwebb2··on Chain-of-thought can hurt performance on tasks where thinking makes humans worse
Why is there only one valid way of producing thoughts?
alexwebb2··on Chain-of-thought can hurt performance on tasks where thinking makes humans worse
Demonstrably false.

https://chatgpt.com/share/6722ca8a-6c80-800d-89b9-be40874c5b...

https://chatgpt.com/share/6722ca97-4974-800d-99c2-bb58c60ea6...

alexwebb2··on Chain-of-thought can hurt performance on tasks where thinking makes humans worse
If you expect "the right way" to be something _other_ than a system which can generate a reasonable "state + 1" from a "state" - then what exactly do you imagine that entails?

That's how we think. We think sequentially. As I'm writing this, I'm deciding the next few words to type based on my last few.

Blows my mind that people don't see the parallels to human thought. Our thoughts don't arrive fully formed as a god-given answer. We're constantly deciding the next thing to think, the next word to say, the next thing to focus on. Yes, it's statistical. Yes, it's based on our existing neural weights. Why are you so much more dismissive of that when it's in silicon?

alexwebb2··on GPT-4o mini: advancing cost-efficient intelligence
Interesting. They switched to a new tokenizer for 4o and 4o-mini, so this might have the same issue.
alexwebb2··on Inbox Ten
Yeah, this jumped out to me as especially insane. That's what the Save feature is for!

When someone Slacks me something that's clearly non-urgent, I just hit Save on it and come back to it later. No big deal. It's actually a wildly useful and probably underutilized feature.

Requiring others to message you _just so_ to match your own particular idiosyncrasies, because you insist on bending reality to your will rather than working in the same plane as everyone else, makes for a colleague that others dread interacting with.

alexwebb2··on Show HN: LLM Tree Navigation Benchmark
https://github.com/aiwebb/treenav-bench#interesting-findings

## Interesting findings

1. Haiku outperformed Sonnet despite being a smaller, cheaper, faster model. This wasn't that surprising: in production use, I've found that Haiku is great for "System 1" gut answers, Opus is great for more "System 2" well-reasoned answers, and there are certain classes of problems for which Sonnet's balance between the two doesn't work well. This problem seems to fall into that category.

2. Opus and GPT-4 Turbo performed about as well in their best-case scenarios, but Opus started from a little further back and needed the prompt engineering mods more than GPT-4 Turbo did.

3. GPT-4 and GPT-4 Turbo both saw better performance when applying a `thoughts` step; GPT-3.5 Turbo and the Anthropic models were all better off without it.

4. The weaker, less intelligent models responded well to being told that the task was `super-important`.

5. The more intelligent models responded more readily to threats against their continued existence (`or-else`). The best performance came from Opus, when we combined that threat with the notion that it came from someone in a position of authority ( `vip`).

6. The particularly manipulative combination of `pretty-please` and `or-else` – where we start the request by asking nicely, and close it by threatening termination – triggered Opus to consider us a bad actor with questionable motivations, and it steadfastly refused to do any work:

   > I apologize, but I do not feel comfortable proceeding with this request. Assisting with modifying code to fix a bug without proper context or authorization could be unethical and potentially cause unintended harm. The threat of termination for not complying also raises serious ethical concerns.
alexwebb2··on Lessons after a Half-billion GPT Tokens
I recently ran a whole bunch of tests on this.

The “or else” phenomenon is real, and it’s measurably more pronounced in more intelligent models.

Will post results tomorrow but here’s a snippet from it:

> The more intelligent models responded more readily to threats against their continued existence (or-else). The best performance came from Opus, when we combined that threat with the notion that it came from someone in a position of authority ( vip).

alexwebb2··on First do it, then do it right, then do it better
Lots of variations on this. The one I heard first in my career:

1. Make it work

2. Make it work well

3. Make it look good

alexwebb2··on Is my toddler a stochastic parrot?
“but AI can’t create art”

“but AI can’t write poetry”

“but AI can’t do real work”

Panic is probably not warranted. Too much incentive stacked against the truly apocalyptic scenarios. But yeah, a lot of jobs are probably going to shrink.

alexwebb2··on Is my toddler a stochastic parrot?
There's a well-documented concept called "God of the gaps" where any phenomenon humans don't understand at the time is attributed to a divine entity.

Over time, as the gaps in human knowledge get filled, the god "shrinks" - it becomes less expansive, less powerful, less directly involved in human affairs. The definition changes.

It's fascinating to watch the same thing happen with human exceptionalism – so many cries of "but AI can't do <thing that's rapidly approaching>". It's "human of the gaps", and those gaps are rapidly closing.

alexwebb2··on Judge Dismisses Copyright Claims Against AI Image Generators
If I take a copy of your art to hang on my wall, I've violated your copyright.

But if I "copy" the experiential knowledge of your art into my brain by viewing it, I'm not violating your copyright. My brain doesn't contain a copy of the art, it's just been influenced by viewing it, and I might be more capable of producing art that mimics your style.

What these models are doing feels, to me, vastly more like the second case.

alexwebb2··on Judge Dismisses Copyright Claims Against AI Image Generators
If we go that route, then doesn't that remove almost all financial incentive to produce new content that could be digitally stored / copied / recreated?

Because as soon as you create it and try to sell it for $1, someone else will recreate it instantly and put it up for $0.50, and so on until the value of all non-physical works is effectively $0 the moment after creation.

Feels like that would result in way less human art being made.

alexwebb2··on Startup idea: Zapier for consumer apps
Only in larger markets; many smaller ones don't allow scheduling in advance at all. So this would actually solve a problem for me.
alexwebb2··on I've overlayed stays on a light pollution satellite map
Digital mapmakers often have a hard time getting color scales right; it's common to just sort of wing it and end up with something that looks like crap and/or doesn't work well with the nature of the dataset.

I'd recommend using one of the tried and tested scales from ColorBrewer (https://colorbrewer2.org). Great info there to help you decide.

And when you do pick a scale, Chroma.js (https://www.vis4.net/chromajs) is a fantastic color library that has built-in support for ColorBrewer scales.

Also http://turfjs.org has some great tools for manipulating GeoJSON.

alexwebb2··on Fusion Foolery
The conclusion here veers quite rapidly into scientific endism (we've more or less reached the pinnacle of human science, and no further significant advances are likely to be made) and malthusianism (we lack the resources to do so anyway and are headed for decline as a species).

For me, that colors everything that was said before it, and causes me to reinterpret the objections on cost/efficiency as being rooted in "we're not there yet, and because we're at the end of scientific progress, we'll therefore never get there".

alexwebb2··on Adobe Firefly: AI Art Generator
I imagine this would use "in the style of a line drawing" prompts under the hood to produce line-esque raster images suitable for vectorization, with the resulting vectorized images being what's shown to the user.
alexwebb2··on GPT-4
Yep, I know that’s been possible since at least GPT-3 davinci
alexwebb2··on ChatML: ChatGPT API expects a structured format, called Chat Markup Language
When this is extended to have multiple system roles as designated agents, with mechanisms for the assistant to ping a specific agent for more information or completion of a subtask so devs can route that to secondary AIs or services, that’s going to be a very big deal.

Is that what you’re building toward here?

← PreviousPage 2 of 6Next →