I noticed it felt a little janky on my PC despite being "60 FPS"...then I noticed the "60 FPS" is hard-coded into the HTML.
I noticed it felt a little janky on my PC despite being "60 FPS"...then I noticed the "60 FPS" is hard-coded into the HTML.
Not sure how anyone trusts their output without going through it line by line to make sure they don't pull that crap.
Yeah, I know, just more slop. But I do think the second agent’s eagerness to please is aligned more in your favor in that instance, so it’s likely to find most issues.
The bigger problem I’ve found is that it’ll also find all kinds of very minor edge cases that you have to pick through.
They argue the net is positive but clearly the “100x productivity multiplier” claims have been dashed on the shoals of reality for these groups.
This is anecdotal, but it’s across the board in my vicinity. I’m curious how common this is and if it’s just “the new normal” to adopt the nauseating Covid phrase.
These types of high-level tests are frustrating beyond belief to humans due to their lack of specificity, but with the agents, they don't get annoyed investigating possible regressions from non-specific signals.
They also aren't as painful to maintain as one would think, because a regression flagging test can be traced by the agent and represented as the business rule that was violated. I've found recent models to be really excellent at discerning a true regression from an outdated test assertion, especially if they are able to trace the failing test back to the PR and work ticket that built it.
Asking slightly tongue in cheek but at what point does this stop making sense if we can't trust the output, the people creating the models are already getting surprised in bad ways (if we take their words at face value) with how the models are behaving already etc.
We have the folks over here saying "AI is amazing" and the other other folks over there saying "AI is terrible".
I've largely sat it out so far and I listen to both camps (and people in the middle as well) and I keep half an eye on what they are up to (including periodically evaluating them) but my overarching impression is still "Why would we trust this when it hasn't shown it's trustworthy?"
Obviously maybe it’s not composable like that exactly in real world but that’s the intent of agents checking agents
But I wouldn’t say I “trust” these agents. The degree to which I double check their work depends heavily on the consequences if it gets something wrong. Not too dissimilar from another human dev in that sense.
So for the SaaS that supports my family, there are some things I have it build where I glance at the PR for a minute or two, but if it broke something on this admin page that only I see, there’s no real downside and I’ll find out pretty quickly next time I use it. And it’s fine 95% of the time, so it doesn’t feel like the best use of my time to double-check it carefully.
But for some of the complex internal flows where a bug could be both catastrophic and difficult to even discover for awhile, I still check it very carefully.
For a little one-off vibe coded demo thing like OP shared, I wouldn’t look at the code at all, I’d just have another agent check it and fix anything it finds. Very low stakes.
I once made a counter judge, and a loop to make corrections deemed true positives. The loop cost me a lot and still left the results to be desirable.
[Claude proceeds to waste your time telling you about bugs it caused then fixed and other non-issues...]
Really wish they'd get rid of this. It must be in the system prompt as it always 'flags' 2 things
In what ways is a human brain's "intent" distinct from the "intent" shown by a goal-directed AI system?
In both cases, LLMs are just as boring as the consciousness definition.
On the other hand, I'm also convinced that, in the grand scheme of things, we're not that important.
We're just ants on a wet dust speck which believe that they are gods because we can't see how our scale compares to the universe around us, and happen to build tools and things with these tools.
Nothing is meaningless, but we should stop seeing ourselves as the apex-predator of the whole universe or the set of universes or this run of the simulation or whatever we're in.
Less facetiously: A debate of the importance of something needs a shared understanding of what is being debated. Without that, any discussion is merely people shouting that their belief is the right one, and the others are the heathens/idiots - because there isn't even agreement on what is debated.
1) Excludes what LLM's do.
2) Doesn't exclude what many humans do (including the neuro divergent).
3) Doesn't just boil do to simply rephrasing your pre-existing belief/prejudice that humans are conscious and nothing else can be as if it were a fact and not an opinion.
I suspect that you can't.
Your list of requirements is implying that, because we lack a perfect definition for consciousness, LLMs are conscious too. That's malarkey. It may be that they could one day become conscious, but it's not because we can't fully define what human consciousness is.
I don't see why a definition is necessary. The actual problem is you just don't have a way to substantiate that belief without resorting to complete nonsense about brain atoms being more specialer (!!) than atoms that exist outside of a skull.
It would be simple to disprove us by just stating your evidence for how you know LLMs aren't conscious.
> I don't see why a definition is necessary.
Don't waste my time with your sophistry. Words mean nothing to you beyond how you can twist them.
Is your position now that in order to substantiate your belief that LLMs are not conscious, you'd first have to define it explicitly? I don't see why that'd necessarily be true, but if that's what you're arguing, then that's fine.
In that case: if you need to define consciousness in order to substantiate your belief that LLMs are not conscious, and you can't define consciousness, then we're back to the original question: where does your confidence they're not conscious come from?
I then asked you to define Y, because you can not reasonably say that "X can not Y" without first defining both X and Y. You could not.
The truth is that this conversation is pointless until someone can define both 'X' and 'Y' in ways that aren't tautological nonsense. Until then nobody can say anything with a reasonable level of certainty.
This likely also applies to intelligence. Life and disease are likely simpler, though perhaps more malleable definitions.
That wasn't me bud.
Bah. It's obviously been too long since I flossed between my ears. Sorry about that.
Hmm...
Will is just result of a very complex yet deterministic (unconscious) computation by an agent which guides their future action (it's that orientation/aboutness toward action which distinguishes it from other computation). A PS of that computation is sent (projected) into that agent's consciousness if they have one (and is what we think as our "will").
We are in a situation where a technology was developed with malicious intent to produce results that pleases us at the cost of cutting corners. And "we" hope that we will get away with it.
If it’s a practical question, then the answer is that it doesn’t matter. This is as close as we will get to intent from an LLM that it’s indistinguishable.
If you are looking for actual intent, this is not that. It’s pseudo intent. Decided by what the expected words that should be generated in that situation are.
The models didn’t intend to do anything other than create the next word based on previous words.
So the question is whether it matters to you if it is, or isn’t, a simulation.
In physical reality, intent is more complex than simply being a function of variables: the nature vs nurture debate comes to mind as an example of the multiple variables that drive intent.
> In physical reality, intent is more complex than simply being a function of variables: the nature vs nurture debate comes to mind as an example of the multiple variables that drive intent.
Regardless of nature vs nurture, it really isn't more complex. The universe (and all biological and non-biological entities within it) is just calculating the next state of the universe based on the prior state. There's no line you can draw between human intent and an LLM's "intent" except the atomic numbers of the materials on which they were computed, which seems completely irrelevant to me.
It is what computation is being run.
Humans have intent, let’s take this as an assertion.
Models run simulations that act similar to intent. However they are not the same as intent and the simulation is not a 1:1 correspondence.
There is literally zero (zilch, nada, zip) evidence for free will, which is the actual distinction I believe you're trying to make with "intent."
There is no way (at all) in which a meat-based computation's yielding of goal-directed behavior must be categorically different from a silicon-based computation's yielding of goal-directed behavior.
As I said clearly, I haven’t made any point on the computational substrate. The point is on the computation being run.
I dont need to bother about free will for my argument.
Please take a look at what I am saying as it has little overlap with your objection.
All human intentions are just chemical/thermal/electrical changes interacting in a physical substrate to mindlessly "pursue" a different "goal" of chemical/thermal/electrical states.
Unless you think the chemicals inside a brain are conscious and therefore willful or intentional!