Boy, it sure would be nice if real LLMs were capable of giving an answer like that.
Boy, it sure would be nice if real LLMs were capable of giving an answer like that.
- it's difficult
- ok fine but how
- it's difficult
- right i'll see that but how
- it's difficult
then it dawned on me this meant get away you fool :DThe captain told us that if we took the bus to Humaitá (a smaller provincial city), we could take a smaller boat that would take us to Manaus. But he warned us that the boat only goes once in three days, and that it would leave soon. The last bus to Humaitá would also leave from Porto Velho promptly.
Despite this flimsy instruction, we didn't see any alternative. So we went. With great luck, we caught the bus and made it to Humaitá (I still have a picture of the boat river transfer the bus took: https://bashify.io/i/CdNcLf).
Our time in Humaitá was surreal. When asked about a boat to Manaus. Everyone told us a different story. There was no harbour (hydroviaria) personal. One person told us the boat to Manaus was named "Caçote". Another person said the boat was named something else. Then someone said it had stopped ferrying years ago. No, we heard from someone else, it would come in 5 hours! Yet another one said it was tomorrow. Someone else felt sorry for us because it had just left. I felt like I was in a (difficult) point and click adventure. There weren't a lot of people in town close to the river, so we ran into the same people from time to time. They would often give different answers to the previous time.
No one was willing to tell us they didn't know. Not a single person out of the 20+ we must have talked to.
In the end, a boat arrived. It went in the direction of Manaus. The captain said that he would only go up to Auxiliadora, and that was still a long way from Manaus. Once again without alternatives (going back to Porto Velho would surely mean we'd miss the flight), we chanced it. In the hopes that getting closer was worth it.
When we arrived at Auxiliadora, it was the smallest inhabited place I had ever been to. Perhaps about 13 houses. Some fishermen. No passenger boat would come for days, they told us. Not to take us to Manaus, neither to take us back. The fishermen had boats, I tried to offer them money so they would take us further. But their day was at an end and they wanted to relax, regardless of what I offered them (we were on a tight budget, but I was desperate enough to offer a significant chunk of a monthly wage, no dice).
Then we found out that on the boat with us, a woman had come who was in a similar predicament as us. She was Brazilian, living in MT and wanting to visit family in Manicoré (which was bigger, and closer to Manaus). Exasperated, she ended up convincing her family to come and pick her up with a speedboat. We hitched a ride. We were very thankful.
When we arrived in Manicoré, I felt like exploring the place. It looked so different from anywhere else I had been, like something out of a movie. But I couldn't. The docks were little more than a collection of wooden jettys (trapiche) that ran everywhere in criss-cross fashion. In order to even get to the quayside we would have to pass through many other boats. In the first one we went through, the captain walked past and I asked him whether he knew of a boat going to Manaus. He signalled where to put my bags. We were leaving.
We reached Manaus in the nick of time.
I love this story, and that time. These anecdotes definitely triggered my memory.
its common playbook for corporate self-development in NA.
I'm a patient person, but it can be frustrating to have to endure 10 minutes of verbal diarrhea that eventually results in a "no" or "I don't know".
I don't know any Spaniards but I do know Filipinos and the confidence projection is a real thing. The Filipino IT guy confidently declared that my OnePlus Android phone wasn't certified for the software he was trying to install and was getting errors. It is a bog standard application that can be installed on any modern Android phone but the level of confidence he projected, just because he didn't know OnePlus as a brand, made me doubt myself until I turned on the critical hat and pushed back a little with alternative approaches, which solved the problem.
But I kid, I have a friend who's the same way. He's an Austrian who grew up in Chicago and was in the army.
I have considered the phenomenon. I somewhat disapprove but I can also see the advantage of always presenting a confident face
They com like that from factory. Hardcoded to never say no.
Lot of confabulation going on in the moist blob of electrochemistry found encased within the hydroxyapatite crystal cage we call a skull. Why is it that things which rhyme, or get repeated a lot, seem more true? How come some of us suffer uncontrollable seizures from flashing lights?
We have to study reason to get any good at it, we absolutely suck at this without training. Our natural state is illiterate, innumerate, and illogical.
The part of our (all animals, not just humans) intelligence that is most magical*, is how efficient we are with few examples, not reasoning.
* Questions about consciousness will have to wait until we can agree which of the 40+ definitions of the word actually answers the question we care about, and then also we figure out how to actually test for whatever that is.
It's not a guaranteed way to control their behavior, but you can more than move the needle.
Steering an LLM with a prompt is way less reliable than steering a car with a steering wheel, but there's still control. It's just not absolute.
Non?? Only those with sh*tty code, surely.
There's nothing inherently non-deterministic about inference.
There's lots of people with lots of opinions, but you (often) have to pay for the good ones.
That's not an unfair take, I think. Again, just IME, they expect too much because the tool is oversold: it does not deliver that well. And we always hear, this new model is so much better, it's tiring.
I think we should all learn to use LLMs but we should still carefully review what they did. And that is what the employers don't quite get: the review still takes a lot of time. So, gains are not 10x but more like... 10%? Maybe 50 for boiler plate. Still gains are there, I guess.
And unfortunately a lot of people will say it’s their reports’ fault for not properly utilizing it (even as they barely use it) because otherwise they would have to admit that they bought a tool without any plan for how to deploy it. So regardless of what is or isn’t a fair take, the results are the same. We are burdened with utilizing a thing whether it is useful or not and the results are generally not what is measured, but rather “are you using it?”
I’m just glad I work at a company that has more reasonable expectations and has been very slowly, thoughtfully rolling it out to individuals at the company and assessing what is and isn’t good for. They are interested in getting me in line, but as somebody in video production to be perfectly honest the use case for Claude is a bit tricky to navigate. We don’t write a lot of scripts and I already have bespoke software for organizing/maintaining footage that isn’t on a subscription basis. The work I’m also doing doesn’t call for these speed-editing solutions that generate tik tok chaff. All our stuff is hours long and it’s high volume. Any video-centric AI service costs an arm and a leg.
I do think it could be useful for writing some terminal scripts and such, but as far as a daily tool we are still scratching our heads and thinking about it. But it’s nice to be able to do that without somebody saying “why aren’t you using it?” every meeting.
I know I'm shouting into the void, but seriously.
I've been trying to work on a new LLM code editor that does just that. When you instruct it to do something, it will evaluate your request, try to analyze the action part of it, the object, subject, etc, and map them to existing symbols in your codebase or, to expected to be created symbols. If all maps, it proceeds. If the map is incomplete, it errors out stating that your statement contained unresolvable ambiguity
I think there is a real benefit here, and it might be the actual next beneficial grounded AI sustainable use in programming. Since I the current "Claude code and friends" are but a state of drunkenness we fell into after the advent of this new technology, but it will prove, with time, that this is not a sustainable approach
LLMs are just generating text, they don't know anything. They can't assess whether there is enough data for an answer. When you add a follow up prompt "This is wrong, why did you lie?" only then is it able to generate text, "I was wrong, I'm sorry," and so forth.
It seems that they are loath to tell anyone “no”, or that something can’t be done, or that an app doesn’t have a feature or can’t be used in a certain way. Especially when a feature has been removed for security reasons.
In fact, it gets so crazy that I simply cannot get a straight answer out of somebody and if I persist in my line of questioning and they become evasive or vague or I just can’t get a straight answer for long enough, ultimately, I suspect that the answer is “no”, and that they're simply not allowed to tell me, and they're paid and trained specifically to avoid uttering the “n-word”.
In my first job, as a network operator, my supervisor admonished me, and said “we must never tell a customer that we don't know something”. He said that we should tell the customer that “I will go ahead and find out for you, and get back to you on that”.
And that is kind of the kind of slippery non-answer I often received in my most recent job, that some manager or supervisor would “look into something” for me and “get back to me”. But the ‘getting back to me’ part never happened, and I began to suspect that it was a platitude meant to satisfy me enough that I would shut up for a while, and stop pressing the issue.
Asimov's Multivac at least had the dignity to wait.
Maybe hackernews is becoming reddit...