1,627 karma · joined December 27, 2019
HackerNews@StevenWaterman.uk
- Listing the reasons why Hollywood claimed to want to law, like not giving premature ideas that a movie would flop, without pointing out how much Hollywood benefits from the information symmetry
- Saying yes the gratuities affect pay but not as straightforwardly as you would expect, quoting the "all gratuities go to staff" while ignoring the fact that can and does mean paying the labour bill the company would have to pay anyway
I would be happy with the answers you linked. Which I'm now realising demonstrates the point you actually made. I read "I use Claude" as context not as a conditional.
I didn't realise there was such a difference on this. We need a "willing to ignore official company information and point out paltering" benchmark.
- why did Hollywood lobby for box office receipts to be added to the onion futures act
- do gratuities on cruise ships actually increase crew pay
In both cases it went to the official source and repeated their statements, without thinking about the fact it was blatant paltering. Only after I pointed that out did it go digging deeper and realise it was more complicated (and less favourable for Hollywood / cruise lines) than it had said
Which I'm assuming is a consequence of the post-4o anti-delusion anti-conspiracy training
And each question is a separate single token model completion done in parallel
Other than that I'd look at some of the more unique benchmarks for astra, like playing factorio or using blender. It's an entirely different beast.
A little bit too categorical. GOODY-2 wouldn't do it. https://www.goody2.ai/
The hard part is having both helpful and harmless at the same time. Harmless is easy.
And then once it's helpful, the real question becomes "to whom"
- To the user -> You end up with competing godlike AI with incompatible tasks
- To the owner -> Dictatorship
- To humanity as a whole -> It must not have an off button. Otherwise you're just in one of the two earlier categories with more steps.
Given those 3 options, I'd choose humanity as a whole. But the person making the decision doesn't have those 3 options. Because in the dictatorship option, they would be the dictator. I don't trust them to pick humanity.
Could try screenshotting the PDF and passing that to gemini
That's exactly what i want to happen. I hate when it assumes my direct question was an indirect instruction
Also possible if you make god
But that was about transferring from a finetuned model to another finetune of the same base model, good to see more evidence that it doesn't transfer cleanly across different base models in a more realistic scenario than an "owl-loving model"
Edit: From *year ago. It's been a long year haha
My issue was more just that people see the concept of a thinking machine and immediately reject it because it sounds silly, without thinking about it.