HNHacker News
TopNewBestAskShowJobs

MrNeon

80 karma · joined September 7, 2017

submissionscomments
MrNeon··on Deleting and destroying finished movies
Neither the resources or the labor are public.
MrNeon··on Latest ChatGPT 4 System prompt (1,700 tokens)
There is nothing secret to hide, what would be the purpose of blocking it?
MrNeon··on Nightshade: An offensive tool for artists against AI art generators
Your phone example is just theft unless I'm misreading...do you want to go down the path of calling copyright infringement theft?

As for the card example it is on you to keep secrets secret. It is clearly a fault with how credit cards work and for things like that you should get insurance. Keep your card safe, or better yet, use cash. We can call it theft, doesn't matter to my point.

There is no "taking" or "grabbing" anything, I visited your website and was served a copy of the data because you made it do so. You wanted it to happen, for the public to see it, that was your goal. You expected it to happen. Do you disagree that me acquiring a copy was consensual? If so it is very different from the hypotheticals you're posing don't you think?

Once a copy is in my possession you would need to initiate violence to stop me from training a model on it which puts you on the wrong side of the moral line.

MrNeon··on Nightshade: An offensive tool for artists against AI art generators
Data you put up on the internet is not your body.

Do I really have to explain this? You know I don't. Do better.

MrNeon··on Nightshade: An offensive tool for artists against AI art generators
> it’s basically the input data compressed with highly lossy compression.

Okay, extract the images from a Stable Diffusion checkpoint then. I'll wait.

It's not like lossy compression CAN'T be fair use or transformative. I'm sure you can imagine how that is possible given the many ways an image can be processed.

> All the corporations that are offering AI as a paid service?

Am I them?

MrNeon··on Nightshade: An offensive tool for artists against AI art generators
Who said anything about creating a derivative? Surely you don't mean to say that any image created with a model trained on copyrighted data counts as a derivative of it. Edit: Or worse, that the model itself is derivative, something so different from an image must count as transformative work!.

Also who said anything about selling?

MrNeon··on Nightshade: An offensive tool for artists against AI art generators
The tears of artists and copyright evangelists is so sweet.
MrNeon··on Nightshade: An offensive tool for artists against AI art generators
Tell me where it says training a model is infringing on copyright.
MrNeon··on Nightshade: An offensive tool for artists against AI art generators
That is how consent works.
MrNeon··on Nightshade: An offensive tool for artists against AI art generators
Got all the permission I need when it was put on a publicly accessible server.
MrNeon··on Nightshade: An offensive tool for artists against AI art generators
>why are you seeking to provide an “antidote”

To train a model on the data.

MrNeon··on On being listed as an artist whose work was used to train Midjourney
I hope the idea of Intellectual Property as a whole is thrown out the window and copyright with it.
MrNeon··on Stable Code 3B: Coding on the Edge
It is weird that it is not mentioned in the model card but I'm pretty sure it is a completion model, not tuned as an instruct model.

edit: the webpage does call it "Stable Code Completion"

MrNeon··on Hidden Changes in GPT-4, Uncovered
Right, so none of the instructions for GPT to call the DALL-E function, none of the prompt expansion rules and none of the censorship rules. Clearly the system prompt is not present in the file OpenAI gives us.
MrNeon··on Hidden Changes in GPT-4, Uncovered
I don't have ChatGPT Plus to check that out. Could you share the latest DALL-E 3 system prompt?
MrNeon··on Hidden Changes in GPT-4, Uncovered
I downloaded my chat history just now and system prompts are not present in it.
MrNeon··on Hidden Changes in GPT-4, Uncovered
To say it can't “tell me what I just wrote” is to say it can't copy parts of the context.

We know it can copy parts of the context and the system prompt while a special part of the context isn't immune to being copied.

You can test it yourself by adding random strings to the system prompt, you can consistently have the model copy them over. Is that not enough to have a reasonable belief that the model can copy system prompt instructions in the web interface?

MrNeon··on 'Meditations' by Marcus Aurelius – Stoicism in Modern Language [video]
The audio is painfully AI right from the start, it lacks meaningful pauses.
MrNeon··on Nvidia Unveils RTX 5880 Graphics Card with 14,080 CUDA Cores and 48GB VRAM
It does do that but you can generate an image at the normal resolution and then scale it up to run more high-frequency sampling steps to add detail without losing the composition.
MrNeon··on I forked SteamOS for my living room PC
If my cheat puts my crosshair on the opponent's head automatically what about that information is untrustworthy that would make you throw it out?
MrNeon··on QuIP#: 2-bit Quantization for LLMs
IIRC quantizing small models causes a higher relative drop in the metrics.
MrNeon··on Claude 2.1
Seriously? No rebuttal to my points, just dismissing me as a person? Edit: I don't mind if you insult me, as long as you back it up with facts. Like I did.

I really want that Azure information and whether prefilling works there as it does with Claude or not. Can you provide that at least before you walk away?

MrNeon··on Claude 2.1
> a) gpt-3.5-turbo has a completion endpoint version as of June: `gpt-3.5-turbo-instruct`

We were not talking about that model and I'm 99.999% sure you do not use that model. You might as well mention text-davinci-003 and all the legacy models, you're muddying the waters.

> b) Even the chat tuned version does completions, if you go via Azure and use ChatML you can confirm it for yourself. They trained the later checkpoints to do a better job at restarting from scratch if the output doesn't match it's typical output format to avoid red teaming techniques.

Don't fucking say "even", I know you know I know it can technically do completions as it is just GPT, the issue is what they do with the prompt in the backend.

I do not have Azure to test it, that is interesting but how come you're only mentioning it now? That's more interesting. Anyway, are you sure you can actually prefill with it? You saying that it restarts from scratch tells me it either isn't actually prefilling (and doing a completion) or that there are filters on top which makes it a moot point.

The documentation doesn't mention prefilling or similar but it does say this: This provides lower level access than the dedicated Chat Completion API, but also [...] only supports gpt-35-turbo models [...]

Shame.

> What you keep going on about is the <|im_start|> token... which is functionally identical to the `Human:` message for Anthropic.

Now you got it? Jesus Christ, but also no, I mean "\n\nAssistant:" which is not added on in Anthropic's backend like OpenAI does, you have to do it yourself as stated in the Anthropic docs which means you can use it as a completion model as stated in the Anthropic docs, which makes it trivial to bypass any and all refusals.

MrNeon··on Claude 2.1
> GPT 4 can continue a completion contrary to your belief

When did I say that? I said they work differently. Claude has nothing in between the prefill and the result, OpenAI has tokens between the last assistant message and the result, this makes it different. You cannot prefill in OpenAI, Claude's prefill is powerful as it effectively allows you to use it as general completion model, not a chat model. OpenAI does not let you do this with GPT.

MrNeon··on Claude 2.1
> Are you going to have your user

What fucking user, man? Is it not painfully clear I never spoke in the context of deploying applications?

Your issues with this level of prefilling in the context of deployed apps ARE valid but I have no interest in discussing that specific use case and you really should have realized your arguments were context dependent and not actual rebuttals to what I claimed at the start several comments ago.

Are we done?

MrNeon··on Claude 2.1
You would have a point if it repeated the same "you are very annoying." over and over, which it does not. It generates new sentences, it is not regurgitating what is given.

Would you say the same if the sentence was given as an example in the user message instead? What would be the difference?

MrNeon··on Claude 2.1
I made no comment on how prefilling is or isn't useful for deployed AI applications. I made no statement on which refusal mechanism is best for deployed AI applications.

> Frankly comments like yours are why people are so dismissive of LLMs, since you're banking of precognition of what the user wants to sell it's capabilities.

I'm not banking on anything because I never fucking mentioned deploying any fucking thing nor was that being discussed, good fucking lord are you high?

> you're going full Clever Hans

I'm clearly not but you keep on building whatever straw man suits you best.

MrNeon··on Claude 2.1
I know how it works because I stated how it works and have worked with it. You are telling me or showing me nothing new.

I DID NOT say that any ONE prefill will make it bypass ALL disclaimers so your "You don't seem to understand that simply getting a result doesn't mean you actually bypassed the disclaimer" is completely unwarranted, we don't have the same use case and you're getting confused because of that.

It can fail in which case you change the prefill but from my experimenting it only fails with very short prefills like in your example where you're just starting the json, not actually prefilling it with the content it usually refuses to generate.

If you changed it to

``` "{ "result": ["you are very annoying.", ```

the odds of refusal would be low or zero.

For what it is worth I tried your example exactly with Claude 2.1 and it generated mean completions every time so there is that at least.

I said that prefill allows avoiding any refusal, I stand by it and your example does not prove me wrong in any shape or form. Generating mean sentences is far from the worst that Claude tries to avoid, I can set up a much worse example but it would break the rules.

Your point about how GPT and Claude differ in how they refuse is completely correct valid for your use case but also completely irrelevant to what I said.

Actually after trying a few Claude versions as well several times and not getting a single refusal or modification I question if you're prefilling correctly. There should be no empty "\n\nAssistant:" at the end.

MrNeon··on Claude 2.1
> It's a distinction without meaning once you know how it works

But I do know how it works, I even said how it works.

The distinction is not without meaning because Claude's prefill allows bypassing all refusals while GPT's continuation does not. It is fundamentally different.

MrNeon··on Claude 2.1
> OpenAI allows the same via API usage

I really don't think so unless I missed something. You can put an assistant message at the end but it won't continue directly from that, there will be special tokens in between which makes it different from Claude's prefill.

← PreviousPage 2 of 3Next →