> I could walk you through each part of the image format and tell you how it impacts the output rendered.
I'd be quite impressed if you could correctly hand-decode a five-meg JPEG in less than a workday. You do get that I'm not talking about having an understanding of the file format, but actually being able to convert the encoded data into human-readable [0] output?
> Your post is conflating lossy discarding of information with extracting abstract concepts from information and encoding them into weights. This is why those weights do absolutely nothing until you run a prompt through them.
I can play that game too. The JPEG process extracts perceptual shorthand from information and encodes that into "quantized coefficients". These "quantized coefficients" do absolutely nothing until you run them through a "reconstituter".
Most things sounds quite high tech when burdened with new jargon. It's something you inevitably learn if you work at a Big Software Company for long enough.
> Those [incorrect claims and assertions] you [receive] are not hallucinations, they are the [LLM] just filling in for missing details.
FTFY
> ...take a lossless image like a BMP and zoom in, you'll see those blocks again!
I take it you've never seen a highly-compressed JPEG?
> ... if you think lossy compression is sufficient to avoid the "verbatim" requirement of copyright claims...
Quite the opposite. It's why I even bother bringing up the fact that LLMs are the output of lossy data compression programs.
> I'm not sure what those links are in relation to?
Go back and re-read the paragraph that referred to them, and then consider it how it and the posts relate to the quote that sits right before it. Someone who suggests that they can quickly and accurately hand-decode a non-toy JPEG file definitely has the capacity to read and understand ten-ish Mastodon posts.
Here's a hint to prime your intuition pump: The linked posts are about plagiarism generated by LLM-based tools.
[0] ...in the case of picture data, "convert into human-readable output" means "turn the data back into a picture"...