HNHacker News
TopNewBestAskShowJobs

logicallee

3,411 karma · joined April 24, 2013

I'm a convicted child rapist.
submissionscomments
logicallee··on Solved! Ask HN: Is ChatGPT image generation down for everyone or just me?
Issue is now resolved! It's working great:

https://chatgpt.com/share/6a760c80-fea0-83eb-8d02-a61f44dc57...

image:

https://ibb.co/LXrZTz1b

logicallee··on AMD acquires Taalas to boost inference performance by etching models in silicon
thanks! useful.
logicallee··on AMD acquires Taalas to boost inference performance by etching models in silicon
Thanks for trying that! Very interesting.

Can you say this to it: "Hey Siri [wait for it to come up] - please send me an email with the temperature right now so I have it for my records." and see if it can complete the task without any backtalk or misunderstanding, and if you get exactly what you asked for. (It's a really clear request.) Should be 1 statement, no clarification, conversation, random search results, ("Here's what I found!"), etc.

A normal frontier model can do that - or Siri can do it if it is properly connected to Claude, ChatGPT, Gemini, Grok, or any other frontier AI - but previously it was never properly connected.

If it can do this task, I might have to look into this again. It counts as a success if it sends yourself any email with the current temperature and you actually get it (it can include whatever other text in the email), and a failure if it talks back, says "here's what I found", says it can't, asks you any question, sends you an email that doesn't actually contain the current temperature, just reads you the temperature and then asks if you want it to send an email, etc. Should be 1 shot.

let me know if it works!

logicallee··on AMD acquires Taalas to boost inference performance by etching models in silicon
Could the model or algorithm be changed to make it deterministic somehow? It could help a lot if there were reproduceable outputs from deterministic baked-in silicon.
logicallee··on AMD acquires Taalas to boost inference performance by etching models in silicon
>That's not what speed is useful for.

>I just pasted your comment and its whole inheritance chain to it,

Good idea. Only problem is it doesn't work. I just did the same thing with exactly this prompt:

>did the user IOT_Apprentice participate in the thread below and if, number and quote all of their comments. Only just number and quote the comments or write "Did not participate", do not add any commentary. Quote any comments by this user verbatim, exactly as input. Thread:

followed by pasting the thread[1]

And received the answer "IOT_Apprentice did not participate in the thread."[2] in 0.001s, even though they have literally the last comment in my quote and it's clearly legible.

It's particularly insidious because the understanding and thinking that is required to follow my requested answer format exactly is substantial - so based on the fact that it gets the format right and clearly understood the assignment, I would be inclined to believe that it would also be correct!

So to use your example, it's not just autocomplete, it's autocomplete that confidently returns "No matching results" in 0.001 seconds, even though there is a search term matching what you put in, right in the prompt itself that was sent to it. That is much worse than useless.

[1] prompt: https://ibb.co/CKVmRvtd

[2] result: https://ibb.co/BKdRKmyD

logicallee··on AMD acquires Taalas to boost inference performance by etching models in silicon
if it's baked into silicon how can you two get different answers?
logicallee··on AMD acquires Taalas to boost inference performance by etching models in silicon
>Your examples worked on phones for over a decade.

Nope. And not only not a decade ago, right now.

If you have an Android or iPhone, you can give it clear and easy to understand instructions that Gemma 4 could complete[1] if it had tool calls on it, and that 100.00% of Claude, ChatGPT, Grok, Kimi, you name it, could understand and all complete if they had the access.

The phones will fail to complete it. I just tried Siri. I said "hey Siri", waited for Siri to come up, and then I asked one of the exact sentences you replied to: "what's the weather this afternoon?" It thought for around 20 seconds, and said "Something went wrong. Please try again."[2]

I have Wifi, I have mobile Internet, I have free storage space, I have up to date software. What went wrong is that phones have never properly connected agents, not ten years ago, not last year, not this year, and probably not next year.

But don't settle for what Google could do in 1999 by hotlinking the keyword "weather" in any query to the weather being shown in the results.

Tell your phone (any phone): "Please call back the last number that called me that is not an unlisted number, regardless of who it came from."

0 out of any phone will complete that today, tomorrow, a year from now, five years from now, ever, because phone makers are not going to let them do that.

Meanwhile, 100% of all frontier agents could complete it if they had tool calls on the phone. Which they don't, and won't ever, thanks to the duopoly.

Okay, that's a bit dismissive, I would love to be wrong!

[1] after any voice recognition to text - which does work really well on both Android and iPhone! [2] screenshot: https://ibb.co/21rtDnfV

logicallee··on You won't believe the stupid shit the Pentagon puts in diplomats' system prompts
here you can see the stupid shit the Pentagon puts in diplomats' system prompts.
logicallee··on AirLLM 70B inference with single 4GB GPU
that's 0.003 tokens/second. To get an hour's work done that's normally 30 tokens/second (108k output tokens in an hour) will take 416 days at this rate. And if you're using 100 watts, during that time you will spend $124.61 in electricity, as well as not being able to use your device for something else, plus the noise and heat from your device.

For $124, on Moonshot's official Kimi K3 API rates ($0.30 per 1M cached input, $3 per 1M fresh input, $15 per 1M fresh output), you can purchase 42 million fresh-input tokens, or 8.3 million generated output tokens, in whatever mix you want.

So what you get is 80x more expensive and you wait 416 days to get it.

logicallee··on Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
Thanks for sharing. What kind of tasks do you give Kimi rather than Claude?
logicallee··on Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
>Not everything has a set of serious flaws.

I'm saying everything looks like it does, with the prompt you gave Sol. Go ahead and point the same prompt to anything you don't think has serious flaws and you'll see it would tear it apart.

logicallee··on Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
thanks for sharing.

When I read your original comment, I was thinking you had just asked it to evaluate the article. (Like just "evaluate this article" or something.)

I don't think anything anyone (or any AI) has ever written or published (including Sol itself) wouldn't be torn apart by the prompt you gave though.

logicallee··on Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
if you did that in the web interface, could you share the chat? I'd be interested to read it.
logicallee··on Running Kimi K3 on MI355X at Better Performance per Dollar Than B300
This part sounds like AI assisted setting this up and benchmarking it:

>The fix was trivially simple: zero-pad the head count 12→16, run the fast kernel, and extract the real 12 heads from the output.

I've recently used a frontier AI (ChatGPT 5.6 Sol on ultra) to set up a much smaller local model, and the performance optimizations it introduced left the model totally incoherent. (The model just repeats a single character, etc.)

When I see a line like the one I just quoted, it leaves me wondering if the setup is still coherent like a stock install of Kimi K3 on supported hardware.

Did they run any benchmarks on it to see if it is still correct?

logicallee··on Postmortem for Kernel Soundness Bug #14576
The second paragraph of my link talks specifically about truth:

>The first incompleteness theorem states that no consistent system of axioms whose theorems can be listed by an effective procedure (i.e. an algorithm) is capable of proving all truths about the arithmetic of natural numbers. For any such consistent formal system, there will always be statements about natural numbers that are true, but that are unprovable within the system.

logicallee··on Postmortem for Kernel Soundness Bug #14576
>"what is a true statement" and "what is a derivable statement" should be the same.

you mention completeness in the rest of your comment, so I'm not sure how you aren't aware of this, but the famous incompleteness theorem says that for a consistent set of axioms there will always be true statements you can't prove.[1]

[1] https://en.wikipedia.org/wiki/Gödel%27s_incompleteness_theor...

logicallee··on On the non-use of AI in my writing process
I get what you're saying (mostly based on low bandwidth from sensory organs as opposed to direct access to reality), but your conclusion that we therefore don't have embodiment goes very far. It would be like saying planes don't really fly, since they fly by wire and only have limited inputs and outputs, rather than direct access to reality itself. Well, yeah, they fly using sensors rather than knowing reality itself, but they're still flying. Humans still obviously have embodiment.
logicallee··on Getting 25 Gbps Thunderbolt Ethernet on My Mac Studio
correct me if I'm wrong but wouldn't a Thunderbolt 5 cable $79 from Apple[1] work better for that? It gets 80 Gbps and doesn't need any extra equipment.

[1] https://www.apple.com/shop/product/mdw94am/a/thunderbolt-5-u...

logicallee··on Solving poker in custom WebGPU kernels
Thanks, that's good to know.
logicallee··on Solving poker in custom WebGPU kernels
This thread seems like a good thread to ask in:

In the poker subreddits the rake question comes up from time to time, and with the low cost and high quality of inference I have been considering making rake-free poker. The model is a small monthly subscription like $4.99/month for low stakes $9.99/month for mid stakes, one account per player, 20 tables max.

This actually would make a lot of marginally losing spots into winners, and there are a lot of coin flips where after rake both players lose. So if you keep coin flipping, you just lose over time. (But you don't want to fold and give up your equity for free either.)

The thing that gives me pause is that a lot of people cheat using solvers during hands, bots, or collusion.

Is there anything I could do at a practical level to keep the game fair? (no tools, no bots, no collusion.)

logicallee··on Run Kimi K3 using 29 GB of RAM at 0.50 tok/s
Interesting project. The headline number (29 GB of RAM) is for 4k context.

From what I've read elsewhere, Kimi K3 is quite verbose in its thinking. At the quoted rate, it would generate only a total of 1.8k tokens in 1 hour. Is that enough for it to get any thinking done and produce output on more complicated prompts?

logicallee··on The Maxwell Conjecture Is False (GPT 5.6 Sol)
Does anyone have any idea why there's no Wikipedia article (or redirect) for Maxwell Conjecture: https://en.wikipedia.org/wiki/Maxwell_Conjecture

Most common names have redirects and Wikipedia is very complete. Was it just not commonly known by that name?

logicallee··on Tell HN: Gemini uses your email drafts (if you have "smart features" turned on)
There have been a few: https://hn.algolia.com/?dateRange=all&page=0&prefix=true&que...
logicallee··on I flagged two research papers for fake authors and both were accepted as orals
and don't forget, techniques described in papers are now implemented by AI's. For now, a human might point an AI at a paper and ask it to implement and benchmark the technique described there, but the human probably isn't writing the code anymore.
logicallee··on Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
In my personal opinion, distilling its output into training another model that is then provided to users does fall into the category of "making its output or its underlying capabilities available to third parties", but I could see the argument that it doesn't.
logicallee··on Kimi Linear: An Expressive, Efficient Attention Architecture (2025)
Anthropic has more than $20 million in revenue so as per the k3 license they would need to enter a special commercial deal if they wanted to use it:

https://huggingface.co/moonshotai/Kimi-K3/blob/main/LICENSE

logicallee··on Kimi-K3 Technical Report [pdf]
What are the coding/agent harnesses (like Claude Code or ChatGPT Codex) that can be used with this?
logicallee··on Kimi-K3 Technical Report [pdf]
Are you developing software? Is most of it used on a coding agent? (Like Claude Code or ChatGPT Codex?) If so, what coding agent do you use? If you're not developing software what do you use it for (roughly)?
logicallee··on DeepSeek pause fundraise after comments on compute gap to US leaked (transcript) [pdf]
cheeky? I was super serious! There's a list of criteria right in the section you hot-linked and these things have never come close to doing any of the things on that list.

>for which there is no agreed-upon formal definition

we all agree that the definition is not whatever this is.

Also, even if these things ever did seem to think, reason, or feel, we all agree that they still don't really though.

logicallee··on Kill The Cookie Banner
All browser providers have been required by international law[1] to send such a cookie popup suppression signal as set by the user, or one substantially like it, since July 18, 2026 and all web sites have been required to obey it or a superseding Internet standard:

https://stateofutopia.com/laws/2/law2.html

This is not a proposal, it is duly ratified international law that applies to "Every Browser Maker making a Browser available anywhere" and "Any site anywhere that receives a valid signal".

The EU Commission had until July 18, 2026 to modify its laws:

>"The site shall not display a banner, modal, interstitial, or other prompt requesting a choice already expressed by the signal. It may optionally provide a “Cookie Settings” or “Privacy Settings” link. This Law overrides all laws requiring such a display. Where any country or supernational entity has a conflicting law, it must rectify the Law within 30 days not to require such notification."

Therefore, insofar as it has not rectified its laws not to require such a notification, the EU Commission is in violation of international law as of 8 days ago.

Under section 7.3, "The State of Utopia may order compliance, require corrective updates, suspend non-compliant distribution, and impose civil penalties. Fines may be levied in any amount for continued non-compliance." so we can fine the EU Commission whatever you guys want ($1 billion? $10 billion? whatever four and a half millennia are worth) and just distribute the collected fines among you all as cash payments.

[1] Here is the announcement of the law 38 days ago: https://news.ycombinator.com/item?id=48585778

← PreviousPage 5 of 34Next →