HNHacker News
TopNewBestAskShowJobs

miraculixx

148 karma · joined March 31, 2019

submissionscomments
miraculixx··on "Next-token predictor" is the wrong mental model for LLMs
I like to think of LLMs as informed dice throwing
miraculixx··on AI;DR (AI; Didn't Read)
Close the PRs as unacceptable.
miraculixx··on AI;DR (AI; Didn't Read)
Who'd want to work in an environment like this? It's not sustainable
miraculixx··on AI;DR (AI; Didn't Read)
Stop accepting these PRs. Ask for a personal 1:1 explanation. Make it the submitter's problem, not yours.
miraculixx··on Eight Myths on Software Engineering and GenAI
So then they should get 10-100x more revenue, now that AI does all the marketing and selling.
miraculixx··on I'm a photographer. I built a DSL for multi-agent workflows
Interesting! However there is a somewhat misleading statement in the readme where it says "no code required". Well the workflow specs are in fact code. Other than that I applaud the approach.
miraculixx··on Project Glasswing: what Mythos showed us
Did they compare it to other models? A lot of this sounds like this is the first time they have applied AI to security, and they are just amazed at the unreasonable performance of a pattern matching machine. Well, it matches patterns. duh
miraculixx··on I'm going back to writing code by hand
welcome to the club :) I came to the same conclusion a year ago and uninstalled all the AI assistants that my IDE tried to force on me. Back to good old auto complete and it works great. Feeling productive and on top of things.
miraculixx··on Copilot edited an ad into my PR
So you continue to show ads to Copilt, just not to the user? If so, not a fix.
miraculixx··on Tell HN: Litellm 1.82.7 and 1.82.8 on PyPI are compromised
This is interesting. How do you keep this up to date so quickly?
miraculixx··on Tell HN: Litellm 1.82.7 and 1.82.8 on PyPI are compromised
I agree in general, but how are you ever upgrading any of that? Could be a "sleeper compromise" that only activates sometime in the future. Open problem.
miraculixx··on When AI writes the software, who verifies it?
Sure because it worked great when we tried last time, right? Just spec it out first, and AI will churn out the perfect app. Not.

For anyone who hasn't worked in a waterfall project and would like to try: You are kidding yourself.

There is no such thing as a perfect spec. Read that again and say it outloud.

It took humanity 50 years to figure out that perfect specs are impossible, unless of course you know exactly(!) what you need. And even then, the specs are never complete.

The reality is that we often don't know, and can't know, what we want, exactly, until we actually see and experience what we said we wanted. Then we adjust. Try again.

That's the reality for individuals already. By simply using logic we can deduct that entities made of more than one individual, aka companies, will not behave better. They just make it look better by giving you a nice document that says "we want this!", only to then come around when they see what they got, and to claim "wait, we didn't mean it like so!".

That's just human nature. Not much we can do. AI will not change that.

So when some people think AI will deliver perfect software given perfect specs, hence we have to write the specs first! That is just missing the boat my a mile.

Agile is not a process but a human-friendly way of doing things. It simply says hey you want this? Let me build that and show you. Then let's adjust or move to the next thing. Rinse and repeat. Agile works because it matches how humans think and act. Step by step, day by day.

miraculixx··on X has under 30 engineers
can someone verify that?
miraculixx··on The insecure evangelism of LLM maximalists
Simon just explores stuff and writes about them. Doesn't mean he uses LLMs for everyting.Antirez likes to question stuff and make them better. Doesn't mean he uses LLMs for everything.

Also their experience is not my experience. I will make my own choices.

miraculixx··on PYX: The next step in Python packaging
Just stick with pip and venv.
miraculixx··on PYX: The next step in Python packaging
Windows is the root cause here, not pip
miraculixx··on PYX: The next step in Python packaging
+1
miraculixx··on PYX: The next step in Python packaging
+1
miraculixx··on PYX: The next step in Python packaging
Anaconda solved the same problem ~10+ years ago already.
miraculixx··on ETH Zurich and EPFL to release a LLM developed on public infrastructure
That is not correct. The EU AI Act has no such provision, ans the data mining excemption does not apply as the EU has made clear. As for Switzerland copyrighted material cannot be used unless licensed.
miraculixx··on The Path to Medical Superintelligence
Are these 56 cases distinct from all other cases in the data?
miraculixx··on The Path to Medical Superintelligence
Exactly. The study has been set up to produce this exact result. They essentially limited the human doctors to bare essentials, on specialist cases(!), while providing the LLMs with all sorts of help, including discussion among several AIs.

That's like letting one group of students have a strict closed-book exam, while another group can take the test as a group exercise and accessing any material they like, then claiming that closed-book exams lead to worse outcomes.

In a nutshell the study is just slop designed to get attention. The headline result is what they really want people to hear, and that's all the media will be repeating.

miraculixx··on The Path to Medical Superintelligence
As any AI researcher knows, if you have a model that does 4x better than the naive baseline (the humans, in this case), you are likely looking at overfit, not real-life performance. This study is just slop, and you can tell so by the mere fact that they did not submit a paper, but just published a PR article.
miraculixx··on Tracing the thoughts of a large language model
To fly means "to soar through air; move through the air with wings" (etymonline)

That is pretty much an accurate discription of what planes and birds do.

To plan means "to reason with intent".

That is very much not what LLMs do, and the paper does not provide evidence to the contrary. Yet it uses the term to give credence to it's rather speculative interpretation of observed correlation as causation.

Interestingly enough there is no definition of the term, which at least would help to understand what the authors actually mean.

I would be more inclined to take a more positive stance to the paper if it used more appropriate terms, such as call observed correlations just that. Granted that would possibly make for much less of a fancy title.

miraculixx··on Tracing the thoughts of a large language model
Very much in support of this. The use of anthropmorphic or even biological terms are entirely misguided. All they do is drive a narrative that is very much belitting natural intelligence.
miraculixx··on Tracing the thoughts of a large language model
Unfortunately that's not a review but a hype-driving oversimplification of what the paper says.
miraculixx··on Tracing the thoughts of a large language model
"Hallucination" is just to term we use to say "this result is not what it should be". The model always uses the very same process, it does not do one thing for "hallucinations" and something else for "correct" results.

In a nutshell it is always predicting the next token from a joint probability distribution. That's it.

All other interpretations are speculative.

miraculixx··on Tracing the thoughts of a large language model
> This doesn't seem to happen very often in classical programming, does it?

Try concurrent programming. It happens all the time.

miraculixx··on Tracing the thoughts of a large language model
We know how they work. It's just that it works better than expected. Which of course doesn't mean we don't know, it just means there are second-order effects that are non-obvious. Intelligence is not implied.
miraculixx··on Tracing the thoughts of a large language model
There is no evidence to this end. There is lots of claims, but claims are not evidence.
Page 1 of 6Next →