HNHacker News
TopNewBestAskShowJobs

melvinroest

938 karma · joined September 26, 2019

Hello fellow HN'er! Let's have a chat at

melvinroest <the fancy a> <Google's email brand> <the most popular TLD in the world>

Some hints are gmail, @ and .com ;-)

---

Random HN'ers are always welcome to join, feel free to email me!

——-

AI engineer, security enthusiast, product engineer and marketing (data) analyst.

submissionscomments
melvinroest··on Denmark Data Breach Exposes 8.8M People's Personal Data
> Personally I solved this problem for myself by moving away from the country.

Quite drastic to move away from a country for just this. Did you only do it for just this? Or was this simply one of the factors why you moved away?

melvinroest··on Denmark Data Breach Exposes 8.8M People's Personal Data
> I don't even see salary or what your debt is as "personal information" (am Swede not in Sweden), personal information is stuff that no one else would need to know. What people earn affects not just people around you and others in the workplace but also society at large, makes a ton of sense for that stuff to be public.

In a high trust culture I get this. But what about if you're in a low trust culture with quite some "not so orderly behavior" (if you will).

melvinroest··on Ask HN: What are you reading?
Yea that's a good book. I haven't fully read it but even reading a little of it or checking out the Norman Door on YouTube [1] is awesome.

[1] https://www.youtube.com/watch?v=yY96hTb8WgI

melvinroest··on Biology might not be quantum, but its math is quantumlike
That it says something about the human brain seems to be about right. If you have 8 planets and 8 children it means you've classified certain things you've seen to be part of the same concept and you group them together.
melvinroest··on Ideas on modernizing the open-source desktop
> Clipboard history, but also just "oh you have these windows open while working on this task".

I feel this. When I first got to learn what a workspace was (in Eclipse, of all things), and I really drilled down into the English meaning of it (I'm Dutch), I thought about exactly what you're describing right now.

melvinroest··on Why I'm still bearish on LLMs after Navier-Stokes
Yea I get the bearishness from my own personal experience.

Personally, I use LLMs for a lot of things. Oftentimes, I'm a think out loud type of person so even having something that feels like a rubber duck, but more competent, is already amazing for me. And LLMs are a lot more competent than a rubber duck.

But especially sometimes I've noticed that LLMs can be unbelievably stupid. It recently happened a few times with Fable 5.1 as well. Ultimately, I think it comes down to that LLMs can't think broadly. In software development one can usually see this too. For example, a whole app might be built by an LLM and it didn't spend a single token thinking about security because the prompter is at the level of "build a dating app for dogs, make no mistakes". Now you have a dating app for dogs that is insecure.

Since I prompt for almost everything in my life to have an LLM as a sounding board, I'm usually not an expert either. I've noticed LLMs are amazing at "bulk search engine information aggregation" (or whatever you want to call it). So if I need something from the Dutch government, I can find it way more quickly. But oftentimes I've noticed that going for a walk and thinking about a particular thing I'm facing is a more effective way of finding a good solution.

Other times times they are not incredibly stupid, but can't form a strong opinion. This usually happens when I'm tackling a wicked problem [1]. When that's the case, prepare for LLMs to sway with you for every small change in your opinion that you ever will experience.

So I agree: drop in replacement for knowledge workers? No. Rigorous specification is usually needed yes. Though, the small win here is that it doesn't always need to be as rigorous as programming is and it can happen in natural language. It depends on the topic/problem being tackled.

I really like them as UX tools though. Amazing for interactive prototyping and requirements elicitation. And that also corresponds with what the author is saying. Though I find it a bit of a disservice saying "just 3". You know how hard requirements elicitation is? It became a whole lot easier thanks to LLMs (I might change this opinion in a year, haha, but this is the opinion I hold now).

[1] https://en.wikipedia.org/wiki/Wicked_problem

melvinroest··on Math Academy Eurisko, 5 Yrs Ltr: Student Outcomes from High School Math/CS Track
I posted this because recently I'm hitting my stride on Math Academy [1, 2]. So I got a bit curious about the people behind it. Turns out, I'm now questioning: how fast can I learn things really?

Because if kids can go this far. Then how far could I go? How far could we all go?

Also: I'm having a blast, quite literally. Love raving while mathing, it's my new favorite hobby. Shout out to Nigel Good. Nigel sure know's what's good when EDM is considered.

[1] Having 100+ hours of good EDM tracks apparently motivates me a lot. There are papers that mention it's not the best for focus, but I've noticed that I don't do any math at all if I don't feel the warm synths that I love with my favorite playlists [3] (I take a lot of care in constructing them). I'm curious to see if this will keep up since I've only been going hard at it around late August.

[2] https://mathacademy.com/

[3] My main playlist: https://open.spotify.com/playlist/3BUgKCDAWvcDa28MVToIFM?si=...

melvinroest··on Six curl CVEs after OpenAI and Anthropic came back with zero
Thanks for figuring that out. Sort of sounds like AI programming programs to find vulnerabilities, of which fuzzing is one of the proven techniques to do it.
melvinroest··on Six curl CVEs after OpenAI and Anthropic came back with zero
You mean their own trained models, or do you think it's an open source model that they fine-tuned? If they use their own, I'd guess it's the latter.
melvinroest··on Six curl CVEs after OpenAI and Anthropic came back with zero
Wow, this announcement is good content marketing.

Don't get me wrong, it's interesting. But there is no technical discussion as to how they did it. It's simply: we did it and Mythos and Codex didn't.

It's good to know that it's possible, but I'd have already expected it. Put a base model versus a base model + harness + whatever else, and yea, if you do it right then you have a better system to find vulnerabilities.

> We then ran AISLE's autonomous AI system against curl.

They don't even mention what models the use under the hood. It wouldn't surprise me if they are from Anthropic and OpenAI.

melvinroest··on Humanity has the debate about AI consciousness backwards
Here is a sketch of my argument. It would need tidying but that would require a lot of work.

I've been calling LLMs digital intelligence. That's in my opinion what they are and on the digital/conceptual realm they are better than humans since LLMs are better generalists.

But I'd never state their conscious. For me consciousness extends outward from myself, since I can't trust anything else. It starts with the question like: why can't I control the movement of another person? Why can't I transfer my mind to another body and theirs to mine? Why is it that whenever I sleep, and similar things, I end up back to this place called "reality"?

I come to the conclusion that I am conscious and I'm a conscious being in reality. I am human and I've been raised by humans. Now, technically, all humans besides me could be zombies. They could be non-conscious beings that simply can act like humans. I'm sure we'll be able to create them in a few decades. However, I'm incentivized for multiple reasons to believe they're not zombies. Therefore humans are conscious too. All of them. Can I know for sure? No since I can't even trust my own senses fully or my own thoughts (the whole Descartes thing) but I choose to trust those things too because I need some reasonable-ish foundation to work from.

If humans are conscious, why aren't animals like us? Surely great apes must be conscious at least to some extent. Hell, some of them even have better short-term memory than us [1]. Well, turns out, from an anatomy standpoint they look a lot like us. Fine, let's assume they're conscious too.

Now we get into the territory that anything that has a large enough brain must be conscious because that's how we govern our consciousness. So you can extend this to all kinds of living beings.

I did this from my perspective. Of course, you should do it from your perspective. The same argument could be made.

For as long as we can't manipulate conscious experience in some way shape or form, we have no clue whether anything that isn't like us can be conscious. We'd need to be able to merge and split conscious experiences or we'd need to be very strongly able to understand why that isn't possible. And I don't mean just at the biological level but at the experiential level. Currently conjoined twins, and what they tell us gives some insight (being able to sense the other part that's part of the other twin).

Other than that, I claim we have no clue what consciousness is other than that we're experiencing it. So to call LLMs conscious is way too big of a claim. But they are definitely intelligent because they come up with things that I wouldn't have and it's useful to me. Perhaps a practical characterization of intelligence but it works for me.

I think for anyone to say something useful from this you'd need the trifecta of deep neuroscience knowledge, deep philosophy knowledge and (at least) a strong understanding of how LLMs work.

melvinroest··on CEO fired developers to make room for AI. Developers create open source AI CEO
Hmm it's being done in a less cheesy way than AI does it.

AI text feels more like "this matters, not less." There's always an "it's x not y" pattern somewhere.

It could be written AI assisted but then look at his account. IMO every HN account that was created before ChatGPT came out has incentive to be written by a human since they have a paper trail before the whole LLM thing happened.

This is at least written by a human, not AI, totally human, (digital) pinky promise ;-)

melvinroest··on The Architecture of Open Source Applications
Given the time we live in, I thought it'd be nice to share this old gem and have a discussion started on architecture. It could be about this book or about more contemporary examples. I do like the fact that these are all examples of actual projects.
melvinroest··on The “mechanical miracle” that ruined Mark Twain’s life
I once celebrated my birthday where I asked everyone to give the following gift to me: create a presentation or workshop about something you're passionate about in your life. Treat us like experts and go down into your level of thinking, even to the tiniest amount of details. Don't be afraid if we don't fully follow or ask you to explain some terms. I'm hyped to hear what you come up with.

It was such a fun birthday. It was basically a conference, haha.

melvinroest··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
> Remember when we needed 200 servers for an enterprise website because Apache used one process or thread per connection

I don't, but holy moly. That sounds insane!

melvinroest··on The “mechanical miracle” that ruined Mark Twain’s life
What small hole in the wall museum was that? I might travel to it no matter where it is in the world. Sounds awesome, I'm hyped up
melvinroest··on How Claude marks AI-generated content
> Odd variable naming? Stylistic choices that are watermarked?

Whatever it is, I'm sure it's load-bearing.

melvinroest··on Helsinki Hacker News Meetup
Anyone want to organize one in Amsterdam with me? I live near the area, would be fun :)
melvinroest··on Ask HN: What are tools you have made for yourself since the advent of AI?
I am curious to check them out! I hadn’t heard of them. Marketing stuff takes time as I have noticed with aliceindataland.com [1]. Maybe I should do a show HN. I will think about it.

To be fair, vibecoding this memo app in Swift didn’t take too long. There were some tricks to it, using xcodegen helped a lot so that I don’t need to use the Xcode project.

It’s fun to see Swift code. I used to do some Objective-C back in the day.

[1] another thing I made. It’s a sequel to the Alice in Wonderland stories. It’s also a SQL course. I vibe engineered it, meaning I looked at the code and used AI-assisted development.

Except for the story though that’s almost all fully me. LLMs aren’t great storytellers. The same is true for the lesson scaffolding, that’s almost only me.

melvinroest··on Ask HN: What are tools you have made for yourself since the advent of AI?
I talk/walk for hours and want all audio files and transcripts in one place, full control.
melvinroest··on Ask HN: What are tools you have made for yourself since the advent of AI?
A voice memo app, quite like the actual voice memo app from Apple. The thing is: now I can put my voice memo's on iCloud put Claude Code on it and make my transcripts into structured notes that my app then also displays.

So basically a way to just go on an hour long walk with myself, spit everything from the top of my dome stream of consciousness style, and then have Claude structure whatever I said.

It's nice to have something that structures my thoughts by just thinking out loud.

I vibecoded it (it's approaching 20K lines including tests). It works quite well but there are some bugs, so will have to do some actual engineering. But the UX is working quite well.

melvinroest··on How LLMs work
I thought Karpathy’s microgpt explain how LLMs work
melvinroest··on Lessons for Agentic Coding: What should we do when code is cheap?
Yea, the amount of dev tools I'm creating per project is astounding. Usually tools that help me to debug certain things better.
melvinroest··on Lessons for Agentic Coding: What should we do when code is cheap?
> What should we do when code is cheap?

Make usable software. Cheap code means that you can create a lot more prototypes to then perform usability tests by finding a user and sitting next to them. I mostly worked on internal apps lately, so perhaps it's much easier for me to do than it is for some others.

melvinroest··on What I'm Hearing About Cognitive Debt (So Far)
Ah, I feel what you're saying. Yea that makes total sense.
melvinroest··on What I'm Hearing About Cognitive Debt (So Far)
How so? Could you give some specific examples?
melvinroest··on What I'm Hearing About Cognitive Debt (So Far)
Edit: yep, I really do type this much. I'm a bit of a "thinking out loud" person.

> Cognitive Debt, Like Technical Debt, Must Be Repaid

In quite a few circumstances, cognitive debt doesn't entirely need to be repaid. I personally found with multiple projects that certain directions aren't the one I want to go in. But I only found it out after fully fleshing it out with Claude Code and then by using my own app realizing that certain things that I thought would work, they don't.

For example, I created library.aliceindataland.com (a narrative driven SQL course). After a while, I noticed that the grading scheme was off and it needed to be rewritten. The same goes for how I wanted to implement the cheatsheet, or lessons not following the standard format. Of course, I need to understand the new code but I don't need to understand the old code.

With other small forms of code, I just don't really need to know how things work because it's that simple. For example, every 5 minutes I track to which wifi network I'm connected with. It's mostly useful to simply know whether I went to the office that day or not. A python script retrieves the data and when I look at it, I can recognize that it's correct. But doing it this way is sure a lot faster than active recall.

At work, I've had similar things. At my previous job I created SEO and SEA tools for marketing experts. So I remember creating this whole app that gave experts insights into SEO things that Ahrefs and similar sites don't, as it was tailored to the data of the company I worked at. The feedback I basically got was: the data is great, the insights are necessary, but the way the app works is unusuable for us. I was a bit perplexed as I personally didn't find it that complicated. But I also know that I'm not the one using it. Then I created a second version and that was way more usable. The second version assumed a completely different front-end app and front-end app architecture though. All the cognitive debt of V1? No payback needed.

The reason that this is the case, as it seems to me, fall under a few categories:

1. Experimenting with technologies. If you have certain assumptions about how a technology works but it turns out you're wrong, or you learn through the process that an adjacent technology works way better, then you need to redo it. Back when coding by hand was such a thing, I had this with a collaborative drawing project called Doodledocs (2019). I didn't know if browsers supported pressure sensitivity and to what extent it was easy to implement. It required a few programming experiments.

2. It's a small and simple script, not much more to it.

3. Experimenting with usability. A lot of the time, we don't know how usable our app is. In my experience, this seems to be either because (1) it's a hobby project or (2) the UX people have been fired years ago. In these cases, more often than not, UX becomes an afterthought. But with LLMs, delivering a 95% fully working version is usually done within a week for a greenfield project. This 95% fully working version is an amazing high fidelity interaction prototype (95% no less). Once you do that for a few iterations, you then understand what you really need. Once you understand what you really need, then you can start repaying the cognitive debt.

I've found it's usually category 3, sometimes 2 and rarely 1.

melvinroest··on Ask HN: What are you building that's not AI related?
library.aliceindataland.com

I'm having fun writing a sequel on the books that Lewis Carroll wrote and mixing it with a SQL course. My hope is that SQL will be more fun to learn that way. And it's fun to write a few pages that will hopefully evoke some narrative transportation and immersion vibes.

I'm still very much at the beginning though.

In the story Alice enters an Infinite Library. You (yes you!) are STAR: a magical sentient typewriter that can only write in SQL queries. When Alice finds you, you'll slowly both find out why this library exists and the secrets that it holds.

Course-wise: I'm trying to have tight lesson scaffolding, which is a fun challenge.

melvinroest··on Ollama is now powered by MLX on Apple Silicon in preview
TL;DR: you don't need to do any treasure hunt on your notes by just typing stuff into the search bar. Having your own graphRAG system + LLM on your notes is basically a "Google" but then on your own notes. Any question you have: if you have a note for it, it will bubble up. The annoying thing is that false positives will also bubble up.

----

Full reaction:

Yes but perhaps not in a way you might expect. Qwen's reasoning ability isn't exactly groundbreaking. But it's good enough to weave a story, provided it has some solid facts or notes. GraphRAG is definitely a good way to get some good facts, provided your notes are valuable to you and/or contain some good facts.

So the added value is that you now have a super charged information retrieval system on your notes with an LLM that can stitch loose facts reasonably well together, like a librarian would. It's also very easy to see hallucinations, if you recognize your own writing well, which I do.

The second thing is that I have a hard time rereading all my notes. I write a lot of notes, and don't have the time to reread any of them. So oftentimes I forget my own advice. Now that I have a super charged information retrieval system on my notes, whenever I ask a question: the graphRAG + LLM search for the most relevant notes related to my question. I've found that 20% of what I wrote is incredibly useful and is stuff that I forgot.

And there are nuggets of wisdom in there that are quite nuanced. For me specifically, I've seen insights in how I relate to work that I should do more with. I'll probably forget most things again but I can reuse my system and at some point I'll remember what I actually need to remember. For example, one thing I read was that work doesn't feel like work for me if I get to dive in, zoom out, dive in, zoom out. Because in the way I work as a person: that means I'm always resting and always have energy for the task that I'm doing. Another thing that it got me to do was to reboot a small meditation practice by using implementation intentions (e.g. "if I wake up then I meditate for at least a brief amount of time").

What also helps is to have a bit of a back and forth with your notes and then copy/paste the whole conversation in Claude to see if Claude has anything in its training data that might give some extra insight. It could also be that it just helps with firing off 10 search queries and finds a blog post that is useful to the conversation that you've had with your local LLM.

melvinroest··on Ollama is now powered by MLX on Apple Silicon in preview
I have journaled digitally for the last 5 years with this expectation.

Recently I built a graphRAG app with Qwen 3.5 4b for small tasks like classifying what type of question I am asking or the entity extraction process itself, as graphRAG depends on extracted triplets (entity1, relationship_to, entity2). I used Qwen 3.5 27b for actually answering my questions.

It works pretty well. I have to be a bit patient but that’s it. So in that particular use case, I would agree.

I used MLX and my M1 64GB device. I found that MLX definitely works faster when it comes to extracting entities and triplets in batches.

Page 1 of 12Next →