HNHacker News
TopNewBestAskShowJobs

comboy

6,618 karma · joined January 22, 2010

hn@comboy.pl
submissionscomments
comboy··on Book review: There Is No Antimemetics Division
It's not a well-written book. It's an interesting book (more like a story).
comboy··on Components of a Coding Agent
Program defines the exact computer instructions. Most of the time you don't care about that level of detail. You just have some intent and some constraints.

Say "I want HN client for mobile", "must notify me about comments", you see it and you add "should support dark mode". Can you see how that is much less than anything in any programming language?

comboy··on Components of a Coding Agent
The spec manually crafted the user is ideal.

It's just that we're lazy. After being able to chat, I don't see people going back. You can't just paste some error into the specs, you can't paste it image and say it make it look more like this. Plus however well designed the spec, something like "actually make it always wait for the user feedback" can trigger changes in many places (even for the sake of removing contradictions).

comboy··on Components of a Coding Agent
Hey, you seem to have similar view on this. I know ideas are cheap but hear me out:

You talk with agent A it only modifies this spec, you still chat and can say "make it prettier" but that agent only modifies the spec, the spec could also separate "explicit" from "inferred".

And of course agent B which builds only sees the spec.

User actually can care about diffs generated by agent A again, because nobody wants to verify diffs on agents generated code full of repetition and created by search and replace. I believe if somebody implements this right it will be the way things are done.

And of course with better models spec can be used to actually meaningfully improve the product.

Long story short what industry misses currently and what you seem to be understanding is that intent is sacred. It should be always stored, preferably verbatim and always with relevant context ("yes exactly" is obviously not enough). Current generation of LLMs can already handle all that. It would mean like 2-3x cost but seem so much worth it (and the cost on the long run could likely go below 1x given typical workflows and repetitions)

comboy··on Apple approves driver that lets Nvidia eGPUs work with Arm Macs
GPUs can do graphics too?
comboy··on Claude Code Unpacked : A visual guide
It's hard to tell how much it says about difficulty of harnessing vs how much it says about difficulty of maintaining a clean and not bloated codebase when coding with AI.
comboy··on Claude Code Unpacked : A visual guide
I mean, tools change, but I'd be happy to hear if any tool can create that by just saying create "Claude Code Unpack" with nice graphics. or some other single prompt. It likely was an iterative process and it would be lovely if more people started sharing that, because the process itself is also very interesting.

I've created some chinese characters learning website and I took me typing 1/3 of LoTR to get there[1]. I would have typed like 1% of that writing code directly. It is a different process, but it still needs some direction.

1. https://hanzirama.com/making-of

comboy··on Ollama is now powered by MLX on Apple Silicon in preview
As things stand today even when doing research tasks, time spent by model is >> than fetching websites. I don't see it changing any time soon, except when some deals happen behind the scenes where agents get to access CF guarded resources that normally get blocked from automated access.
comboy··on I put all 8,642 Spanish laws in Git – every reform is a commit
Add CI to check if new laws don't contradict with any existing ones.
comboy··on Apple says no one using Lockdown Mode has been hacked with spyware
*that we know of
comboy··on Hong Kong police can now demand phone passwords under new security rules
Haha, here's some random AI generated content:

    At least 225 judges have ruled in more than 700 cases that the administration's mandatory immigration detention policy likely violates the right to due process[1] The Fifth Amendment's Due Process Clause generally requires those having federal funds cut off to receive notice and an opportunity for a hearing, which was not provided in many of DOGE's spending freezes[2]
(there's more but what's the point)

1. https://www.justsecurity.org/107087/tracker-litigation-legal...

2. https://www.cbpp.org/research/federal-budget/many-trump-admi...

comboy··on $500 GPU outperforms Claude Sonnet on coding benchmarks
Yeah, good tests are associated with cost. I'd like to see benchmarks on big messy codebases and how models perform on a clearly defined task that's easy to verify.

I was thinking that tokens spent in such case could also be an interesting measure, but some agent can do small useful refactoring. Although prompt could specify to do the minimal change required to achieve the goal.

comboy··on $500 GPU outperforms Claude Sonnet on coding benchmarks
Not really related, but does anybody know if somebody's tracking same models performance on some benchmarks over time? Sometimes I feel like I'm being A/B tested.
comboy··on Schedule tasks on the web
People are loading huge interpreted environments for stuff that can be done from the command line. Run computations on complex objects where it could be a single machine instruction etc. The trend has been around for a long time.
comboy··on Claude Code Cheat Sheet
lol, yeah

> We’ve rewritten Claude Code’s terminal rendering system to reduce flickering by roughly 85%.

tells you all you need to know

and I keep running it remotely through tmux, that explains so many things

edit: if they are writing it in react anyway (sic!) maybe we could at least get a web interface, skipping mapping it to terminal output part ..

comboy··on Claude Code Cheat Sheet
Wow /insights is genuinely useful, perhaps CLI should be pushing that as a tip, if one has enough sessions, instead of keep nagging me about the frontend developer skill which I already have installed

In general CLI could be more reliable and responsive though, it's a text based env yet sometimes feel like running windows 95 on 386dx

It seems clear from the insights that some model is marking failure cases when things went wrong and likely reporting home, so that should be extremely valuable to Anthropic

comboy··on GitHub appears to be struggling with measly three nines availability
I think stability and reliability have vastly improved over the last years in general (not necessarily talking about gh specifically)

It's just that everybody is using 100 tools and dependencies which themselves depend on 50 others to be working.

comboy··on OpenCode – Open source AI coding agent
OpenX is becoming a bit like that hindu symbol associated with well being..
comboy··on Push events into a running session with channels
Claude getting clawed.
comboy··on Nvidia NemoClaw
> a state sponsored threat actor

your CPU, your OS, CPU and firmware on your motherboard chips, ethernet, wifi, HDDs (btw did you know your sim card has JVM?), your browser, all your networking equipment in between, BGP and all the root certs and I'm just scratching the surface

the ballpark is on anther planet

comboy··on AI coding is gambling
Fascinating how HN is torn about vibe coding still. Everybody pretty much agrees that it works for some use cases, yet there is a flamewar (I mean, cultured, HN-type one) every time. People seem to be more comfortable in a binary mindset.
comboy··on AI coding is gambling
> That is extremely stupid. What does that ban get you?

confidence in firing coders I presume..

comboy··on Give Django your time and money, not your tokens
Me too, the problem is that it's hard to come up with tools that are needed but not made yet, and we don't want to end up with https://malus.sh/index.html
comboy··on Give Django your time and money, not your tokens
Perhaps we should start making LLM- open source projects (clearly marked as such). Created by LLMs, open for LLM contributions, with some clearly defined protocols I'd be interesting where it would go. I imagine it could start as a project with a simple instruction file to include in your project to try to find abstractions which can be useful to others as a library and look for specific kind of libraries. Some people want to help others even if they are sharing effectively money+time rather than their skill.

Although I'm afraid big part of these LLM contributions may be people trying to build their portfolio. Some known project contributor sounds better than having some LLM generated code under your name.

comboy··on I beg you to follow Crocker's Rules, even if you will be rude to me
Alright, my original comment was wrong (as was the parent). I still stand by my opinion that it is not practical though.
comboy··on Ageless Linux – Software for humans of indeterminate age
Eshittification (by Cory Doctorov) is a shitty book but it does explain how that dynamic works.
comboy··on I beg you to follow Crocker's Rules, even if you will be rude to me
Applying them to only one side of the conversation doesn't seem practical.
comboy··on I beg you to follow Crocker's Rules, even if you will be rude to me
That's your worldview. Crocker's rules is that you don't have to take receiver feelings into account you just communicate efficently.
comboy··on 1M context is now generally available for Opus 4.6 and Sonnet 4.6
Not sure if it's a common knowledge but I've learned not that long ago that you can do "/compact your instructions here", if you just say what you are working on or what to keep explicitly it's much less painful.

In general LLMs for some reason are really bad at designing prompts for themselves. I tested it heavily on some data where there was a clear optimization function and ability to evaluate the results, and I easily beat opus every time with my chaotic full of typos prompts vs its methodological ones when it is writing instructions for itself or for other LLMs.

comboy··on Your phone is an entire computer
Your sim card is an entire computer.
← PreviousPage 5 of 34Next →