It has particularly good observability (the ability to see the full content of every prompt and response and tool call) compared to other harnesses, that's the main thing that stood out to me.
Weird. I don't use the official harnesses by OpenAI or Anthropic but with all others i tried there wasn't a single one that hides any of that. What harnesses did you compare to?
Have you tried pressing ctrl+o?
I think what he meant was that it gives you good overall observabiity, over your entire context, kind of how langfuse or these platforms do, but it does so locally. A good philosophy i picked up is how they treat the transcript as the backbone of the product.
It uses an append-only design for the context which rules out all kinds of cache-invalidation bugs that pop up regularly in other harnesses.
All LLMs work the same to a certain degree, it's a matter of personal preference, token cost and customization. I asked DSH to build a news aggregator for me and it built a really nice app from 30 news sources via rss feeds for less than 50 cents and presto, reading tech news has never been so gratifying
literally every harness can do that tho, that doesn't really answer the question.
i read somewhere that it can write a plugin to change its font size
Pi can do this too.
what an age we live in
I feel like I live in another reality from others here; I read people saying things ‘while their jaws are hanging open’ (not said here literally but the feel is the same ; I see it in other threads on HN literally here though); why are we so chuffed with stuff that’s now been one shot for maybe already a year, but definitely the last 6 months? Literally everyone who tried LLMs know this and yet it seems a huge surprise to people here that it is so easy now? And that’s opposite of the people, also very much on HN, who say AI is shit at coding and will not replace humans because humans have to fix its bad code.
Honestly some of these comments and posts just seem like propaganda or shills just trying to push DeepSeek.