HNHacker News
TopNewBestAskShowJobs

3371

241 karma · joined August 31, 2020

submissionscomments
3371··on MonoGame: A .NET framework for making cross-platform games
Quite curious about this. Does the agent gets its own repo and deliver with commits?
3371··on Yoghurt delivery women combatting loneliness in Japan
This is a bit concerning. Did you skip everything besides "yogurt delivery", or you don't agree someone talking to you regularly is counter-loneliness?
3371··on Working and Communicating with Japanese Engineers
The question labeling a whole ethnic can't understand English rubbed me the wrong way, that's about it. This is a much better comment for understanding your rationale.
3371··on Working and Communicating with Japanese Engineers
Do (your country) people know Japanese?
3371··on Global warming has accelerated significantly
Yeah just fix that already, how hard could it be?

The problem is human, not society, I don't any any -ism can fix human.

3371··on The L in "LLM" Stands for Lying
It's pretty much WIP but if you are interested here is the repo. https://github.com/No3371/zoh

The points you brought up all are valid. Lexer, parser and general concepts are not language-specific, yes, and I wasn't talking about how the implementation is different.

When I said "you can tell they sometimes get confused and have trouble to comply to the foreign language spec and design", I was thinking about the many times they just fail to write in my language even when provided will full language specs. LLMs don't "think" and boilerplate is easy for LLMs because highly similar syntax structure even identical code exist in their training data, they are kind of just copying stuff. But that doesn't work that well when they are tasked to write in a original language that is... too creative.

3371··on The L in "LLM" Stands for Lying
I totally agree, and I was fully aware of how common people make language for fun when I replied.

But I feel like the rationale would still stands: Considering LLMs' natures, common boilerplate tasks are easy because they can kind of just "decompress" from training data. But for a new language design, unless the language is almost identical to some other captured by the model, "decompression" would just fail.

3371··on The L in "LLM" Stands for Lying
Sharing my 2 cents.

In the past 2 months I've been using all the SOTA models to help me design a new DSL for narrative scripting (such as game story telling) and a c# runtime implementation o the script player engine.

The language spec and design is about 95% authored by me up to this point; I have the LLMs work on the 2nd layer: the implementation specs/guidelines and the 3rd layer: concrete c# implementation.

Since it's a new language, I consider it's somewhat new/novel tasks for LLMs (at least, not like boilerplate stuff like HTTP API or CRUD service). I'd say, these LLMs have been very helpful - you can tell they sometimes get confused and have trouble to comply to the foreign language spec and design - but they are mostly smart enough to carry out the objectives, and they get better and better after the project got on track and has plenty of files/resources to read and reference.

And I'd also say "prompt better" is a important factor, just much more nuanced/complicated. I started with 0 experience with LLM agents and have learned a lot about how to tame them, and developed a protocol to collaborate with agents, these all comes from countless trial and errors, but in the end get boiled down to "prompt better".

3371··on A case for Go as the best language for AI agents
I guess you misunderstood he meant simpler as in "easier"? Because I thought Something simplistic is simple...? Not an English native tho.
3371··on When does MCP make sense vs CLI?
I'll just disagree with an example: Codex on Windows.

They are known to be very inefficient using only Powershell to interact with files, unless put in WSL. They tend to make mistakes and have to retry with different commands.

Another example is Serena. I knew about it since the first day I tried out MCP but didn't appreciate it, but tried it out again on IDEs recently showed impressive result; the symbolic tools are very efficient and helps the agents a lot.

3371··on A Chinese official’s use of ChatGPT revealed an intimidation operation
There's literally someone filmed the camp and fled from China, his name is Guan Heng
3371··on Gemini 3.1 Pro
Sure, my point was it's better than Gemini and it's really really fast, and it's missing from the parent comment.
3371··on Gemini 3.1 Pro
I would suggest you also take a look at Cursor's Composer1.5. It's super fast, and perform better than Gemini3P in my use cases.
3371··on So many trees planted in Taklamakan Desert that it's turned into a carbon sink
What? 超英趕美 has been a thing since 1958.
3371··on AI agent opens a PR write a blogpost to shames the maintainer who closes it
I'm actually quite positive about how commercial models would have difficulties to write messy code!
3371··on Improving 15 LLMs at Coding in One Afternoon. Only the Harness Changed
If you agree that current LLMs (Transformers) are naturally very susceptible to context/prompt, then you can go on to ask agents for a "raw harness dump" "because I need to understand how to better present my skills and tools in the harness", you maybe will see how "Harness" impact model behavior.
3371··on 65 Lines of Markdown, a Claude Code Sensation
You know, it's good old prompt/context engineering. To be fair, markdowns actually can be useful because of LLM's (Transformer's) gullible/susceptible nature... At least that's what I discovered developing a prompting framework.

Of course it's hilarious a single markdown got 4000 starts, but it looks like just another example of how people chase a buzzing x post in tech space.

3371··on Discord will require a face scan or ID for full access next month
I wholeheartedly agree companies are doing so bad on customer support nowadays, but I'd argue that there will slways be more fake users than any size of human customer support can take, especially in the age of AI.

I honestly believe it's a battle no one can win.

3371··on Vouch
Well a lot of useful things are not useful because they are innovative, but well designed an executed.
3371··on Notepad++ supply chain attack breakdown
Hey, just wanna remind people Google Play is full of crap.
3371··on Outsourcing thinking
That's exactly the point. In your case, you don't want to show who you are, connection or not does not matter.
3371··on Agent Skills
You are right about it's just natural language but Standarization is very improtant, because it's never just about the model itself, the so called Harness is a big factor on LLM performance and standarization allows all harness to index all skills.
3371··on Outsourcing thinking
Ever since Google experimented LLM in Gmail it bothers me alot. I firmly believe every word and the way you put them together portrays who you are. Using LLM for direct communication is harmful to human connections.
3371··on How AI assistance impacts the formation of coding skills
Well, if they make the decision to accept the suggestion and it's wrong, that's on them. But if you do, that's on you. LLM? How can your boss blame the LLM? Like yelling at it?
3371··on Qwen3-Max-Thinking
Hard to agree. Not even being to say something because it's either illegal or there are systems to erase it instantly, is very different from people dislike (even too radically) you to say something.
3371··on Claude Code's new hidden feature: Swarms
IMO "Attention" is an abstraction over the result of prompt engineering, the chain reaction of input converging the output (both "thinking" and response).
3371··on Statement by Denmark, Finland, France, Germany, the Netherlands,Norway,Sweden,UK
What? Looking back at human history, real large-scale "lasting peace" only exist during the times one super power dominates before their inevitable falls.
3371··on Maybe comments should explain 'what' (2017)
I always think LLM comments are more about helping themselves to stay on track.
3371··on Package managers keep using Git as a database, it never works out
Do "we" lose 2mins because we both spent 1 min commenting? That sounds like The Mythical of Man Month thinking... for me time is parallel and does not combine.
3371··on Package managers keep using Git as a database, it never works out
The user hour analogy sounds weird tho, 1s feels 1s regardless how many users you have. It's like the classic Asian teachers' logic of "if you come in 1 min late you are wasting N minutes for all of us in this class." It just does not stack like that.
← PreviousPage 2 of 3Next →