HNHacker News
TopNewBestAskShowJobs

WiSaGaN

1,027 karma · joined September 18, 2015

submissionscomments
WiSaGaN··on DeepSeek v4.1 Flash
This is definitely not on par with GPT-6 astra. Not with GPT-5.6 sol either. But probably will set as a new baseline for modern API based LLM because it's so cheap.
WiSaGaN··on Our decision on Cursor following its acquisition by SpaceX
It's one of the things you need to do if you want your company to later become the only company in the world. They also promised that they would treat you nicely afterwards after they get what they want.
WiSaGaN··on Qwen 3.8-Flash-Next releasing tomorrow (125B a6B)
You probably meant Qwen3.8-35B-A3B. But judging from some of the words from their team, it seems unlikely unfortunately.
WiSaGaN··on DeepSeek-v4-flash-vision-exp
That would've been a very strange arrangement.
WiSaGaN··on OpenAI’s head of ethics leaves less than a year after joining
It's unlikely HugginFace colluded given they specifically cited they used open models to save them from the hacking.
WiSaGaN··on How Claude marks AI-generated content
My guess is that they will later "reveal" some "violations" but provide little evidence citing proprietary algorithm.
WiSaGaN··on AI's top startups are barely publishing their research
Top startups in China publish a lot?
WiSaGaN··on PGSimCity - How PostgreSQL Works
I assume this is done with the help of AI? I have done similar vibe project that explaining catastrophic forgetting with the help of AI. It's just so satisfying now that if you really want to learn something, you can always do it with the help of AI. Before, it's hard to find good resources, now it's not a problem, the issue now becomes one's own focus and agency.
WiSaGaN··on DeepSeek pause fundraise after comments on compute gap to US leaked (transcript) [pdf]
U.S. policymakers believe that even if the gap is small—like six months to a year—whoever reaches AGI first (whatever that means) could gain such an overwhelming advantage over their perceived adversary that it would effectively kneecap them. (You can look at the kinds of things they mention—cyber, WMDs—to get a sense of what they mean.) Jensen Huang disagrees and has said AI is a marathon.
WiSaGaN··on Show HN: Trust – Coding Rust like it's 1989
Maybe I should start a project rewriting pctools 5.0 in rust!
WiSaGaN··on The Claude Code Source Leak: fake tools, frustration regexes, undercover mode
Yeah, this seems like a personal choice, which does work out given the current result.
WiSaGaN··on Day 1 of ARC-AGI-3
Harness is fine. I think people here are arguing what provided here to take the test is not harness.
WiSaGaN··on Detecting and Preventing Distillation Attacks
This violates the ToS, but I don't think it's distillation. Distillation requires knowing the logits, which current API does not provide. This is just synthetic data generation. Anthropic definitely knows the difference.
WiSaGaN··on Google restricting Google AI Pro/Ultra subscribers for using OpenClaw
Google has gigantic power over its users. Consider that for some reason, Google banned your gmail account, which you are using for large number of logins for different essential services.
WiSaGaN··on Consistency diffusion language models: Up to 14x faster, no quality loss
I think diffusion makes much more sense than auto-regressive (AR) specifically in code generation comparing to chatbot.
WiSaGaN··on Improving 15 LLMs at Coding in One Afternoon. Only the Harness Changed
Great article. To me, this highlights a key question in the era of rapidly advancing machine intelligence: if we know machine intelligence is progressing, what is more valuable to build for? As humans, we still find many tools useful even when doing knowledge work. For instance, a calculator. Sure, a smart person can perform calculations in their head, but it’s much easier to teach everyone how to use a calculator, which is 100% reliable in its intended domain.

In this era, we should build these kinds of tools for problems we know are straightforward ones you can’t get smarter than, even as intelligence continues to advance. Using tools like "bash" or command-line interfaces originally designed for humans is a good initial approach, since we can essentially reuse much of what was built for human use. Later, we can optimize specifically for machines, either accounting for their different cognitive structures (e.g., the ability to memorize extremely long contexts compared to humans) or adapting to the stream-based input/output patterns of current autoregressive token generators.

Eventually, I believe machine intelligence will build their own tools based on these foundations, likely a similar kind of milestone to when humans first began using tools.

WiSaGaN··on Gemini 3 Deep Think
Yes, agentic-wise, Claude Opus is best. Complex coding is GPT-5.x. But for smartness, I always felt Gemini 3 Pro is best.
WiSaGaN··on Prism
OpenAI has a former NSA director on its board. [1] This connection makes the dilution of the term "PRISM" in search results a potential benefit to NSA interests.

[1]: https://openai.com/index/openai-appoints-retired-us-army-gen...

WiSaGaN··on Floating-Point Printing and Parsing Can Be Simple and Fast
Rust's `serde_json` recently switched to use a new library for floating string conversion: https://github.com/dtolnay/zmij.
WiSaGaN··on The microstructure of wealth transfer in prediction markets
A market maker needs a premium to provide liquidity. If all else is equal, why would they take on execution time risk? This is a universal feature of continuous-trading Central Limit Order Books (CLOBs), not something unique to prediction markets.
WiSaGaN··on Reproducing DeepSeek's MHC: When Residual Connections Explode
I guess I am asking how we know Gemini and Claude relies on the additive residual stream. We don't know the architecture details for these closed models?
WiSaGaN··on Reproducing DeepSeek's MHC: When Residual Connections Explode
How do you know "GPT-5, Claude, Llama, Gemini. Under the hood, they all do the same thing: x+F(x)."?
WiSaGaN··on CLI agents make self-hosting on a home server easier and fun
I have a similar experience when I found out that claude code can use ssh to conect to remote server and diagnose any sysadmin issue there. It just feels really empowered.
WiSaGaN··on New information extracted from Snowden PDFs through metadata version analysis
I think it's likely someone already discovered this. It's just that info is not broadcasted to people who want to comment on this thread.
WiSaGaN··on Ed25519-CLI – command-line interface for the Ed25519 signature system (2024)
I can't find the source. Anyone can point to it?
WiSaGaN··on How uv got so fast
This argument falls apart when you look at Rust and Cargo. uv is literally trying to be "Python's Cargo." The entire blueprint came from a flagship FOSS project.

Rust's development used a structured, community RFC process—endless planning by your definition. The result was a famously well-designed toolchain that the entire community praises. FOSS didn't hold it back; it made it good.

So no, commercial backing isn't the only way to ship something good. FOSS is more than capable to ship great software when done right.

WiSaGaN··on MiniMax M2.1: Built for Real-World Complex Tasks, Multi-Language Programming
It is now. But the limit on $20 plan is quite low and easy to use up.
WiSaGaN··on 996
Investment banks in Hong Kong were almost exclusively western back in the days with very few ethnic Chinese in senior management.
WiSaGaN··on 996
China is not the birthplace of so called '996'. Long before tech scene in China, there are a lot of investment banks doing that in HK especially for junior analysts. Calling 996 a China thing is just orientlalism. Everything bad is Chinese, everything good is western.
WiSaGaN··on LLMs tell bad jokes because they avoid surprises
That's true. You would think LLM will condition its surprise completion to be more probable if it's in a joke context. I guess this only gets good when model really is good. It's similar that GPT 4.5 has better humor.
Page 1 of 11Next →