HNHacker News
TopNewBestAskShowJobs

zbendefy

263 karma · joined June 18, 2020

submissionscomments
zbendefy··on Muse Glimmer: 30B-parameter model optimized for always-on local agent workflows
Similiarly I wonder why we dont run our own email server despite the sensitive data there.
zbendefy··on Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA
Dont have much experience how well they perform but quantized models can run on ~4 H100 or A100 which should be true for kimi k3 and qwen 3.8 as well.
zbendefy··on The state of open source AI
I think they didnt have the amount of debt /negative cashflow as they do now.
zbendefy··on Kimi K3: Open Frontier Intelligence
I dont have estimates on the cost of running models, but I think openai and anthropic are running on subsidized prices. At actual prices it might be worth it in the future.
zbendefy··on Qwen 3.6 27B is the sweet spot for local development
What harness are you using?
zbendefy··on How JPL keeps the 13-year-old Curiosity rover doing science
Maybe we would get a microphone on mars. Just kidding i know air pressure is vastly different, but still it would be cool to listen to ambient sound from there
zbendefy··on How LLMs work
Good point!

There is also the case for Markov chains being theoretically able to do these if tuned well. Or even SAT problem.

zbendefy··on OpenClaw deletes Summer Yue's emails
This reminds me of stuff like interns wiping the production servers.

The solution in the comments there is always "have backups" and "why can an intern do that stuff? why arent policies set".

But I dont think these can be applied to something like openclaw agents

zbendefy··on Raspberry Pi Drag Race: Pi 1 to Pi 5 – Performance Comparison
I did the same with an rpi3, not sure if I used this guide but it seems good:

https://www.raspberrypi.com/news/printing-at-home-from-your-...

zbendefy··on Zen-C: Write like a high-level language, run like C
i think so. The biggest hurdle with new languages is that you are cut off from a 3rdparty library ecosystem. Being compatible with C 3rd party libraries is a big win.
zbendefy··on Cameras and Lenses (2020)
Tangential, but Flash had a nice side effect that the "app" could be exported in a self contained way via SWF.

Exporting this site for example in a future proof way is not that obvious. (Exporting as pdf wont work with the webgl applets, exporting the html page might work but is error prone depending in the website structure)

50 years from now, flash emulators will still work on swf files, but these sites might be lost. Or is there a way to archive sites like this?

zbendefy··on No Graphics API
>The user writes the data to CPU mapped GPU memory first and then issues a copy command, which transforms the data to optimal compressed format.

Wouldnt this mean double gpu memory usage for uploading a potentially large image? (Even if just for the time the copy is finished)

Vulkan lets the user copy from cpu (host_visible) memory to gpu (device_local) memory without an intermediate gpu buffer, afaik there is no double vram usage there but i might be wrong on that.

Great article btw. I hope something comes out of this!

zbendefy··on 1GB Raspberry Pi 5, and memory-driven price rises
What has changed now in the memory landscape/ai workload in the recent months compared to summer or spring?
zbendefy··on A time-travelling door bug in Half Life 2
there is also a vr mod for HL1 as well
zbendefy··on I almost got hacked by a 'job interview'
My takeaway is that sandboxing should be more readily available, and integrated into the OS.

I used sandboxie a while ago for stuff like this, but afaik windows has some sandbox built into it since a few years which I didnt think about until now.

zbendefy··on Homeowner baffled after washing machine uses 3.6GB of internet data a day (2024)
Why not get the wifi enabled fridge and just not hook it up to your router?

Genuinely asking because I plan to do this once I have to get new appliances, is there something missing that way?

zbendefy··on What to do with C++ modules?
Thing is (correct.me if Im wrong), that if you use modules, all of your code need to use modules (e.g. you cant have mixed #include <vector> and import <vector>; in your project). Which rules out a lot of 3rd party code you might want to depend on.
zbendefy··on From XML to JSON to CBOR
How different is CBOR compared to BSON? Both seem to be binary json-like representations.

Edit: BSON seems to contain more data types than JSON, and as such it is more complex, whereas CBOR doesn't add to JSON's existing structure.

zbendefy··on Debian switches to 64-bit time for everything
its 12 years not 22.

An embedded device bought today may be easily in use 12 years from now.

zbendefy··on 3-JSON
offtopic: why is the Copyright © icon shake like crazy at the bottom of the page?

Edit: Oh I guess it seems to be intentional, I clicked around and I like the rgbcube site map.

zbendefy··on Mark Zuckerberg says social media is over
This is such a good analogy. Awereness about social media shluld be like awereness about junk food you consume.
zbendefy··on South Korea Is over [video]
Is it wrong?
zbendefy··on Valve releases Team Fortress 2 code
Correction, I didnt remember correctly as those are actually there in the code release, here is the r200 (radeon 8500) renderer specific code:

https://github.com/id-Software/DOOM-3/blob/master/neo/render...

zbendefy··on Valve releases Team Fortress 2 code
Not just that, they had specific renderer backends, one for GeForce, one for GeForce3, one for Radeon 8500 that they had to cut out as they used proprietary information or code perhaps.
zbendefy··on How do modern compilers choose which variables to put in registers?
Aren't registers fixed by x86_64, while cache is a CPU hardware specific thing (e.g.: newer cpus have more cache than older ones, bit register count is fixed 8 on x86 and 16 on x86_64)?

So I think the compiler can work with registers at compile time but cant work with an unknown structure of cache

zbendefy··on Why DeepSeek had to be open source
No, the full R1 model is ~650GB. There are quantized version that quantize it down to ~150GB.

What you can run locally are the distilled models, that is actually LLama and Qwen weights further trained on R1's output

zbendefy··on Why DeepSeek had to be open source
I dont think they rent gpus for $5million because its cool and want to show the world...
zbendefy··on Why DeepSeek had to be open source
Note: you are probably running a distilled version of R1, which is actually LLama or Qwen further trained on the input/output of R1.

The full R1 is huge (~700GB), altough there are still quantized versions, the smallest one is around 150gb (1.58bit)

zbendefy··on OpenAI Furious DeepSeek Might Have Stolen All the Data OpenAI Stole from Us
Also without the "attention is all you need" paper from google
zbendefy··on Diffusion models are real-time game engines
I think some state is also being given (or if its not, it could be given) to the network, like 3d world position/orientation of the player, that could help the neural network anchor the player in the world.
Page 1 of 3Next →