HNHacker News
TopNewBestAskShowJobs

pferdone

257 karma · joined May 30, 2016

submissionscomments
pferdone··on H3-metal – Native MiniMax-H3 inference for Apple Silicon
1) It would still run on the Mac's GPU.

2) Since it's unified memory, you won't have 96GB available.

3) I offered a solution that is usually recommended to the "gpu poor", if he's concerned with how much memory he would need.

4) I stated, that people already pointed out how he should be fine and that "gpu poor" doesn't apply to him.

5) "gpu poor" depends on what model you are trying to use. If you want to run Kimi or GLM you are still "gpu poor" even if you have an RTX Pro 6000 with 96GB of VRAM.

pferdone··on H3-metal – Native MiniMax-H3 inference for Apple Silicon
a friend told me there's a reddit called: unstable diffusion
pferdone··on H3-metal – Native MiniMax-H3 inference for Apple Silicon
you should have a look at https://github.com/deepbeepmeep/Wan2GP which is the goto tool for "gpu poor", although as people below already pointed out you should be fine with comfyui's standard setup aswell
pferdone··on OpenAI says its AI went rogue and launched 'unprecedented' cyber-attack
Public doesnt know about Huggingface. ChatGPT (OpenAI) says it‘s dangerous. They must know.
pferdone··on America pays workers just 27% of what its wealth allows – the worst in the OECD
You didn't even read your own article.
pferdone··on Steam Machine launches today
What does identity or sexuality have to do with it you fucking moron? These are people! People play games.
pferdone··on Gemma 4 12B: A unified, encoder-free multimodal model
But do I have the option to run it 'text only'?
pferdone··on Show HN: KVBoost – chunk-level KV cache reuse for HuggingFace, 5–48x faster TTFT
slop
pferdone··on The last six months in LLMs in five minutes
The consensus right now is that Qwen3.6 in its 27B and 35B-A3B versions is better for coding whereas Gemma4 is stronger when it comes to OCR, audio transcription and the likes. Margins are slim though and the harness at these model sizes is the most important factor.
pferdone··on ZAYA1-8B matches DeepSeek-R1 on math with less than 1B active parameters
That‘s so cool man! Congrats!
pferdone··on SANA-WM, a 2.6B open-source world model for 1-minute 720p video
First video with the guy walking the mountain in snow has consistency issues with the cave entrance. Which is "expected" at this model size?!
pferdone··on SANA-WM, a 2.6B open-source world model for 1-minute 720p video
Who wrote your comment?
pferdone··on ZAYA1-8B matches DeepSeek-R1 on math with less than 1B active parameters
I can see that and I don't know your setup, but there are people pushing >70t/s with MTP on a single 3090, with big contexts still >50t/s. 64k is not a lot for agentic coding, and IIRC 128k with turboquant and the likes should be possible for you. r/LocalLLM/ and r/LocalLLaMA/ are worth a visit IMO.

EDIT: just found this recipe repo, may wanna give it a go: https://github.com/noonghunna/club-3090

EDIT-2: this can also shave off a lot of context need for tool calling -> https://github.com/rtk-ai/rtk

pferdone··on I cancelled Claude: Token issues, declining quality, and poor support
pi.dev as well
pferdone··on Qwen3.6-Plus: Towards real world agents
> unlike almost all qwen models

Almost all means there have been ones before that were not open. So, no contradiction there.

pferdone··on Qwen3.6-Plus: Towards real world agents
They said in the last paragraph[0]:

"[...] In the coming days, we will also open-source smaller-scale variants, reaffirming our commitment to accessibility and community-driven innovation. [...]"

[0] https://qwen.ai/blog?id=qwen3.6#summary--future-work

pferdone··on Z-Image: Powerful and highly efficient image generation model with 6B parameters
It‘s mainly due to system requirements that Flux.2-dev doesn’t get same usage as Z-Image. A 5090 needs about a minute to generate an image with a basic workflow with Flux.2-dev. But prompt adherence and scene/character consistency in edit mode is (way) ahead of Qwen-Edit-2509 if you ask me.
pferdone··on Transparent computer monitor designed to protect your vision
I also have a HUD in my car and I can read it just fine, even in bright sunlight.
pferdone··on Grok 3 claims its system prompt includes censorship about Musk/Trump
I feel tired reading about him and it doesn’t even phase me anymore. It‘s just another thing I add on to the pile. Maybe it’s part of the plan to go numb to everything he does.
pferdone··on Omni SenseVoice: High-Speed Speech Recognition with Words Timestamps
I mean they make a bold statement up top just to paddle back a little bit further down with: "[…] In terms of Chinese and Cantonese recognition, the SenseVoice-Small model has advantages."

It feels dishonest to me.

[0] https://github.com/FunAudioLLM/SenseVoice?tab=readme-ov-file...

pferdone··on Mpv – A free, open-source, and cross-platform media player
only next frame
pferdone··on The Guardian digital design style guide
The Guardian, especially for their podcasts, is the only news website I am paying and have ever payed for. And I pay more willingly than any other newspaper would get from me for their paywall stuff. It‘s that valuable for me to support this approach.
pferdone··on Cooperative C++ Evolution – Toward a TypeScript for C++
Where have you heard that? Because Typescript is the defacto standard right now. Projects like bun and deno cement that status further.
pferdone··on Language Learning with Netflix
Language learning via hearing comprehension of content not produced in the target language is almost impossible, because the subtitles never match.

However there‘s s difference between CC (close captions) and subtitles, with the former being the verbatim representation (including sfx, music etc.) in my experience.

I already commented [0] on this 2 years ago.

[0] https://news.ycombinator.com/item?id=27420959#27435311

pferdone··on Honey consumption improves blood sugar and cholesterol levels, study suggests
I dont know anyone who thought input == output. Diets like keto clearly show the opposite.
pferdone··on Organ donations, transplants increase on days of largest motorcycle rallies
> the US cares about you not endangering other people

Through their rigorous license process to operate road vehicles…yeah right.

pferdone··on QualityScaler: Image/video deeplearning upscaler with any GPU
It looks pretty good, but especially with the Spiderman comparison you can notice a little smudging and it basically loses the pattern on the suit.
pferdone··on We improved React loading times with Next.js
I mean good for them for improving their user‘s experience. I am just surprised that maintaining your own webpack config seems like such challenge. Enabling chunks to split code to improve loading times or dynamic imports, it‘s not exactly rocket science. Yes DX and all improved as well, but CRA is such a bad choice for production code. It gets you going fast(er), but as soon as you have to tweak it, it seems like devs jump to the "next" framework that already has those tweaks builtin…until they hit the next roadblock. Thus you never learn what actually makes it all work.
pferdone··on What can we learn from leaked Insyde's BIOS for Intel Alder Lake
You are right, it isn‘t of any value to me on HN. Additional information on the topic or a discussion with arguments is. Now you know what is.
pferdone··on What can we learn from leaked Insyde's BIOS for Intel Alder Lake
I couldn’t hide my dislike for it and commented. Maybe it was unnecessary to draw comparisons to reddit and it’d be better to just state HN rules. I will try to consider this next time. Thank you.
Page 1 of 4Next →