HNHacker News
TopNewBestAskShowJobs

the_pwner224

2,440 karma · joined December 30, 2018

submissionscomments
the_pwner224··on Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
Huh I didn't see those instructions when I tried it last week. Must not have looked closely enough. I do remember it not having NPU support (confirmed by other people) back at the Gemma 3 launch a while ago.
the_pwner224··on Do you even need a database?
SqliteBrowser will let you open up your tables in an Excel-type view. You can also edit directly from the GUI. Still not as frictionless as a plain text file, and I'm not sure how good the search functionality is, but it lets you skip having to write any SQL.
the_pwner224··on Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
> Google’s engineers likely gave up on trying to compile custom attention kernels for Apple’s proprietary tensor blocks iirc.

The AI Edge Gallery app on Android (which is the officially recommended way to try out Gemma on phones) uses the GPU (lacks NPU support) even on first party Pixel phones. So it's less of "they didn't want to interface with Apple's proprietary tensor blocks" and more of that they just didn't give a f in general. A truly baffling decision.

the_pwner224··on Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
I tried it and it was unusably slow at ~5-6 TPS. 26A4B gets close to 40 TPS which is faster than you can read, and still pretty quick with reasoning enabled.
the_pwner224··on Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
I have a 128 GB Strix Halo tablet (same as the other commenter here with the Framework Desktop). I'm using the larger Gemma 4 26B-A4B model (only 28 GB @ Q8) and it's been working great and runs very fast.

It's a 100% replacement for free ChatGPT/Gemini.

Compared to the paid pro/thinking models... Gemma does have reasoning, and I have used the reasoning mode for some tax & legal/accounting advice recently as well as other misc problems. It's worked well for that, but I haven't tried any real difficult tasks. From what I've heard re. agentic coding, the open weight models are ~18-24 months behind Anthropic & Google's SOTA.

Qwen 3.5 122B-A10B should just fit into 128 GB with a Q4/5 and may be a bit smarter. There's apparently also a similar sized Gemma 4 model but they haven't released it yet, the 26B was the largest released.

the_pwner224··on Google Gemma 4 Runs Natively on iPhone with Full Offline AI Inference
If you're looking to buy new hardware, also consider the Asus Rog Flow Z13. It has the same chip as the Framework desktop and is ~20% cheaper ($2,700) for the 128 GB spec while coming in a tablet/laptop form factor. It's capped at a slightly lower power but Strix Halo scales down very well in TDP - I never even use the max power mode on my Z13 because you don't really get any extra perf.

The only downside is that I suspect the Framework would be a decent bit quieter under load (not that this thing is abnormally loud). As well as you're limited to a single M.2 2230 internal SSD slot in this (I believe Micron recently launched a 4 TB model, but generally you'll max out at 2 TB without using an external enclosure).

I don't have anything against the Framework, I'm sure it's a great machine, but the Z13 is an incredible portable all-in-one device that can handle everything from general PC use to gaming to tablet/entertainment to LLMs & high perf.

the_pwner224··on Good Sleep, Good Learning (2012)
That part didn't make any sense to me either. Yes, the natural circadian cycle in a vacuum is slightly over 24 hours, but exposure to light keeps it synced to the normal 24-hour day. If you free run sleep your cycle should stay locked to 24 hours, just like it has always been with our ancestors who lived without artificial light.
the_pwner224··on Google removes "Doki Doki Literature Club" from Google Play
"Violent"? Do you consider news reporting to be violent too? This isn't remotely in the league of all the shooter games you can find on the store.
the_pwner224··on OpenClaw’s memory is unreliable, and you don’t know when it will break
Please elaborate
the_pwner224··on OpenAI backs Illinois bill that would limit when AI labs can be held liable
> Is there anything we can do to push back against and discourage the externalization of costs onto others?

On a societal scale, no. Occasionally this works in some individual cases. Like the online outrage over SOPA/PIPA 15 years ago.

But when entity X can gain $$$$$$ (or power) from doing an action, and that action costs everyone only $ (or a minor bit of inconvenience or ideological righteousness), then the average person has very little incentive to take time out of their day-to-day life to fight it.

Meanwhile the entity will do whatever it takes to get the $$$$$$/power because they have a huge incentive. This is the same mechanism that allows democracies to be eroded, as we're seeing right now in the US.

the_pwner224··on OpenAI backs Illinois bill that would limit when AI labs can be held liable
That whole fiasco actually soured me on Anthropic. They were clearly super desperate to take blood money. "Anthropic has much more in common with the Department of War than we have differences."
the_pwner224··on Running Gemma 4 locally with LM Studio's new headless CLI and Claude Code
AMD Strix Halo. Available in the Framework desktop, various mini PCs, and the Asus Rog Flow Z13 "gaming tablet." The Z13 is still at $2700 for 128 GB which is an incredible deal with today's RAM prices.

There's also the Nvidia DGX Spark.

the_pwner224··on OpenClaw privilege escalation vulnerability
You have yet to answer the original question - what do you actually do with OpenClaw? A concrete example of something that actually happens, not a system architecture description.
the_pwner224··on Qwen3.6-Plus: Towards Real World Agents
Agreed
the_pwner224··on Qwen3.6-Plus: Towards real world agents
I had the exact opposite reaction. I stopped using OpenAI/Google a while ago due to privacy and moved to local Qwen, now I'm considering using Alibaba cloud. You know Google and OpenAI are going to share everything with the US government and Western ad networks. But with Alibaba, who cares if the CCP & Chinese ad networks have a comprehensive profile on me? From a pragmatic perspective it's much better for (outcomes related to) privacy.
the_pwner224··on We haven't seen the worst of what gambling and prediction markets will do
It was extremely obviously not some intentional 500 IQ plot to keep crypto illegitimate...
the_pwner224··on Our commitment to Windows quality
Bill definitely wouldn't approve of the current Windows quality. This email by him (2003) is very interesting. It looks like he was powerless to stop the degradation.

https://web.archive.org/web/20080626154537/http://blog.seatt...

https://news.ycombinator.com/item?id=227045

the_pwner224··on Our commitment to Windows quality
I feel like most interns would be smart enough to know that you should lazy load these metrics. It's incredible that MS put this into production.
the_pwner224··on Google details new 24-hour process to sideload unverified Android apps
> My solution is educating about smartphones and computers first.

98% of people literally do not care and/or are too dumb to understand. You could force them at gunpoint to sit in the education class, and give them a simple basic quiz afterwards, and they'd get half the answers wrong. They will continue to not even read what's on their screen, and just click the big highlighted button every time they see one.

the_pwner224··on Google details new 24-hour process to sideload unverified Android apps
Go to your uBlock Origin settings and enable the annoyances/social filter lists.
the_pwner224··on Kagi Small Web
I switched about a year ago. At the time it did seem like a step up from Google results. But there's been an increasing prevalence of low quality results. Blogspam, AI websites, etc. Obviously not blaming Kagi here, web search has gotten hard recently.

Is Kagi still better than Google? Probably, I don't really know because I don't use Google anymore. But at this point I feel like I'm with them out of inertia more than being an avid supporter. One of these days I'll re-evaluate Google and decide whether to switch back or not.

It does occasionally surface interesting results from small sites that you wouldn't get on Google. I do find that to be useful.

Kagi definitely isn't a bad search engine by any means. Honestly if you haven't used it, try the 100 search free trial on one device. Maybe you'll like it. This feels more like a general decline of the open web.

the_pwner224··on The bureaucracy blocking the chance at a cure
.
the_pwner224··on Can I run AI locally?
Arch with KDE, it works perfectly out of the box.

I configured/disabled RGB lighting in Windows before wiping and the settings carried over to Linux. On Arch, install & enable power-profiles-daemon and you can switch between quiet/balanced/performance fan & TDP profiles. It uses the same profiles & fan curves as the options in Asus's Windows software. KDE has native integration for this in the GUI in the battery menu. You don't need to install asus-linux or rog-control-center.

For local AI: set VRAM size to 512 MB in the BIOS, add these kernel params:

ttm.pages_limit=31457280 ttm.page_pool_size=31457280 amd_iommu=off

Pages are 4 KiB each, so 120 GiB = 120 x 1024^3 / 4096 = 31457280

To check that it worked: sudo dmesg | grep "amdgpu.*memory" will report two values. VRAM is what's set in BIOS (minimum static allocation). GTT is the maximum dynamic quota. The default is 48 GB of GTT. So if you're running small models you actually don't even need to do anything, it'll just work out of the box.

LM Studio worked out of the box with no setup, just download the appimage and run it. For Ollama you just `pacman -S ollama-rocm` and `systemctl enable --now ollama`, then it works. I recently got ComfyUI set up to run image gen & 3d gen models and that was also very easy, took <10 minutes.

I can't believe this machine is still going for $2,800 with 128 GB. It's an incredible value.

the_pwner224··on Can I run AI locally?
Strix Halo you can get at least 120 GB to the GPU (out of 128 GB total), I'm using this configuration.

Setting the kernel params is a one-time initial setup thing. You have 128 GB of RAM, set it to 120 or whatever as the max VRAM. The LLM will use as much as it needs and the rest of the system will use as much it needs. Fully dynamic with real-time allocation of resources. Honestly I literally haven't even thought of it after setting those kernel args a while ago.

So: "options ttm.pages_limit=31457280 ttm.page_pool_size=31457280", reboot, and that's literally all you have to do.

Oh and even that is only needed because the AMD driver defaults it to something like 35-48 GB max VRAM allocation. It is fully dynamic out of the box, you're only configuring the max VRAM quota with those params. I'm not sure why they choice that number for the default.

the_pwner224··on Your phone is an entire computer
> But you can unlock it without needing Google.

Well akshually.... the bootloader is initially not unlockable. You must connect the phone to the internet. Within a few minutes a background process will reach out to Google servers to check whether it was purchased outright or with a payment plan. It will only enable the bootloader unlocking toggle after this step. Phones bought with a carrier contract won't be unlockable until paid off.

In those initial few minutes (/ before you connect it to the interwebs), the bootloader unlock option in the developer settings & fastboot will be disabled.

the_pwner224··on Can I run AI locally?
Yep, I have a 13" gaming tablet with the 128 GB AMD Strix Halo chip (Ryzen AI Max+ 395, what a name). Asus ROG Flow Z13. It's a beast; the performance is totally disproportionate to its size & form factor.

I'm not sure what exactly you're referring to with "Only Apple has the unique dynamic allocation though." On Strix Halo you set the fixed VRAM size to 512 MB in the BIOS, and you set a few Linux kernel params that enable dynamic allocation to whatever limit you want (I'm using 110 GB max at the moment). LLMs can use up to that much when loaded, but it's shared fully dynamically with regular RAM and is instantly available for regular system use when you unload the LLM.

the_pwner224··on Cardiorespiratory fitness is associated with lower anger and anxiety
For those of us who don't have a Polar strap, can you explain at a high level how your app works? Based on what the page says, seems like something about using R-R interval to estimate where you area on the Meyer wave cycle?

I have a different heart rate monitor (Amazfit smartwatch, mine has their latest sensor that matches the higher end Garmin watches for accuracy, it can be used as a Bluetooth device or you can develop software to run on it directly). What topics/keywords should I look into if I want to develop the equivalent application for my hardware?

the_pwner224··on Blue light filters don't work – controlling total luminance is a better bet
I see. Not a Safari user myself. On Firefox & Chrome Dark Reader can force its own dark theme even if the site doesn't provide one. Like Noir does on mobile Safari.
the_pwner224··on Blue light filters don't work – controlling total luminance is a better bet
"Dark Reader" does the same thing on desktop Firefox/Chrome/etc. (& mobile Firefox, maybe also available on mobile Safari?).
the_pwner224··on GrapheneOS – Break Free from Google and Apple
> Being able to configure multiple VPNs at once, e.g. for Tailscale, ad filtering, blocking HackerNews during times when I should be doing something more productive

AdAway (in F-Droid) can block with /etc/hosts (no VPN involved) if you have root. The hosts blocking still works even when connected to a VPN. Aside from loading ad domain lists into /etc/hosts, it also allows you to specify custom domains to block - I personally have Reddit and HN in there :)

Page 1 of 26Next →