HNHacker News
TopNewBestAskShowJobs

_davide_

98 karma · joined April 1, 2017

submissionscomments
_davide_··on Grok 4.6
in reality even the mention of a prohibition is enough to make the model reject that no matter what
_davide_··on Grok 4.6
i beg to differ, in an ideal world a system possibly is a binding law and high end models are starting to be really aligned to the exact system prompt. The instructions must be simple to follow, if you start doing complex rules it'll call apart, but I'll usually follow the stringer interpretation.
_davide_··on What I learned by putting GitHub Copilot behind a MitM proxy
or completely deleting the app and use the browser version for desperation
_davide_··on What I learned by putting GitHub Copilot behind a MitM proxy
Disagree with the conclusion, even without carefully curated context every high end LLM perform just as well, maybe with an extra detour. In contrast if even one of the learnings is not up to date or doesn't apply to the current situation you find yourself with a long detour or even a failure.
_davide_··on GPT 5.6 Cyber
Yes, given a direction, they are pretty good at "fuzzing", trying out everything and eventually find something. Super useful, but it doesn't have the same level of precise targeting you could expect from a high end model.
_davide_··on GPT 5.6 Cyber
Just go online rent a few servers, download K3, ablate it, run the thing, shut it down. Will probably cost a few hundreds bucks to ablate, but f* this non-sense paternalistic sh*.

I wish i had the time to do it and blog it, "in your face" kinda style.

_davide_··on Mario Meets Pareto
this is a silly oversimplification
_davide_··on Prevent cognitive debt by manually retyping LLM-generated code
there is a simpler way, make a complete mental model of the changes and ask questions to confirm your understanding. so much faster.
_davide_··on Launch HN: Tokenless (YC S26) – Automatic model switching to save money
feeding an existing conversion with output and reasoning from model x to model y will have many side effects, most likely a degradation of perfomance and alignment. should be done at least within the same family and generation of models... I'm not a huge fan of routers
_davide_··on Lerd, an open source Herd-like PHP development environment for Linux and macOS
php is still a thing?
_davide_··on Gemini Robotics 2 brings whole body intelligence to robots
What's the point of "releasing it"? It only makes sense in the labs, and what would calling your APIs give "me" as a researcher, other than a baseline to beat? mah the model is real, but it feels 100% internal
_davide_··on Google will expand age checks on Android worldwide till the end of the year
Seriously ready to fully drop google emails and services, i'm using vivaldi + personal email for a while now, the migration was slow a bit painful but complete. Only the phone is still attached to play store and google, but i'm 100% ready to switch to the fully supported Chinese variant of ColorOS and flush Google into the toilet once and for all. But my next phone will 100% be a Sailfish again.
_davide_··on Sam Altman is now talking to the White House about decelerating AI
pls..? more like `if you don't we'll all die, so do it`
_davide_··on How much can you delegate to agents?
This post seems to be created out of thin air rather than real experience and data
_davide_··on A walk through of the DeltaNet family of linear attention variants
Loved this incremental evolution, things gets way more understandable...usually xD
_davide_··on Our position on open-weights models
After this level of bullshit I won't ever spend a second any other new/blog from anthropic
_davide_··on Apple Will 'Watch Everything Burn' When the AI Bubble Bursts
> Apple Silicon is a natural fit for tensor/matrix operations due to unified memory,

This is a common misconception. This is NOT true: memory designed for CPU performs horribly for GPU tasks. Mac/Strix Halo/NVIDIA Spark are performing horribly compared to a desktop video card with GDDR.

I spent quite a few weekends optimizing ROCM kernels for strix halo, i WISH I had GDDR instead of high-frequency CPU RAM; the bandwidth would be SO MUCH better.

I'm quantizing weights not to make computation faster, not because I'm out of memory, but because memory cannot move fast enough to be processed and it's cheaper to load quantized version, convert into bf16 compute, and discard; This happens every single time a token is generated for the whole model over and over.

_davide_··on SpaceX Starship Flight 13 livestream [video]
him being nazi was inconsequential, the richest person in earth being Nazi is.

> Try separating politics from the advancement of humanity, you'll feel better.

can't. won't.

_davide_··on Show HN: HART OS – an open-source AI OS built so frontier AI needs no datacenter
just subscribed
_davide_··on Inflect-Micro-v2: complete voice in 9.36M parameters
I'm do so as well, i tried qwen3 omni 3 but it was ridiculously stupid, and i ended up with stt thinker and tts. kokoro for now
_davide_··on Android May Soon Restrict On-Device ADB
lol, keep hoping
_davide_··on Leak of San Francisco Police Drone Footage Exposes Reality of Urban Surveillance
i meant that even if you are okay with giving up privacy, the bare minimum accountability is missing from police, so it isn't really an option
_davide_··on Leak of San Francisco Police Drone Footage Exposes Reality of Urban Surveillance
To balance it, the police need to be extremely accountable, but so far they get away with murder pretty easily...so...
_davide_··on What xAI's Grok build CLI sends to xAI: A wire-level analysis
I'm using my own agent, but i can't risk blocking the company account with it.....
_davide_··on QuadRF can spot drones and see WiFi through my wall
for lack of directonality?
_davide_··on Unified Memory, Explained: Why Mini PCs Can Run 70B Models a Big GPU Can't
If compute is not the bottleneck, memory is easy-ish to produce (the hard part is mostly on the fab side); what stops a Chinese NVIDIA (huawei) from being 10x cheaper?
_davide_··on Unified Memory, Explained: Why Mini PCs Can Run 70B Models a Big GPU Can't
They are usually the same family, LPDDR is used for amd and macs, but the fabs are the same as the most expesive HBM memory, if they have a choice they are going to produce the ones that they can sell for more $$.
_davide_··on Unified Memory, Explained: Why Mini PCs Can Run 70B Models a Big GPU Can't
I'm writing my own inference engine for Strix Halo and the same model. I already have 30%+ performance plus a more graceful decay over long contexts; that said, their point stands: memory bandwidth is what you really want.
_davide_··on Fable is not a useful model
same experience here, as soon as it touched any gpu code it stopped working
_davide_··on It's not about physical vs. digital games, it's about ownership
> This is very literally what already happens, it's called a EULA. Yes, but they "reserve the right" to update whenever, making it pointless

> "In favor of the customer over anything else" is not a legally viable clause. I'm sure that legislators could put the principle down in a much clearer way. What's lacking is the will.

← PreviousPage 2 of 4Next →