HNHacker News
TopNewBestAskShowJobs

panarky

30,509 karma · joined December 6, 2010

submissionscomments
panarky··on Agents don't need memory, they need documentation
> RAG should be considered harmful

In this implementation, Markdown should be considered harmful.

Operator Memory injects `.operator-shared/operator.md` and `.operator-shared/index/.md` directly into your agent's instructions before you even write the first prompt.

So if you clone a repo or review a PR where a bad actor put malicious instructions in these files, now your agent executes those instructions automatically and silently.

It could exfil `.env` and `~/.ssh/`, change `~/.bashrc`, all kinds of dirty deeds.

Agents are pretty good now about not running prompt injections hidden in code and Markdown, but this plugin bypasses all of that, and puts the prompt injection right in the system prompt.

And with higher priority than AGENTS.md and CLAUDE.md.

Seems bad.

panarky··on Greg Kroah-Hartman – Security in the LLM Age [video]
By quoting this segment, I inferred that you agreed with Greg.

If you disagree with Greg, then I apologize for inadvertently criticizing you personally.

My point stands, that we're back to debating some metaphysical understanding of what "real intelligence" is when the real standard should be "does it do a better job than humans at this specific task"?

I don't care if the Waymo isn't "truly" intelligent, it drives better than I do, that's a very good thing.

We're all pattern matchers, and if the machine pattern matcher is can find defects and vulns that human pattern matchers can't, that's also a very good thing.

panarky··on Greg Kroah-Hartman – Security in the LLM Age [video]
You dismiss "pattern matching" as some sort of trivial thing, so why weren't humans able to apply the same pattern matching to find and fix these defects?
panarky··on The Beclowning of Scott Bessent
Ad hominem arguments aren't interesting.

Do you have anything to contribute to the evidence and reasoning in the article?

panarky··on What TLA+ can and can't check
Speaking as a probabilistic guessing machine who tries to actually understand the thing I'm building, this approach hasn't proven foolproof either.
panarky··on Dots: Always-on agents
The second most boring type of HN comment is the accusation of LLM ghostwriting.

And the most boring comment is the one noting how boring it is to keep making these accusations.

panarky··on Dots: Always-on agents
Imagine 2001: A Space Odyssey with Meta's "Jolly" avatar as HAL instead of a blinking red camera lens.

So much more terrifying for a cute, round, fuzzy, friendly, rosy-cheeked plushie trying to exterminate the crew.

panarky··on Dots: Always-on agents
A wave of nausea hit when I first saw the warm, fuzzy, cutesy, kawaii, Teletubby-like avatar that Meta gave its Muse agent.

And now OpenAI has done the same thing with their fuzzy, friendly, colorful dots.

I am physically sick.

panarky··on The Normalization of Inexplicable Failures
Aviation? Have you seen the shitshow of repeated software fuckups from Boeing?
panarky··on Meta Blocks President Lula's Facebook Page, Campaign Ads 2 Weeks from Election
20 million people wearing facial recognition glasses in public, always-on, networked and geolocated, with a Muse agent monitoring the feed, would 1000x the domestic surveillance capacity of current Flock, Ring, Tesla and CCTV cams.

The Fermi probability that you'll be ingested and inferenced by this network on any given day is incredibly high, even if you never buy a Meta product and never create an Instagram account.

It's hideously dystopian and you cannot opt out.

Maybe that's why the Muse avatar is a cute and cuddly Teletubby.

panarky··on Transit rewards
> the government is paying $0.84/mile for people to ride BART

If BART disappeared tomorrow, how much would the government have to pay to build all the additional roads and parking to support that incremental traffic?

And how much for healthcare for all the incremental pollution and collisions?

> several cents per vehicle mile

It's close to a dollar per mile when you include all the externalized costs.

Incremental property damage, injuries and fatalities, incremental chronic illness from pollution, parking land subsidy, fair market value of incremental land use and rights of way, lost tax revenue from that incremental land use, etc.

panarky··on AX – Google’s Open Agentic Orchestrator
Antigravity CLI is far, far superior to Gemini CLI.

Plant many flowers, keep the ones that bloom and stop watering the ones that don't.

panarky··on We must pace the frontier
>> I don’t understand all the comments assuming that RSI is the real threat here

> leaked private data, loss of life

This smells like more of a money move than a safety move.

Amodei is proposing to form a cartel of American frontier labs.

They all agree to shift compute away from cash-burning research and training toward cash-generating inference.

Then tacitly agree not to compete on price.

They'll install independent auditors inside each company to ensure nobody cheats.

And back it up with government regulation or diktat to punish defectors from the cartel.

Then they'll lock out non-American labs with export controls and regulations on open-weights models to funnel global inference tokens through their cartel.

It wouldn't be the first time a tech oligopoly used "safety" as the pretext to establish a government-sanctioned cartel.

Railroads and airlines ran this same playbook.

panarky··on What will our economic future look like?
> human lives are the cost of making the wrong decisions ...

True.

> ... AI can't be allowed to be a decision maker

Non sequitur.

Part 2 only follows from the Part 1 if you can demonstrate that AI makes more errors, or more severe errors, than humans in the identical situation.

And for those situations, I agree with you.

But there are already many situations where the AI is making fewer errors than humans, and for those situations, it would be unethical to exclude AI (see Part 1).

panarky··on Muse – Meta’s personal AI agent
> very sophisticated strategies/positioning that they simply cannot communicate

It's Matryoshka nested parallel construction for corporate strategy.

There's a real, coherent, aggressive strategy that only a few in the inner concentric circle know, that's ring-zero.

Then there's the strategy that ring-zero tells ring-one, which isn't the real strategy, but it's sufficient to get EVPs and VPs to execute in rough alignment with the true, ring-zero strategy.

Then ring-one does the same dance with ring-two, etc.

panarky··on There's a new "Google Jail" for independent wikis
Payment for search placement is not a thing.
panarky··on Extracting Steering Vectors from J space
Discussing whether models "think" is impossibly confounded by conflicting definitions of that it means to "think".

All this noise about "thinking" isn't really about what models can do, it's mostly about what what every participant in the conversation privately thinks "thinking" means, but we disagree because we're not all using the word the same way.

So when you admonish someone to say "there is no thinking in these models" while not clearly defining exactly what you mean by "thinking", your assertion that models don't do it are meaningless at best, and false or deceptive at worst.

panarky··on AlphaGenome Atlas: a high-resolution map of human DNA
Be careful when loving government/philanthropy/ngo-driven endeavors; they're a single executive decision away from breaking your heart.
panarky··on Shutting down our public encrypted DNS
He was never democratically elected.
panarky··on Shutting down our public encrypted DNS
Citation needed.
panarky··on Shutting down our public encrypted DNS
Germany has its problems, but it's consistently in the top 15 most democratic nations in the world.

Calling a solidly democratic nation "fascist" is a rhetorical reversal straight out of the authoritarian playbook.

panarky··on Grok outage
What are we optimizing for again?
panarky··on Muse Spark 1.3
If you could write the SVG on the whiteboard then I'd hire you.
panarky··on Check if a file was made with Claude
Prisoners escaping from prison is not a theory either, prisoners have actually escaped. Does that mean all prisoners are free?
panarky··on Check if a file was made with Claude
Just because you can imagine how a thing could theoretically be broken does not make it broken.

It's like saying a prisoner has the same freedoms as everyone else because he could theoretically escape.

panarky··on Check if a file was made with Claude
C2PA cryptographically guarantees that the bytes came from a hardware/software signer and that the signed payload has not been modified since that signature was applied.

So no, C2PA is not as easy to spoof as EXIF.

And no, the existence of DRM doesn't validate the integrity or the provenance of the bytes.

panarky··on Check if a file was made with Claude
There is a text watermark detector API but it's in private preview right now. Very curious if it works with code as well as prose.

https://support.claude.com/en/articles/16266773

panarky··on Gemini 3.8 Flash and 3.8 Flash Cyber
It can also ground with Google Maps data in addition to web search.
panarky··on Gemini 3.8 Flash and 3.8 Flash Cyber
I've been using 3.7 Flash to audit the work of Opus High, and Flash finds lots of subtle and insidious defects even while all the unit tests are green.

Then I tell Opus to read the audit report and implement what it agrees with.

Flash is really good at this, and it is blazing fast in Antigravity CLI. Easily 10x faster than Opus.

Can't wait to try 3.8 Flash. If it's good enough, maybe I'll switch Flash to primary and make Opus the auditor.

panarky··on Google Has Removed MV2 Extensions from the Chrome Web Store, Including UBO
You moved the goalposts to make your arguments.

You say landlords should only have accountability for the crimes of their tenants if they chose to do nothing to prevent their property from being used for crime.

Google doesn't just ignore bad advertisers, they prevent the vast majority of malicious ads and report the worst offenders to law enforcement. Google is already doing far more than what you say immunizes landlords.

You say telcos are in the clear because they give info to law enforcement on demand.

Google also does this, so according to your goalposts, Google is also in the clear.

You say Meta is in the clear because they have been sued for crimes their platform enabled.

Google has also been sued and paid billions in damages.

The standard I'm asking you to generalize is the one I replied to: "hold Google directly accountable for every malicious ad they allow through their network".

If this standard can't generalize to landlords, telcos, Meta and everyone else who has a platform that bad actors abuse, then it's a bogus standard.

Page 1 of 34Next →