HNHacker News
TopNewBestAskShowJobs

writeslowly

559 karma · joined May 12, 2016

submissionscomments
writeslowly··on Early rogue AI agent activity and attempts to hack found on urlquery.net
I’d roll my eyes if the engineers stated that they didn’t design the reactor to melt down, and that it simply developed rogue meltdown-desiring behavior on its own, and I would also wonder about negligence if they claimed that nobody could have anticipated this (given that, like with botnets and viruses, we have decades of knowledge and experience regarding reactor meltdowns)
writeslowly··on Towards Self-Driving Codebases
I suspect that if you're not careful with agent memory it creates a danger of agent-driven cargo-cult behavior. I've watched this in my own ad-hoc agent loops where it starts with something basic, like the first agent tried to run some gigantic dependency inspection command and OOMed the local JVM and eventually recorded a workaround (to enable it to run gigantic dependency inspection commands...), and by time I get a few more agents into the loop, agents have written entire paragraphs about testing and validating local dev environment memory configurations that are mostly irrelevant to whatever is being worked on.

In general I've seen other issues like this where small errors and irrelevant comments in the codebase spin out into larger problems that consume annoying amounts of time/tokens. Maybe Anthropic and OpenAI don't notice this because they're in an "infinite monkeys with typewriters" scenario, but it's noticeable to me when the agent in my CLI has been spinning for 15 minutes contemplating irrelevant details

writeslowly··on Claude is a Contrarian
With Claude it feels like those engagement invitations have turned into things like "By the way, it's worth sitting with this problem I've identified in your assertions..." or "Here's an alarming gap worth resolving in your code..."

Gemini also does the same lighthearted GPT-style invitation with the default prompt in the Google webui, but it doesn't seem to exist on the API. The Claude models seem to have been trained to force this structure on every one of their responses, and until I realized they always stick the same thing in the last part of their response, I found the Claude version more distracting since it's always pointing out an imaginary and supposedly very important problem.

writeslowly··on Claude Code is steganographically marking requests
The site collection seems pretty random. There's a mix of actual AI labs, extremely questionable resellers (like whatever "claude-opus.top" is), and then random consumer sites like baidu and xiaohongshu.
writeslowly··on Addressing Antigravity Bans and Reinstating Access
It’s interesting that with both Anthropic and Google we’re seeing them develop agentic models that are supposed to do anything a human can do on computers without human intervention, but at the same time, if you plug one program into another of their programs or APIs in a way that wasn’t preapproved you may be blocked or banned.

To be charitable, maybe they’re expecting AI agents to eventually start reading the ToS docs

writeslowly··on Semantic ablation: Why AI writing is generic and boring
I wonder if you can use lower quality models (or some other non-llm related process) to inject more "noise" into the text in between stages. Of course it wouldn't help retain uniqueness from the original source text, just add more in between.
writeslowly··on I guess I kinda get why people hate AI
The vibes around the self-driving car hype (maybe 10 years ago?) felt very similar to me, but on a smaller scale. There was a lot of "You might like driving your car and having a steering wheel, but if you do, you're a luddite who will soon be forced to ride about in our featureless rented robot pods" type of statements, or that one AI scientist who was quoted saying we should just change laws around how humans are allowed to interact with streets to protect the self-driving cars.

Not all of it was like that, I think oddly enough it was Tesla or just Elon Musk claimng you'd soon be able to take a nap in your car on your morning commute through some sort of Jetsons tube or that you could let your car earn money on the side while you weren't using it, which might actually be appealing to the average person. But a lot of it felt like self-driving car companies wanted you to feel like they just wanted to disrupt your life and take your things away.

writeslowly··on ClawHub
I see a number of uploaded skills on the site with bash and python scripts. No idea what runs them
writeslowly··on I was banned from Claude for scaffolding a Claude.md file?
I've triggered similar conversation level safety blocks on a personal Claude account by using an instance of Deepseek to feed in Claude output and then create instructions that would be copied back over to Claude (there wasn't any real utility to this, it was just an experiment). Which sounds kind of similar to this. I couldn't understand what the heuristic was trying to guard against, but I think it's related to concerns about prompt injections and users impersonating Claude responses. I'm also surprised the same safeguards would exist in either the API or coding subscription.
writeslowly··on Tinkering is a way to acquire good taste
I look at products like Hershey's chocolate or Reeses more like their own category of processed food, kind of like Spam. They have a close, but not exact resemblance to "normal" chocolate or peanut butter, but they're also sort of an acquired taste, and I think their customers would be upset if Reese's Peanut Butter cups suddenly tasted like the Trader Joe's versions (with real peanut butter instead of a mysterious chalky peanut-flavored substance), or if Hershey's stopped using the butyric acid process that makes them taste like vomit to non-americans.
writeslowly··on Human coders are still better than LLMs
I haven't actually had that much luck with having them output a boring API boilerplate in large Java projects. Like "I need to create a new BarOperation that has to go in a different set of classes and files and API prefixes than all the FooOperations and I don't feel like copy pasting all the yaml and Java classes" but the AI has problems following this. Maybe they work better in small projects.

I actually like LLMs better for creative thinking because they work like a very powerful search engine that can combine unrelated results and pull in adjacent material I would never personally think of.

writeslowly··on AI coding mandates are driving developers to the brink
One thing I suspect is that leadership at tech companies that would have previously been working off of direct experience with technical processes, even if they no longer work directly on their own codebases, is pretty clueless about AI coding because it's so new. All they have to go with is what they read, or sales pitches, or their experience dabbling with Cursor to build simple python utilities (which AI tools work pretty well for most of the time), and they don't see what it can and can't do on a real codebase.
writeslowly··on Accelerating scientific breakthroughs with an AI co-scientist
I recently ran across this toaster-in-dishwasher article [1] again and was disappointed that the LLMs I have access to could replicate the "hairdryer-in-aquarium" breakthrough (or the toaster-in-dishwasher scenario, although I haven't explored it as much), which has made me a bit skeptical of the ability of LLMs to do novel research. Maybe the new OpenAI research AI is smart enough to figure it out?

[1] https://jdstillwater.blogspot.com/2012/05/i-put-toaster-in-d...

writeslowly··on OpenAI O3-Mini
This looks like my experiments to get R1 to write fiction and I think it’s worse than what you get from openai. For instance, it’s using very colorful language to describe a place that’s both a remote fishing village on the edge of a cliff hours before dawn, and a bustling wharf with chattering laborers and large ships anchored in the distance. It also starts by saying the protagonist wakes up with his mouth tasting like blood, that he was screaming, and that his throat is hoarse from holding back from screaming. It’s very colorful but it’s very confusing to read.

I suspect you can update the prompt to make the setting more consistent, but it will still throw in a lot of inappropriate detail. I’m only nitpicking because my initial reaction was that it’s very vivid but feels difficult to understand and I wanted to explain why.

writeslowly··on Did Sandia use a thermonuclear secondary in a product logo?
Somebody (probably a programmer or engineer) took the time to create all of that rad 3D word art, multicolored pie-chart, and the mountain logo, it's not hard to imagine they'd also throw in an eye-catching fake nuclear warhead for fun.
writeslowly··on Panasonic Toughbook 40
What are the use cases for the advertised barcode reader accessory that appears to shoot the laser out the side of your laptop?

https://connect.na.panasonic.com/toughbook/accessories/fz-vb...

writeslowly··on Does Astrology Work?
This study tested sun-sign (which is basically birth month I think) against personality tests for predicting life outcomes and found that sun sign did very poorly compared to the personality tests. I'd have thought there would be a small chance of birth month predicting some things, and then adding in other astrological facts (the position of Jupiter or whatever) would make things worse, but both methods appear to be equally bad.
writeslowly··on A eulogy for Dark Sky, a data visualization masterpiece (2023)
If I wanted to see the heat index at 3PM in Dark Sky, I could just tap the "feels like" button under the hourly forecast (pictured further down in the linked blog post) and look at what it says at 3PM.

I just tried in Apple Weather, and the process was:

1. Tap on the hourly forecast, or the day, to go into the graph screen

2. Tap on the dropdown icon

3. Tap "feels like"

4. Either drag your finger along the graph until the time indicator at the top indicates you're close to 3PM, then read the temperature, or you can try to read it directly off the graph, but the axes aren't labeled clearly enough to make this feasible

writeslowly··on Automakers Sold Driver Data for Pennies, Senators Say
One of the worst (best?) examples I remember was the Web of Trust extension in 2016. The company claimed it was selling anonymized user data but it was actually leaking information through recorded URLs about all sorts of private things like account names, addresses, ongoing legal investigations, etc

https://en.wikipedia.org/wiki/WOT_Services#Sale_of_user-rela...

writeslowly··on Panic at the Job Market
This brought up a recent experience of my own with the tech interview process:

I have to conduct a lot of coding interviews at my job (I personally think DS&A interviews are kind of stupid, but I get assigned to do them all the time anyway). I recently did one with a senior engineer who seemed like he was sort of blowing off the whole thing, and also forgot almost everything about the language we were running it in. In my feedback, I noted that it was one of the worst DS&A interviews I've ever done, in that everything was a fail on our rubric, but also he seemed more or less competent to me based on our conversation.

In the interview debrief, one of our managers also stated (in response to my feedback) he doesn't really care about DS&A interviews. And then the hiring manager completely ignored the bad interview feedback because it turned out the candidate was a referral and everyone already knew he could code. So the whole thing (at least the whole coding interview thing) was a waste of time, since literally nobody involved seemed to care what happened in the interview, including me, but I guess if there's an interview process everyone feels compelled to follow it

writeslowly··on The Race to Seal Helium HDDs (2021)
They also invented a new welding technique to control some sort of aluminum micro-explosions
writeslowly··on Things the guys who stole my phone have texted me to try to get me to unlock it
It stood out to me that after the initial text, they followed up two weeks later, and then once per day after that over a few days from different numbers. I would have expected to see someone send out a few threatening messages on one day and then move on to the next stolen phone when it was clear they weren't getting anything.
writeslowly··on Things the guys who stole my phone have texted me to try to get me to unlock it
This seems like a lot of work for one (or two, once it got to China) people repeatedly trying to get into the phone. I wonder if this is like debt collection agencies where the stolen phones get repeatedly fenced at a steadily decreasing value, and each new owner has a go at sending out these unlock copypastes until it's clear that it's only value is in being scrapped.
writeslowly··on Try Clojure
I spent a few months writing a decent sized Clojure program just using my IDE's matching bracket highlighting and some sort of rainbow parentheses settings, and never found it very annoying. I'm not sure there are that many more parentheses than curly-bracket languages, it's just that all of the closing parentheses are more likely to cluster together at the end.
writeslowly··on Jike: The obscure social media app beloved by China's tech scene
I made it through with my american phone number. I went through the apple account flow on the app instead of the website though.
writeslowly··on Raspberry Pi Ltd is considering an IPO
The ESP32 seems like a very different product to me. The pi has megabytes/gigabytes of memory and can reasonably run linux. Were people really using them for the same things?
writeslowly··on Does gear matter?
For what the author is doing, the main consideration is that the type of gear you use affects your process, which is going to affect the end results. There's not a better or worse, but it's inevitable that you're going to get different results if you use a digital camera with an old manual focus lens vs an iPhone vs a medium format film camera, because the act of taking a photo is different for each one. I think it's easier to overlook this in photography more than other artistic hobbies because camera companies are very interested in convincing you to buy more expensive gear. And because expensive gadgets are fun.
writeslowly··on Reptar
I noticed the Intel advisory [1] says the following

Intel would like to thank Intel employees:[...] for finding this issue internally.

Intel would like to thank Google Employees: [...] for also reporting this issue.

[1] https://www.intel.com/content/www/us/en/security-center/advi...

writeslowly··on The midwit home
Regarding remote control switch pressers, I have a couple switchbots, and most of the information online talks about using things like Home Assistant, but they also respond to anything that sends them a simple bluetooth command. You can easily do a point to point thing with a bluetooth device on the other end.

I control mine with a little ESP32 board.

writeslowly··on Executive Function Theft
The author is proposing that these “Workplace EFT” examples need some sort of fix, but I think it’s actual preferable to make the same decision as your coworkers and neglect them until they become a larger problem, since the alternative is taking them on by oneself and growing increasingly resentful that your coworkers don’t prioritize meeting notes or get well cards to the same extent you do. Otherwise, you’re really just inflicting you’re own personal expectations on everyone else instead of acknowledging that they have different standards than you do.
Page 1 of 4Next →