HNHacker News
TopNewBestAskShowJobs

simsla

597 karma · joined September 24, 2015

submissionscomments
simsla··on Software factories and the agentic moment
My read of it was "by today", aka cumulative. But you're right that it can also be read as "just today". The latter is an absurdly strong statement, I agree.
simsla··on Software factories and the agentic moment
It doesn't say 1k per day. Not saying I agree with the statement per se, but it's a much weaker statement than that.
simsla··on Sometimes your job is to stay the hell out of the way
> Wolves don’t care if they are seen or not.

Well... As one of those supposed 10x engineers, that's not quite true

It's true that the intellectual satisfaction is my main driver, but I'm also quite vain. Appreciation and respect (especially from peers, who cares about an All Hands) add juice to the battery.

That's just me though.

simsla··on Antirender: remove the glossy shine on architectural renderings
It's probably just prompt based. Actual fine-tuning for these kind of use cases is getting less common than it used to be.
simsla··on Claude Code daily benchmarks for degradation tracking
Probably, but with a small sample size like that, they should probably be taking the uncertainty into account, because I wouldn't be surprised if a lot of this variation falls within expected noise.

E.g. some binomial interval proportions (aka confidence intervals).

simsla··on A macOS app that blurs your screen when you slouch
This was me, and now I have horrific back pain almost every week. Fix what's broken before it breaks you.
simsla··on Are we all plagiarists now?
Depends if it's sufficiently transformative or not.
simsla··on GPTZero finds 100 new hallucinations in NeurIPS 2025 accepted papers
Pointing out these errors isn't wrong. But making the leap to "therefore: AI hallucinations!" without substantiating those accusations is.
simsla··on Building an AI agent inside a 7-year-old Rails monolith
Yes, and a YouTube video more than a text article. etc. etc.

It's a tool. The main question should be: is it useful? In the case of AI, sometimes yes, sometimes no.

simsla··on AWS CEO says replacing junior devs with AI is 'one of the dumbest ideas'
I don't think that's the same. I spitball crazy ideas, but my core knowledge/expertise is sound, and I try not to talk out of my ass. (Or I am upfront when I'm outside my area of expertise. I think it's important to call that out once your word starts carrying some weight.)

A product manager can definitely say things that would make me lose a bit of respect for a fellow senior engineer.

I can also see how juniors have more leeway to weigh in on things they absolutely don't understand. Crazy ideas and constructive criticism is welcome from all corners, but at some level I also start expecting some more basic competence.

simsla··on We Induced Smells With Ultrasound
Good / bad / unclassified.

It makes sense for unclassified to smell worse than good, and it'd probably be the biggest category by a long stretch.

(Pure speculation.)

simsla··on A $1k AWS mistake
You could set a cloudwatch cost alert that scuttles your IAM and effectively pulls the plug on your stack. Or something like that.
simsla··on Why I'm Learning Sumerian
Those models are terrible. I tried one of the "best" ones and it told me the US Constitution was AI generated.
simsla··on Laptops with Stickers
My laptop at Amazon was also covered in stickers, although I shied away from the more politically charged ones.
simsla··on ChatGPT terms disallow its use in providing legal and medical advice to others
I think it's simpler than that, and they're just trying to avoid liability.

With all of their claims about how GPT can pass the legal/medical bar etc. I wouldn't be surprised if they're eventually held accountable for some of the advice their model gives out.

This way, they're covered. Similarly, Epic could still build that feature, but they'd have to add a disclaimer like "AI generated, this does not constitute legal advice, consult a professional if needed".

simsla··on China has added forest the size of Texas since 1990
Damn, I goofed, thanks for calling that out. Can't edit the comment anymore unfortunately.
simsla··on China has added forest the size of Texas since 1990
European here.

One of the problems is that many of our society's systems are predicated on a growing population. Social security and pensions, for example, are structured not unlike a pyramid scheme: for every old person we should have more than one working young person. People take more than they give. Fixing that will be painful, but possible.

More worrying is how many countries' birth rates have fallen below the replacement rate. Some SE Asian countries are interesting case studies here (Japan, S Korea), but it's not looking good, and much of western Europe is heading in the same direction. Maybe the worry is overblown and populations will eventually stabilize at a lower point, but currently it seems like a declining population will just add to the stressors that are putting people off from having children, so it could just as well keep snowballing.

All that's to say, I don't worry too much about over/underpopulation, but I do worry about a shrinking population.

simsla··on Show HN: Sober not Sorry – free iOS tracker to help you quit bad habits
Is there any way to get notified if/when an Android app becomes available?

Great work, app looks great!

simsla··on How to inject knowledge efficiently? Knowledge infusion scaling law for LLMs
There's no inductive bias for a world model in multiheaded attention. LLMs are incentivized to learn the most straightforward interpretation/representation of the data you present.

If the data you present is low entropy, it'll memorize. You need to make the task sufficiently complex so that memorisation stops being the easiest solution.

simsla··on You did this with an AI and you do not understand what you're doing here
Agreed.

I've found some AI assistance to be tremendously helpful (Claude Code, Gemini Deep Research) but there needs to be a human in the loop. Even in a professional setting where you can hold people accountable, this pops up.

If you're using AI, you need to be that human, because as soon as you create a PR / hackerone report, it should stop being the AI's PR/report, it should be yours. That means the responsibility for parsing and validating it is on you.

I've seen some people (particularly juniors) just act as a conduit between the AI and whoever is next in the chain. It's up to more senior people like me to push back hard on that kind of behaviour. AI-assisted whatever is fine, but your role is to take ownership of the code/PR/report before you send it to me.

simsla··on Bulletproof host Stark Industries evades EU sanctions
Elon was three years old when the first Iron Man comic book came out.

EDIT: and the movies are pretty faithful to the comic books.

simsla··on I'm absolutely right
I was just thinking about how LLM agents are both unabashedly confident (Perfect, this is now production-ready!) and sycophantic when contradicted (You're absolutely right, it's not at all production-ready!)

It's a weird combination and sometimes pretty annoying. But I'm sure it's preferable over "confidently wrong and doubling down".

simsla··on Everything is correlated (2014–23)
This relates to one of my biggest pet peeves.

People interpret "statistically significant" to mean "notable"/"meaningful". I detected a difference, and statistics say that it matters. That's the wrong way to think about things.

Significance testing only tells you the probability that the measured difference is a "good measurement". With a certain degree of confidence, you can say "the difference exists as measured".

Whether the measured difference is significant in the sense of "meaningful" is a value judgement that we / stakeholders should impose on top of that, usually based on the magnitude of the measured difference, not the statistical significance.

It sounds obvious, but this is one of the most common fallacies I observe in industry and a lot of science.

For example: "This intervention causes an uplift in [metric] with p<0.001. High statistical significance! The uplift: 0.000001%." Meaningful? Probably not.

simsla··on Knuth on ChatGPT (2023)
The problem is that ChatGPT doesn't really know letters, it writes in wordpieces (BPE), which may be one or more letters.

For example, something like "running" might get tokenizef like "runn"+"ing", being only two tokens for ChatGPT.

It'll learn to infer some of these things over the course of training, but limited.

Same reason it's not great at math.

simsla··on Self-Signed JWTs
The problem I see with (1) is that it becomes a little bit too easy to regenerate public keys and circumvent free tier metering.
simsla··on Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic
In general? In the past I've known ASS to be used a lot for things like anime, but less for live action shows.
simsla··on Complete silence is always hallucinated as "ترجمة نانسي قنقر" in Arabic
At least for English, those "fansubs" aren't typically burnt into the movie*, but ride along in the video container (MP4/MKV) as subtitle streams. They can typically be extracted as SRT files (plain text with sentence level timestamps).

*Although it used to be more common for AVI files in the olden days.

simsla··on Let me pay for Firefox
I've been in organisations with great developers but no leadership. It's a shit show.
simsla··on Woman takes 10x dose of turmeric, gets hospitalized for liver damage
Damage is usually done in aggregate. Leaded gasoline didn't have people dropping like flies, but still caused significant damage.

Although it seems this needs more research, I'd be wary dismissing it out of hand just because people haven't been having an acute reaction.

simsla··on Slack's 57MB 404 page
Mostly worked at bigger companies. Once joined a company with pretty shit colleagues/management. Left after three months.

If you have mobility, it's worth shopping around for a decent place to work.

← PreviousPage 2 of 7Next →