HNHacker News
TopNewBestAskShowJobs

paimapi

2,077 karma · joined July 31, 2026

"Once it's flattened, it could never be rebuilt.”

“Because why? Got any salt?”

“Because all the available surface metals and coal have already been mined,” said Crake.

“Without which, no iron age, no bronze age, no age of steel, and all the rest of it. There's more farther down, but the advanced technology we need for extracting those would have been obliterated.”

— Margaret Atwood, "Oryx and Crake"

you can reach me at hn.record316@passmail.net :)

submissionscomments
paimapi··on Why does Opus 5 feel worse to work with?
so why don't people want to work for them? they don't get paid enough? what if they were paid more? what if they were FTE with all the benefits? what if AI projects were nationalized and trainers were funded by grants? what if we increased the NIH budget, made peer review and journals far less exploitative of researcher's time, and got rid of academic middle management, focusing mostly on paying more towards actual research and academia?

definitely a utopian vision that is not likely to happen in our current reality but I like to imagine better worlds that are possible. as LeGuin once said, "We live in capitalism, its power seems inescapable — but then, so did the divine right of kings. Any human power can be resisted and changed by human beings."

paimapi··on Why does Opus 5 feel worse to work with?
sure, that's very much the 'just let people like things' argument where literal white supremacists can enjoy Rage Against the Machine in spite of the music literally being in total opposition of their ideology

everyone's free to enjoy media however they want, with whatever level of interpretation they like. I provided the BB references as short examples, they aren't meant to represent the definitive diagetic experience of the show. if you have a different view, great. if it was triggering for you to hear 'toxically masculine', also fine but... might be something worth self-examination on given that it's very much also a clinical and academic term [0] as much as it is one steeped in the artificially manufactured culture wars by people who don't want to change their anti-social behaviors

I would also say that understanding Scar as a villain is the sixth grade reading level understanding of the character. and you're free to stay at that level of understanding. someone who wants to engage more critically might map the character to Claudius, comparing and contrasting how they're characterized given the context of the audience for Disney and Shakespeare, and appreciate the character that way, as a standardized trope utilized throughout all other forms of media. they may even get a tattoo of Scar, symbolizing their dive into the analysis

my point is not that people should or shouldn't engage critically. it's that this practice of critical engagement, of being skeptical and analytical provides you with the skills to not be a total sucker who falls for the latest manufactured fad that someone with a strong theoretical understanding of semiotics and social capital created (ie most modern marketers). the pertinent example being how people engage with AI - seeing it either as a specialized tool with a set of flaws that need to be accounted for and checked against or as just some kind of authoritative voice because it sounds smart and so-called smart people like Elon are terrified of it and AGI. which, again, if you prefer the latter engagement all power to you but the chances of you taking some really bad advice forward is not negligible

[0] https://www.wi.edu/news-Shifting-the-Conversation-From-Toxic...

paimapi··on Why does Opus 5 feel worse to work with?
I see it as a problem of context which maps to my theoretical understanding of LLMs. models trained on large data sets will probabilistically veer towards the median in all aspects - reasoning, assumptions, environments, etc. specialized context about your specific codebase's solutions don't exist unless you add them in, either in the prompt, as a skill or rule, or more generally in the harness via memories, tests, etc (though ideally a combination of all of the above). without that the LLM will suggest the median solution for the median codebase according to some ephemeral, unqualifiably trained understanding of best practices

it makes me wonder if the solution that businesses/users need to implement is just the same solution to everything since the beginning of time ie standardization. skills/harnesses/agents.md/etc maintained by codeowning teams that must be invoked for AI-assisted code changes on ABC part of the codebase, these existing as replacement for the bevy of other documentation required for the days of hand-written code. a company-wide orchestration skill knows how to search and pull down the relevant .mds, cleans it as cruft at the end of a session, every merge with a short changelog saved to a corpus somewhere with a TOC + appendix that an LLM can navigate to and read for context, major changes in the logic documented in the working skill doc, all of it generally automated but requiring HITL vetting

this wouldn't fully solve the problem of subject matter expertise but it seems like it would remove a lot of the friction for new employees and other teams with dependencies on your work or with whom you have dependencies

paimapi··on Stop sending me huge PRs; a rant
is there a reason why there's no standardization in orgs in terms of skills/harnesses/etc for AI-assisted development? for example, a rule of 'you must invoke ABC skill that contains all of the context for this part of the codebase if you plan on making changes there' with the codeowning team dedicated to maintaining it both for their own use and for the use of other teams that have up or downstream dependencies
paimapi··on Why does Opus 5 feel worse to work with?
I do appreciate the thoughtful response to a really long wall of text lol. and yes, I agree - I think that'll be the lesson that society is going to take probably far too long to learn, to not see everything as a nail that AI can hammer at. a lot of tech companies are in essentially a 'fuck around and find out' phase with AI taking over code review, testing, etc. combined with the expectation of shipping 3X the amount of code, we've enshittified the entire SDLC. and so we have near-daily incidents, data leaks, etc, something that I was able to measure and report on at my old place of work to, well, no avail

it's the old tortoise vs hare parable, I think. go fast, make a bunch of mistakes, get too arrogant, and you lose out. your forum might be slightly lower engagement now while people are caught up in the latest fad but your rules are proactive for a future where average people hopefully realize that you can't trust an LLM that has zero context, no real harness and determinstic tests to speak of, and a propensity towards probabilistic rabbit holes that result in hallucinations. at least that's the kind of space I'd look for now and largely why I've given up on a lot of other forums

paimapi··on Cursor is now a part of SpaceX
or the name of Elon and Grime's actual child
paimapi··on Why does Opus 5 feel worse to work with?
there's a wide array of assessments when it comes to reading comprehension. the one you refer to, the GRA, sets the 'sixth grade level' as whether or not a reader understands the author's main points, is able to answer conceptual questions related to the text, and then apply those to relevant situations. beyond this level is the ability to essentially be skeptical of a text and to know how to critically analyze it. so if your comprehension level stops before this you get 'big words in complex sentence structure sounds smart and right so it is smart and right' even if the reasoning and process is poor

it makes me think about how people engage with movies and television - as passive, plot-and-character driven consumption (eg I hope Walter White survives) with no critical analysis of how and why the writers added ABC thematic element (eg Walter White as a motif of a toxically masculine narcissist with specialized knowledge as a larger critique how mass media tends to valorize their male leads in the same vein as many other prestige shows at the time like Mad Men), and the larger, downstream sociocultural impact that piece of media has on how people see the world (eg people who now have the Heisenberg tattoo, unironically)

there's been some musings on why this the case like Hofstadter's Anti-Intellectualism in American Life - the valorization of obedience and trust in hierarchy and the state are net wins if you're an institution that seeks to increase it's power, whether religious or governmental. I was talking about this with a few friends the other day and it's a dismal future reality where not only did we make anti-intellectualism normalized and politically legitimate in the USA (eg Fox News, clickbait articles, and all the other forms of yellow journalism that have emerged), we now have tools by which individuals can even further remove themselves from having to critically engage with thoughts, feelings. I heard a story about how someone scanned a group activity at a baby shower into ChatGPT and had it answer for them instead of, well, socially interacting with the other guests and forming a memory of the moment with their friends

the counterargument to that might be that Claude/ChatGPT/etc have more epistemic rigor than your average American (sure) but the sycophancy of modern day LLMs is an actual danger that enables more harm than good. it does seem as if Claude is the only one interested in guarding against some small amount of it (though to the detriment of people just trying to get work done. as an aside, I get the feeling Mythos was intended to be the bespoke enterprise solution without the guardrails but the Anthropic marketing department or some power-hungry department lead made it about how dangerous/effective it was from a security perspective which threw a wrench in things). but then I think about people like my parents asking ChatGPT which specific house to buy in their retirement only to later find out the house was sold weeks ago, or just in bad condition, or in a neighborhood where the housing value has already reached equilibrium, it makes me think about how it's not enough and the future is bleak

I'll also say that I think Claude sounds the way that it does because it, like many other LLMs, are RLHF trained largely by lowly paid gig-workers, many of them ESL speakers. if their trainers were, for example, dedicated and highly trained academics, scientists, and other researchers, you'd likely see a lot more concise and more importantly skeptical reasoning and responses. but that won't happen in our current reality of capitalist-driven development so we get encoded solutions like MoE that still largely depend on the messy, imprecise RLHF training at baseline

in the right hands, I do think AI is a wonderful tool. one of the first things I did with it was to create a research skill that reviews white papers from the lens of someone who knows how to read/interpret research methodology, is aware of things like p-hacking, and deterministically assigns weight according to the hierarchy of evidence. even still, I'll still read the studies because there's so often nuance that's missed if the sub-agent read only a search snippet but that takes effort, time, and the practiced knowledge of critical analysis to even want to do it

paimapi··on Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index
cool

here's Stanford HAI's graph on the carbon emitted from model training per model:

https://spectrum.ieee.org/media-library/chart-showing-estima...

note that Grok's training, thanks to its portable gas generators that are magnitudes less efficient than even other integrated, permanent gas turbines, means the training for this model is dramatically less efficient than models like DeepSeek

a lot of the CO2 emission debate on AI is overblown but it's accurate for Grok

paimapi··on Georgia police officers fired after Flock camera misuse
"it's not the gun, it's the person that holds it" is not a philosophy that's tracked with me given the probabilistic, population-wide outcomes. the same applies here - tools should build in strict data and privacy protections as part of their UX. you lower the probability of abuse wherever you can because you know a bad actor will find a way, regardless, but the point is to add enough friction that even the worst of the bad actors has a difficult time getting what they want
paimapi··on Woman pulled over twice after Flock-linked software connected her to homicide
a significant portion of police training literally teaches them to fear your regular citizen. they're taught that anybody may be armed and they face the chance of death at every traffic stop and are shown dashcam footage of the worst incidents where this did happen without the probabilistic context

see https://revealnews.org/podcast/what-cops-arent-learning/

paimapi··on Georgia police officers fired after Flock camera misuse
I think the point we disagree on is the idea that accountability by a singular department is an indication that it's sufficient for vendor to avoid regulatory action. in terms of privacy protections, the strategy should be defense in depth. I think about the audits that happen in health tech for eg on the software side, both at the request of regulatory agencies and their clients

any platform that enables people to be stalked, harassed, tracked, and so on should face the same level of scrutiny in the form of auditing and open access to internal policies around data sharing

paimapi··on Illinois just passed a law that puts Linux on the hook for age verification
they're both grounded in fearmongering that's led to cultural panics and utilize the rhetoric of 'protecting children' as the sell

https://www.law.georgetown.edu/gender-journal/wp-content/upl...

"Overwhelmingly, state legislators behind recent anti-trans legislation have argued that they were introducing these bills with the intent to protect children – whether that be “protecting children” from “profoundly unethical and morally wrong” medical practices"

https://www.eff.org/deeplinks/2026/06/eff-gov-pritzker-veto-...

"Much of H.B. 5511 is modeled after controversial legislation passed in California (A.B. 1043) and New York’s Stop Addictive Feeds Exploitation (SAFE) for Kids Act"

there's also the rampant astroturfing

https://www.theguardian.com/world/2023/jun/30/christian-hate...

"The surge in funding to the ADF, which has been termed an “anti-LGBTQ hate group” by the Southern Poverty Law Center, saw it record revenue of $104.5m in 2021, according to filings with the Internal Revenue Service."

ADF's main funders being the Servant Foundation or a single billionaire family's pet foundation (https://en.wikipedia.org/wiki/Servant_Foundation) and the National Christian Foundation which receives large donations from multiple extremely wealthy individuals like David Green, co-founder of Hobby Lobby (https://readsludge.com/2019/03/19/americas-biggest-christian...)

this is very much like the funding of these 'privacy' acts by Meta, Alphabet, etc - lots of money directly into the hands of local legislators across multiple state legislatures with the ultimate goal of providing institutional power over individuals, one just has a slightly more Machiavellian economic logic foregrounding it but the strategy, tactics, and overall erosion of civil liberties make for very similar outcomes

paimapi··on Illinois just passed a law that puts Linux on the hook for age verification
it depends on whether or not there's someone who sees it as a crusade (rare), sees it as a way to get political clout (very common) and/or is getting a lot of lobbying money (extremely common [1]). see, for eg, the anti-trans bills that are being passed

this is the same rhetorical and political strategy, that there are 'dangerous' people who will exploit your children so please vote for me, the person who cares the most about children and will go after the 'dangerous' people

[1] https://themarkup.org/privacy/2021/04/15/big-tech-is-pushing...

paimapi··on Georgia police officers fired after Flock camera misuse
it's also possible that we're only seeing a few select departments fire their officers. there are no laws around this as far as I'm aware so enforcement is purely departmental policy. how many sheriff's offices and police departments do these audits and how many of them act on abuses?

you should need a warrant to monitor the products of mass surveillance and it should be as narrowly targeted towards a suspect as possible. or just not have any of the mass CCTV placements everywhere in the first place

paimapi··on Depression has tripled in the last 15 years. Arthur Brooks about the cause
my partner is a psychiatrist (that's the one with an MD)

in our household we treat all overly broad, trendy statements like this one with a huge grain of salt

the difference between rates of diagnosed major depressive disorder (MDD) vs the past has much more to do with a combination of the diagnostic criteria being refined [1] over the years and the increase of screening [2] during your annual checkup (something physicians were recommended against doing as late as the 90s). you know how your PCP asks you now how you're feeling, if you've had thoughts of self-harm, had any loss in appetite, or have given up your hobbies? that's the one

if you need to blame it on a social trend, there's a sharp correlation between the rates of suicide in the USA [3] with the transfer of wealth from the bottom 90% of people to the top 1% [4]. as they say, correlation is not causation but it is one of the best indications that it might be. smarter and much better trained people than I can probably do the in-depth research and analysis needed to make that claim

[1] https://www.ncbi.nlm.nih.gov/books/NBK519712/table/ch3.t5/ [2] https://pmc.ncbi.nlm.nih.gov/articles/PMC3291670/ [3] https://www.cdc.gov/suicide/data/index.html [4] http://blog.oxfamamerica.org.s3.amazonaws.com/politicsofpove...

paimapi··on AI psychosis is the new leadership blind spot
leadership as a category functions on low-to-minimal subject matter expertise. you don't want c-suite merging code into production, they should be handling the business interests and direction based on their training and general, high-level, trickled-up understanding of how teams are cohering on strategy

that generalized, trickled-up understanding is also something that's much more easily replaced by an LLM. it was easy for me to spin up a program governance skill paired with a deterministic set of tests in a few weeks. it crawls team channels, project trackers, etc, identifying date drifts, updates, and missing milestones. add an analytical layer of why and how it all fits together in the roadmap and you've functionally obsoleted a lot of the non-technical bureaucrats whose only purpose is to tell someone what other people are doing - from Director to C-Suite, from C-Suite to shareholders

of course you want to keep the people around with the technical depth to construct a roadmap but when was the last time senior leadership in your org actually did any of the substantive work on roadmaps? it's all team leads with the leadership layer serving as the arbitrary rubber stamp on the plans someone else created

when it comes to knowing that there's a dependency on the component you're altering the data contract on, the LLM is useless unless there is pre-existing context or unless you command the LLM to burn your pile of tokens on it tracing every bit of logic in your codebase. unless you want to do nothing but this level of discovery all of the time, you'll need an engineer with lived experience of working in the codebase. here this averaged understanding is useless - you need context, you need granular expertise, all things an LLM fundamentally lacks for your specific codebase and your specific requirements

paimapi··on AI psychosis is the new leadership blind spot
there was an engineering director at my previous company (and who was former Amazon) who wrote his first PR with Claude

it had 30k lines of diff and was a complete refactor of how one of our core features worked

the staff eng he tagged to code review it politely declined to do so and given that this director was let go fairly recently, I'd like to imagine that the PR will exist as a fun surprise for any future maintainer running cleanup

paimapi··on AI psychosis is the new leadership blind spot
"74% of executives said they had more confidence in AI’s advice than in that of colleagues or friends, and 44% said they would defer to its reasoning over their own insights"

it's sad that c-suite believes the advice given by an LLM for a Platonically ideal software stack that looks nothing like the actual implementation moreso than their own ICs. it just reads as more evidence that LLMs are better at replacing leadership than anybody else

never understood why someone would trust Claude's take on your codebase after it's read a 2% snippet of it over the advice of someone who has years worth of experience, who knows all the weird kludges and other quirks that were shipped MVP and never truly tackled as debt, whose done risk assessments and mapped dependencies and has the rough outline of all ADRs involved

I think it's always been true that bad leadership tends to become out-of-touch with the reality of the business over time especially if they don't know how to cultivate good working norms, favoring sycophants instead. with LLMs, that trend seems like it'll just get worse and worse

paimapi··on New Mexico court orders Meta to pay $567m over harms to children’s mental health
this is true and the knock-on effect of this is that other states with intrepid AGs and consumer protection bureaus now have a case law example to cite in their arguments

the hope is that if other AGs pursue these measures that they do their due diligence of creating a case that FB's corporate lawyers won't be able to defeat in subsequent trials. otherwise case law will ultimately let large, unaccountable corporations like Meta win-out in perpetuity

so this is a great battle won but there's a long road ahead for the larger legal strategy. and I do hope it's very long - the current Supreme Court ideological makeup does not favor consumers or citizens at the moment and it'll take decades for the insidious work of the Federalist Society to be mitigated

paimapi··on Nashville uses eminent domain to block data center near zoo
my state power commission just passed a plan to build out 9.5 gigawatts of daily production that will be purely generated by gas turbines [1]. this is at the behest of paid lobbyists for the energy sector and commissioners who, according to open records, have received upwards of six figures from the energy company that is granted a monopoly over our energy here [2]

this is after a strong push for clean energy options like a nuclear plant, solar, and wind by, well everyone, even rural communities who don't want the ambient pollution of those gas plants. but given this 'demand' for data centers, that transition to clean energy is all but abandoned. each of those turbines have an estimated lifetime use of 25-30 years [3] so we'll be emitting multiple times the amount of carbon we do today until at least the 2060s and all of it funded predominantly by taxpayers

if you're not a believer in what people now call 'communist' politics such as believing that anthropogenic climate change is real, public health initiatives save lives, and vaccines are effective then I suppose you'd find this all very good and silly

regardless, claiming anti-data center organizing comes from astro-turfing is exactly the opposite of what's happening. if you went to a single hearing or were engaged with your local community politics, you'd see as much

[1](https://www.canarymedia.com/articles/utilities/georgia-power...) [2](https://energyandpolicy.org/georgia-psc-incumbents-majority-...) [3](https://rmi.org/resources/you-might-be-paying-for-a-worthles...)

paimapi··on Sycophantic AI Decreases Prosocial Intentions and Promotes Dependence (2025)
a therapist should, if they follow their training, not be a sycophant whatsoever. CBT, for eg, is a multi-step process including Socratic self-dialogue, reframing practices, etc to guide the person towards the development of insight and healthier coping mechanisms
paimapi··on Amazonian civilization had estimated 3M people in 3% of forest area
the Native Americans were highly literate, they just didn't record granular details to ledgers because they didn't have strictly hierarchical institutions that utilized this data for controlling their population

there is so much easily searchable resources for this and though it might not be your intent, your comment reads as willfully ignorant for being so broad and bought into the type of 16th century Puritan Eurocentrism that many of the articles above discuss in depth

* https://en.wikipedia.org/wiki/Mesoamerican_writing_systems * https://www.science.org/content/article/roots-mesoamerican-w... * https://en.wikipedia.org/wiki/Ojibwe_writing_systems * https://www.britannica.com/topic/Indigenous-languages-of-Nor...

paimapi··on Amazonian civilization had estimated 3M people in 3% of forest area
there are recorded incidents: https://project1492.org/small-pox-blankets/

the controversy is mostly around whether or not it was a systemic, pervasive form of biological warfare. the evidence base is fairly small but needs to be considered in context of Native American cultural habits around durable recordkeeping (eg trade ledgers) vs colonial, and if colonists would record such an act

there was also active ethnic cleanses and genocides committed by the US government repeatedly throughout US history too and those are very well documented, all dotted with stories as bad or actively and intentionally worse than biological warfare

https://news.uoregon.edu/content/historian-examines-native-a...

"He was speaking in opposition to a treaty that proposed the Cherokees sell off 20 million acres of homeland—a large portion of present-day Kentucky and Tennessee. This tension exploded with the commencement of independence hostilities in July 1776; some Cherokee leaders sided with the British, and in response the US charged thousands of colonial troops with “the utter extirpation of the Cherokee Nation.”

https://sites.uab.edu/humanrights/2017/04/17/indian-removal-...

"Throughout American history, the treatment of indigenous Native Americans has violated numerous articles of the United Nations Universal Declaration of Human Rights. These violations resulted in the loss of numerous Native American homelands, the Cherokee being only one example, and the genocide of numerous other smaller tribes since the beginning of European colonization. This is largely due to Eurocentric ideals, like the natural law of the Puritan worldview, which elevates the status of European peoples over that of indigenous, Native American peoples through a biased worldview. This mindset is so pervasive and powerful that it still prevails today, evidenced by modern films and television that paint Native American tribes as savage, ignorant and of ill intent toward the “white man”, and the policies of the current United States government."

see also: https://en.wikipedia.org/wiki/Native_American_genocide_in_th... or https://en.wikipedia.org/wiki/List_of_Indian_massacres_in_No... or https://en.wikipedia.org/wiki/Denial_of_genocides_of_Indigen...

paimapi··on 200 Milliseconds
is it bad that while this does look like a really cool method to explain a complex topic, my instinct on reading mic-droppy, RLHF AI prose is to be dismissive? there's just something about the persistent mic drops and this-not-that writing that feels so cheap

I think it's because the explainer is passive, there's no interiority, it tells and doesn't show. plus, stylistically, if this were rewritten in second person (like most explainers are) it would make it a heck of a lot more readable:

'You order a coffee on your coffee shop's tablet. 211.4 ms later you see 'Order Confirmed'.

There's a world of complexity behind that status confirmation. Let's see how it all works.' etc

← PreviousPage 10 of 10