89 karma · joined November 12, 2013
I have found including this in my AGENTS.md to be quite transformative in this regard:
> Always use an aggressive red/green TDD-approach. It is critical to remember that in the red phase, things like module/exports import failures due to trying to import file paths that don't yet exist, exports that don't yet exist, etc. is not valid TDD. For valid TDD, the test cases must actually run. For this, you must create stubs of the expected modules and exports in the red phase, so that the test cases actually run and fail on the test case assertions themselves. In some cases, when using this approach, once in a while some of the red phase test cases might "incidentally" pass, and this is ok. Before running red phase tests you should always make predictions about the number of test cases you expect to fail/pass -- this count is not the number of test files or test suites, but rather the number of test cases. By performing these red phase expected counts of passing/failing test cases, it will help you catch errors in your prior reasoning quickly and efficiently.
> Always use a proof-driven, scientific method-based approach to validate hypotheses, assumptions, and conclusions: define the smallest falsifiable hypothesis, create or identify a reproducible failing case, gather direct evidence, make the smallest targeted change, and then re-run the same proof to confirm the issue is fixed. Avoid speculative fixes, broad rewrites, or changing multiple variables at once. When possible, preserve the reproduction as a regression test before implementing the fix. Consider that when gathering evidence, additional logging and durable files can be very helpful.
> The strict TDD and proof-driven approaches described above could be described as "proof-driven development". Try to internalize and generalize these concepts, as they are broadly applicable.
That is not sustainable for me. I understand that Anthropic can't subsidize all of us forever, but this will force me away from claude which is unfortunate. A number of the people I work with are similarly affected.
99% of my max plan usage is non-interactive, and this post-June 15 pricing will far, far exceed what I can afford. I assume this applies to a great deal of us.
After like day 2 my workflows would take 10-15 minutes past their trigger time to show up and be queued. And switching to the self hosted runners didn't change that. Happens every time with every workflow, whether the workflow takes 10 seconds or 10 minutes.
> The images were initially believed to have been obtained via a breach of Apple's cloud services suite iCloud, or a security issue in the iCloud API which allowed them to make unlimited attempts at guessing victims' passwords. Apple claimed in a press release that access was gained via spear phishing attacks.
I also found it notable that the source for the above unlimited password guessing password guessing is an Apple press release that states no such thing.
Also interesting was that all sources in that article suggesting anything about unlimited attempts describe to an app or script (unclear which) called iDar, which the only source to actual name iDar claims that it reports success 100% of the time, regardless of its actual success in guessing the password.
I've no love for Apple. Maybe it's true. But the evidence presented in this wiki article is weak.
[1] https://www.mintpressnews.com/meet-ex-cia-agents-deciding-fa...
[1] https://www.mintpressnews.com/twitter-hiring-alarming-number...
1. Would you agree that reasonable people exist who, right or wrong, believe that the psychological toll of SM on children, while not in the same universe as the physical toll of cigarettes, still manages to cross the line of "this is sufficiently harmful that the government must make parenting decisions"?
2. Are you open to the possibility that, in a hypothetical future with sufficient empirical data, those beliefs might be shown to be accurate?
I don't have children / a horse in this race, I'm more exploring your position about the role of government.
A rather high percentage of pages are far too much for a GPT prompt!
The first thing I did was fall back to a headless browser. Let it sit for 5 seconds to let the page render, then snatch the innerText.
But 5-10% of sites do a good job of showing you the door for being a robot.
I wanted to try and solve those cases by taking a screenshot of the page and using GPT-4 visual inputs, but when I got access I realized that 1) visual inputs aren't available yet and 2) holy crap is GPT-4 expensive.
So instead what I do is give a screenshot service the url, get back a full-page PNG, then I hand that off to GCP Cloud Vision to OCR it. The OCRed text then gets fed into GPT-3.5 like normal.
I can't speak to the veracity of that claim, but as the post points out, the past several years has shown us that it doesn't matter, not in the least.
The author goes on to say how it feels inevitable that we'll see a Bannon or Stone type use this technology to create fake scandals.
I'm more worried about the grass roots efforts. Crowdsourced conspiracies like QAnon. Now they'll have more capable tools to radicalize people.
What I found particularly striking about them was how much they reminded me of both neurons and larger brain structures, as well as some of those newer, ML-assisted FMRI imagery.
Probably just coincidence and wishful thinking, but it instills a sense of daydream-like wonder all the same.
Half the time it's brought up, FTL is offered as the solution. Which as best we can tell is fundamentally impossible.
That squishy or otherwise organic bodies are generally unable to travel interstellar distances has always seemed to me to be the simplest solution.
Assuming intelligent life is out there, surely there are civilizations that have destroyed themselves and so on. But lack of FTL travel would be a common constraint, regardless of all other scenarios.
For prior art, see what happened to Aereo. A little-known fact is that YouTube TV started with the exact same strategy. But Google, of course, had more money and lawyers to get over the hump.
It was explained to me that the reason YouTube didn't offer a karaoke feature -- despite having licenses to a lot of lyrics -- is that karaoke is considered a separate license.
Even if you license both the recording and the lyrics, combining them into a karaoke feature isn't on the table by default.
I personally have no idea how accurate that is, but this was the scoop I got from those in a position of authority.
If it is indeed true, then perhaps that is a factor in Apple's decision not to use that word.
That said, like many others in these comments, I'll be waiting at least 10 years to understand the long term effects, if any.
The last time I was in Sweden, ordinary folks were expressing ideas like "well maybe not everyone deserves free healthcare".
Like many western nations, economic opportunities and quality of social services is dipping. Standard practice for right wing parties to blame immigrants (or refugees, etc.).
I got an early taste of his mental illness. He started maybe 5 songs, each of which he would cut short in the middle to rant about the sound being off. He was completely unhinged and rambling each time. After 20 minutes he simply walked off stage and that was that.
Ever since then, I've found that his illness has been very evident in the art itself. It precludes me from enjoying it, and there's been a number of times I've felt terribly sad seeing these signs celebrated by those who don't see the connection (which is not to suggest that they should).
All that said, I've never even considered this perspective, so thank you for sharing it. It makes a mountain of sense, and makes the whole situation that much sadder (and more complex).
> Look, I'd love to stop CP distribution in America! Really, I would! But Google's encryption policies are preventing law enforcement from intercepting pedophile communications now, today.
It's the same "think of [vulnerable group]" type of statement.
> You’ve probably heard how Pornhub can’t accept credit cards anymore.
That was only temporarily true due to illegal UGC. It lasted less than a month. Once pornhub removed their unverified UGC they were allowed to take credit cards again, or at least Visa.
One would think that given the nature of this post, they ought to have their facts straight.
Credit cards aren't anti-porn. They are anti-unmoderated UGC porn, for obvious reasons.
It's understandable that this does present a problem for sites like Tumblr. But the way it is presented is misleading at best and in part based on a false premise.
Edit: Typos.
We can easily do that. Most people can't. This disdainful attitude reeks of privilege.