270GB of source code from The New York Times leaked to 4Chan
twitter.com
twitter.com
https://boards.4chan.org/g/thread/100843783
Magnet link:
https://files.catbox.moe/jsowk8.txt
(Archived: https://archive.is/N3M8F)
--
It looks like most peers are at 86%, so should have confirmation pretty soon.
https://x.com/stackdiary/status/1799180195892535441
From the linked article:
> those attempting to download the data via torrent are reportedly stuck at 85% and unable to complete the download... A few hours later, the person responsible for the leak provided another torrent link, which separated the files as opposed to making them a single archive. This new torrent has three folders called nytimes, nytm, and TheAthletic. The two files inside the TheAthelic folder are iOS.tar and android.tar; we verified that they reflect the source code of The Athletics mobile apps.
My guess is that there will be also be some interesting tooling in there that they used for data-mining massive numbers of scans/pdf's snagged from disinterested/hostile bureaucracies.
> ripgrep is faster
I've only worked at monorepo companies, and when I see the "monorepo vs. multiple repo" debates, I always picture in my mind that we're arguing about 1 vs. maybe 5 or 6 repos--like a repo for each major project. But thousands of repos, one for every little nugget??? That is totally wild. Is this an actual industry practice?
Source code for Wordle.
The dedication to homely, friendly, inoffensive, kind, folksy words needs to stop.
It's 4chan. Where conspiracy theories fall flat, they'll just spin more conspiracies about a cover-up. The flow-chart process of conspiracies is designed to never disappoint the believers.
So... Explicitly fabricated conspiracies shall not be conflated with "conspiracy theories"? Whatever attachment you have to either would make you a conspiracy theorist.
Conspiracy theories are when you make that kind of thing up.
A conspiracy is just one or more people working together. Putting a weird tone on that word is...weird.
Whatever attachment you have to using words wrong - would make you a fool.
They would have better open sourced their code in the first place.
Though I would not want to use that git repo!
Fuck
NYT regularly creates interactive stories and those creative developers have no idea how to deal with assets. I deal with that category often and our little projects reach half a gig unless I catch it in time. No PRs, just commit all, merge from remote and push.
It’s a lot.
[1]:https://x.com/ballmatthew/status/1774413955940639199/photo/1
The person who tweeted/xed/whatever doesn't know what he's talking about.
270 gigs of data? Entirely possible.
270 gigs of source code? YeahNo.