HNHacker News
TopNewBestAskShowJobs

joshuakelleyds

19 karma · joined March 11, 2024

submissionscomments
joshuakelleyds··on AI companies destroy physical books – let's scan rare books before it's too late
I keep seeing headlines, videos, etc and the recent copyright court case, Anthropic v. Bartz (1.5 billion dollars) gives the best context around this. I encourage everyone to read the full thing, but here are some excerpts:

> Anthropic spent many millions of dollars to purchase millions of print books, often in used condition. Then, its service providers stripped the books from their bindings, cut their pages to size, and scanned the books into digital form — discarding the paper originals. Each print book resulted in a PDF copy containing images of the scanned pages with machine-readable text (including front and back cover scans for softcover books). Anthropic created its own catalog of bibliographic metadata for the books it was acquiring. It acquired copies of millions of books, including of all works at issue for all Authors. Anthropic may have copied portions of Authors’ books on other occasions, too — such as while copying book reviews, academic papers, internet blogposts, or the like for its central library. And, Anthropic’s scanning service providers may have copied Authors’ print books along the way to delivering the final digital copies to Anthropic. But neither side here specifically raises legal issues implicated by any such copies. Nor will this order

Also the summary:

> To summarize the analysis that now follows, the use of the books at issue to train Claude and its precursors was exceedingly transformative and was a fair use under Section 107 of the Copyright Act. And, the digitization of the books purchased in print form by Anthropic was also a fair use but not for the same reason as applies to the training copies. Instead, it was a fair use because all Anthropic did was replace the print copies it had purchased for its central library with more convenient space-saving and searchable digital copies for its central library — without adding new copies, creating new works, or redistributing existing copies.However, Anthropic had no entitlement to use pirated copies for its central library. Creating a permanent, general-purpose library was not itself a fair use excusing Anthropic’s piracy.

https://copyrightalliance.org/wp-content/uploads/2025/06/Bar...

joshuakelleyds··on The Mojo language (by Modular, now Qualcomm) is now open-source
I would add quick note to this as title is misleading

* It was partially open-sourced before this. There were a lot of cool things they open sourced before like MAX for large scale LLM serving which was outperforming VLLM, Dynamo, etc on a lot of models. (super valuable GPU kernels). This is why Qualcomm acquire them imo.

* Chris (also created swift) talked in the past the reason for not fully open-sourcing was more because he wanted to get all the core design decisions right. He said this was a big thing Swift got wrong as it scaled too quickly being fully open source at the beginning.

Mojo is an awesome language, I've used it a lot as a Swift/Python lover. A couple things though

* If you want to understand Mojo spend 10x the time in MLIR before. It's just a fancy MLIR wrapper (good thing)

* They still haven't lived up to the python "superset" promise and that's the big thing preventing bigger adoption.

* https://www.spheron.network/blog/modular-max-mojo-gpu-cloud-...

joshuakelleyds··on FreeInk: Open ecosystem for e-readers
Winterbreak covers most of them >2012, https://kindlemodding.org/kindle-models.html
joshuakelleyds··on Fleet: Hierarchical Task-Based Abstraction for Megakernels on Multi-Die GPUs
Nice read! Cool to see more GPU programming models that expose chiptet topology instead of treating it as a flat execution. Reminds me a little of Cerebras' CSL although far from as extreme as that.