I agree it would be very cool if they were open sourced, but the article is talking about the KV cache technique which if I understand correctly is open source (or at least the white paper is published).
I would say that there are very few people on earth that I trust less than Sam Altman, Dario Amodei and Elon Musk. Also my own government claims to have used Anthropic models to bomb a girls' school in Iran. If you combine US regulations with sociopathic private companies, you get into a worst case scenario for humanity imho. Again, I'm thankful to China or any other entity pushing open models, local models and even distribution of this technology.
I'm also grateful to the Chinese labs for providing workarounds for the walled gardens that the US based AI companies are attempting to create.
Does anyone know if there are any distillation datasets available? I'd love to see these distributed on BitTorrent. I think it's critical that AI be democratized and not isolated in the hands of a few private companies.
I wouldn't call it "half-baked" when it's a well known playbook regularly leveraged by large companies.
> At best you'll get legal restrictions for particular US industries which are especially risk-aware. The CEOs of those industries will counter-lobby to be able to use whatever model they want.
Companies won't spend the money and time to lobby to use different models and Anthropic knows it.
> Since there's plenty of public attention on this issue, achieving meaningful regulatory capture will be difficult.
I don't have high hopes. This is also why Anthropic is pushing the "regulate or ai will kill you" angle.
I think this is very clearly a regulatory capture play. Anthropic/OpenAI/X see that there's very little moat around training (especially with distillation) so they want the government to build the moat for them.
It's also worth remembering that there is zero percent chance that entities like the US military are going to be pacing anything. What Amodei and his ilk are aiming for is a highly regulated industry where they control the political barriers and the ability to sell SOTA model access to state actors that have a monopoly on violence. It's the worst possible situation for consumers and citizens. Thankfully I don't think they can put the cat back in the bag and Chinese and other models will keep progressing as a counterbalance to the techno-fascism Anthropic is aiming for.
You can get a binary compiled by the author or a trusted source and none of the dependencies can change out from under you. This isn't possible with an interpreted language where the dependencies are resolved (often from dubious places like npm) at install and update time.
I think these are both examples that help prove my point. Neither of these tools benefit the user by being in Python and distributing a runtime just for a CLI tool.
It's not just the complexity. You're also vulnerable to supply chain attacks via NPM. It's also performance as you don't need the entire javascript runtime just for a CLI.
Whether it was deliberate or not, I do not think it's a good choice. When agents can write in any language, there's no reason to pick the wrong tool for the job. At this point Javascript/Typescript belongs only in the browser. It's the suboptimal choice for every other environment. Especially for a command line tool. Even if the back-end is written in Typescript (also not the best choice imho), the clients need not be in the same language.
I don't understand why this is written in Typescript. This is a great example of how agents can write code (I'm sure they wrote `cf`), yet having fundamental computer science knowledge is still critical. Do not force your users to manage the dependencies of your cli. Do write your cli in a compiled language. Understand the reason for those decisions and tell your agents to use the correct architecture.
You shouldn't have to do this and I won't. I'll vote with my dollars and never buy another device from them. It's too bad because I genuinely liked it as well, but it's not worth the invasion of privacy and normalizing sending this type of information to any consumer company.
I had an Oculus Quest and it was a really great device. After the "Meta" rebrand, the quality rapidly decreased and now it wants me to upload my ID to Meta to keep using it. Absolutely not! Zuckerberg already has way too much information on me, uploading my state issued identification to them is such an insane ask, I will never do it. While this looks like interesting hardware, the company behind it has unacceptable and user-hostile practices, so I'll pass.
There are so many crimes the US has committed that it still has not accounted for. On protests, I think there were a lot during the Biden era but the government and Zionist entities cracked down hard on them, especially the university protests. That obviously happened to an even more extreme degree during Vietnam (like the Kent State shootings), but the protests were even more intense then due to the draft. I think we'd see something similar if Americans started getting drafted to fight in Iran or anywhere else.
I totally agree, but the US has done this for a long, long time. Elsewhere in this thread we were talking about the My Lai massacre, where the people involved also escaped punishment.
Welp, it's now blocking me from doing extraordinarily mundane tasks because of "safety". I've been an Opus fan for a long time, but this instantly made me cancel my subscription and move to OpenAI (which I also assume will screw me soon enough). Chinese models are almost there for my needs, and I can't wait to switch to them and never look back.
The SOTA models now work really well in my codebases, but that's only been since Opus 4.5/4.6-ish. Prior to that, and with current local models, they simply couldn't work holistically and would just thrash around. Now I feel as if SOTA are approaching my coding levels if not surpassing it. I still need to guide on architecture, but I can see that going away within the next year or so as well.
I so want this to be true, but for the kind of coding I do (not Flask apps), it's definitely not the case. Like I said, SOTA models just barely, barely work for me. My projects are usually 100k-1M lines of Rust or Go.
Just to be clear, I'm specifically talking about coding. I think local models can help with productivity today, just not coding.
I'm also a huge fan of local models and think it's absolutely imperative that they continue to advance so we can move off of the Anthropic/OpenAI hosted models. It's important to accurately asses where we are in that journey though.
I believe they can currently be used productively for non-coding tasks (classification, light summary)... but they definitely are not even close to SOTA when it comes to software development.