438 karma · joined February 14, 2023
Why couldn't excluding Anthropic be done a different mechanism than a supply chain risk. Why wouldn't this be a standard part of an agreement or a request and simply refuse to renew or cancel a contract rather than being designated a supply chain risk. If some third party contractor for an unrelated non military reason wanted to use Claude as part of their process, it seems perfectly allowable.
If this was a remotely standard way of operation, why did a fair amount of corporate America sign briefs concerned with the retaliatory aspect.
Parts of the action seemed wholly retaliatory as well since Pentagon officials certainly used it as a threat. Why can't a US company have views contrary to the policies of the US government? Certainly the US government is free to not do business with them, but this designation affects everybody doing any kind of indirect business with the US government which is a rather long chain. If this becomes a legitimate mechanism, how might we distinguish caring about a secure supply chain and simply wiping out a company that disagreed with a pro war attitude? If a machining shop had a policy against manufacturing weapons at all, could they be blocked from making server racks for Microsoft or perhaps light fixtures for the Department of Labor? There are plenty of areas that are, again, non military? I don't think it's credible to claim that Anthropic will deliberately sabotage operations.
The hope with many of these problems in math is that in trying to prove that, we get some additional insight into why it blew up that could be applied elsewhere to more general PDEs that cannot be easily controlled.
I think the observation from Tao and many others is that when humans solved these problems, the additional insights into intuition and theory building came for free since humans can give expository on what they found hard or what was their own intuition. This is much more difficult or tedious to extract from an AI model. Even when people did have access to the chain of thought, it wasn’t always very helpful to figure out what was the exact thing that made it all click. This is even more difficult how that the CoT are hidden but I would think the sort of difficulty of extracting the key ideas for a human might be worse now with more advanced models.
There’s a long term aspect to this too where we have historically used these problems as markers for the other parts of mathematics but if AI can solve it all, then suddenly this signal is not very meaningful.
Maybe to bring it closer to home. If an oracle just gave you P \neq NP, then this would be generally uninteresting since this was already expected. There’s a deeper question of why that needs to be answered. However, one would hope that creating such a separation would give us tools that allow us to create lower bounds on a lot more problems we do care about and perhaps some bigger insight onto what makes a problem intrinsically hard or easy. These long term considerations are helpful but are definitely more vague. The remarkable part is that AI is separating the part about proving theorems and the “free” insight you get.
There's unlikely to be any engineering applications since even if the solution can be approximated, you still need to set up the initial conditions but at that point you can also drive pressure in other ways.
Regarding local models, Gemma, GPT-OSS, Nemotron, and Inkling (maybe upcoming Muse series as well) all are fairly decent options at a variety of hardware costs.
Regarding your concerns, I don’t see why it is unreasonable to simply renounce citizenship later or just simply double check your travel itinerary during the late stages of a pregnancy. This seems like at best an unfounded concern.
That being said, age based restrictions isn't a fine grained control over the system as perhaps one would like but that also would be inherently more complicated to think about from a legislative perspective (e.g. how fine grained and how to categorize possible dangers) and user control perspective when it looks like a lot of parents are looking for a blunt generic button that basically goes "this is agreeable with general practices". This seems more or less how real systems are gated.
The other issue is that both present privacy challenges but this just a little more so from a fingerprinting perspective. Presumably you need quite a few bits to completely specify the filter whereas age is only a ~1.58 bit field in the CA model. Not really sure how much this matters when there are so many other signals for fingerprinting and we should probably make fingerprinting from it illegal but just some thought.
> Instead, what we get proposed is a system that cares very much about how old you are, and not one bit about the things that one's guardian understands one needs to be protected from.
Regarding your linked comment, I think it's a bit strange to say that if legislators really did care about child safety they would mandate fine grained controls instead. I'm not sure what additional fine grained factors you may be thinking of precisely, but we already use age as a gate in real life for many things we consider dangerous so it's quite natural for legislators to transpose those. Our laws already very much care about how old you are.
However, it's also unclear to me if this is directly coming from a directed political ideology from the firm itself or a more general "let's do what the government wants so as we can publish this stuff". Those imply two different ways about thinking of the model and whether we can sort of containerize the issue. I think if a firm like Huawei were to publish a model, these concerns would be significantly more vocal. For better or worse, many of these political questions are also distant to many users on this site.
On the other hand, many people on this website live in regions that are directly affected by Musk's constant political activism. It's hard not to be when he was such an active part of an administration that controls a global superpower and continues to push his view via X. The DeepSeek owners, by contrast, are not to my knowledge constantly calling for Taiwan to be invaded.
I do think if Musk was less politically active and less personally involved with his companies, there would be less discussion of Musk's politics. People, for better or worse, are willing to put aside political discussion, in the "everything is political" sense, that may be more loosely linked.
It is simply in the case of Musk that this tension boils over and legitimately becomes impossible. There is perhaps some kind of Singer-style argument about how this is some form of hypocrisy but as a practical matter, I don't think it's reasonable to ask people to turn down their political discussion around someone like Musk.
https://futurism.com/grok-looks-up-what-elon-musk-thinks
To your narrow point, it's very obvious that Musk influences the bot to share his views. For example,
https://www.nbcnews.com/tech/tech-news/elon-musks-ai-chatbot...
If your claim is that somehow I should not be concerned about Elon's politics with regards to the model itself, then this seems wrong.
Anyway, to the broader point of whether or not the we can avoid discussion about the Musk's politics and talk about the politics of the model as if it were independent of him, this also seems difficult. It is impossible to ignore because the man has made himself the face of every one of his companies and is an obviously political figure unlike any other company and has politics that are definitely characterized as more radical. This makes the political component basically impossible to ignore unlike any other company.
The next time the current American administration issues an executive order on AI, should the conversation always be limited to the technical merits of the executive order?
I don't think you need to somehow get personally offended by every Tesla on the road but it seems ridiculous to ask people to not be political about a such obviously political figure.
https://arxiv.org/abs/2604.21691
There's of course empirical results and relatively weak theoretical results like the UAT but I also don't think that answers your question fully, especially since it seems impossible to definitively answer questions that the industry seems to betting on like whether or not there is a lower bound to their error rate or whether hallucination as a problem can be solved. We have much stronger ideas of what linear regression is doing relative to what LLMs are doing.
1. Does that mean the same thing in the ToS?
2. How valid are these requests?
On a more practical level, forcing them to go to court might not be much better. If this went to a FISA court, those are essentially rubber stamps and give nearly 100% approval.
https://www.washingtonpost.com/technology/2024/09/25/elon-mu...
https://www.washingtonpost.com/technology/2024/09/25/elon-mu...
I think there's quite a bit of variance in model performance depending on the scaffold so comparisons are always a bit murky.
Right now, it's not even clear how to create parental controls at a reasonable level so there's no clear path for what to do or how to respond.
I think this is a reasonable balance without being invasive as there's now a defined path to do reasonable parenting without being a sysadmin and operators cannot claim ignorance because the user input a random birthday. The information leaked is also fairly minimal so even assuming ads are using that as signal, it doesn't add too many bits to tracking compared to everything else. I think the California bill needs a bit of work to clarify what exactly this applies to (e.g. exclude servers) but I also think this is a reasonable framework to satisfy this debate.
I've seen the argument that this could lead to actual age verification but I think that's a line that's clearly definable and could be fought separately.
> A Meta employee (Jake Levine, Product Manager) contributed $1,175 to ASAA sponsor Matt Ball's campaign apparatus on June 2, 2025. Source: Colorado TRACER bulk data.
> No direct Meta PAC contributions to any ASAA sponsor across Utah, Louisiana, Texas, or Colorado. Source: FollowTheMoney.org multi-state search.
While it is true that Meta has funded groups that advocate for age verification, a lot of them also appear to have other actors so it's not like this is some pure Meta thing as some of the other commenters are suggesting.
This presents the problem of governments being able to gatekeep speech which I am quite uncomfortable with but maybe there's some safeguard within the eIDAS proposal that makes this idea incorrect?