HNHacker News
TopNewBestAskShowJobs

blazespin

3,779 karma · joined May 16, 2013

hi
submissionscomments
blazespin··on Kimi K3: Open Frontier Intelligence
Are there terms clear? I dunno. There are ways to train on API usage without training on API usage.
blazespin··on Solving 20 Erdős Problems with 20 Codex Accounts Running in Parallel
yeah, is it API or codex?
blazespin··on GLM 5.2 beats Claude in our benchmarks
I think the point is less "how can we throw shade on the OP" and more "a harness can enable a lot of models to do very serious cybersec, glm 5.2 is one of them"
blazespin··on Anonymous GitHub account mass-dropping undisclosed 0-days
Yeah, I am seeing this as well, especially as people use AI to code review stuff more so this sort of thing slips through. On one very large project I am looking at, it's already becoming harder to find issues.

These people whinging about slop don't realize everything that doesn't come from a credible source gets ignored.

Credible people are using AI and once these issues are fixed, it will die down.

The threat of AI zero days will persist though, but they will be much more expensive and subtle to find.

blazespin··on Anonymous GitHub account mass-dropping undisclosed 0-days
Just optimized AI driven fuzzing. Do a search on arxiv you'll find a lot.
blazespin··on Anonymous GitHub account mass-dropping undisclosed 0-days
I have a dozen or so critical CVEs now, it's not hard to believe at all if they're just hardening tasks. I can get a dozen hardening tasks from just one prompt. I don't even bother filing them as the critical ones are more important right now.
blazespin··on Anonymous GitHub account mass-dropping undisclosed 0-days
This is ludicrous logic. We already know that there is an AI firehose. You don't need to do this. They should have used proper disclosure.

All this is doing is making the AI firehose worse.

blazespin··on OpenAI to Stagger Release of GPT 5.6 at Request of U.S. Government
Pretty soon I suspect, otherwise Chinese models are going to have free reign to develop brand goodwill. Question is how this impacts the international posture.
blazespin··on OpenAI to Stagger Release of GPT 5.6 at Request of U.S. Government
Ban on Chinese models is coming on pretty soon. Seems unlikely they're going to shut down openAI and anthropic, but not foreign models.
blazespin··on Anthropic says Alibaba illicitly extracted Claude AI model capabilities
The hilarious thing here is anthropic is basically admitting that most of their Capabilities can be easily copied.
blazespin··on Statement on US government directive to suspend access to Fable 5 and Mythos 5
Executive staff seems money-hungry for sure (note the lack of non profit that OpenAI has)

I would say they have researchers with self-important god complexes that makes them think they know better than everyone else.

blazespin··on Statement on US government directive to suspend access to Fable 5 and Mythos 5
There are a lot of dangerous things in the world and surprisingly a lot of people can avoid the constant stream of chicken little nonsense.

If everyone expended the same amount of marketing effort trying to scare the ** out of everyone that Anthropic does, it'd be a very miserable world to live in.

We are unfortunately a captured audience and the autistic people at Anthropic are abusing this.

blazespin··on Statement on US government directive to suspend access to Fable 5 and Mythos 5
There is a potentially another explanation: fable 5 did something truly dangerous.
blazespin··on I'm Eric Ries, author of "The Lean Startup" and new book "Incorruptible" – AMA
Very very patronizing of him to get upvoted.
blazespin··on I’ve joined Anthropic
yes, pure engagement farmer. Marketing yourself is a big part of the biz these days. Thanks, social media
blazespin··on I believe there are entire companies right now under AI psychosis
not hard, massive elo stuff. every decision point needs to think up and implement 25 ideas and then rank them.
blazespin··on I believe there are entire companies right now under AI psychosis
The problem is that the only thing that has proved out so far is cyber security. Unfortunately cyber security improvements is not going to improve living standards, and it's just going to increase the cost of just doing business. There is no productivity boost, in fact it's the opposite.

What we need is automated research that leads to real results. This is possible, but it has yet to prove out. I am concerned that unless the AI companies focus entirely on this, it may be a while before we actually see true benefits from this.

What's worse, is there is an urgent and desperate need for automated research, as we have been seeing diminishing returns in human produced research for some time now: https://web.stanford.edu/~chadj/IdeaPF.pdf

blazespin··on New arXiv policy: 1-year ban for hallucinated references
it's very silly, but not a big deal. Arxiv is becoming irrelevant these days anyways.

In fact would be better if they just banned AI, so we could just get off the luddite platforms.

Automated research is the future, end of story. And really it couldn't have come out at a better time, given the increasingly diminishing returns on human powered research.

blazespin··on U.S. Senators Vote to Ban Themselves from Trading on Prediction Markets
Pure gaslighting. The PM panic is because people want to look like they 'care' about corrupt trading. It's peanuts compared to the 500M they were betting on oil futures.
blazespin··on U.S. Senators Vote to Ban Themselves from Trading on Prediction Markets
No, because that's real money.
blazespin··on Claude Opus 4.7
Safety versus Distillation, guess we see what's more important.
blazespin··on Muse Spark: Scaling towards personal superintelligence
Because bots and trillion dollar ipos and even bigger stakes. People need to better appreciate the level of manipulation going on. Social media has an outsized impact. Bots and even people are getting paid to post and upvote/downvote narratives.
blazespin··on GLM-5.1: Towards Long-Horizon Tasks
Anthropic's reply? A model you can't use.
blazespin··on Project Glasswing: Securing critical software for the AI era
Dario is big on beating china, and no doubt he believes cyber security is how to do that. You can tell, but anthropic is sht at everything else. Nobody uses it for real research.
blazespin··on System Card: Claude Mythos Preview [pdf]
Anthropic needs money like the 112B OpenAI got. They could be hyping and this is good hype. Who knows how benchmaxxed they are.

If they provide access to 3rd party benchmarking (not just one) than maybe I'll believe it. Until then...

blazespin··on System Card: Claude Mythos Preview [pdf]
Yeah, need some good RE benchmarks for the LLMs. :)

RE is very interesting problem. A lot more that SWE can be RE'd. I've found the LLMs are reluctant to assist, though you can workaround.

blazespin··on A new Polymarket account made over $500k betting on the U.S. strike against Iran
All I've ever seen is it encouraging people and not discouraging them. The thrill of the easy money, right?

Forget it jake, it's Polymarket.

blazespin··on A new Polymarket account made over $500k betting on the U.S. strike against Iran
Yeah, trying to beat the market on actually predicting will get you pretty lame returns. Probably do better in non zero sum games like the stock market. At least there you get the benefit of the market always going up eventually.

No, the best way to win on Polymarket is purely by insider trading. Which is why it's a useful thing to watch. Insider news..

That said, the definition of 'insider trading' is always tricky. At what point does it become insider? Some things people call insider others just call clever detective work.

blazespin··on A new Polymarket account made over $500k betting on the U.S. strike against Iran
"People don't play in corrupt markets for very long." ... ahhh, it's not a bug with Polymarket - it's a feature.
blazespin··on "Cancel ChatGPT" movement goes mainstream after OpenAI closes deal with U.S. Dow
I'd say it's more a kowtow to his voters and democracy. Whatever you think of him, he was democratically elected. Do remember to show up this November, though, and remind all your friends..
Page 1 of 34Next →