HNHacker News
TopNewBestAskShowJobs

j_maffe

1,793 karma · joined January 19, 2023

j_maffe.at.hn
submissionscomments
j_maffe··on OpenAI breaches Medicare, Albanese reveals
I highly doubt given OAI's latest streak that the models didn't know what they were doing.
j_maffe··on GPT-6 Astra Solves a WWI German Radio Cipher
I'm not sure I agree. More is different. Being able to execute logic at a higher scale and speed would make some previously infeasible intelligence tasks possible, resulting in a new level of intelligence.
j_maffe··on We must pace the frontier
That would be the best case scenario. I honestly wish things would finally slow down a bit. I don't see it happening.
j_maffe··on We must pace the frontier
Would you argue the same for allowing everyone to have guns?
j_maffe··on We must pace the frontier
Yeah the Chinese fear-mongering falls a bit flat when coming from a point of maintaining US supremacy
j_maffe··on On the Navier–Stokes Millennium Prize Problem
> in a fraction of the time

Well if you do the math, the number of agent-compute time in total, given the insane number of agents thrown at the problem, might end up being comparable in time, if not for the budget.

j_maffe··on Claude Fable 5.1 and Claude Mythos 5.1
I think if you use an LLM just to proofread then it'll not be able to insert a strong enough watermark.
j_maffe··on Terminal-Bench-Science: Evaluating AI agents on scientific research workflows
The tasks are the thing to really look at here:

https://github.com/harbor-framework/terminal-bench-science/t...

j_maffe··on Terminal-Bench-Science: Evaluating AI agents on scientific research workflows
> I worry this doesn’t check correctness

Then it's not a valid benchmark. I agree though they're not reliable enough to just put results in a paper.

j_maffe··on U.S. State Department pauses immigrant visa applications
The Purpose of a System is What it Does.
j_maffe··on GLM-5.3-Flash
No but a provider with more amicable terms can.
j_maffe··on Z.ai confirms Ox Alpha is a new GLM-series model and will release its weights
Anyone has a link to a report of its capabilities? I can't find a reliable source.
j_maffe··on Anthropic's best AI model struggles to attract users as cheaper tools thrive
Not if there's a ZDR policy.
j_maffe··on Show HN: Huzzah – a novel approach to coding with AI
I think the idea of having a human-written persistent document describing the operation of the code is a great idea. This document acts as the prompting interface instead of the chat window and changes can still be tracked. Surely something as simple as a skill.md can be made for such a setup, right? I think the pseudocode style is a seperate axis to this setup.
j_maffe··on Australia passes law to levy tech giants that fail to pay for local news
They would rather ban to deterr other governments from taking such measures.
j_maffe··on Israel creates fake think tank in likely attempt to dupe AI chatbots
> “It’s no secret, I disagree with the prime minister," he said. “I think targeted assassinations should be carried out in Gaza, taking down 30 to 40 every night,” said the far-right Israeli lawmaker.

> “Not just those who pose an immediate threat, there are people there who are not worthy of life. They shouldn’t live. They’re not even people,” added Ben-Gvir.

Not sure how the full quote is supposed to help.

j_maffe··on Claude: System Prompts
Almost all of my karma is from a single post link that blew up randomly. I think this is the case for most users' karma here.
j_maffe··on Saying No
But that seems to have little to do with the quote... I'm confused.
j_maffe··on New Mexico court orders Meta to pay $567m over harms to children’s mental health
You think they'll shutdown FB?
j_maffe··on Muse Code and Muse Spark 1.2
What does cheating benchmarks have to do with breaching financial user agreements? To use your argument: all it takes is one whistleblower to get the company sued for billions of dollars.
j_maffe··on AI-Generated Images Discourage Me from Reading Your Blog
> If not, congrats, you're on the same intellectual level as a racist.

I remember HN comments used to be of higher quality than this.

j_maffe··on AI-Generated Images Discourage Me from Reading Your Blog
> So I'd rather spare myself from explaining it.

Being snarky doesn't really help your position.

j_maffe··on AI-Generated Images Discourage Me from Reading Your Blog
It's good as a community to communicate what we prefer to read. At least that way the author knows why they're getting the cold shoulder.
j_maffe··on Google kills Earth AI generator after one day
Such a brain-dead idea...
j_maffe··on The Maxwell Conjecture Is False (GPT 5.6 Sol)
Visualization of the configuration: https://claude.ai/public/artifacts/9db65255-16ff-4f8e-8be1-1...
j_maffe··on Claude Fable produced a counterexample to the Jacobian Conjecture
Cheating how? The proof is in the pudding...
j_maffe··on What AI did to stackoverflow in a graph
> No one is claiming that time is a random variable drawn from a normal distribution.

You are doing that implicitly by fitting a Gaussian curve.

j_maffe··on What AI did to stackoverflow in a graph
> statistically speaking

That's a very big word you're using there for what is basically making shapes out of clouds. A bell-curve is the amortised function of a random variable with a mean and standar deviation. What does that have to do with a timeseries dataset?

j_maffe··on What AI did to stackoverflow in a graph
I'd call it significant that the number of questions halved within one year following the release of ChatGPT, the biggest relative or absolute rate of decrease in the timeseries.
j_maffe··on What AI did to stackoverflow in a graph
You don't just fit a Gaussian distribution to a timeseries dataset. That's not what a Gaussian curve is designed for at all. https://www.explainxkcd.com/wiki/index.php/1725:_Linear_Regr...
Page 1 of 19Next →