2,072 karma · joined October 2, 2018
"A Google employee familiar with model development said there is "large consensus" internally at the company that Gemini 4 is at the frontier."
So this is as informative and specific as an article with completely imaginary sources would be.
What did they promise their investors who invested billions?
Couldn't I simply give a Chinese friend my key on Open router?
It's also a rhetorical move I've seen several times here...
If you really want to compare apples to apples you need to test Gemini models against other models using the same third party search harness.
Otherwise you are largely measuring how much computation the model provider is allocating to a search harness.
Anecdotally I'd rate Gemini behind Claude and OpenAI models at fiction and I can't find any benchmarks showing Gemini is the clear winner at this task.
If I like the novel, an alternative version where things happen differently would be fine.
The problem isn't that it's different. The problem is quality.
A lot of Isekai stories are crap but you can still rank them in terms of the author's ability or inability to have a creative POV that elevates the material.
And no I don't think asking the A.I. to write the next 2000 words would require it to have the entire novel planned out. Not all writers even outline in advance.
If so I'm hoping we can track them down and have them tell us if they think this is AGI.
This is like saying AI developers should be qualified to evaluate the effacy of a model without testing the model.
A major correction would be a bummer but we were never entitled to these abnormal gains in the first place.