368 karma · joined December 2, 2016
Google is desperate. They haven't been performing in half a year. It's clear their researchers have been forced to integrate existing benchmarks into their training.
These numbers are meaningless. Shame on them.
But yeah, weird timing after a meh release and their competitor killing it right now.
Add visual understanding, add reasoning and bring down the size to run on my computer. That's when it will be interesting.
So many people that don't understand the tech jumped on the hype train because "it cannot hallucinate" and else. It's crazy.
Anyway, it's fine for us to just disagree here. You just sound like someone that likes to argue for the sake of arguing, and just like to do apps, I don't have time for that.
- Cross platform frameworks are starting to become an anti pattern these days
- Hello world examples are losing relevance when I'm not the one writing code
- A To Do list is just the absolute least impressive thing you can showcase
From a marketing perspective, I want to showcase something that other frameworks cannot do and "this only took 10 lines of Python code" is a very weak sales pitch in 2026.
Truly, if this is the biggest criticism left, they should be celebrated. While in reality, all of this has a bitter aftertaste.
So weird.
Isn't that exactly what almost everyone would do given that they wanted to see how capable their model is and the tense competition they have with Anthropic right now? Stealing impressive headlines from your competitor is pure gold.
- They threw compute on a problem another team/company was rumored to have solved to see what their secret model could do.
- The texts I read do make it seem like OpenAI wanted to talk and share credit generously.
- Imagine working on a frontier math problem with someone at Anthropic and not only do you use Codex but also through a non-business account that allows training on your data.
- Timeline-wise, if they mainly used GPT 5.6 it's unlikely any meaningful data made it into an model that's being internally validated right now.
Move 37 comes to mind.
It looks more like Google execs losing their mind and pressuring researchers to put DeepSWE directly into the training set.
Anyway, it's over for Perplexity. They never had a great a product and the only reason for using them, was when they offered Pro accounts for free. Many people joined. Me included. But with a "meh" product and the general AI business not being very sticky, they lost quite harshly.
I thought they might be able to make money as a search api/index, but this article closed the book.