2,886 karma · joined March 5, 2012
The idealism that has been sucked out of the tech industry. It was so (naively) hopeful at one point, and now the arms race and profit-maximization has eroded it all. Your observations really resonate with me.
I'm surprised I hadn't heard of the Long-Term Stock Exchange, it seems like a much healthier direction for the market.
I think the open web needs to come back, but in a fair way for everyone, giving readers control over their feeds while also sending traffic and comments back to the original sources. Not quite sure how to do that yet.
Frequent words I see from GPT: "shape", "seam", "lane", "gate" (especially as verb), "clean", "honest", "land", "wire", "handoff", "surface" (noun), "(un)bounded", "semantics" (but this one is fair enough), and sometimes "unlock"
It feels like AI really likes to pick the shortest ways to express ideas even if they aren't the most common, which I suppose would make sense if that's actually what's happening.
> In our internal evaluation, DeepSeek-V4-Pro-Max outperforms Claude Sonnet 4.5 and approaches the level of Opus 4.5.
• Reasoning: Through the expansion of reasoning tokens, DeepSeek-V4-Pro-Max demonstrates superior performance relative to GPT-5.2 and Gemini-3.0-Pro on standard reasoning benchmarks. Nevertheless, its performance falls marginally short of GPT-5.4 and Gemini3.1-Pro, suggesting a developmental trajectory that trails state-of-the-art frontier models by approximately 3 to 6 months. Furthermore, DeepSeek-V4-Flash-Max achieves comparable performance to GPT-5.2 and Gemini-3.0-Pro, establishing itself as a highly cost-effective architecture for complex reasoning tasks.
• Agent: On public benchmarks, DeepSeek-V4-Pro-Max is on par with leading open-source models, such as Kimi-K2.6 and GLM-5.1, but slightly worse than frontier closed models. In our internal evaluation, DeepSeek-V4-Pro-Max outperforms Claude Sonnet 4.5 and approaches the level of Opus 4.5.
While they're some months behind closed SOTA (though benchmarks put them close), I wonder if Deepseek 4's longer context capabilities and kv-cache advantage will make up for this
If that's true, then the value comparison is not so positive for Codex any more
[1] https://old.reddit.com/r/codex/comments/1sgxy71/so_did_they_...
Yet my banking app (here in Singapore) doesn't let me block any prior authorizations. It feels like the payment networks don't want to make it too easy to cancel periodic payments? Which isn't surprising, of course, but it feels like something I'd change banks for.
"write a summary handoff md in ./planning for a fresh convo"
and it's generally good enough), but maybe a skill like you've done would save some typing, hmm
My ./planning directory is getting pretty big, though!