5,208 karma · joined June 23, 2017
"Expanded projects, tasks, and custom GPTs"
Do you have custom GPTs in Codex?
This is just the generic ChatGPT plan comparison, but under the Codex url.
I have no idea what this means. Also, the spam of top-level AUDIT-*.md files doesn't inspire confidence :)
You're right for Anthropic though, their 5x/20x is mostly for 5h limit.
The winner is Claude
They're not providing a local extension with the same performance at the time - it's only offered on their cloud services.
The local version https://github.com/planetscale/lead is mainly just for testing the syntax, it doesn't have the same perf characteristics.
> A WORLD, FOUR BILLION YEARS IN THE MAKING
> DEEP TIME · CONTINENTS IN MOTION
> Distance becomes an ocean.
> Continents drift apart. Shallow seas spread across land that will one day be dry.
> EXPLORE 4.54 BILLION YEARS OF EARTH HISTORY
> EARTH THROUGH TIME
> Scroll to travel through time Drag to orbit
> THE PAST IS ANOTHER WORLD. IT IS ALSO OURS.
> ILLUSTRATIVE TRANSITION
> Ga = billion years · Ma = million years / Event spacing is not linear
It gets to the point where if you ask a model to make a tiny toy "OS", it'll write stuff like "REAL MODE. REAL MULTITASKING. DRAG WINDOWS / TAB TO SWITCH" directly in the background of the desktop of said toy OS.
But it hasn't.
C2PA is file metadata and can be trivially stripped away, unlike hidden watermarks, e.g. SynthID.
> POWBlock is shareware. Good shareware we hope, because we only gate the optional "enterprise-level" super-nerd stuff behind a paid license and give 100% of the main features for free. But for that reason the source code is closed and private. That being said, we're not scared of code reviews. If you're a bona fide security researcher willing to sign an NDA and publicly vouch for what you find, we can let you look at it.
> Reach out to us and/or drop some crypto. A license key is a flat 50 US dollars. If you buy as an individual operator, the key is valid for your lifetime for as many copies as you want, any where you want, just as long as you are the sole owner/renter of the physical servers you deploy on. Corporate/government entities need 1 license key for each US State and/or non-US country they plan to deploy it in, but can do unlimited deployments in any such place that they have a license.
If someone's interested in a real Windows 7 lookalike, take a look at https://gitgud.io/aeroshell/atp/aerothemeplasma, it's extremely close.
See for yourself:
- Kumander: https://www.kumander.org/img/images/kumander.png
- AeroShell: https://gitgud.io/aeroshell/atp/aerothemeplasma/-/raw/Plasma...
- https://arxiv.org/abs/2311.14455
It also requires a root launcher that runs code from the user ~/.hp1008 dir, so security is weakened.
It also requires a root launcher that runs code from the user ~/.hp1008 dir, so security is weakened.
As I understand, the code part is in https://github.com/Tiger3807861189/J-Space-Cognition-Suite-V... but there's nothing there that would make V4 Pro perform as well as Fable, it's mostly prompting - that's not even a proper implementation of "J-Space".
If you check on OpenRouter, some other providers serve V4 Flash at seemingly cheaper normal input/output tokens rates, but with a huge caveat: they have at least a 5x increase of the cache hit cost of the official API, some have a 10x+. No provider comes close to Deepseek's old low cache prices, and cache is 90%+ of what matters in agentic sessions.
Closest comparison:
- Deepseek: $0.14/$0.28 with $0.0028 cache hit cost for official API
- DeepInfra: $0.08/$0.18 (cheaper base rates!) with $0.016 cache hit (almost 6x!! Deepseek's current cache cost)
Another great example is Kimi K3, official API is $3/$15 and the cheapest provider on OpenRouter is $2.8/$14, only a tiny difference.
Query: HN
Result:
{ "function_calls": [ { "name": "lock_door", "arguments": { "door": "front door" } } ], "reasoning": "User wants to lock the door. No specific door mentioned, so use 'front door' as default.", "confidence": 0 }
I'd expect it to at least ignore (call no tools) for the queries that it doesn't understand. And it seems like it does do that, just not consistently.