Not as good as Astra or Fable 5.1 on this test as far as I can see. I wonder if any benchmark exists for artistic taste, visual sophistication etc. I think your Pelican test does touch on these aspects of a model and is useful for developers trying to build rich digital experiences (includes games, interactive websites and apps). These benchmarks are subjective so it may not be easily established and will have polarized reactions before it gains legitimacy. May even need human judgement layers adding to the cost of running it.
I am not seeing the connection. Whenever I hear capital and elites, it’s a clear red flag around lack of understanding. And just laziness about money. I think a lot of these voices would change if they just had 20% of [edited: discretionary] income auto-invested in equities from an early age.
I just feel apps like linear are increasingly getting in the way of full send agentic development where sub agent orchestration is done through agent to agent messaging, work trees, on demand git restructuring and epoch specific coordination plains, often .md files. The smaller the human component of total product development gets, the more this may be the case.
Any real world experience with Grok Ultra $300 monthly subscription vs Claude Code Max in terms of overall built work mileage, or general token limits?
I use SelfControl for Mac and set the down time to around 15 days. No Youtube no news sites etc. Except HN. It's helped a lot, after a week you can notice measurable change in attention and habit and productivity. SelfControl lets you add 48 hrs to your blocked time repeatedly, so you can get to a 15 day mark pretty easily through repeated mouse clicks.
Waiting for my coding agent to have zero idea what I mean when I say build for iPhone duo (due to knowledge cutoffs in pre-training; happened when Liquid Glass came out previously).
I don’t quite get the assumptions here. If c were actually 5 km/h, causal influences would propagate much more slowly. With G unchanged, planets and stars at their current masses and sizes would lie within their Schwarzschild radii. Any rock greater that 1.5km in radius is a blackhole. With c reduced to 5 km/h, the black-hole threshold would correspond to a surface escape speed of just 5 km/h. Holding the other relevant constants fixed would also radically alter atomic structure and chemistry. It seems like a lot would change before we could be around to observe the cool redshifts. Is the reduced c being applied only to the visual effects?
Oof. In the AGI scenario, why not preserve a deduplicated set of specimens like the unique set of individuals or groups at the top of their field in all walks of life ranging from sciences, sports and arts. But then there’s the cultural value of the global population in their shared norms and collective patterns. This too would be unique and valuable to an AGI vs not having it. Unless we were an existential threat to it, the AGI would keep us around for things beyond petty resource scarcity.
“ It’s important to note that there are several key differences between the workspace we identified in Claude and the global workspace model in humans. The brain’s workspace is sustained by recurrent loops—signals cycling back through the same circuits over time. In contrast, Claude’s workspace evolves over a single pass through the network, with the network’s depth playing the role that time plays in the brain. In this sense, Claude’s internal workspace processing is time-limited relative to humans’ (though it can compensate for this constraint by “thinking out loud” using its scratchpad).”
You can’t have mass immigration from mainly economic migrants from the third world, under funded police forces, and a legal system built for high trust, highly educated, fairly homogeneous populace at the same time; and expect things to be all happy days.
Those rear tail lights don’t sit right with me. I know there’s probably some aerodynamic reason behind it but Jony, those aren’t the proportions that just work. Steve wouldn’t approve this. And I feel Jony was always partly Steve when Jony was at his best.That said the issue is the asymmetric black negative space below and above the red circles. This is mostly fixed if you get the Luce in black or very dark gray.
I can relate to this. Gemini 3 doesn’t know a thing about iOS 26 or Liquid Glass. It constantly assumes this is some custom view that I want it to develop and ends up building something out the previous gen apis like ultrathinmaterial.
Many sub-Saharan African populations, such as Bantu-speaking West Africans, exhibit relatively lower genetic diversity compared to the Khoisan people, who typically have light brown skin. The Khoisan lineages diverged from those leading to Bantu and other sub-Saharan groups around 100,000–150,000 years ago, making them one of the most ancient human ancestries.
The UI is better - they box the specific types of actions the orchestrator agent takes with a clear categorization. The standard quality of life shortcuts like type a number to respond to an MCQ are present here as well. They use specialized sub agents such as one with big context window to find context in the codebase. The quotas appear to be much more generous vs CC. The agent memory management between compacting cycles seems to have a few tricks CC is missing. Also, with 3.0 Flash, it feels faster with the same level of agency and intelligence. It has a feature to focus into an interactive shell where bash commands are being executed by the orchestrator agent. Doesn't feel like Google is trying to push you to buy more credits or is relying on this product for its financial survival - I suspect CC has some dark patterns around this where the agents runs cycles of token in circles with minimal progress on bugs before you have to top up your wallet. Early days still.
The UI is better - they box the specific types of actions the orchestrator agent takes with a clear categorization. The standard quality of life shortcuts like type a number to respond to an MCQ are present here as well. They use specialized sub agents such as one with big context window to find context in the codebase. The quotas appear to be much more generous vs CC. The agent memory management between compacting cycles seems to have a few tricks CC is missing. Also, with 3.0 Flash, it feels faster with the same level of agency and intelligence. It has a feature to focus into an interactive shell where bash commands are being executed by the orchestrator agent. Doesn't feel like Google is trying to push you to buy more credits or is relying on this product for its financial survival - I suspect CC has some dark patterns around this where the agents runs cycles of token in circles with minimal progress on bugs before you have to top up your wallet. Early days still.