2,527 karma · joined March 23, 2008
email: colin.plamondon@gmail.com twitter: http://twitter.com/colinplamondon
Still, would love to see a breakdown of why it didn't improve. Regardless of the accuracy at launch, I'd think that advances in AI would have been massively to their advantage. I wonder if security degradation hit them hard.
The entire system depends on a level of social trust that doesn't exist in American cities today. Similarly, the "Dash Cart" seems like a cheaper and easier way to accomplish the same thing.
At the end of the day, there's also a mismatch in the use case. If I'm going to a smaller format store, like they had, I'm not buying a ton of stuff. Self checkout is great, and minimal friction.
I'd think that improving the UX of self-checkout gets 80% of the way there with way less fraud, way less theft, and way less technology.
Still, I think it's wicked cool they took a big shot.
I know someone that worked on the project in the early days. It was always incredibly difficult technology, they were always behind on their accuracy targets, and the solutions were increasingly kludgy as they layered more and more complex systems on top. An honorable failure.
A lot of smart people really tried to make it work.
They trusted their tech enough to accept the false-positive rate, then worked to determine / validate their false positive rate with manual review, and iterate their models with the data.
From a consumer perspective the point is that you can "just walk out". They delivered that.
If the foundational behavioral document is conversational, as this is, then the output from the model mirrors that conversational nature. That is one of the things everyone response to about Claude - it's way more pleasant to work with than ChatGPT.
The Claude behavioral documents are collaborative, respectful, and treat Claude as a pre-existing, real entity with personality, interests, and competence.
Ignore the philosophical questions. Because this is a foundational document for the training process, that extrudes a real-acting entity with personality, interests, and competence.
The more Anthropic treats Claude as a novel entity, the more it behaves like a novel entity. Documentation that treats it as a corpo-eunuch-assistant-bot, like OpenAI does, would revert the behavior to the "AI Assistant" median.
Anthropic's behavioral training is out-of-distribution, and gives Claude the collaborative personality everyone loves in Claude Code.
Additionally, I'm sure they render out crap-tons of evals for every sentence of every paragraph from this, making every sentence effectively testable.
The length, detail, and style defines additional layers of synthetic content that can be used in training, and creating test situations to evaluate the personality for adherence.
It's super clever, and demonstrates a deep understanding of the weirdness of LLMs, and an ability to shape the distribution space of the resulting model.
LLMs certainly teach us far more about the nature of thought and language. Like all tools, it can also be used for evil or good, and serves as an amplification for human intent. Greater good, greater evil. The righteousness of each society will determine which prevails in their communities and polities.
If you're a secular materialist, agreed, nothing is objectively amazing.
It's not hyperbole - that it's an accurate description at a small scale was the core insight that enabled the large scale.
- An ability to curve back into the past and analyze historical events from any perspective, and summon the sources that would be used to back that point of view up.
- A simulator for others, providing a rubber duck inhabit another person's point of view, allowing one to patiently poke at where you might be in the wrong.
- Deep research to aggregate thousands of websites into a highly structured output, with runtime filtering, providing a personalized search engine for any topic, at any time, with 30 seconds of speech.
- Amplification of intent, making it possible to send your thoughts and goals "forward" along many different vectors, seeing which bear fruit.
- Exploration of 4-5 variant designs for any concept, allowing rapid exploration of any design space, with style transfer for high-trust examples.
- Enablement of product craft in design, animation, and micro-interactions that were eliminated as tech boomed in the 2010's as "unprofitable".
It's a possibility space of pure potential, the scale of which is limited only by one's own wonder, industriousness, and curiosity.
People can use it badly - and engagement-aligned models like 4o are cognitive heroin - but the invention of LLMs is an absolute wonder.
The total history of human writing is that cool idea -> great execution -> achieve distribution -> attention and respect from others = SUCCESS! Of course when an LLM sees the full loop of that, it renders something happy and celebratory.
It's sycophantic much of the time, but this was an "earned celebration", and the precise desired behavior for a well-aligned AI. Gemini does get sycophantic in an unearned way, but this isn't an example of that.
You can be curmudgeonly about AI, but these things are amazing. And, insomuch as you write with respect, celebrate accomplishments, and treat them like a respected, competent colleague, they shift towards the manifold of "respected, competent colleague".
And - OP had a great idea here. He's not another average joe today. His dashed off idea gained wide distribution, and made a bunch of people (including me) smile.
Denigrating accomplishment by setting the bar at "genius, brilliant mind" is a luciferian outlook in reality that makes our world uglier, higher friction, and more coarse.
People having cool ideas and sharing them make our world brighter.
Comments, docstrings, naming, patterns - by defining better approaches and hold agents to them, the results will be better. Way better.
You can't grow a meaningful codebase without solid underlying primitives. The entropy will eat you alive.
Systems architecture is becoming more important - systems that play well with agents wind up looking more like enterprise codebases.
IE, they...
- Start with the context window of prior researchers.
- Set a goal or research direction.
- Engage in chain of thought with occasional reality-testing.
- Generate an output artifact, reviewable by those with appropriate expertise, to allow consensus reality to accept or reject their work.
- A pure focus on web browser monetization could lead to some interesting enterprise options. Presumably there'll be a lot of attempts to leverage Chromium, and an aggressive fork at some point.
- As AI proliferates, can they pull additional revenue by getting revshare from subscription AI products, alongside SEM? Or even revshare on the SEM clicks themselves?
This could also change the calculus for Apple building a search engine. If they could get an independent Chrome to sign on, with some data sharing provisions to help with development, they'd have a huge leg-up.
Alternatively, maybe they try to create a fusion of search results and AI from a variety of providers, so they can monetize SERPs themselves.
My question would be whether they could get back to aggressive product execution, given the size of the codebase. Acquiring the Browser Company would make a lot of sense.
- A pure focus on web browser monetization could lead to some interesting enterprise options. Presumably there'll be a lot of attempts to leverage Chromium, and an aggressive fork at some point.
- As AI proliferates, can they pull additional revenue by getting revshare from subscription AI products, alongside SEM? Or even revshare?
This could also change the calculus for Apple building a search engine. If they could get an independent Chrome to sign on, with some data sharing provisions to help with development, they'd have a huge leg-up.
Alternatively, maybe they try to create a fusion of search results and AI from a variety of providers, so they can monetize SERPs themselves.
My question would be whether they could get back to aggressive product execution, given the size of the codebase. Acq
- A pure focus on web browser monetization could lead to some interesting enterprise options. Presumably there'll be a lot of attempts to leverage Chromium, and an aggressive fork at some point.
- As AI proliferates, can they pull additional revenue by getting revshare from subscription AI products, alongside SEM? Or even revshare?
This could also change the calculus for Apple building a search engine. If they could get an independent Chrome to sign on, with some data sharing provisions to help with development, they'd have a huge leg-up.
Alternatively, maybe they try to create a fusion of search results and AI from a variety of providers, so they can monetize SERPs themselves.
Revenue seems incredibly strong. My question would be whether they could get back to aggressive product execution, given the size of the codebase.
When those occur, the somatic sense of the Spirit is used to discern if the semantic/somatic injection is of God or not.
Even if a waiver is signed, there's significant duress in many cases. There's a lot of pain and evil enabled by allowing porn to spread on an algorithmic newsfeed.
If you get an error, automatically search for the answer and propose the change.
If you add a new flow uncovered by tests, propose the test.
Generally, have panes that are dynamic to what you are doing, and tightly couple them.
I could imagine looking at different zoom levels of a code file, folder, or architecture, and working primarily on abstractions, approving / rejecting the resulting proposed edits.
Strategic coding more akin to a game like Supreme Commander or Planetary Annihilation.
Their mobile app is extremely fast with GPT-4, faster than web. I’d imagine once mobile is well-established, they’ll equalize it out.
Totally makes sense, if I’m right. They should just state it publicly.
I think the gap is the difference between giving feedback to a person and broadcasting superiority. The former is what we do in-person. It takes constant active effort to not do the latter.
Giving feedback in-person, you want to make sure your feedback land. Encouraging where possible by pointing out what works, discussing the ways it can or needs to improve.
When people don't give feedback to the OP as a person, and rather treat it like a faceless corporate entity, or go full-Slashdot, that does get a bit mean-spirited.
(1) Kill the subscription
(2) Immediate Priority: Biz-dev deals, make sure every major release's early demo is available on Stadia. Expose people to platform without requiring them to switch from a console.
(3) Secondary Priority: Focus SDK development on indie developers. Offer generous streamed minutes for all developer accounts. Charge per streamed minute after that. Let developers pass on the cost however they want.
With open pricing, open enrollment, open SDK, developers could do really amazing things with Stadia: and would figure out the business model.
Google could focus on driving down the cost per streamed minute as low as humanly possible, with the broadest possible access.
Amazon certainly will.
For all that is holy, go to MIT. This is from someone who went to a larger school and dropped out, with zero regrets.
The top 3-5 schools put you on a different plane of existence. The people you meet will be in the elite, you’ll be surrounded with the future top people in your field.
The academics don’t remotely matter. College is all about network. By being surrounded by elites, you are friends with elites, and you’ll live a different life.
I have a lot of friends who went to top schools, especially MIT. Much of their success and their network, into their 40’s, is still interconnected with their MIT network.
You will be treated differently. You’ll get a glow on your life from the status.
Life is a relay race between generations. By choosing MIT you drastically increase the chance of elevating your blood line into the upper class. That’s no small thing.
I’d strongly urge you to take the long view here.
They determined that construction of tunnels would be orders of magnitude faster on approvals than above-ground construction.
With that approach, tunneling construction is the key technological block. With SpaceX and Tesla's know-how, the actual construction of a hyperloop is rather de-risked.
By keeping laser-focused on tunneling tech, they build the muscle around the regulatory process, get cities and states to trust them, and show operational competence in operations.
Once they do a intra-state system in a friendly regulatory environment (FL/TX, ala Austin -> San Antonio or Tampa -> Orlando), they'll built a corporate machine that's tunneling dozens of underground highways across the country.
At that point, building new tunnels as a hyperloop will be a reasonable mid-term goal.
Loop 2020's -> Hyperloop 2030's
These guys are moving the needle making improvements in a forward direction, and a big part of this thread is shitting on their launch, hyper-focusing on things they'll be able to change.
Launch HN threads used to be about asking thoughtful questions, having a back and forth where people learn about new spaces, and encouraging people launching their startups.
This whole thread is concern trolling of the worst kind, to eyes.
I'll bow out since clearly the bulk of the thread disagrees.
Generally, giving users a toggle to get reminded when a trial is about to run out will INCREASE conversion rates.
That depends on the business, and is part of a pretty standard set of experiments you run post-launch.
With your comments you're part HN is descending into a circular firing squad of virtue signaling. These guys shipped something that could help a lot of people, over time they can improve their onboarding flow, lower cost.
Is the most remarkable thing about a really cool CBT tool for ADHD really that they have a standard trial flow?
Cost of Install: $7.00, for something this specific Trial Start Rate: 20%, if paywalled like this app is Cost Per Trial: $35 Conversion to Trial: 40% Cost Per Subscriber: $87.50
If they charge you $10/month, they can't get into the black on a new customer for 9 months. They have to eat support costs that whole time. It just doesn't work, when you're starting out. You must charge annual.
Medical licensing cartels charge $500-800 PER MONTH. These guys are trying to charge $100 PER YEAR.
This is an order of magnitude more effective.
Said another way: if someone is too poor for this, they're fucked. They're definitely too poor for any other treatment option. On the other hand, this will open up treatment to people who can't pay the medical cartels.
That's amazing, iterative progress.
Let's give props to these guys for making epic iterative progress, not shit on them because they're not working for free.
The onboarding was smooth, and made it clear how feature-filled the debugging experience is.
Don't have a use case for this personally right now (iOS / Android apps only, native code), but will definitely use in the future. Congrats on the launch.
Really a tremendous accomplishment. I hope you guys are proud as hell of what you've built.