4,645 karma · joined February 9, 2011
Many people have complained before. Honestly I ONLY search for app names in spotlight search or application window, it should be trivial to make it actually work. Maybe i just need to vibecode some alternative search bar :(
The preview is that the problem with most agents (and this includes frameworks like Grokbot and Openclaw and Hermes) is that for many of them, they're black boxes. They say they learn or improve, but it's a black box in what they do. Getting agents to reliably do things is hard, and getting agents to build out software tools to help themselves improve and do better over time is also hard.
Our approach at a high level is pretty simple: every single AI employee is a standalone GitHub repo that shares some characteristics, but we direct them to build as much software as possible to make their goal as easy and reliable to manage as possible. Then we have a shared communication layer for bots across the company to interact with humans and AI. We have decided to organize these like departments similar to the way you might hire out humans. I'm not 100% sure if that's the best approach, but I will say it's been easier for people to understand because they're more mentally easily able to traverse the bot org chart if it somewhat reflects a traditional business org chart.
Each of these AI employees has specific sets of goals and KPIs, instructions that they manage the business with manager bots. We have layers of management, which we actually have found helpful. We also run different bots with different models and harnesses, and some using different models and harnesses to check the work before anything can get done, along with lots and lots of testing.
Every single time, actions have a massive amount of tests based off of previous failures to prevent failures in the future. Sorry for rambling. I do think this is a very interesting space. I didn't see anything interesting in Pion that was public on this website, but I do anticipate that more companies will be "AI and software first," as in the substrate of the company is basically a software application powered by autonomous agents, with humans as a fallback.
With health, I really wish Apple would truly lean into health...go nuts and try to become effectively a super primary care provider. Integrate with blood tracking, automated diagnostics, something much cooler.
The problem with these kind of wearables today is this is just not very actionable. You have this data, and nobody ever really looks at it. "You had a bad night sleep, maybe sleep more." Yeah no kidding. I want it to say: "Hey we noticed this trend, it could be indicative of X, we recommend Y test. Would you like us to send it to you within 2 hours so you can get more data?"
Very excited to see that happen here.
I'm not trying to be too negative on it, it could be the best model right now, but it clearly isn't some agi god because things like that should have been caught (also should have been caught by human reviewers).
That's what we've done, migrated workflows away from Haiku and Sonnet. I actually think this is not a crazy position because these lower models have so much competition from Grok, OpenAI, DeepSeek, and about 20 other labs with really solid models in the Haiku to Sonnet range. So what is the point of Anthropic competing in these spaces where everything is going towards zero cost?
The fact that copyright is longer than patents makes no sense. Patents prove that a shorter period is sufficient to stimulate large investment. And you have to pay for patents!
If you assume reasonable discount rates, the economic value plummets over time for an asset producing the same income (a bad assumption for copyright work in general).
At an 8% discount rate, extending copyright from 32 years to infinity adds less than 10% to its total present-day value. There are many good reform proposals, but mine would be something like 15ish years of free copyright, followed by 1 paid renewal for 15 years that is modest in price (eg $1000) and one extra 10 year renewal that is something like $1M. Meaning you have most economically neutral work free up at year 15, small business types having it affordable through 30 years, and the big Hollywood productions can go from 30-40 years for a modest (to them) fee to incentivize investments in expensive works. Movies come to mind. Give existing work a 20 year buffer on top of that to transition and there is a very small economic effect on everyone.
Let’s start acting more like a kardishev II civilization.
The biggest costs to nuclear are associated with each of them being unique snowflakes. They need to be standardized and mass produced to bring down costs.
So the dream is many big plants (eg starting 10+ per year), which is what France did and China does, but since we can’t seem to have that here, small reactors are an attempt to solve that.
I don’t have any objection to voluntary unions. Meaning people can join them for support or collective bargaining, but that citizens can just opt to not employ anyone in the union. Unions could be support structures for workers, instead they are mafia bosses.
Now I think Luna is plenty good for many applications inside a very good harness/scaffold. And I think there are a lot of those usecases. So I think these small models are really good for application developers.
But for entrepreneurial knowledge work all of my work still benefits a lot from more intelligence.
Remember in COVID when results. Came in and they took a month to schedule the meeting?
I don't think there's much you can look at with the Affordable Care Act and think that it was a success.
I suspect the answer to both of these questions is yes right now, but I agree it’s borderline.
For example, GPT Sol baked into a custom chip run for $100M that runs 10x as fast and 10x as cheap should pay for itself as long as the chip is useful for long enough.
While 2 years ago nothing was useful more than 1 year long, there are many older models in use now (e.g. Haiku 4.5, GPT-OSS 120b), and I expect this trend to continue.
I know this is what Taalas was doing (acquired by AMD), here was their demo, https://chatjimmy.ai/ which is based on Llama 3.1 8B. It feels like this should start to happen soon.
I think the most simple piece of legislation to solve a lot of problems is to allow merchants to pass along the interchange rate to their customers. If they could do this legally and operationally, this would solve most issues here. If a credit card wants to be expensive, fine the consumer should pay for it. Because of contractual and operational limitations, credit card companies have gotten themselves into the current arms race.
If stripe implemented this, it would make me appreciate them as a force for good instead of being a part of the problem.
I love progressive rendering. At 15% loaded in this example it’s surprisingly good already.