HNHacker News
TopNewBestAskShowJobs

mchusma

4,645 karma · joined February 9, 2011

Founder & CEO SignNow Founder & CEO tidy.com CTO HotMic
submissionscomments
mchusma··on GPT-5.6 Luna vs. GPT-6 Astra: Is a $1.20 Model Good Enough for Code Review?
We had a two human PR requirement until recently we dropped it. It was slowing us down too much now the human developer creating the future is obviously writing it all with AI so they need to check it then depending on the feature and it’s use it requires a PR but it’s not universal and we’ve stepped up our automated test Tan X what it used to be it’s been so far fewer bugs better delivery
mchusma··on iOS 27, iPadOS 27, and macOS 27
Apple's search is so comically bad. For example, I can click on my applications, and I have an app called Chess. I can often type letters C-H-E-S-S and have it not show the Chess app. This is inside the applications folder!

Many people have complained before. Honestly I ONLY search for app names in spotlight search or application window, it should be trivial to make it actually work. Maybe i just need to vibecode some alternative search bar :(

mchusma··on Pion, an agent designed to run any company autonomously
There was surprisingly little information on how they actually do this, but we run our business with a large number of, what we call "AI employees" in addition to regular employees, and they act in interesting ways. We've been building out orchestration tools to handle this, and we'll probably do a write-up or blog post on it soon. May even open source some of it.

The preview is that the problem with most agents (and this includes frameworks like Grokbot and Openclaw and Hermes) is that for many of them, they're black boxes. They say they learn or improve, but it's a black box in what they do. Getting agents to reliably do things is hard, and getting agents to build out software tools to help themselves improve and do better over time is also hard.

Our approach at a high level is pretty simple: every single AI employee is a standalone GitHub repo that shares some characteristics, but we direct them to build as much software as possible to make their goal as easy and reliable to manage as possible. Then we have a shared communication layer for bots across the company to interact with humans and AI. We have decided to organize these like departments similar to the way you might hire out humans. I'm not 100% sure if that's the best approach, but I will say it's been easier for people to understand because they're more mentally easily able to traverse the bot org chart if it somewhat reflects a traditional business org chart.

Each of these AI employees has specific sets of goals and KPIs, instructions that they manage the business with manager bots. We have layers of management, which we actually have found helpful. We also run different bots with different models and harnesses, and some using different models and harnesses to check the work before anything can get done, along with lots and lots of testing.

Every single time, actions have a massive amount of tests based off of previous failures to prevent failures in the future. Sorry for rambling. I do think this is a very interesting space. I didn't see anything interesting in Pion that was public on this website, but I do anticipate that more companies will be "AI and software first," as in the substrate of the company is basically a software application powered by autonomous agents, with humans as a fallback.

mchusma··on The case against JPEG XL
TIL that AVIF supports progressive decoding, cool!
mchusma··on I spent $220 on Google app ads and 60% of the installs were robots
Google display ads are completely unusable for me (not app ads). So much flagrant fraud. Like even the most simple of AI models could detect the fraud. Very disappointing.
mchusma··on Feeling Sad about AI
I have a different perspective. Most of the world is software. The US Constitution is software. Companies are software. With AI, we can build so much more software, its incredible! How these systems are architected are complicated. Sure, if you really loved coding at an extreme low level, that may go away. But I think there is 10x more software engineering going on my company (and in my personal life). Things change, but things are more fun than ever.
mchusma··on Growing proof that autonomous cars save lives
It is correct to compare to average drivers because that is ultimately what they are going to replace. AVs are coming for the whole market, and if they are significantly safer than the average driver the average driver shouldn’t drive. I intend to stop driving when possible.
mchusma··on Apple Watch Series 12
I was actually quite impressed with the AI note-taking and other features, and I could see myself using them, so I did like that.

With health, I really wish Apple would truly lean into health...go nuts and try to become effectively a super primary care provider. Integrate with blood tracking, automated diagnostics, something much cooler.

The problem with these kind of wearables today is this is just not very actionable. You have this data, and nobody ever really looks at it. "You had a bad night sleep, maybe sleep more." Yeah no kidding. I want it to say: "Hey we noticed this trend, it could be indicative of X, we recommend Y test. Would you like us to send it to you within 2 hours so you can get more data?"

mchusma··on Google DeepMind Releases AlphaGenome Atlas
This has Demis written all over it. There is a great video of him with AlphaFold chatting with the team about releasing some results, and he asked something like “what if we just do them all?”

Very excited to see that happen here.

mchusma··on LibreOffice breaks download records after declaring it has no AI features
Can you control it agentically? If so it’s has all the ai features I would want. I typically don’t care for built in ai features.
mchusma··on GPT-6 Astra
The games on mobile safari were broken. Buttons all misaligned in the kart racer one, the spaceship thing froze for a while, then kind of loaded but maybe not? Wasn't super compelling.

I'm not trying to be too negative on it, it could be the best model right now, but it clearly isn't some agi god because things like that should have been caught (also should have been caught by human reviewers).

mchusma··on Launch HN: Nori Robotics (YC S26) – A low-cost humanoid robot for development
This is cool! As someone who's built based off of old Roomba platforms and other models, I feel like this looks pretty fun and a good price point. What do you think the next version will target? Will you stay educational and low price point? Targeting some other usecases?
mchusma··on Claude Fable 5.1 and Claude Mythos 5.1
I think the signal from Anthropic is pretty clear between Haiku not getting an update in a year and the Sonnet issues this year. They don't care about low intelligence models. You should go elsewhere.

That's what we've done, migrated workflows away from Haiku and Sonnet. I actually think this is not a crazy position because these lower models have so much competition from Grok, OpenAI, DeepSeek, and about 20 other labs with really solid models in the Haiku to Sonnet range. So what is the point of Anthropic competing in these spaces where everything is going towards zero cost?

mchusma··on EFF to Courts: Don't Rewrite Copyright over AI Hype
I’m just not convinced that most artists would change their behavior under no copyright, so somewhat agree. But the main issue is length of time.

The fact that copyright is longer than patents makes no sense. Patents prove that a shorter period is sufficient to stimulate large investment. And you have to pay for patents!

If you assume reasonable discount rates, the economic value plummets over time for an asset producing the same income (a bad assumption for copyright work in general).

At an 8% discount rate, extending copyright from 32 years to infinity adds less than 10% to its total present-day value. There are many good reform proposals, but mine would be something like 15ish years of free copyright, followed by 1 paid renewal for 15 years that is modest in price (eg $1000) and one extra 10 year renewal that is something like $1M. Meaning you have most economically neutral work free up at year 15, small business types having it affordable through 30 years, and the big Hollywood productions can go from 30-40 years for a modest (to them) fee to incentivize investments in expensive works. Movies come to mind. Give existing work a 20 year buffer on top of that to transition and there is a very small economic effect on everyone.

mchusma··on EFF to Courts: Don't Rewrite Copyright over AI Hype
The government could always pay for results too, like $10B for something that does this. I don’t love it either and am not positive it is better than patents.
mchusma··on Launch HN: Almanac (YC S26) – AI that knows your company
I would agree that there are a lot of people in this space, both products and roll your own. We are doing our own stuff here in our company right now, and could be interested. But for me, I didn't feel like i know enough to actually assess and take the next steps. Good luck!
mchusma··on Europe's summer drought is so extreme that desertification is a growing threat
This is all solvable with abundance. We could put together nuclear power with solar and desalination and carbon removal.

Let’s start acting more like a kardishev II civilization.

mchusma··on Smaller reactors bring nuclear power closer to fulfilling its promise
> If it's so much cheaper to build multiple small reactors, just build one big plant with 24 small reactors.

The biggest costs to nuclear are associated with each of them being unique snowflakes. They need to be standardized and mass produced to bring down costs.

So the dream is many big plants (eg starting 10+ per year), which is what France did and China does, but since we can’t seem to have that here, small reactors are an attempt to solve that.

mchusma··on Gemini Omni 1.1 Flash
You realize his examples are public sector unions. Which mean the “employer” you are referring to is the citizens? Not some evil corporation, government mandated monopolies which then get taken over to serve their union members.

I don’t have any objection to voluntary unions. Meaning people can join them for support or collective bargaining, but that citizens can just opt to not employ anyone in the union. Unions could be support structures for workers, instead they are mafia bosses.

mchusma··on Small Models Have Arrived
I have an agentic workflow and Luna just always gets stuck, SOL and grok 4.6 don’t. I like Luna in theory I just find not much practical work for it yet in coding type work.

Now I think Luna is plenty good for many applications inside a very good harness/scaffold. And I think there are a lot of those usecases. So I think these small models are really good for application developers.

But for entrepreneurial knowledge work all of my work still benefits a lot from more intelligence.

mchusma··on FDA approves first in class targeted therapy for metastatic pancreatic cancer
FDA is just too slow. Whole I’d like to do more radical reform, at a minimum this stupid long wait should be tossed. If someone hits a target endpoint auto approval pending further review or something

Remember in COVID when results. Came in and they took a month to schedule the meeting?

mchusma··on Study reveals UnitedHealth's profit margins four times what it claimed [pdf]
Some of this has to do with limits from the ACA (Affordable Care Act), which limited the margins of insurance companies. It creates incentives for higher premiums, but also these types of gains, which is just bad for everybody.

I don't think there's much you can look at with the Affordable Care Act and think that it was a success.

mchusma··on FDA authorizes first wearable device that monitors ketone and blood sugar levels
Wild numbers. I already thought GLPs were huge but this makes me think it’s even larger.
mchusma··on OpenAI Jalapeño: Better than Nvidia Blackwell
You are correct. I think this is the bull case. It seems like this would be useful right now for some things (eg moderation).
mchusma··on OpenAI Jalapeño: Better than Nvidia Blackwell
Maybe! (1) Would SOL level intelligence be useful 3 years from now? 5 years? (2) would dedicated chips be the most affordable way to run this model in 3-5 years?

I suspect the answer to both of these questions is yes right now, but I agree it’s borderline.

mchusma··on Starbase, LA
Yeah, the Data Center NIMBY crowd is making these look real good right now.
mchusma··on OpenAI Jalapeño: Better than Nvidia Blackwell
I think they talked about this being general purpose chip but I would think that Anthropic/OpenAI are at the scale now they could bake LLM weights into chips themselves.

For example, GPT Sol baked into a custom chip run for $100M that runs 10x as fast and 10x as cheap should pay for itself as long as the chip is useful for long enough.

While 2 years ago nothing was useful more than 1 year long, there are many older models in use now (e.g. Haiku 4.5, GPT-OSS 120b), and I expect this trend to continue.

I know this is what Taalas was doing (acquired by AMD), here was their demo, https://chatjimmy.ai/ which is based on Llama 3.1 8B. It feels like this should start to happen soon.

mchusma··on Why some US restaurants are banning tips
I’ve never met someone who didn’t think tipping was too widespread and too high. You say a majority of people are happy with the state of US tipping?
mchusma··on How credit card rewards became a $9.2B wealth transfer
One note on patio11’s opinion on this is that he really overweights the ongoing work and innovation required for electronic payment processing. It WAS a great novelty and deserves to have made a lot of money for 30 years. But the reason they make so much money today is monopolistic low behaviors to lock in their advantages. It’s not a free marlet because of deals over time, some of the most famous of which are their prohibition on charging different rates for cards or even disclosing the rates on cards.

I think the most simple piece of legislation to solve a lot of problems is to allow merchants to pass along the interchange rate to their customers. If they could do this legally and operationally, this would solve most issues here. If a credit card wants to be expensive, fine the consumer should pay for it. Because of contractual and operational limitations, credit card companies have gotten themselves into the current arms race.

If stripe implemented this, it would make me appreciate them as a force for good instead of being a part of the problem.

mchusma··on Firefox intent to ship: JPEG XL
I am excited for this! A great practical format.

I love progressive rendering. At 15% loaded in this example it’s surprisingly good already.

← PreviousPage 2 of 34Next →