HNHacker News
TopNewBestAskShowJobs

george_max

142 karma · joined January 20, 2026

submissionscomments
george_max··on Cf: The Agentic CLI for the Cloudflare API
It feels heavy, the emojis are overused, and it is probably the only CLI tool I've felt as "clunky". Maybe that and Claude Code TUI.
george_max··on Show HN: HN.watch – Videos of all Hacker News posts
Haha, I should have read more thoroughly! Thanks
george_max··on Show HN: HN.watch – Videos of all Hacker News posts
Nice concept, but why not generate one video per post rather than re-generating the video for each click per user? The result is going to be the same, and it will be significantly cheaper to run.
george_max··on Ollaya – Ollama for open-source, Jev-style decision models
Has anyone actually seen better or the same results with Laya compared to Jev? From my experience, Laya performs significantly worse. It's less confident and often makes wrong decisions with more complex queries.
george_max··on Ollaya – Ollama for open-source, Jev-style decision models
I am fairly confident if Jev-style decision models are seen as prominent (which, they seem to be), Ollama will support them. Surprised the team hasn't implemented this already.
george_max··on How accurate have Ed Zitron's AI skeptic predictions been?
Help me understand where Meta and Alphabet's issues are. They seem to be doing quite well on paper -- nothing anomalous in terms of profitability or scaling recently. Layoffs are to be expected when AI can increase productivity.
george_max··on Claude Fable 5.1 and Claude Mythos 5.1
Agreed. The area I think will become more prevalent in the future for organizations are cost per intelligence -- effectively efficiency. An unoptimized model that costs 90x more than another that is only 10-15% less intelligent is something I would say is not a good deal.
george_max··on Claude Fable 5.1 and Claude Mythos 5.1
"Price. Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token. This is because we’re reducing our pricing on cache reads (where the model reads inputs that have already been processed and stored). For highly agentic work, the savings will often be much larger—up to approximately 45%."

They show this off, but artificial analysis contradicts the statement. Fable 5 cost $3.14 per task, while 5.1 cost $3.69 -- around a 15% jump in pricing.

https://artificialanalysis.ai/

These, IMO, are marginal improvements for a more expensive model. I stopped using Claude ~3 months back; its outputs are too jargoned, it makes architectural decisions that are not right, and it's incredibly pricey for what it is. Each decision it makes, it acts as if a problem as major as world hunger has been solved. And the overly verbose code comments, strange commit descriptions, duplicate code, and slop it generates -- which I know is not specific to Fable -- is just too much for me.

I found the best is to use something like Deepseek V4 Flash -- with a fast TPS provider -- and work on the code myself. For agentic work with computer use, GLM 5.3 flash with Hermes Desktop works well.

george_max··on Claude Fable 5.1 and Claude Mythos 5.1
This is just cache reads. In real usage it costs 15% more than Fable 5 -- all for marginal gains.

https://artificialanalysis.ai/

george_max··on How Europe is killing makers and micro-entrepreneurs
I run a small hardware startup, InfiShark Tech -- we sell affordable battery-powered cybersecurity hardware worldwide. Europe has been by far the hardest market for us.

The costs in this article are understated for products like ours. This is not just for packaging -- they have separate compliance schemes for electronics (WEEE). If selling products with batteries, you must comply with separate battery regulations. If you pass a certain threshold or contain specific chemicals within your product, you must also comply with the REACH regulation. They work almost exactly the same the article describes, but often with higher costs.

Additionally, you must have CE certification for your hardware product. There is no threshold. This means going to a lab and testing RF capabilities, and for wireless / Bluetooth devices it can cost $1,500-$8,000. You can self-certify but it is risky.

We pay 800-2,000 EUR per country annually for the authorized representative (AR) and producer responsibility organization (PRO) fees. The AR's task is to hold some documents and provide them to the government when requested -- but that's it. We can do this for free. There's no reason the costs should be this high -- and I see minimal reason why they should exist in the first place. But we must comply.

It is important to understand which regulations and which fees to pay for the product. Therefore it is important to use something like PRONEXA -- a "One-Stop-Shop for simple compliance management".

With them, we paid ~800-1,000 EUR per country as a one-off registration fee. Then we pay 975 EUR annually per country to keep compliance. Note that these costs are ONLY to PRONEXA, not to each AR or PRO. They are effectively the middleman between us and the ARs / PROs.

The fee stability is quite poor, too. We paid the 2025 annual fees roughly a month before year-end and were charged another full annual fee in February 2026. In November 2025, we paid ~500 SEK for Sweden, and in Feb 2026, we were invoiced ~15,000 SEK for that same line.

With all of this burden we only pay the countries ~0.5-50 EUR annually for the actual environmental fees. This only makes sense for the government if that number jumps up to the thousands.

I think this contributes to lack of innovation in Europe. Small startups can't ship there, and you are limited to large companies overseas. The fix is dead simple -- just introduce a threshold, or charge a recycling fee upfront with each parcel coming in, depending on the HS code. What they did with IOSS is great. You register with one country, and all parcels under 150 EUR can be sent to all of Europe while being VAT compliant. They could implement the same thing for WEEE, battery regulations, packaging, and REACH. Register with one country and cover the rest.

We are barely break-even on some countries, and others we are totally underwater. We decided to do it as a future scaling opportunity, but this first year is quite tough. We are debating if we should continue selling to Europe at this stage.

george_max··on Incident with Github.com
Could GitHub be affecting other sites, or maybe this is something else?

Did Anthropic or OpenAI fail to sandbox their models again?

https://downdetector.com/

george_max··on Slightly reducing the sloppiness of AI generated front end
Most of them don't look amazing, but I like the GTK version the most -- even though it looks slightly outdated.
george_max··on Open source AI must win
If humanity is over-reliant on frontier labs' models to perform work, the result is a dependence on the actual intelligence of these models -- not on human intelligence. This could be a small reason, on top of many others, why investors are throwing hundreds of billions of dollars a bit "carelessly" to these labs. It's fascinating seeing the models do the "hard work" (the deep, challenging thinking) for you.

The conundrum which tricks me though - is this a net negative or a positive? If humans are less intelligent, but their output is 2-3 times more intelligent (with AI), what's the result? At what point do we, as humans, stop comprehending anything and give all intelligent work to the neural nets?

And if that does happen, could we live in a society where no work, or at least a significantly less amount of work, is needed? To me, it seems like a dystopian net positive.

It might seem far-fetched to ask these, but I think these questions are getting more prevalent by the day.

george_max··on US Government directive to suspend access to Fable 5 and Mythos 5
Seems like we're starting to get reliant on the intelligence of these models to keep our outputs less "sloppy". Effectively an IaaS (Intelligence-as-a-Service). With the U.S. putting the suspension on Fable 5, we might be stuck with slop.
george_max··on Open source AI must win
Agreed. The only "issue" is that commercial products will always be ahead, with less friction for most users. This ultimately results in most people using these over open-weight variants. Users might not even be aware that the open-model variants exist. Similar to Windows / MacOS and Linux.
george_max··on Open source AI must win
I do believe that if OpenAI and others release an open-weight model that is better or on par with their frontier variants, it might ruin their primary business model.

That is, of course, unless they develop their own hardware specifically to run this open model. But, that does ruin the point of open models.

george_max··on Open source AI must win
With open-weight AI, there might not be an incentive to put large sums of capital towards training / research. There might be a donation fund of some sorts, but it certainly won't reach the level of fundraising that the frontier labs are receiving.

Because of this, I think it might not be possible to have AI *only* open-weight; major players like OpenAI, Anthropic, Google will likely stay for good, with better models than open-source versions.

I think it might look something like Photoshop & GIMP, with Photoshop being a frontier lab, and GIMP being the open-weight model. GIMP is decent for many different image editing workflows, but Photoshop is just better.

I would definitely prefer to have an open-weight model better than frontier labs'. Though I don't think it's possible.

george_max··on Statement on US government directive to suspend access to Fable 5 and Mythos 5
> Warns users about how dangerous and powerful Mythos Preview is

> Restricts model to large corporations

> Release information about how Fable / Mythos 5 is stronger than Mythos Preview, give access to every user for a limited time via subscriptions

> Users jailbreak model

> U.S. suspends Fable / Mythos use

Who didn't see this coming?

I wonder what this means for the future of AI models. Either we'll see worse guardrails than what was there for Fable 5 (for me, it was a unusable at times), or the models just stop getting better from here.

I think it's that the guardrails will be more strict, which is unfortunately not good news.

george_max··on LLMs are eroding my software engineering career and I don't know what to do
I see many comments saying, "AI can't do X with 80-100% accuracy; therefore our professions are in good hands."

While I don't want to sound overly pessimistic, the models are improving at a rapid rate. If asked ~3 years ago where the state of the models are today, it would sound like sci-fi if answered, "the models are creating full MVP apps in ~30 minutes with one prompt".

The hurdles the models are facing now, like reducing hallucination rates, ensuring compliance, and keeping a clean codebase, do not seem far away from being resolved IMO. Fetching specific information is already partially done with various MCP servers / RAG.

I am, of course, a bit worried about the future of software engineers. If these quirks are resolved, where do their professions fit in the industry? Delegating tasks to the AI model? Unfortunately, this does not require years of expertise, which is a double-edged sword. Reviewing AI's output? Ask it to explain each line not understood.

I think we will see more waves of larger layoffs, similar to how human computers were replaced by digital computers. To some, doing complex mathematical calculations mentally is a fun task / challenge, but it is ultimately significantly slower and more error-prone than calculating with a computer. In the same way, I think hand-crafting code will be seen as a fun "challenge" and AI will be seen as the "modern-day calculator".

george_max··on Artificial intelligence is not conscious – Ted Chiang
But is the model aware of the training? Unless you hook the model up to an MCP server, or something similar, and have it analyze the RL changes, it will not know if it has changed or not. Even if it is real-time RL, it is not aware of the previous state.
george_max··on Pwnd Blaster: Hacking your PC using your speaker without ever touching it
Not sure what would count as a vulnerability if this does not.
george_max··on MacBook Neo is so popular that Apple doubled production
Framework recently released their Laptop 13 Pro; their goal is to match Macbook quality while being fully open. The key issue is the price.

I definitely agree with your statement -- it's unlikely the price to performance and quality ratio on Macbooks are going to be outperformed soon.

george_max··on ESP32-S31
It's good news Espressif is using RISC-V over something like Xtensa; it's much more open and flexible. Excited to play around with this
george_max··on Axios compromised on NPM – Malicious versions drop remote access trojan
With all the recent supply chain attacks, I'm starting to think it's only a matter of time before all of us are victims. I think this is a sign to manually check all package diffs or postinstall scripts.
george_max··on PCB devboard the size of a USB-C plug
Very nice. I am wondering -- why have a devboard this small?
george_max··on CasNum
Haha, cool project and README!