HNHacker News
TopNewBestAskShowJobs

itsmeduncan

52 karma · joined February 24, 2010

Founder & CEO Whelk: Private AI · Your models. Your agents. Your network.

work: getwhelk.com blog: itsmeduncan.com code: itsmeduncan

submissionscomments
itsmeduncan··on The Economics of Open-Weight Inference
The finding that flips the usual story is the sparse-MoE reversal: on gpt-oss-120b the A100 comes in cheaper per output token than the H100, because the workload is latency-tolerant and the MXFP4 experts fit in 80GB. Chip age stops predicting cost once you decouple "newest silicon" from "cheapest inference" and let price-elastic work route to whatever's efficient.

The point about no salvage value: the paper is really pricing a rental service, not the devices, and even there the 5-year A100 term holds 80% of the one-month mark. The reason there's no clean resale market yet isn't that old GPUs stop earning. It's that the workloads keeping them earning (batch, RL rollouts, long-horizon agents) are the ones that tolerate latency, and the people running those tend to rent rather than buy secondhand.

The part the paper understates: rented A100s aren't the biggest pool of already-paid-for, latency-tolerant capacity. That would be the Apple Silicon and flagship phones sitting idle, where the capex is fully sunk and there's no hourly meter at all. Same economics, one step further. The deciding factor is who gets to make the deployment decision, and open weights are what move that decision to the operator.

itsmeduncan··on Heretic removes restrictions from language models
The most compelling use in this thread isn't edgy content, it's the boring legitimate work hosted models refuse by default: reverse-engineering a camera you own, a PoC for a CVE on your own network, decoding a protocol to connect your own software to a printer you bought. A policy layer tuned for the median user turns into a wall for the person doing real security or repair work on their own hardware, and a local model with no refusal layer just does the job.
itsmeduncan··on macOS 27: Workaround to avoid downloading AI models and save storage
There's a distinction worth drawing here: "on-device" and "local" are not the same thing, and this thread is really about the gap between them. A model Apple pushes to your disk that you can't inspect, swap, or remove, that runs when the OS decides, is on-device in the literal sense and gives you none of the control that makes local AI worth wanting. Local, in the sense people actually mean it, is that you chose the weights, you chose when it runs, and you can delete it. Forcing a 14-29GB model you can't uninstall is the inverse of that. The storage complaint is real, but the deeper one is that "we put AI on your device" quietly became "we put our AI on your device."
itsmeduncan··on Hister: A private search engine for the pages you visit and the files you keep
[flagged]
itsmeduncan··on GitHub is having trouble counting things
My scheduled for 8am PT GH actions workflows don't run until anywhere between 6-8 hours later. I gave up, and run them locally now. It is just to create a tag, and push to a specific place. A super simple problem that was solved for so long simply with GH Actions that just doesn't work anymore.
itsmeduncan··on Giving up on smart rings
I was an early adopter of Oura, and have recently stopped wearing it. I've reverted to the Apple Watch & EightSleep for all of the tracking I could ever need. I found the experience with Oura to be overly optimistic numbers wise, as well as borderline dark patterns in the app to drive my anxiety around getting good scores.
itsmeduncan··on Why does Opus 5 feel worse to work with?
This is interesting take. How do you think the split will show up? Harnesses/models will be designed for human consumption, and those for machine consumption? I'm building getwhelk.com and it is a cool thought-experiment for me.
itsmeduncan··on Nvidia Nemotron 3.5 Lightning and NeMo Switchyard
Me too. I think there are a few waves we can ride here. Let's collaborate?
itsmeduncan··on Claude Code is leaking real email address as a User-Agent string in curl command
Agreed, and run your own inference too. Open source/weight models are getting so great.
itsmeduncan··on Improving GPT‑5.6 Sol in ChatGPT, expanding GPT‑5.6 Luna access for free users
How long can OpenAI (and Anthropic) incentivize, and keep costs artificially down for tokens to depress investment in private/offline LLMs? It's interesting to watch the cost curves, and burn to see where it stops being venture subsidizes, and the frontier open-weight models become as good.
itsmeduncan··on A Founder’s Guide to Understanding Investors
It looks like the article was deleted.
itsmeduncan··on Ask HN: What are good software architecture interview questions?
A great "how would you architect X" is Uber. Lot's of hidden complexity.
itsmeduncan··on AWS Snowmobile – Massive Exabyte-Scale Data Transfer Service
There is an example in the blog post about DigitalGlobe.

https://aws.amazon.com/blogs/aws/aws-snowmobile-move-exabyte...

itsmeduncan··on Music for Programming
Plus one to di.fm. They've been around forever and have great channels.
itsmeduncan··on Silicon Valley startups rein in spending and prepare for layoffs
Do you work for GCE? You've made this exact same comment twice on the same link. Do you have any examples or links that back up the 50% cheaper point? As well as the trade-offs for switching to GCE? We run on GCE, Heroku, and AWS.
itsmeduncan··on AWS Device Farm adds support for iOS
"Pricing is based on device minutes, which are determined by the duration of tests on each selected device. AWS Device Farm comes with a free trial of 250 device minutes. After that, customers are charged $0.17 per device minute. As your testing needs grow, you can opt for an unmetered testing plan, which allows unlimited testing for a flat monthly fee of $250 per device."

From the FAQ[1] which albeit was a bit buried.

edit: It is also at the bottom of the landing page now.

1: https://aws.amazon.com/device-farm/faq/

itsmeduncan··on The Abandoned Facebook Tech That Now Helps Power Apple
Apple uses Riak for this.
itsmeduncan··on PagerDuty raises $27.2M in Series B led by Bessemer
Congratulations to them! It's a pretty fantastic tool. We use it for incident management for our customer care department as well as engineering.

But why does the link go to the comments...?

itsmeduncan··on OS X Yosemite public beta starts tomorrow
I am running beta 4 (14A298i) on a mid-2013 11" MacBook Air. It sometimes loses the WiFi connection, but reconnects automatically. I've noticed the laptop will restart itself with PowerNap turned on, and plugged into an external display and external drive. Turning off PowerNap seems to fix it. Everything works pretty well. Handoff is a little funky, but you'll probably not be running iOS 8.
itsmeduncan··on ShopKeep’s Point Of Sale Software Rings Up $25 Million
Director of Engineering at ShopKeep here. Happy to answer any questions. We're also hiring (like everyone). Checkout our careers pages[1].

[1] http://www.shopkeep.com/careers

itsmeduncan··on Mobile-Payments Startup Square Discusses Possible Sale
Director of Engineer at ShopKeep here. Happy to answer any questions about the technical side of things. Some recent ShopKeep news: https://news.ycombinator.com/item?id=7641525
itsmeduncan··on T-Mobile Turns An Industry On Its Ear
I am on the 10th floor.
itsmeduncan··on T-Mobile Turns An Industry On Its Ear
I was an early adopter when T-Mobile rolled out their $50 unlimited everything plan in NYC. The pros outweigh the cons, but they do sometimes give me pause. I don't have service in 90% of my apartment in the middle of Manhattan. It is a newish(2001) building, but my fiancee has full service from Verizon. Outside of NYC the service is fairly spotty compared to bigger networks. T-Mobile with an unlocked 5S is priceless for travel. I just got back from a trip to Japan, and the Philippines. You don't need a SIM card. The phone joins the network it supports, and T-Mobile texts you your limits. 200mb of 3G data, unlimited text messaging, and cheap phone calls were included in my standard unlimited plan with no configuration. I pay $77.76 a month for unlimited everything, and 2GB of tethering data. I suggest it in a bigger city or if you travel internationally and need to stay connected. I give pause to people in the suburbs. I am sure it will get better though.
itsmeduncan··on Amazon Kinesis Available For All Customers
We just wrote something very similar to this using Kafka which replaced a very similar system you have. Do it. We would have used Kinesis if it was around.
itsmeduncan··on The inside story of Bombardier’s $4-billion gamble on a super quiet jet
That is correct. The short story here is the FAA simply can't keep up with the innovation so they certify a group of people that can certify a repair as safe, and effective. It can than be used, while the FAA eventually full approves the change. The DER assumes full responsibility for the repair, and it's safeness. There have been stressful Christmas eves when approving the duct taping of a loose part of a V2500 so a plane full of troops can make it home for Christmas. Don't worry, it was a single flight one direction approval (but it happens all the time)
itsmeduncan··on The inside story of Bombardier’s $4-billion gamble on a super quiet jet
I know this article is focused on noise, but the most important part about this engine is the efficiency. The fuel savings alone from having the high pressure section running at peak RPM in any situation is incredible. Billions and billions of dollars in fuel savings for a modern fleet.

Think of it as an orbital reduction gear more than a transmission that allows two different sections of the same axel to spin at differen speeds.

Checkout fuel burn vs. other engines. [1] The maintenance cost ratios across an entire fleet versus the fuel savings make it an absolute no brainer for regional fleets.

Source: My father a DER at MTU Aero Engines who developed the geared turbofan with Pratt & Whitney. DERs develop and sign off repairs in accordance with FAA regulations. He was also certified is EASA repairs.

1- http://en.m.wikipedia.org/wiki/Pratt_&_Whitney_PW1000G

Edit: Double the

itsmeduncan··on Hosted server status pages for startups
We are using this at ShopKeep[1]. The co-founder was nice enough to let us get it set up for free, and start paying for it later once our issues settled down.

[1]: http://status.shopkeep.com

itsmeduncan··on Puma, a fast concurrent web server for Ruby
We (ShopKeep) are using Puma for our thin web-services around our platform. Specifically, we send all of our data to a two nodes load-balanced running Puma for our analytics aggregation. It's pretty amazing how a single instance running Puma replaces an entire cluster of Unicorn workers.

p.s. We are also hiring Ruby, and iOS folks. Contact information is in my profile.

itsmeduncan··on Full MongoDB database dump of the Blippex search engine
This will be fun as a list of places to try out 0-day exploits on.
itsmeduncan··on Ask HN: Who is hiring? (April 2013)
ShopKeep (http://www.shopkeep.com)

Full time in NYC, or starting after May 1 in SF

We're looking for engineers to come work with us on our iPad point of sale system. iPad point of sale? That's just a cash register, right? Boring you say? Hellz no.

We're working on an API-centric web application and a native iOS app. We have a burgeoning data product and have started work on a payment gateway. We do front-end and back-end and web and mobile and data and security. It's a managed chaos of technologies. We work like horses, argue like lovers, and play like children.

Most of the effort is in Ruby on Rails, Javascript, and Objective-C. Someone with experience in JRuby would be an awesome addition.

Most of us are full stack, but we didn't start here that way. Right now, we are especially in need of senior Ruby on Rails engineers, great front end people, and another DevOps person. Find out more about us here[1]. Email alex [at] shopkeep dot com for more information.

[1] - http://www.shopkeep.com/about

Page 1 of 2Next →