HNHacker News
TopNewBestAskShowJobs

momojo

453 karma · joined January 4, 2022

submissionscomments
momojo··on 5x faster Edge Functions: V8 isolates to Firecracker MicroVMs
1. What's the node + aws story? Did you stick with aws or stay but try a different serverless tech? 2. Is lambda/serverless overhyped? I'm worried about reaching for it too soon.
momojo··on OpenAI is well positioned to fast-follow Jev
I'd also add that they're hoping Jevon's Paradox also leads to a whole new segment of users who would have never reached for a classifier in the first place, given the barrier to entry.
momojo··on Kev: Tiny Jev-like family of decision models built on top of Qwen3.5
> "[I don't understand what it's actually doing under the hood, or why we should be excited.] Jev feels more and more like a glorified if/else block"

You should be excited because there's so many low-hanging classification problems (e.g. "Is this email spam?", "Is this yelp review happy or sad or neutral?") that a "cheap" LLM is overkill for. It really should be this cheap, and now Jev is the first to do it well and do it at scale.

If I was explaining it to my mom, I'd say "Classifying 1000 yelp reviews used to cost $50. Now it costs 5¢, at similar accuracy"

momojo··on Grim Fandango Puzzle Document (1996) [pdf]
This is one of those masterpieces that I'm surprised hasn't been scooped up by Netflix and turned into a movie.

Its not that I don't like it as a game, but it's just so chock full of character.

The premise alone is so compelling; in the Land of the Dead, even dead men chase money. Even with the literal afterlife a train ticket away, you have conmen and good guys living (or dying?) like there's no end of the line.

Love it. One of those rare pieces of art that continues to live in my head rent-free.

momojo··on A misalignment of AI in mathematics
I don't think this is a perfect analogy, but i think it's closest to the original articles declaration.

I wish they had stayed this as the very first paragraph.

Maybe I'm too stupid to grok the original declaration. Does anyone have a better summation?

momojo··on Desert Ant Labs: local, fast models that run on device
Cellpose[0] and Stardist[1] are the two you'll see most heavily run and talked about on the forums[2]. They're classic CNN's, but it just feels like they (and any modern models) are swept up in "Go out and buy a 4090.But these guys will give you plenty of mileage before you have to reach for the bigger ones.

[0] https://github.com/mouseland/cellpose

[1] https://github.com/stardist/stardist

[2] https://forum.image.sc/

momojo··on Desert Ant Labs: local, fast models that run on device
> The world ships more than a billion capable phones, tablets, and laptops a year, most with a chip built for exactly this work, paid for and idle most of the day. Run the model there and the economics flip: no per-call cost, no round-trip, and nothing leaves the device.

This. I run small models (>50MB) for bio-imaging/biotech applications, it feels like every README implies that you need a discrete GPU to get started. While some do, many, especially the most useful ones, do not. Sure it matters if you're also going to do fine-tuning, but I believe your typical user just wants to detect some nuclei and get some cell-body ratios.

The laptop on your desk won't be running Meta's SAM, but it has more than enough compute to crunch 100's of your H&E slides overnight.

momojo··on Mercury 2.5
Anyone here use Mercury 2.0? Curious what your experience with the model is.
momojo··on Gemini 3.8 Flash and 3.8 Flash Cyber
Same. Love oneshotting or sanity checks. Which fortunately is a lot of my workflow (lot of long tail stuff fits in one prompt).
momojo··on Malleable software = solid bases and custom code
Love this take. In the bio-imaging space, Napari is a great example of this. Wonderfully solid base, but extremely extensible since its just python all the way down.

There's something wonderful about having my coworker walk up with an issue, and being able to bang out a Napari plugin that solves their exact problem before lunch.

My order of 'tool escalation' usually goes: - Can I solve their problem from napari's inline terminal? - Can I solve it with a one-off script? - Can I solve it with a one-off script that creates a one-off plugin interface? - Should I add the plugin to our company-wide repo since this problem seems to occur a lot?

momojo··on FDA authorizes first wearable device that monitors ketone and blood sugar levels
Did you get the grant? Are you under NDA? What was the innovation?
momojo··on FDA authorizes first wearable device that monitors ketone and blood sugar levels
I'm in biotech and if I had a dollar for every engineer I've met who dreamed of or worked-at-a-place-that-tried or even tried-in-their-freetime to make a non-invasive glucometer, I'd have at least enough to buy a coffee, which is a lot.

There is an ever-growing bodycount in the NIO-GM graveyard [0], but I too hope that one day, it'll get figured out. My old roommate and good friend was T1 and monitoring one's glucose and remembering not to eat too much/little is half your life.

[0] https://pmc.ncbi.nlm.nih.gov/articles/PMC8655290/

momojo··on Show HN: A techno machine in one HTML file, with verifiable renders
Ditto on the lovelyness of single-file, zero-dep, standalones. In my day job I'm stripping out hundreds of megabytes from one of our main docker containers, but running into little projects like this that make me smile. Beautiful
momojo··on Numba in the Browser: Unlocking a New Scientific Python Stack in JupyterLite
Why not use Numba? In my mental model they achieve the same win, generating GIL-free, bytecode-level perf.

I'm not super familiar with JAX though.

momojo··on Numba in the Browser: Unlocking a New Scientific Python Stack in JupyterLite
Good question.

I'd stick with the venv if it's heavy duty crunching and you do it often.

However, if this is a one off or doesn't need heavy compute, and you don't mind waiting a little longer, use the notebook.link.

momojo··on Numba in the Browser: Unlocking a New Scientific Python Stack in JupyterLite
This is awesome! For those here not familiar with Numba, it helps bridge that performance vs ergonomics tradeoff that's always existed when you reach for python over a lower level but faster lang like C or C++.

Sure you could write Cython but then you have to have a build step and make wheels for every platform you and python version. Sure numpy has gotten faster over the years but you're still hampered by the GIL.

Numba is a little magic because you get to write stuff that feels like numpy, but get literal bytecode perf.

BUT there's a cost to this, which u learned the hard way when I imported a color map extension for matplotlib recently.

I thought i was going to be importing a couple megabytes at most. But Numba+llvmlite alone is almost 100MB!

This might be a drop in the bucket in some applications but for a color map library that has only two hot paths that need to be JITed, it's excessive.

Overall though, love this achievement, and i love what's being done for in-browser (aka local-first) scientific computing!

momojo··on The mathematical beauty of hyperbezier curves
I love bezier curves. In community-college, it was the first time I ever encountered a subject that made me want to go do more research on my own. One of my core memories is toiling away for multiple nights when the rest of the house was asleep on my 2015 Macbook Pro, writing janky p5.js code and pressing refresh on the browser page over and over until suddenly, I started see real, beautiful curves blossom from my control points.

I just went back to dig up some old sources[0], and I can't believe this post is almost a decade old now. This guy's explainers and animations were leagues beyond any other resource I could find through Google searc at the time.

[0] https://jamie-wong.com/post/bezier-curves/

momojo··on Why does Opus 5 feel worse to work with?
Reminds me of this: https://www.geoffreylitt.com/2025/07/27/enough-ai-copilots-w...

I think this is such a great reframing. It makes so much sense; I need an AI that acts more as a HUD and gives me superpowers, not just a copilot that can tell me when I've misspelled a word.

momojo··on NP-overrated
Do you have any examples of the second class?
momojo··on What Happened to HackerOne?
The brokenness of man? Sin?
momojo··on DeepSeek V4 Flash 0731
Did anyone else experience a change in verbosity? I've been playing around with an agent that holds your hand in a Jupiter notebook and it felt like it started writing essays versus nice, concise, helpful paragraphs like before. My gut was correct because I checked my Deepinfra usage and it was almost a 2x out-token usage for every in-token.

Not a huge deal since it's still cents per session, but my bigger issue was the weird change in tone. It became a lot more pretentious and over-explanatory.

Heavy prompt reworking helped but maybe that's just the cost of being better at coding and ARC-AGI?

momojo··on Kitesurf: Agent-first browser that runs in V8 isolates
I'm intrigued! There's so much movement in the sandbox space. Can someone tell me how a V8 sandbox compares to say Fireworks or if they've chosen one over the other? I think, from a technical standpoint, that it's neat that we already have an entire sandbox in the browser (albeit with a bit of chrome) but someone tell me why it shouldn't be used that way
momojo··on AMD acquires Taalas to boost inference performance by etching models in silicon
I'm sure life would find a way. I'd love to see what kind of power-harnesses people have to come up with to steer 16k tps QPU's (Qwen Processing Units) productively.
momojo··on AMD acquires Taalas to boost inference performance by etching models in silicon
I don't have a great answer but you pose a great question.

Obviously a CTO is not going to walk away from the technology just because it's not good enough. That much more incentive for someone to create a powerful enough harness that can direct that power safely and productively. Like a nuclear core, we'll need to come up with the graphite rods and water tank. And if tokens are essentially free, why not, for every million tokens, spend 10x tokens on code review, testing, etc?

momojo··on AMD acquires Taalas to boost inference performance by etching models in silicon
My 2 cents to for your point:

- Deepseek V4 Flash is impressively capable. Sonnet still beats it out by a thin margin, but the real kicker is that a typical session with Sonnet at current API costs is ~$2. The same session with Deepseek is 2 cents (ha). Its even allowed me to consider offering free-with-limits API usage on my own app. - Taalas (or competitors) have a lot going for them. If anything I feel like they need to join hands with these smaller model makers and converge in 2028

momojo··on AMD acquires Taalas to boost inference performance by etching models in silicon
There's certainly incentive to do so. And its only an engineering problem haha.
momojo··on Show HN: Maple-Preview – Ternary 20B MoE running at 120 tok/s on a iPhone
At this point I think Apple just needs to simply not do anything stupid and these small model makers are going to hand them models
momojo··on DeepSeek V4 Flash 0731 Intelligence, Performance and Price Analysis
I use DeepInfra. Their whole catalogue is openai compat. US based.
momojo··on Show HN: Bento - An entire PowerPoint in one HTML file (edit+view+data+collab)
Nice. Did it end up working out?
momojo··on Show HN: Bento - An entire PowerPoint in one HTML file (edit+view+data+collab)
Did you chew on monetization at all? I'm building something similar but for bio-imaging scientists, where I've packed everything I can into a single HTML (and even some modern segmentation models like Cellpose3 and StarDist that fit under Cloudflare's 25MB free tier limit haha.

Maybe I'm thinking too far ahead but I'm hoping maybe I can cover the (future?) compute costs by charging for the heavy duty ML that would have to be run non-locally but IDK. This isn't critical yet but would love to hear your musings.

Page 1 of 6Next →