HNHacker News
TopNewBestAskShowJobs

antupis

1,048 karma · joined December 29, 2014

submissionscomments
antupis··on Cloudflare K2: serverless event streams
Yup, and if someone is looking for a non-AI startup idea, take a use case with lots of data and a read-heavy workload, and do pretty much what Turbopuffer did for text and vector data. That seems like a very good way to go.
antupis··on AI companies in race to demonstrate their model most threatening to humanity
But issues is that for the valuations you need models that work without human expert and Opus 5.5 isn’t there even easy stuff like programming.
antupis··on Claude Opus 5.5 Intelligence, Performance and Price Analysis (Max)
Pretty much this, I don't just see how ASI could happen with current LLMs without big theoretical breakthrough so that's why 'phasing the frontier' and 'distillation attacks'.
antupis··on GPT-6 Sol and Luna
Flash thinks much more so it’s pretty much line with Sol for performance. That said I like flash coding style much more than OpenAi models.
antupis··on AI coding has made CI a bottleneck, so we reworked ours to keep up
I think it's more about money and scaling, bus factor is pretty big if run very lean organization eg whatsapp 2014 and that point all those developers are kinda your cofounders and probably start asking much bigger piece of pie. With small teams you kinda trade scaling and availability to velocity. It's much easier to ship but running oncall 24/7 with small team is just nightmare.
antupis··on Qwen 3.8 Omni Flash
I think we are starting be on that territory that regular software development is suffering, current models are great for benchmarks and one-shots but in daily development models are too eager and try to force patterns like excessive tests in every turn.
antupis··on Why I'm still bearish on LLMs after Navier-Stokes
I think automation is coming but it will be way more gnarly than frontier labs want public to believe. Value is just too big, when you can automate most of eg customer support it will create huge savings and same time customer satisfaction will get better.
antupis··on Pion, an agent designed to run any company autonomously
we aren't anywhere near lights off software factories, and writing software is easiest domain for modern llms, you can verify results rather easily, plenty of training data etc.
antupis··on Everyone should slow down AI development except for me
I think it’s more that pushing frontier is extremely costly and there is no free lunches in same way as 2024.
antupis··on RTK reports token savings, but our cost benchmarks disagree
My main issue with rtk is that rtk randomly messes modification and agent start polling same tool continuously.
antupis··on Google will buy half the electricity from one of Finland's nuclear power plants
Sweden main issue is that there is no transmission capacity between North and South, like now prices are 2x at south.
antupis··on Large language models develop novel social biases through adaptive exploration
Yup and this propagates those clumsy if this_new_code_branch: actual_code_that_matters else: old_legacy_code_that_should_not_be_there
antupis··on Mistral raises €3B
Yup, new SOTA models especially with high/xhigh/max reasoning too often overengineer solutions, good for benchmarks that usually measure task completion, bad for normal development where you don't want 'rewrite in rust and 1k LOC unit tests' style solutions when agent does mundane bug fixes.
antupis··on Mistral raises €3B
Mistral just needs to good enough category think all those flash models or Qwen3.8 27b which they sadly aren't at the moment, that plus being European lab will mean that they will have very nice business. Even now these SOTA models feel too overkill for most tasks.
antupis··on OpenAI’s head of ethics leaves less than a year after joining
Big corporations there is politics on play and too often you see CABs and other bureaucratic stuff instead of checkboxes for gated releases, 2FA etc.
antupis··on LFM2.5 2.6B model competitive with 4x larger models
I have noticed that these cheaper and faster models are very great for Ops-work. Luna max is beast when you use some stronger model to write detailed instructions/run book what to do and when to stop.
antupis··on Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs
But it was Google fault that they really cannot productize all that innovation that happened DeepMind. Google probably now fumbled with world models and this exodus will create next giant in that space.
antupis··on A $500 RL fine-tune of a 9B open model beat frontier models on catalog review
Speed play also you can get much faster responses with 9b model.
antupis··on Renting a sewing machine from the library
There is lots of English/Swedish books in average Finnish library.
antupis··on Claude: Elevated errors across many models [resolved]
It might be harness or prompting style. Personally I use opencode and my prompting style is very plain and terse . Where tasks are very small. Opus and Sonet too often are too verbose and go tangent. Where GPT5.5 is much stricter.
antupis··on Claude: Elevated errors across many models [resolved]
Personally I prefer GPT 5.5 writing style over Opus 4.8. It’s much more no nonsense and information denser.
antupis··on AI coding at home without going broke
I think hard part is that outside it takes 1-3 months to see if it’s race car. Especially in begin both things look pretty same.
antupis··on Open source AI must win
Every machine nowadays runs Linux in some form and Postgres is the default database.
antupis··on Microsoft and OpenAI end their exclusive and revenue-sharing deal
Thing is that distillation is so easy that it would also need large scale regulatory capture to keep smaller competitors out.
antupis··on An AI Vibe Coding Horror Story
Generally why build your own CRM? ERP and other resource planning systems I get becouse you can tailor made those to your back office. But for CRM you need mostly reliability.
antupis··on Marc Andreessen is wrong about introspection
I think it is more that some people just can’t do introspection, it might even be that they don’t have inner monologue.
antupis··on Microgpt
Humans need way less data. Just compare Waymo to average 16 year-old with car.
antupis··on Lessons you will learn living in a snowy place
Difference is pretty big if it’s icy like breaking 100 meters vs 10 meters. Especially if there’s wildlife like reindeers/moose’s you are going to do emergency breathing semi regularly.
antupis··on AI doesn’t reduce work, it intensifies it
It’s more about operational resilience and serving customers than product development. If you run early WhatsApp like organisation just 1 person leaving can create awful problems. Same for serving customers especially big clients need all kinds of reports and resources that skeleton organisation can not provide.
antupis··on Karpathy on Programming: “I've never felt this much behind”
Or if you like learning new stuff. Personally that has been best part of being programmer.
Page 1 of 19Next →