HNHacker News
TopNewBestAskShowJobs

chis

1,783 karma · joined November 13, 2016

submissionscomments
chis··on Secondhand book sales are booming. Is it because of AI?
So reading more I guess the judge narrowly ruled that destructive scanning is legal because of the one to one replacement, but didn’t rule on whether or not non-destructive scanning would be legal. It’s still not tested if scanning in that way would count as transformative fair use.
chis··on Secondhand book sales are booming. Is it because of AI?
Whatever judge ruled that it was legal to scan and use books if and only if you DESTROY a copy of the book was truly a moron. The law gets bent by judges all the time to be practical and fair, they should have used this opportunity to give a more sane ruling.
chis··on Why does Opus 5 feel worse to work with?
I think Fable is the beginning of Anthropic switching to training models as agent-first, tool second. It’s certainly the best model if you want something to work autonomously without supervision and don’t care to read the code. The code and writing is ugly but it can complete huge tasks and fix its own work.
chis··on Why does Opus 5 feel worse to work with?
Have you tried Sol? Just curious. I don’t wanna sound like a shill, just feels like every generation it’s important to reevaluate models and pick the best again.

I do find myself returning to 4.6 for casual conversation - asking it to help explain some science/engineering or news to me.

chis··on Why does Opus 5 feel worse to work with?
Anthropic is lucky that they've built a lot of loyalty over the last year that they can burn through right now. I see people talking about switching back to Opus 4.8 rather that using 5.6 Sol, which is wild.

My current approach is to occasionally use Fable for high-intelligence tasks but use Sol as the translator and clean-upper afterwards, and otherwise just use Sol for everything. Fable sometimes says the most insane shit, both unreadable and just completely missing the point, and refuses to back down when questioned. It's mentally exhausting to work with and I can't trust it.

chis··on Hello, me. It's been a while
It’s very difficult to intentionally find time to be bored. It’s sort of the opposite of an activity, and the brain hates it. One passable option is to take a 20 minute walk with no stimulus beyond the scenery. Journaling regularly. Meditation provides a similar experience but feels like a different lane.

For me I do find a similar experience to the OP. It’s easy to feel yourself going a little crazy getting sucked along the current of modern life. Flowing from one activity to the next, looking at one’s phone in the between, never any chance for an extended thought except when laying down for bed. If you don’t break out of this mode it’s easy to let weeks go by without noticing or taking much conscious impulse to shape your day.

chis··on Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index
You really have to be able to hold both opinions at once about Elon. Yes, Tesla and Spacex are unbelievable companies and perhaps nobody else on earth could have made them happen - so he’s a genius. But also he repeatedly promised autopilot, Tesla Semis, the space data center idiocy, DOGE, one million robotaxis by 2020 - so he’s a shuckster.

My current headcanon is that he remains an incredible leader and make-things-happen-er, but also that you should just assume all his announcements are complete bs and ignore them.

chis··on Managing AI Coding Costs at Scale
It’s funny how different everyone’s experience is with this stuff. To me the diminishing returns are more around not going crazy with prototyping or running with xmax thinking all the time. I haven’t found it hard to stay under the usage limit of one $200/mo Claude and one $200/mo Codex subscription.

If my company told me yeah we’ve decided you don’t get Fable or Opus 5 because it’s too pricey, you gotta use GLM whatever, I’d be displeased.

chis··on Where .env Went Wrong
https://www.pangram.com/blog/third-party-pangram-evals

You can also just try it yourself I guess really what convinced me was how it perfectly agrees with my own judgement.

chis··on Where .env Went Wrong
This has been pretty well studied. Pangram is about as good at an expert human with a 98% detection rate and <2% false positive, for flagging AI generated text.

https://arxiv.org/pdf/2501.15654

chis··on Where .env Went Wrong
92% of this text is detected as AI.

It may be time for Hackernews to integrate a Pangram detector into the UI, similar to what substack is doing :)

chis··on U.S. used 'virtually all' of its long-range precision missiles during Iran war
It might actually be for the best to stress the system in this way. It's not practical to keep a military based only around "stockpiles" that you're scared to use up. Modern combat looks like long wars of attrition where two countries compete to see who can burn through more missiles so you need to keep the actual supply chain active.

Hopefully they're able to learn lessons from Ukraine as well and pivot towards cheaper missiles and autonomous vehicles

chis··on Be skeptical of OpenAI's rogue hacker agent story
> AI managed to escape using standard and well documented script kiddie methods.

I think truly we don't know enough to say this. OpenAI says their AI found a 0-day exploit in some proxy software they were using but don't give a ton of details. On the Huggingface end we know a little more, they say the AI spun up tons of sandboxes and tested different exploits until it found one that worked.

chis··on DARPA, U.S. Air Force fly AI-controlled F-16
Absolutely terrifying. I'm sure they will soon find that AI pilots defeat human ones in dogfights 100% of the time, since they have faster reaction times, can tolerate higher G forces, and can be perfectly RL trained in the flight sims we already have.

It's really incredible to me that people aren't protesting in the streets over this stuff. Just in the last month we have * AI that goes rogue and hacks billion-dollar companies * AI solving long-standing math problems without any meaningful human help * AI controlled autonomous drones in Ukraine and F-16s in America * AI now represents over 50% of GDP growth in America and yearly capex will soon exceed the size of the $1 trillion US military budget

Is there any line where the public will become deeply concerned? I mean it really all reads like a sci-fi plot with a bad ending at the moment.

chis··on German AI consortium releases Soofi S, an open 30B model that tops benchmarks
It's funny, my reaction was the exact opposite. Details like this show that they're fundamentally unserious and focused on the wrong things. Imagine if Germany, when developing their automotive industry, spent all their time focusing on reusing the waste heat from production to heat homes instead of just building great cars. They probably would not have sold many cars!
chis··on 1,300 Beautiful Wildlife Illustrations from the 19th Century Now Restored
https://www.c82.net/blog/making-of-naturalists-library

You are correct! Apologies for not doing enough reading myself.

chis··on 1,300 Beautiful Wildlife Illustrations from the 19th Century Now Restored
I admire the effort but it's hard to get excited about looking at some ancient illustrations which have been partly filled in by AI. I want to see the work of an actual specific artist, and be briefly transported to the times they lived in.

I have this old book of the Audobon bird illustrations and those are truly incredible. Back in the day there was a public audience for high quality, expensive art prints in books and they spared no expense.

chis··on Why skilled workers come to Germany and then leave again
It is an interesting divide. "German" is both an ethnicity and a citizenship, and it's possible to become one but not the other. "American" on the other hand is purely a citizenship, and so it is possible to become an American after immigrating.
chis··on The Nationwide Backlash Against Cameras Watching Your Car
I have high hopes that America will be one of the few countries on earth to resist the tendency towards a surveillance state. Just because it’s such an individualistic anti-government culture in many parts.

There are so many reasons why adding cameras helps with policing, safety, public order. But it has to be resisted on principle because the government can’t always be trusted and rules aren’t always right.

chis··on America Is Headed Toward the Infinite Workweek
AI is automating all the easier tasks in people’s jobs, leaving them to spend 8 hours a day on the hardest problems which AIs cannot yet solve.

Software engineers are probably already familiar with the feeling of burnout from thinking too hard. The reality is very few people can work on the hardest problems they’re capable of for 8 hours a day.

Writing routine Python code for some system you know well is not that mentally taxing. Managing an agent that rapidly finishes tasks but needs careful review and big-picture planning is much more exhausting, and has higher returns on intelligence and deep careful thought.

I think this points towards the opposite conclusion of the OP. It’s not realistic to expect 8 hours of hard work out of a knowledge worker. Remote work naturally allows this transition, as employees can work a bit less but still overachieve with AI.

(I hate AI. Just observing the world we live in)

chis··on Claude Fable 5
Hackernews not blindly hate on AI challenge: impossible
chis··on Expanding Project Glasswing
But GPT-5.5-Cyber is also not released publicly?
chis··on Microsoft builds MacBook Pro rival with NVIDIA-powered Surface Laptop Ultra
This can’t run any models that cost $25/mtok lol. I think the fastest model it’ll reasonably run will be GPT-OSS 120B which costs $.05/mtok.

This is a laptop for CUDA devs and AI larpers.

chis··on Microsoft builds MacBook Pro rival with NVIDIA-powered Surface Laptop Ultra
This won’t be able to run any of the cutting edge models. And the models it can run can be served from cloud providers for very cheap - like <$1 per million tokens for the latest deepseek.

It’d take many years to break even on your $6000 investment, meanwhile better and better models will come out that the DGX can’t run.

chis··on Microsoft builds MacBook Pro rival with NVIDIA-powered Surface Laptop Ultra
Speculation is it’ll cost at least $6k for 128GB
chis··on Microsoft builds MacBook Pro rival with NVIDIA-powered Surface Laptop Ultra
It's also shockingly twitter-nerd-coded. "The cure for token anxiety", it advertises. To be honest it's hard to see why anyone would buy this product so maybe they decided to take a wild swing with the marketing. The only use is people who really, really, want to run models locally vs getting a much cheaper and higher performance result from a cloud host.
chis··on I made my phone slow on purpose
I think this is a great idea. Wouldn't have guessed this would be possible so I looked into how it'd actually be implemented.

I guess this is done on the device as a VPN via Apple's NetworkExtension config. But instead of a normal VPN where traffic goes through a server, the app just locally applies rules based on the app the packet came from and then routes them normally to their destination.

chis··on Claude Opus 4.8
I think it's probably too soon to say. I certainly still feel that large coding tasks are getting better and better with each model. I'd guess lawyers, doctors, etc feel similarly.

It feels like the only way to push the limits of newer models is with really long context questions that require reasoning. Any short request will naturally just be within the distribution of all the recent models so there isn't a performance difference there.

I think the near future is looking like a bunch of business-critical tasks that scale infinitely with better reasoning, all being done on whatever the most advanced model is at a high cost. Trading stocks, running a business, looking for tax dodges, writing high-performance code. These are all things where there's a tangible return on each jump in reasoning.

chis··on Cursor Introduces Composer 2.5
AI slop detected, you're under arrest
chis··on A recent experience with ChatGPT 5.5 Pro
I feel like this is one of the most advantaged times in history in terms of regular citizens having access to cutting edge tools.

Looking online it seems like the low end estimate might be $30k a year for such math researchers? And ChatGPT pro or whatever you want will run $100 a month, and should be coverable by grants. I’m quite sure matlab alone cost more in the past

← PreviousPage 2 of 11Next →