HNHacker News
TopNewBestAskShowJobs

sheeshkebab

1,384 karma · joined January 27, 2016

submissionscomments
sheeshkebab··on GPT-6 Astra on robot arms
so we finally got our "phd level ai" maybe, semi consistently move blocks around at the level of a 1 year old child.

good fucking job everyone, congrats.

sheeshkebab··on Memory prices climb 500% in 12 months
writing native gui code is atrocious, add to that llms being horrible with ui dev and you end up with with what you describe. but at least if we get more non-ui stuff written in more efficient languages there is going to be some benefit.
sheeshkebab··on Memory prices climb 500% in 12 months
llms make it trivial writing code in rust, without even knowing rust. There is hope…
sheeshkebab··on Gloomberb
It’s pretty bad except that any other way is even worse.
sheeshkebab··on DeepSeek V4 Pro 0813
This. The same goes for “skills”, skill type “subagents” and other bullshit - powerful models don’t need any of that anymore I noticed.
sheeshkebab··on Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs
The idea that agi would help us cure all the human deceases, esp house built into our generics or developed over millions of years gotta be the top indicator of our own stupidity.
sheeshkebab··on Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs
What? psychologists are for that (or drugs and alcohol).
sheeshkebab··on Changes at Google DeepMind: Demis Hassabis from CEO to Chair, Jeff Dean departs
What’s with fascination with agi/superintelligence? It seems people working on it never had kids and just want to compensate for that. It’s a really horrible thing to try to “solve”.
sheeshkebab··on US Treasury undertakes historic intervention in yen market
Treasuries are $ denominated -> sell them for $$, buy yen back, yen/usd drops. or have Uncle Sam buy yen.
sheeshkebab··on AI's top startups are barely publishing their research
good.
sheeshkebab··on Kimi K3-256k
Try doing a refactoring of some sort or larger new feature using just an agent on a moderately sized codebase, 256k will be compacting every few minutes, and result will be unusable.
sheeshkebab··on If digital computers are conscious, they are conscious at the hardware level
My cat is more conscious than whatever latest frontier model is out there, and it runs on a little bit kibble and water.
sheeshkebab··on Startup founders urge Trump not to shut off Chinese open weight AI
Barrels of tokens sold by thieves that copied every book sure looks like selling to me.
sheeshkebab··on Startup founders urge U.S. government not to shut off Chinese open weight AI
Someone going to public library, copying every book out of it and then selling them as their own are thieves.
sheeshkebab··on Zuckerberg says AI agent development going slower than expected
there are plenty of orgs where writing more code is not a good thing, in fact it's the last thing they want. but yet these orgs would still employ 80%+ of all developers, not necessarily to write code though.
sheeshkebab··on Asian AI startups launch Mythos-like models
unless they launched 10t param models, or figured out some amazing new way to compress as many params into say 100b, I doubt it's anywhere near "mythos level". and I have no idea how many params mythos has but that was just some hear say.
sheeshkebab··on Only 16 Percent of Americans Think AI Will Have a Positive Impact on Society
it’s not been useful anywhere else - no self driving cars, no laundry helpers - just some idiotic animatronic carcasses that are barely able to walk around, and semi autonomous killer bots on ukraine frontlines.
sheeshkebab··on Open source AI must win
Qwen models are actually very competitive with frontier models, and you can run them on your local computer. Gotta have a decent graphics card and by that time the current cost of the rig may not justify it over paying $100/month for cloud model but it’s all out there.
sheeshkebab··on Statement on US government directive to suspend access to Fable 5 and Mythos 5
Well, it was great while it lasted - I had fable build me a bunch of stuff this week that opus was just screwing up too much and could never finish. Good thing there are plenty of choices now even if US gov fucks up US AI.
sheeshkebab··on AWS Bedrock to require sharing data with Anthropic for Mythos and future models
Pro/max subs are not as flexible as bedrock in api use and don’t seem to run the same models either - often times they are notably dumber (quantized I guess) than bedrock equivalent.
sheeshkebab··on Claude Fable 5
I’ll ask it to write me some win32 ui crap when I get hands on it, it will need all its brainpower to get that idiocy right.
sheeshkebab··on MiMo-v2.5-Pro-UltraSpeed: 1T model with 1000 tokens per second
Opus regularly bitches and wines to me how long something will take and that I should think before asking it to do it. But then it does it anyway in 15 minutes.
sheeshkebab··on How LLMs work
considering they work with any architecture/configuration given enough compute, just more or less efficiently - then maybe it's fundamental, in the same sense as why electricity works...
sheeshkebab··on MAI-Code-1-Flash
I like gpt oss - great model even if not too smart.. runs on my laptop at over 100ts has a certain tone that I like over all these qwens stuck up their asses.
sheeshkebab··on Liquid AI reveals 8B-A1B MoE trained on 38T
27b is slow as molasses vs 35b on local stuff I have (m5 max). Mtp doesn’t make any difference either.
sheeshkebab··on Constraint Decay: The Fragility of LLM Agents in Back End Code Generation
…but they reason well enough given enough context (using their matmuls).
sheeshkebab··on Electrobun 2.0 will be decoupled from Bun due to the Rust rewrite
One of these days you’ll learn about “enterprise” code
sheeshkebab··on Electrobun 2.0 will be decoupled from Bun due to the Rust rewrite
Mist of human written codebases are unusable for llm dev by that definition.
sheeshkebab··on Gemini 3.5 Flash
It’s great laptop to mess around with llms, it won’t replace claude opus or even sonnet.
sheeshkebab··on Gemini 3.5 Flash
You can only run heavily quantized models on all 3/4/5 rtx gpus (with 32gb or less vram) - and you probably want moe versions like Qwen 35b for this to run at speed somewhat comparable to Claude. It’s still not there to be honest but getting there. Personally I mess around with llama.cpp on m5 max with 128gb - it’s a decent setup to try various medium sized things, and runs llms surprisingly well without quantization, at least the moe models.
Page 1 of 19Next →