HNHacker News
TopNewBestAskShowJobs

c0rruptbytes

451 karma · joined May 3, 2017

submissionscomments
c0rruptbytes··on Strands Harness
OMP has a lot of candy that raises token cost compared to vanilla pi
c0rruptbytes··on Claude Code reads AGENTS.md only when telemetry is on [fixed]
mods seem like a grasp at all the pi and dsh users
c0rruptbytes··on Markdown in /src
MDs in subdirectories is how Anthropic recommends it too. If you have a CLAUDE.md in a subdirectory and an agent starts working in there, it's appended
c0rruptbytes··on GPT-6 Sol and Luna
have you tried using lower efforts?
c0rruptbytes··on GPT-6 Sol and Luna
they're already on bedrock
c0rruptbytes··on Grok 4.7
as someone who is limited by amazon bedrock support at work (no idea why we got stuck with the worst one) - grok is literally the only budget-ish model option, so nice to see it updated, Sol and Opus are just too rich for my blood. Luna is good but so slow at getting things done (tps wise it's fast)
c0rruptbytes··on We must pace the frontier
> This argument doesn't work for countries which don't care what their citizens want, like China.

you didn’t have to use China as an example, the US clearly does not care what its citizens want as the most popular policies are never even discussed or proposed in congress

meanwhile, China destroying their housing market to decommidify it so everyone can have housing…they seem to care about their people more

c0rruptbytes··on iPhone Duo
the iphone mini unfortunately did not kill nor do most small iphones
c0rruptbytes··on DeepSeek launching v4.1 flash cheaper and more capable than v4 pro
are they releasing the weights too?
c0rruptbytes··on Nvidia to acquire Hugging Face
open models don’t imply local models… it just allows more people to run them, most likely with nvidia GPUs
c0rruptbytes··on Show HN: Running 104GB Qwen3.8-Flash-Next on 48GB Mac with at ~12 tok/s
so many inference project, omlx already supports all of this and has a 1000 people trying to optimize it constantly
c0rruptbytes··on My local model setup on an M4 Pro Mac Mini
m5 max really fixed pp with the better matmul support, im sure the m5 ultra will be even crazier

the sparks have much slower memory bandwidth is the trade off

c0rruptbytes··on OpenClaw 2.0, Accidentally
the rr suite seems much better for that
c0rruptbytes··on The Hugging Face incident and the road ahead
OpenAI measures their internal token usage in “rolexes” - it’s literally a flex to be a token burner

i can imagine insane amount of capital is wasted on these two companies compared to the efficiency elsewhere

c0rruptbytes··on Apple introduces M6 and M5 Ultra
it's for me
c0rruptbytes··on New Mac Studio with M5 Max and M5 Ultra
The 512GB could run GLM 5.3 which is Opus level
c0rruptbytes··on OpenAI: GPT 5.6 Sol price reduction (until at least Nov 21)
Sol is closer to Fable than Opus - I like SlopCodeBench the most as a test - https://github.com/humanlayer/advanced-context-engineering-f...

You add requirements and make previous tests invisible to see how pigeon brained the model is - Sol and Fable seem to rank the same as Opus tends to fall behind

c0rruptbytes··on Anthropic's best AI model struggles to attract users as cheaper tools thrive
the writing style is so easy to fix, output styles is documented in claude code and you can change it

still don’t think anthropic models are worth the money

c0rruptbytes··on NanoGPT Speedrun Frontier
auto research the new cool kid on the block - look at https://mlx.fast
c0rruptbytes··on A week of using Codex more than Claude
Pi by itself is more than capable, OMP is okay but you really don't need much for a great harness (these models are RL trained to hell to be a coding agent, sometimes less is more)

I run a lot of SlopCodeBench - https://github.com/michaelasper/benchmarks

Fable/Sol/GLM 5.3/Kimi are its league (in that order) Deepseek/Opus is solid Qwen 27B is the floor - there's no reason to use Sonnet/Terra/Haiku

For everyday activity - I don't think you need to be using Sol (xhigh) for everything - unless you're made of money - I've found using Luna from OpenAI to be more than enough - it'll outreach to Opus/Sol when it needs to

Haven't had access to Gemini 3.7 but we're getting it at work soon, will give it a go!

Codex CLI is pretty bare bones in a bad way (at least Pi is extensible). Claude code is vibeslopped to the extreme

c0rruptbytes··on The New MCP Roadmap
easier to gate MCP tools? you can allow/deny tools very easily
c0rruptbytes··on Ask HN: GitHub employees what's going on? Why?
need competent people to hire competent people first - it seems like their only directive is to raise $MSFT not make good tech

(talking about leadership, not my lovely msft engineers reading this)

c0rruptbytes··on GPT-5.6 Sol Pricing Cut by 50%
plenty of companies offer ZDR and are just as random as OpenAI and Anthropic in their age
c0rruptbytes··on MathCode, Mathematical Coding Agent
looks nice...time to turn it into a pi extension
c0rruptbytes··on Claude: System Prompts
Pi is excellent for this, its system prompt is tiny
c0rruptbytes··on Grok Bot
they probably should kill the grok branding...
c0rruptbytes··on H3-metal – Native MiniMax-H3 inference for Apple Silicon
wow antirez does not sleep
c0rruptbytes··on Docker Sandboxes – Disposable, isolated sandboxes for AI agents
Reminds me of sandboxy - https://github.com/apple/containerization/tree/main/examples...

Also if your thing doesn't work with `pi` out of the box, then low effort

c0rruptbytes··on Born Against, or why hobby programming communities are against LLM usage
i would call myself a tinkerer but still don’t have a necessity for 3)

it’s more akin to 3d printing to me, i get the design all setup and let the machine do 3) and then get to play with it in 4)

c0rruptbytes··on Ten advances in mathematics and theoretical computer science
i think i agree, we are going to have /more/ math and now need /more/ mathematicians (we are seeing https://vibemathed.com/)

these LLMs are great are generating arguments but they don't ask questions, we will need mathematicians to shepherd them into more discoveries

i really want to see open weight models crack some breakthroughs

Page 1 of 3Next →