HNHacker News
TopNewBestAskShowJobs

Escapade5160

161 karma · joined July 18, 2023

submissionscomments
Escapade5160··on Vomit: Clean up Claude 5's token output with a separate LLM
GPT-5.6 models in codex are great at not writing like this. Cheaper, huge usage limits, readable output, and quality code. It's the better system now.
Escapade5160··on GLM-5.3 Artificial Analysis Benchmarks
Sol is an underappreciated model. Dropped Claude today and went to codex. None of that god awful prose Claude used for me any longer.
Escapade5160··on Why does Opus 5 feel worse to work with?
It's horrendous. It constantly scope creeps, will attempt to use admin overrides, assume you are incompetent, speak in half thoughts, and just blatantly ignore instructions. Opus 4.6 was peak for Anthropic. Its a lot better outside of Claude Code but it's still annoying. I vibed out a CLI tool to do packet capture for TUI applications to get an idea of why Claude Code makes it worse. The amount of additional unnecessary context and tool bloat that goes with your sessions is crazy. The memory system ships so much extra info about what you did yesterday that I think it's misguiding the model.

Link to the app for those interested: https://github.com/citizen-123/cli-capture

Escapade5160··on uBlock Origin Is Giving Up the Fight to Keep Ads Off Facebook
Nah, we will find a way to scrape those conversations and distribute elsewhere.
Escapade5160··on Pixel Watch 5
Pebble is the ultimate smartwatch because it knows what it is. Its a watch first and foremost, readable in any lighting conditions, waterproof, battery lasts for weeks. They made a watch first then added only the QOL features that most people need.
Escapade5160··on Emergent Introspective Awareness in Large Language Models
If it's context dependent is it really introspection?
Escapade5160··on Everything you do is being recorded
As someone who regularly attempts to give up a cell phone, it is not optional. A not insignificant number of things require it. It is the expectation society is now built around and therefore it is not an option. A car without telemetry is for the time being still optional. I take very good care of my old vehicles hoping I never need a new one.
Escapade5160··on The Hacker's Renaissance (2025)
Human effort has hit an all time low.
Escapade5160··on Our position on open-weights models
Perhaps they saw it crystal clear.
Escapade5160··on Show HN: Abralo – Free, easy way to run several Claude Code agents in one window
Cmux.... Or Orca.
Escapade5160··on Every new car sold in the European Union must include a driver monitoring camera
Fuck this
Escapade5160··on GLM 5.2 and the coming AI margin collapse
Most large orgs do not need to train end users. They just need to add glm-5.2 to their router and their in house harness will pick it up. Then slowly limit usage on anthropic models and people will swap willingly. It's a simple /model command in every harness.
Escapade5160··on Claude Sonnet 5
At that price you should just use glm-5.2. You get an Opus class model for 1/3 the cost.
Escapade5160··on Micro-Agent: Beat Frontier Models with Collaboration Inside Model API
So be it.
Escapade5160··on Herdr: Agent multiplexer that lives in your terminal
I'm a power user and it's not an issue for me. I just use Kitty with multiple windows and panes. I can jump between using hotkeys with no problem.
Escapade5160··on HPV vaccine cuts risk of dying from cervical cancer before 30
Can we stop calling vaccines jabs?
Escapade5160··on I hacked into the worst e-bike and fixed it [video]
It's in YouTube's best interest to only show users content they're interested in. Replace the word algorithm with users and you'll have a more accurate representation of how YouTube actually works. The reason those videos didn't get the love they deserved is because they're niche content, the 10 minute review videos appeal to a wider audience and therefore gain more traction.
Escapade5160··on Grit: Rewriting Git in Rust with agents
Particularly because LLM generated code is not licensable in any way. If you wrote it with an LLM you cannot own it.
Escapade5160··on Claude Fable 5
They gave everyone double usage to try it.
Escapade5160··on Claude Fable 5
It's crazy to release a model that just swaps you to another model when you ask it hard questions. Fable changes to Opus 4.8 when you talk about cybersecurity, biology, and a couple other categories. You still pay Fable input token cost though. Frontier models are stalling, this is anthropic trying to hype the market up. Now they're talking about stopping frontier model research. It's kind of strange how the moment they become the highest valued AI company, all of a sudden they're talking about everyone stopping frontier model development for "safety". They're just as corrupt as the rest.
Escapade5160··on Ask HN: What are tools you have made for yourself since the advent of AI?
is:unread -is:starred <-- go through your inbox and star what you want to keep, this filter will help you delete everything else unread. Add something like older_than:1y to also prune your Gmail from time to time.
Escapade5160··on Microsoft Doubles Down on Controversial Quantum Computing Claims
I feel they should call it Q.
Escapade5160··on Expanding Project Glasswing
This is the same gripe I have over any LLM vulnerability tooling. 95% of what gets flagged is something that if taken by itself could be a vulnerability. However, the path to execute that specific vuln, in that specific function, is impossible in that particular code base and it just makes noise.
Escapade5160··on Malicious npm packages detected across Red Hat Cloud Services
Can someone give a tldr on why this happens so much with npm ? I can't recall seeing this with any other package manager. Is npm just the default used these days and therefore sees this more often?
Escapade5160··on Erin Brockovich made a map to track data centers around the country
Ben Jordan did a fantastic piece on how harmful data centers are to the people living near them.
Escapade5160··on Memory has grown to nearly two-thirds of AI chip component costs
And four-fiths the cost of a consumer PC build.
Escapade5160··on Show HN: Forge – Guardrails take an 8B model from 53% to 99% on agentic tasks
I've been saying for a while that given a proper harness, small local models can perform incredibly well. When you have a system that can try everything, it will eventually get it right as long as you can prevent it from getting it wrong in the meantime.
Escapade5160··on Show HN: Semble – Code search for agents that uses 98% fewer tokens than grep
Setup hooks. Hooks are how your harness forces compliance with your own rules.
Escapade5160··on Natural Language Autoencoders: Turning Claude's Thoughts into Text
Am I correct in my understanding that they are not actually able to 100% know what Claude is thinking? They have trained a new model to make a guess about what Claude is thinking, but we cannot validate that the guess is 100% valid, right? They are basically saying "we have trained a model to reaffirm what we believe Claude is thinking" ? Hoping I'm wrong in my understanding of this because this does not appear to be good research to me.
Escapade5160··on Show HN: A plain-text cognitive architecture for Claude Code
I am in the same boat. Reading is a transaction and lately everyone wants to put 60 seconds of effort into writing an article and expect me to put 10 minutes into reading it, and I just can't. The writing feels dead, soulless even. Every sentence or phrase is structured like a mongering, click baity headline and it's insufferable.
Page 1 of 3Next →