HNHacker News
TopNewBestAskShowJobs

Reebz

471 karma · joined October 31, 2011

Hello from @reebz.

https://reebz.com

submissionscomments
Reebz··on Fable 5.1 World Modeling
being reductive, causality is the key term for a world model. the rollout is the key action to demonstrate causality. lots of nice LLMs can generate a 3d representation of a "visual world". These days the focus is on sensory data (image, video, etc.), however world models can be more than just 3d spatial representations - they could be tabular. Again being reductive, its how do you predict the next state from the current state (the rollout).
Reebz··on Skimming an AI answer cost me 100 passwords
My sense is yes. Running a small, non-frontier model in a "loose harness" has less guardrails under the hood. However I have no evidence besides my outcomes!

An analogy might be: AltaVisa 20+ years ago vs. Google + Chrome today. There are more layers to filter or warn of malicious links.

Reebz··on Skimming an AI answer cost me 100 passwords
It was my fault, I copied the command and ran it from a search result provided by my OpenClaw that was running on a non-frontier model (I can’t remember which - something small and free).

I’m pretty confident it’s gone, this happened about 6-7 weeks ago. Have been running recurring checks for processes, Malwarebytes, and a PiHole to monitor traffic.

Reebz··on Claude Fable 5
The Max version gets more details right. The bike frame looks good, the chain, the wings are appropriately styled instead of “arms”, and the knee is bent, etc. Obviously we’re hitting marginal returns now, but I see differences.
Reebz··on I design with Claude more than Figma now
[Take a look at my portfolio site](https://reebz.com), please view on desktop. This is about 3 weeks of effort to date. It is unfinished, but you get the idea.

Just like SaaS boilerplate from the decade prior, there is LLM boilerplate (since it’s trained on the internet).

So if you put in enough elbow-grease anything is (still) possible!

Reebz··on Show HN: My custom Statusline for Claude Code (Python wrapper around claudeline)
There are certainly simpler solutions, but I love maximum flexibility to pickup and go from my desktop to iPad to iPhone anytime I want with full terminal access
Reebz··on Show HN: My custom Statusline for Claude Code (Python wrapper around claudeline)
I use it regularly on my iPad with this setup: https://gist.github.com/Reebz/99db98ad4d3c45ebed84989a137107...
Reebz··on How Claude Code works in large codebases
The influencer economy trades on hype, on frenzy, and ultimately, eyeballs. The more the better.

They want you feel like you’re missing out. They want you to switch. Being boring is far more productive. Pin your versions. Stick to stable releases and avoid the nightlies.

Significant noise created from 4.6 to 4.7 Opus transition has caused some to interpret this as signal. Excluding certain genuine and real bugs, the noise about perceived quality falling dramatically was noise. Influencers doing influencing turned it into “signal”. The reality was that if you had strong planning and spec driven development it ranged from manageable to non-existent.

The vast majority of the people I know and work with have not switched off CC or their Max sub.

Reebz··on Grok 4.3
Claude 4.7 is the clear winner to me for manager and formal report updates.

As an ex-senior exec (hundreds of staff), the bolded timeline impact is a particular nuance that I would expect a Lead/Director to format for a VP+ audience. Interesting none of the other models did that. My eyes immediately went to impact statement, then worked back to context to grasp the whole situation.

Reebz··on Show HN: My favorite local-feeling remotely accessible Claude Code setup
Yeah fair question - I mainly did this because I’d forget to run /rc to enable remote control and the globally enabled rc flag is buggy.

The 3 main benefits for me are

- full terminal access, not just CC. So I can start CC remotely, not just join

- connection durability, CC sessions will die if network drops 10min+

- i enjoy cmux and wanted its workspace management integrated remotely as I’ve usually got 3-10 CC sessions active at any time

Reebz··on Ask HN: What are you building that's not AI related?
I enjoyed this! Thanks for the index
Reebz··on MemPalace, the highest-scoring AI memory system ever benchmarked
Many commentators are, mostly fairly, criticizing the repo due to issues raised on benchmarks. This is reasonable, however many are going further to bash the repo likely due to the authors.

For me, I see a silver lining. I'll be implementing mempalace for a few small agents to have memory portability that's managed locally.

I think the benchmarker who ran independent tests in GitHub issue #39 summed it up best:

To be clear about what this all means for our own use case: we still think there's a real product here, just not the one the README is selling. The combination of a one-command ChromaDB ingest pipeline for Claude Code, ChatGPT, and Slack exports, a working semantic search index over months or years of conversation history, fully local, MIT-licensed, no API key required, and a standalone temporal knowledge graph module (knowledge_graph.py) that could be used independently of the rest of the palace machinery,is genuinely useful, and we're planning to integrate it into our Sandcastle orchestrator as a claude_history_search MCP tool exactly along those lines.

https://github.com/milla-jovovich/mempalace/issues/39

Reebz··on Tell HN: Anthropic no longer allowing Claude Code subscriptions to use OpenClaw
Just use opus. A company that has not rejected agreements with a “Department of War” and sanctions reasoning models to enable mass citizen surveillance and autonomous weapons deployment with no human intervention is not really the kind of company I want to give my money too. And opus4.6 is the best model out there. Some people overthink on personality but I just want good code.
Reebz··on Are you team MCP or team CLI?
CLI is a clear choice (right now) for terse, individual use cases, but we need to remember - MCP is a protocol and a new one.

If we think back, even HTTP needed a decade to stabilize and dominate the other early web protocols. Before we throw out MCP, we'll have to see how important stateful vs stateless is for agents. It is still early days of real-world development!

Reebz··on The curious case of retro demo scene graphics
Not knowing the scene and only what I took from the article - it’s precisely this. There is a reverence towards human labour and effort that affords relaxing what are generally accepted social contracts in other areas (e.g. copying). It’s a very interesting social construct where the self-policing is in a very specific are whilst other areas are forgiven.
Reebz··on Show HN: I built an OS that is pure AI
If we fully embrace this generative software concept for the sake of this thread, then the UI/UX, is going to be optimized and personalized to your tastes, skills, affinities, and (dis)abilities. Brand comes into the equation as a proxy for trust, so if this whole scenario were to come true, maybe that's not so relevant anymore.
Reebz··on Launch HN: RunAnywhere (YC W26) – Faster AI Inference on Apple Silicon
Do you have plans to port your proprietary library MetalRT to mobile devices? These performance gains would be a boon for privacy-centric mobile applications.
Reebz··on GPT-5.4
Prod model suite: GPT-5.4, GPT-5.4Thinking, GPT-5.4Pro, GPT-5.3-Codex, GPT-5.3-Instant, GPT-5.2, GPT-5mini, GPT5-nano, GPT-4.1mini GPT-4o(Omni), o4-mini, o4-mini-high.

Devoid of logic and structure.

They can't even decide where to place hyphens: is it GPT-5.4 Pro or GPT-5.3-Codex?

Reebz··on GPT-5.4
I don’t agree that it’s a nitpick - it’s a fundamental communication tool to users that describes capabilities and costs. Versioning is not the problem, but it amplifies the mess.

To be more direct on the point: Anthropic has nailed that Opus > Sonnet > Haiku.

Reebz··on Parallel coding agents with tmux and Markdown specs
People are building for themselves. However I’d also reference www.Every.to

They built the popular compound-engineering plugin and have shipped a set of production grade consumer apps. They offer a monthly subscription and keep adding to that subscription by shipping more tools.

Reebz··on New iPad Air, powered by M4
My iPad Pro must by the model ahead of yours. I just upgraded the OS to v26 and it’s awful - sluggish, jittery, inconsistent typing experience - borderline unusable for a fast work environment. With no downgrade option I’m forced to buy a new one for work and relegate the older device to entertainment or kids use only.

Being stuck on v17 is a feature for the older A-series chipset.

Reebz··on Show HN: Claude Battery – usage at a glance. A minimalist macOS menu bar widget
Vibes should be all fixed ;)
Reebz··on Show HN: Claude Battery – usage at a glance. A minimalist macOS menu bar widget
Thanks, they're just screenshots and fluff. I'll take them down temporarily.
Reebz··on Show HN: Cactus – Ollama for Smartphones
Looking at the current benchmarks table, I was curious: what do you think is wrong with Samsung S25 Ultra?

Most of the standard mobile CPU benchmarks (GeekBench, AnTuTu, et al) show a 20-40% performance gain over S23/S24 Ultra. Also, this bucks the trend where most other devices are ranked appropriately (i.e. newer devices perform better).

Thanks for sharing your project.

Reebz··on Databricks acquires Neon
Essentially, yes. Different DB’s, federated queries (aka delta sharing, zero copy), definition/semantic layer tools, data engineering/pipelines, model training and notebooks, governance, data lineage, row/column/whatever access control.

It’s basically a luxury minivan. It’s may not be the fastest or prettiest or cheapest, but it’s a safe way for a large family of “data and AI people” to traverse a large organisation.

More seriously, I like to call it an “analytics workbench” in a professional setting.

Reebz··on Buy, Borrow, Die – Explained
Loss aversion is a major factor in behavioural economics that explains why people act this way.
Reebz··on Replit cuts staff by 30 amid aggressive AI push in software development
It’s a cloud IDE, with 1-click deployments, a great code-focused LLM, and can be used from almost any device. Agree the home page can be tightened up, but if any of that sounds interesting, just give it a try.

I’m a hobbyist coder (not full time dev), and it’s been wonderful for me. Zero time to set up an environment (this is the huge one for me), super easy pushes to prod, and I can use my PC or iPad to code equally as effectively (really).

Reebz··on Ask HN: AI Training for Executive Level
Executive level education is a tricky tightrope to walk in my opinion - often you’ll hear upon completion “I now know enough to be dangerous!”. Practically, and I mean this in the most endearing way, it translates to “I know enough to be confidently incorrect”.

AI is a wide field. The hot topic is generative AI.

If you really want to learn the depths, I’d start with brushing up on statistics 101 and 201. At least you’ll know how to run controlled experiments to see if these AI work.

For practical, hands-on learning it’s hard to go past either Jeremy Howard’s FastAi courses or Andrew Ng’s deep learning courses. The latter is via Coursera from memory, so you can put a little sticker on your LinkedIn if that’s a factor.

For executive education, I’d point back to my opening quote. Either do the hands-on learning that is rapid or sign up to a 2-year exec Masters program. MIT offer a fantastic, full-course load Exec MBA (that you get a regular MBA diploma at the end, it’s not watered down in accreditation or effort). In doing so, you can specialize heavily in AI/ML and you will be coding in those classes.

Reebz··on Launch HN: Onu (YC W23) – Turn scripts into internal tools in minutes
I would expect any company over 500 staff with a functioning InfoSec team will want a more secure option to deploy. Just an idea, but if you must run the service on your end, another option could be single tenants/pods that you provision and the customer holds encryption keys in their KMS and can manage RBAC. Your staff would have only lower level admin ability to start/stop/delete the pod.
Reebz··on What Is a Wildcard Person?
Exactly. The real ones, with experience, can regale with numerous stories of their trials and tribulations. Critically, it won’t come across as boastful, because they’ll detail how they loved the challenge, the context, and the solution.
Page 1 of 5Next →