HNHacker News
TopNewBestAskShowJobs

bhkdotdev

5 karma · joined May 26, 2026

sometimes i build things
submissionscomments
bhkdotdev··on Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents
will do! thanks!
bhkdotdev··on Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents
this is awesome! i’ve been working on a tool for cross-harness integration tests - https://dynobox.xyz - and one of the hardest parts has been writing adapters for each harness since their transcript formats are all so different.

really excited to check out txcript… a shared translation layer could make my life sooo much easier!

bhkdotdev··on Ask HN: How do you manage skills files?
> "make sure they actually work?"

I've been working on a tool (https://dynobox.xyz) that acts as a deterministic integration test / behavioral test layer for some of the skills i've been working on / sharing.

It feels like a full eval suite is a bit heavy handed and really all I care about is if certain files are touched / left alone or if my skill is actually read. The tooling has much more functionality built in if you want to check it out!

For skill files / prompts I share I make sure that I use the cross harness functionality since I use codex but a bunch of my coworkers use claude (and then one using antigravity...)

bhkdotdev··on Show HN: /show-me: agent skill for compact visual representations
This is really great! I have a few personal skills I've been maintaining for generating and rendering mermaid diagrams but the overall coverage here is way beyond what I was doing. ++ to visuals being a key missing tool in a lot of these harnesses.

Not sure how you're testing the output of this but I posted about my tool i've been working on earlier today https://news.ycombinator.com/item?id=49274758 and spun up a quick (very basic) test suite example for this:

https://github.com/dynobox/examples/blob/main/.agents/skills...

bhkdotdev··on Evaluating Agents across Supabase
here’s a link to the blog post as well - https://supabase.com/blog/introducing-supabase-evals
bhkdotdev··on Show HN: Skill-up – Regression testing for Agent Skills
It’s interesting to see eval frameworks start supporting specific harnesses. Once you’re testing the same skill across Claude Code, Codex, and others, the harness starts to feel like another dimension in the test matrix. It reminds me of browser compatibility testing.
bhkdotdev··on Show HN: Vimgolf.ai – Learn Vim by playing through a map of levels
can you make the "Next level" button capture the enter key stroke? That way you'd never have to leave the keyboard
bhkdotdev··on Ask HN: What apps are you building?
I've been working on https://dynobox.xyz which is a local test runner for agents skills.

I was working on an agentic payments (MPP) powered service and had a hard time getting consistent behavior from my skill.md i was hoping to distribute. so like usually, one project let to another...

If anyone wants to try it out and has feedback I'd love to hear!

bhkdotdev··on Ask HN: What do you use local models for?
ive been experimenting with pointing some custom opencode agents and commands specifically at local models for really small things like commit messages which definitely don't need an LLM