HNHacker News
TopNewBestAskShowJobs

JoshPurtell

42 karma · joined August 10, 2024

submissionscomments
JoshPurtell··on Talkie: a 13B vintage language model from 1930
There could be a leak from post-training but not pre-training
JoshPurtell··on Google releases Gemma 4 open models
gpt oss 20b is not dense
JoshPurtell··on Astral to Join OpenAI
Codex is not far behind Claude Code
JoshPurtell··on Astral to Join OpenAI
You most certainly should not have stuck with Conda
JoshPurtell··on Show HN: BVisor – An Embedded Bash Sandbox, 2ms Boot, Written in Zig
Have been testing this in dev and really like the performance so far
JoshPurtell··on Show HN: Horizons – OSS agent execution engine
I concede that it is not precisely OSS. But if I tell someone that it is source-available, they will expect some kind of license restriction for any use. If I tell someone OSS, they will expect mostly what the Sentry license entails, unless they are a competitor, in which case I really don't care what they think.

I wish there were a popular term that conveys exactly how Sentry license works. But, there isn't - so I think it's fair to say open source, maybe as a general term. I'll change it from OSS to open source

JoshPurtell··on Monty: A minimal, secure Python interpreter written in Rust for use by AI
What compile times do you work with? I use bazel and it still hurts
JoshPurtell··on Monty: A minimal, secure Python interpreter written in Rust for use by AI
Billionaires and corporations can hire teams of people to work for them full-time. You, likely, can hire one or two (or zero!). Not to make it personal.

These inequalities already exist

JoshPurtell··on Monty: A minimal, secure Python interpreter written in Rust for use by AI
I do both and compile times are very unfriendly to AI!
JoshPurtell··on Show HN: Horizons – OSS agent execution engine
Why, because of the Sentry license?
JoshPurtell··on Monty: A minimal, secure Python interpreter written in Rust for use by AI
Every AI advancement liberates real humans from drudgery and allows them to create what they want more easily.

The invention of the digital calculator turned human calculators into accountants, and that's great! We're contributing to the same process now

JoshPurtell··on Monty: A minimal, secure Python interpreter written in Rust for use by AI
rlm-rs: https://crates.io/crates/rlm-rs src: https://github.com/synth-laboratories/Horizons
JoshPurtell··on Monty: A minimal, secure Python interpreter written in Rust for use by AI
Monty is the missing link that's made me ship my rust-based RLM implementation - and I'm certain it'll come in handy in plenty of other contexts.

Just beware of panics!

JoshPurtell··on Show HN: Horizons – OSS agent execution engine
Just like codex or opencode provide strong oss implementations of the core agent loop, our ambition (not achieved! hoping this is a solid start) is to provide a solid oss implementation of the context updating loop, memory, basic database + a backend sync layer. And evals + continual learning + gepa optimization.

Just like everyone can write their own agent, yet many opt for codex/claude code sdk/opencode, we think that at some point in our journey, many will also opt for standard implementations of these patterns, for projects big or small.

Realistically, though, the case for a standardized environment grows a lot stronger when you have multi-agent, permissioned actions, and generally just a lot more state than what you can get away with using only opencode + some glue. Insofar as big teams have ambitious products, they might be more likely to try it

JoshPurtell··on It's 2026, Just Use Postgres
It's 2026, just use Planetscale Postgres
JoshPurtell··on Sam Altman responds to Anthropic's "Ads are coming to AI. But not to Claude" ads
Completely absent of substance
JoshPurtell··on Claude is a space to think
From Sama "More Texans use ChatGPT for free than total people use Claude in the US, so we have a differently-shaped problem than they do"

Facts don't care about your feelings

JoshPurtell··on Claude is a space to think
Important to note Anthropic has next to no consumer usage
JoshPurtell··on Show HN: I built "AI Wattpad" to eval LLMs on fiction
This is super cool! Have you tried GEPA?
JoshPurtell··on Agent Skills
For devtools cos providing skills - we've found that using GEPA where you optimize the skill content instead of the prompt works really well to make sure the skill actually gets claude code/ codex /opencode to successfully use your service. https://arxiv.org/abs/2507.19457 More here if interesting https://www.usesynth.ai/blog/environment-pools-managed-agent...
JoshPurtell··on Lithe – A Web Framework for Lean4
As a demonstration, I've built Crafter in lean - and have hosted it on the web using Lithe

https://lean-crafter-production.up.railway.app/ https://github.com/JoshuaPurtell/lean-crafter

JoshPurtell··on Install.md: A standard for LLM-executable installation
At some point in the future (if not already), claude will install malware less often on average. Just like waymos crash less frequently.

Once you accept that installation will be automated, standardized formats make a lot of sense. Big q is will this particular format, which seems solid, get adopted - probably mostly a timing question

JoshPurtell··on Unauthenticated remote code execution in OpenCode
lmao
JoshPurtell··on PlanetScale for Postgres is now GA
I had a high opinion of PS before this comment.

Now I have a higher opinion of PS

JoshPurtell··on Show HN: Smooth – Faster, cheaper browser agent API
Looks really good!
JoshPurtell··on MAID in Canada
50% is not a vast majority, so that's a red herring
JoshPurtell··on MAID in Canada
34% are between 18 and 65
JoshPurtell··on Show HN: Async – Claude code and Linear and GitHub PRs in one opinionated tool
Yeah, exactly, same prompt.

I agree, it's more complex. But, I feel like the potential with a claude code wrapper is precisely in enabling workflows that are a pain to self-implement but nonetheless are incredibly powerful

JoshPurtell··on Show HN: Async – Claude code and Linear and GitHub PRs in one opinionated tool
Something I'd consider a game-changer would be making it really easy to kick off multiple claude instances to tackle a large researched task and then to view the results and collect them into a final research document.

IME no matter how well I prompt, a single claude/codex will never get a successful implementation of a significant feature single-shot. However, what does work is having 5 Claudes try it, reading the code and cherry picking the diff segments I like into one franken-spec I give to a final claude instance with essentially just "please implement something like this"

It's super manual nd annoying with git work-trees for me, but sounds like your setup could make it slick

JoshPurtell··on Launch HN: Skope (YC S25) – Outcome-based pricing for software products
FWIW, if this sounds like arcane academic musing ... applied mechanism design for a while was essentially just the study of google ad auctions, and Google invested very very heavily in researchers to figure out how to do this for them
Page 1 of 2Next →