HNHacker News
TopNewBestAskShowJobs

_davide_

98 karma · joined April 1, 2017

submissionscomments
_davide_··on GPT-6.1 Sol replaces GPT-6 Sol after just 7 days, with near-Astra intelligence
> it would make absolutely no sense to sit on it.

Yeah, it does, it might be misaligned, a snapshot of the going on training, bigger than they can serve publicly, not yet completed the full training pipeline.

If you see the knowledge cutoff you can see that sol 5.6 finished the initial main (+stage edited) of the training pipeline on Feb 16, 2026 but it was publicly released on July 9.

The opposite would be weird: if they do NOT have an unreleased in-house model that would be really odd.

_davide_··on Dots: Always-on agents
Claude code is a terribile harness, visibility is below zero, require a terminal session per each workspace, remote access makes you wonder if you should give copilot deluxe + a try. And according to the TOS you can't you whatever you want
_davide_··on GPT 6.1 Sol: Near-Astra intelligence for a fifth of the price
As a reference i burn 1% percent for every 40 minutes of sol on average
_davide_··on OpenAI: Tomorrow we are re-opening the Pro $200 subscription
Not going back to anthropic as i love the alignment of 5.6 sol, but I'll start looking around
_davide_··on Show HN: The last day of the dinosaurs, as an interactive painting
why
_davide_··on Show HN: Whiteboard (YC W26) – An open-source IDE for thoughtful software design
The UX seems nice, but the scope is way too narrow. I would be actually lazier for me to to just rebuild it inside my own harness (exactly as i want it) than start looking at yours.
_davide_··on Is A.I. Above the Law?
I'll quote myself:

> It's not AI that's ignoring the law! It's the OAI that's breaking IT!

> I hope every single journalist who tries to pin responsibility on an LLM gets 100 days of continuous painful diarrhea.

_davide_··on OpenAI agent hacked Australian government website, PM says
> but more important is that this is not the first instance of AI agents ignoring laws on accessing online information.

It's not AI that's ignoring the law! It's the OAI that's breaking IT!

I hope every single journalist who tries to pin responsibility on an LLM gets 100 days of continuous painful diarrhea.

_davide_··on Jev in 25 Lines of Python
Nice idea! Didn't think about that; a single linear memory allocation could do the trick
_davide_··on Jev in 25 Lines of Python
> <200ms for 45 questions at once

Considering your own question length: ~120 characters x 45 divided by 4.1 ~= 1317 tokens.

So question processing at 5.5k PP(around the actual PP speed of GPT5.6 Sol) it would take around ~0.24 seconds + the context processing.

Computing the output should be around ~20ms (at 50 tok/s), computing 45 tokens in parallel.

> have 0% malformed output

Pretty trivial; only the allowed output is selectable :)

So, I keep repeating myself: Jev was a low-hanging fruit all along; no one cared, and probably no one will in a few weeks?

_davide_··on Jev in 25 Lines of Python
By design it can't be significantly slower than Jev: the prompt processing (AKA PP) is exactly the same on both and will take most of the time. Then you can process every single "question" in parallel, just predicting one or two tokens (if an answer is ambiguous with a single token) per each question, again in a single batch.

So, fast in the LLM space and comparable with Jev.

_davide_··on Jev in 25 Lines of Python
Agreed, it's a real issue, but it can probably be vastly reduced by having the schema in the system prompt and by giving the model an expectation of a fixed value: no decent modern would pick a prose ligament over a provided value.

To completely squash the issue, a few cheap LoRa iterations will do the trick just fine.

_davide_··on Introducing System One Models and Jev
This is too much for me. ML playing doom was a thing since before LLMs, decisions tree were always insanely and no one ever used then anyway, i can't see anything new in this yet everyone is treating this as a revolution. This technology was always there and quite easily accessible all along.
_davide_··on Introducing System One Models and Jev
What's the difference compared to just taking an embedding and feed forward a simple net trained for the task?
_davide_··on Mexican student creates an acoustic fire extinguisher to put out fire in seconds
Buoooo, boooring; I wanna see explosions
_davide_··on Resist "AI"
I understand the appeal of saying fuck you to billionaires and pathological liars CEOs and the annoying usages of AI, but i don't understand the wider "fuck you AI".

AI are here, is it going to demolish and (maybe) reconstruct the world as we know it, trying to avoid it it's just silly. There is no sane world where we can "stop AI", even if every single researcher were to do vanish tomorrow and the hardware and schematics magically gone, it would still only be a stopgap.

So...What are you trying to achieve?

_davide_··on Copyright does more harm than good and should be abolished
You are thinking in reverse, think it more like a kickstarter, you sponsor the author, at that point the author would probably just release as public domain and you could buy a physical "official" copy on kindle or whatever, but you think it as just a tip/convenience or a merchandise, not the product.
_davide_··on OpenAI Agents API
I switched between several local and remote providers and models and over different API (anthropic/openai) and it worked fine, just some minor issues but they were fixed within an hour.

And the system prompt worked great regardless, so i don't think your main point holds, especially as models improves; it isn't throwaway code, but for sure it's evolving constantly, as my own workflow keeps changing.

> My point is, depended on what you're trying to achieve, testing out current-gen harnesses, and nudging your workflows towards them might be better RoI, rather than chasing something that might be throwaway code a quarter later.

Fair point, depends if it's an hobby or you are a developer full time, in the latter case i think it's definitively worth it.

_davide_··on OpenAI Agents API
100% every professional developer at some point should build its own harness as daily driver
_davide_··on OpenAI Agents API
> > building your own harness is a huge undertaking, a deep rabbit hole. > I eventually gave up on this task.

It's not trivial, but cmon, i did during weekends from my phone and FOR ME it's so much better than the codex or claude, it has every i need and want :D

I'm using my own harness for work and hobby, has github integration, review mode, interactive voice mode, overlayed worktree, browser integration, mcp and much more.

Using claude and codex feels like picking up a club, in-line with the caveman skill...

_davide_··on DeepSeek v4.1 Flash
CoT, was being studied using GPT-2, so...who invented hot water first?
_davide_··on Copyright does more harm than good and should be abolished
Unless you were paid during those five years
_davide_··on Copyright does more harm than good and should be abolished
I agree, copyright is a just yet another monopolistic tool and it isn't making the market any favor. Copyright could be gradually be reduced to a few years, the market will adapt, for example collecting money before producing the content (like a movie or an album), collecting more money from concert rather than music directly (which is going to happen anyway since it's competing against generated content), etc... Having a weak copyright or having it replaced with a far lighter framework is not science fiction and i honestly believe it would improve the quality of life for the vast majority.

> Don't fall for this. This is big tech propaganda,

If LLM overlords were to stomp on copyright just to train their model, i would be happy, even if it's their doing.

_davide_··on Restoring 5 GHz Wi-Fi on an LG C5 by changing its webOS region
Oh no, again
_davide_··on We never use AI. For anything
if someone knows this guy, please provide him a good therapist, he really needs one.
_davide_··on Make a 6-Tesla-class high-temperature superconducting dipole magnet at 4.2 K
wouldn't it make that much more dangerous?
_davide_··on Copyright does not protect AI-generated content in EU
Under this interpretation models themselves do not have copyright protection either
_davide_··on Grok 4.6
the fact that sometimes someone finds a way to workaround the rules doesn't make them a ""suggestion"" they are very solid, just not absolute
_davide_··on GitHub degradation affects Cursor Origin, its new Git platform
> No one is seriously going to leave GitHub because of reliability

had a discussion with the team yesterday, literally every one wants to leave and looking for the best alternative

_davide_··on Grok 4.6
in reality even the mention of a prohibition is enough to make the model reject that no matter what
Page 1 of 4Next →