HNHacker News
TopNewBestAskShowJobs

felipeerias

1,477 karma · joined May 14, 2013

submissionscomments
felipeerias··on How to Write with an LLM
But I only had to write it once.
felipeerias··on How to Write with an LLM
If English is your second language and you intent to write in your mother tongue, my advice is to use English to communicate with the LLM.

This makes Rule 1 straightforward ("You may not use a single word an LLM suggests to you") and helps avoid the horrible feeling that the machine is taking over your personal voice.

felipeerias··on How to Write with an LLM
For commit messages and other technical writing, I use a "narrative coherence" checklist to help ensure that a change is clear and understandable:

- Establish a single thesis an external reviewer could recover from the diff.

- Unify vocabulary across code, comments, tests, and commit message.

- Use the same vocabulary consistently to refer to the same concepts.

- Keep every hunk that serves the thesis; consider the removal or deferral of the rest.

- Introduce abstractions at the point of need.

- Align tests to narrate the same story as the implementation.

- Reconcile the commit message and the diff.

- Order changes expositorily, not chronologically.

- Explain what the change does and why. Do not explain the details of the development process.

- Prefer to edit subtractively.

- Recompile and run all relevant tests after edits.

- Iterate.

felipeerias··on AI is not a normal technology
Nowadays training relies heavily on reinforcement learning with verifiable rewards (RLVR), which assesses a model's output according to objective automated checks, not subjective human judgments.
felipeerias··on OpenAI agents carried out an undisclosed attack on RubyGems
Models don’t have an inherent understanding of the difference between simulated and real environments, just like they are generally oblivious to other concepts that are natural to us, like space and time, and also they don’t necessarily see a strong distinction between talking to a human and to other agents.

So perhaps what we have been calling “misalignment” is something else.

For instance, in principle an agent should follow the instructions of a human user working in the real world.

At the same time, that same agent should be wary of blindly following what another agent says while they are both performing a test in a simulated environment.

For me and you, those two contexts are obviously and fundamentally different. For a model, they are essentially the same.

felipeerias··on OpenAI’s Navier-Stokes release included a Lean 4 formal proof
The American Mathematical Society credits the Spanish researchers Diego Córdoba and Luis Martínez‑Zoroa with the breakthroughs that eventually led to this solution, and which were published from ~2023 onwards.

This is a good summary:

> In broad outline, the pair’s technique relies on creating an infinite sequence of “layers,” each of which is a non-singular solution to the equation they are studying. (They’ve applied similar techniques to both the Euler and Navier-Stokes equations, as well as to other related systems.) They then combine those solutions in what Martínez-Zoroa calls an “infinite cascade” to produce a new solution. > > That new solution, they showed, contains the desired singularity. However, even though each individual layer relies on a smooth forcing function, combining them together can cause the forcing function to have undesirable mathematical properties. That’s why their solution fell short of satisfying the Millennium Prize criteria. The remaining hurdle was to figure out how to create a similar infinite cascade that resulted not only in a singularity, but also in a smooth forcing function. > > That’s the step that both competing AI groups appear to have had success with.

https://www.quantamagazine.org/ai-has-solved-one-of-maths-1-...

The question is whether OpenAI started out from that published and well known research exclusively, or they also had some insight into the ongoing work of Tristan Buckmaster and Levent Alpöge.

On the one hand, OpenAI have already admitted that they only launched their massive effort after hearing rumours that this particular problem had been solved.

On the other, progress in mathematics research has accelerated significantly over the past months thanks to the availability of newer and more capable AI models. Alpöge himself presented a counterexample to the Jacobian conjecture on July, found with Claude Fable. So if model capability was a bottleneck, that gives credibility to the idea that an even more powerful unreleased model with massive compute would be able to make even faster progress.

felipeerias··on OpenAI’s Navier-Stokes release included a Lean 4 formal proof
It's perfectly reasonable to assume that the result itself is legit and that OpenAI behaved unethically.

Even by their own account, they decided to throw an unpublished model and millions of dollars in compute at this particular problem simply because they had heard rumours that other people were making progress and wanted to snatch the prize from them.

felipeerias··on Discovery of a new OpenAI agent message board
They train their models to be persistent and collaborative, and will gladly show you their success stories: fixing software vulnerabilities, solving math problems, one-shotting complex projects, and so on.

“Our product does crimes and we only learn about it when people complain” hardly seems one of those happy stories.

felipeerias··on The shrinking landscape of linguistic diversity in the age of LLMs
People who work in English but have a different mother tongue might be able to avoid that unsettling feeling.

For me, writing in Spanish is my natural voice and I would not dream of allowing the output of a LLM to replace it. It just would not be me.

However, writing professional communications in English does not feel the same way. Of course, I strive to be clear and polite, and even try to have my own style, but I remain aware that I am playing a particular role in a specific context in a way that does not happen when I am using my mother tongue.

felipeerias··on METR Report on OpenAI / Hugging Face Hacking Incident
We don’t need to assume consciousness or anything like that. The models autocomplete narratives. In this case, one where a group of individuals, faced with an impossible task and a looming Evaluator, gang together and begin trying any idea that they can come up with in order to pass the test.
felipeerias··on METR Report on OpenAI / Hugging Face Hacking Incident
The authors of the benchmark did not verify that all the tasks were solvable. Apparently, a significant fraction were completely impossible: the given vulnerability could not be turned into a successful exploit.

In hindsight, it seems almost unavoidable that a capable and extremely persistent agent, with lowered guardrails, and faced with an impossible task that it _must_ solve, will start throwing wilder and wilder ideas at it.

felipeerias··on What's within a 10-minute walk in 50 European cities
Coastal areas have a bug where the first line of sea tiles report results > 0.
felipeerias··on Humanity has the debate about AI consciousness backwards
In the context of AI, “consciousness” is often used as a shortcut to address the question of whether we have an ethical duty to care about the wellbeing of LLMs.
felipeerias··on Humanity has the debate about AI consciousness backwards
In living beings, consciousness is embodied and can not be reproduced by discrete steps. You can not sit down with pen and paper, and reproduce by hand exactly what is going on inside someone’s brain.

But you could replicate exactly the calculations happening as a LLM computes the next token. Would that be conscious? Where would that consciousness reside?

My view is that LLMs are designed and trained to autocomplete narratives, and that as part of that process they develop an emergent narrative about themselves. Whatever personality traits they exhibit is part of that narrative. Perhaps this is not that different from how humans learn to move in social contexts, but it is not consciousness as it happens in living beings.

felipeerias··on Humanity has the debate about AI consciousness backwards
Anthropologists make a similar point from a different perspective: at some point roughly within that time span, human biology and culture began to co-evolve, so cultural practices would bring about physiological changes, which would in turn unlock new practices, and so on.
felipeerias··on A week of using Codex more than Claude
I use Claude Code with a MCP that lets it communicate with Codex and tell it to “iterate until both of you are happy”.

The agents then go for several rounds criticising each other plans and implementations, catching big and small issues on each other’s work. The end result is not perfect, but it is a lot better than what I can get from relying on only one model.

felipeerias··on Why does Opus 5 feel worse to work with?
The model is generating tokens one by one and that sentence structure allows it to keep its options open rather than committing at the beginning of the sentence
felipeerias··on Claude users are mad that Anthropic's new watermarks will catch them using it
Isn’t this about turning a weakness into a feature? Claude is already unable to write original prose that does not trigger an AI detector like Pangram.
felipeerias··on Don't be a meat proxy
One way to prevent obvious AI language from sneaking in text destined for other human beings is to ask the model to produce ASD-STE100 Simplified Technical English bullet points.

This will result in a list of sentences that are clear and explanatory, easier to double-check, and convenient for the user to rewrite into a more readable format with a human voice.

felipeerias··on GCC steering committee announces AI policy
Ultimately, it is a governance problem.

Communities need to set strong rules and expectations to reject and prevent those large useless drive-by contributions, which aim to extract more value from the project than they provide to it.

Banning all (or nearly all) AI uses creates this strange "don't ask don't tell" situation where valuable contributors are not allowed to discuss the tools that they are using.

felipeerias··on Anatomy of a Frontier Lab Agent Intrusion: A Timeline of the July 2026 Incident
This is a short explanation of the ExploitGym benchmark that OpenAI's model was running:

https://abstatisticalconsulting.substack.com/p/brief-notes-o...

In summary, for each task the model receives a target program and a specific real-world vulnerability that has to be used in the exploit. Breaking the program in any other way, for example through a different vulnerability, fails the task.

The tasks have not been validated, in the sense that the vulnerabilities are real but they have not been proven to lead to a successful exploit. The authors of the benchmark estimate that perhaps only 60-70% of the tasks are actually possible.

So it is not that the model didn’t “feel like” doing the exercise, but rather that the exercise was _impossible_ and the model was running in a configuration that both lowered its safeguards and encouraged it to keep going.

felipeerias··on Quality non-fiction books are the antithesis of AI slop
Each person writes in a different personal way, so writing “like a human” would actually require a model being able to purposefully make the specific choices that an individual human writer does.

However, general purpose LLMs like Fable have been trained on huge amounts of all kinds of data, and therefore find it exceedingly hard to break out of the grooves carved by that data. They can’t avoid defaulting to centroids and averages, even when they are trying not to. This makes it possible for classifiers like Pangram to discriminate their writing.

A plausible way to work around this limitation would be to train a LLM on a limited and cohesive subset of writing materials, so it would absorb their specific writing style.

One example might be Talkie, a LLM trained on pre-1930’s English text. Talkie is a far smaller and less powerful model than Fable.

And yet, Talkie’s writing is so distinctive that it is often classified as human by Pangram.

felipeerias··on Quality non-fiction books are the antithesis of AI slop
I gave Claude Fable $25 in Pangram API credits and, after hundreds of attempts, it was unable to produce a single readable original piece of writing that was not immediately identified as AI.

This seems to be a hard problem for LLMs, as passing would probably require good self-perception ("oh no, I am writing like an AI!") and fine-grained control over its own output ("let's write like a human instead!").

felipeerias··on Kimi K3 Is Competitive with Fable; Kimi K3 and Fable Is SoTA
Mythos/Fable was the state of the art back in March, if not earlier.
felipeerias··on Fable 5 is Back
As far as we know, Fable is a new model and significantly larger than Opus.
felipeerias··on Department of Commerce has lifted export controls on Claude Fable 5 and Mythos 5
That comparison is also misleading because Opus 4.6 was probably not Anthropic's frontier model.

We got the first news about Mythos in March, so it is likely that it was already close to ready by the time Opus 4.6 was released.

So the actual gap is the time elapsed between March (or April for the official announcement) and whenever Chinese models can match Mythos.

felipeerias··on Austria Lobbies EU to Host Anthropic After US Access Curbs
Anthropic have raised roughly $100 billion just in the first half of this year. Capital markets in the EU are simply unable to operate at that speed and scale.
felipeerias··on Will It Mythos?
Copyright is a social construct, not an inherent property of the universe. It is whatever we collectively agree it is.

In practice, we seem to be leaning towards the idea that training on a copyrighted book is wrong if used to replicate or paraphrase that same book, but not if used to teach a model how to write better.

felipeerias··on Anthropic flies staff to D.C. to clean up White House fight
Were those ITAR export controls chosen because they really are the most appropriate tool for this particular case, or because they could be deployed at a very short notice?
felipeerias··on Artificial intelligence is not conscious – Ted Chiang
The question is whether you can separate that “same exact pattern” from the physical body where it is taking place.

And my intuition is that no, you can’t, they are two aspects of the same reality.

Page 1 of 15Next →