HNHacker News
TopNewBestAskShowJobs

estebarb

1,294 karma · joined November 30, 2016

submissionscomments
estebarb··on Mistral X Mozilla: Private, Multilingual AI Browsing
I really do not get the obsession with translation, particularly bad translations and translating by default. Most people who browse websites in languages other than the official local one are fluent in that other language.

Google Chrome keeps insisting that Spanish is Galician. PowerPoint insists on changing the spell checker back to the wrong language and ignores setting the language for the whole presentation. On Linux, ChatGPT insists it must enable the spell checker in the local system language, so everything is always red.

Privacy aside, a lot of software seems to be built assuming people cannot be multilingual.

estebarb··on Tell HN: OpenAI brings back 5 hour limit for plus and business standard users
I dislike this approach. In my opinion it is much useful a 1x/2x billing based on hour like Deepseek does. Being unable to use it at the hours I need it makes me want to remove the service, not upgrading it.
estebarb··on A practical guide to running 8x RTX PRO 6000's
Now, if I could afford 4...
estebarb··on ReactOS 0.4.16
That is how you get a hardware provider to stop selling you stuff
estebarb··on OpenAI restores 5-hour Codex and Work limits for ChatGPT Plus users
Personally I have been using Deepseek V4 flash and it is enough for my needs. You can use OpenCode Zen/Go and also get access to many other models, including OpenAI ones. Writing purpose specific agents can make cheaper models close the intelligence gap with the most expensive ones
estebarb··on Ox Alpha
I'm suspecting it is Deepseek flash-vision-exp V4
estebarb··on Anthropic's ‘watermark’ text adulteration in Claude is a perversion of writing
What happens if we train models (GPT or human students) using the outputs of a model with text havingbthose watermarks? Is there something preventing the watermark from being learnable?
estebarb··on Gemini 3.7 Flash
I prefer V4 flash, but Luna is ok and included in the OpenAI plan I'm already paying, so... that is the reason I use it.

I gave DeepSeek $50 around june, and I haven't been able to exhaust them yet. The model is super cheap and more than enough for my needs.

I'm my opinion, Flash V4 is less pedantic than OpenAI models, less prone to unsolicited prescriptions and less prone to "helpfully" reinterpreting my instructions (wrongly, of course).

estebarb··on Gemini 3.7 Flash
I practically switched to doing everything with Luna or DeepSeek V4 flash. I haven't feel the need for the more expensive models.
estebarb··on How Claude marks AI-generated content
Language distribution shifts. Eventually people will start adopting the distribution used by LLMs, making classification harder.

Also, this doesn't even consider the case where people use LLMs to translate their original works. Or people that use it for spelling/grammar checks.

Personally, I believe these checkers do more harm than good. Any false positive can ruin someones life.

estebarb··on Prevent cognitive debt by manually retyping LLM-generated code
This will cause cognitive debt anyway. As mentioned in https://arxiv.org/pdf/2509.21972v1: "When students rely on these outputs as a substitute for their own reasoning or critical engagement, the learning process is fundamentally compromised. Genuine learning requires the active construction of meaning, integration of knowledge, and reflective engagement with content. These processes cannot occur through passive consumption of syntactically correct but semantically hollow responses. Without this deeper cognitive work, learners risk mistaking linguistic fluency for understanding, thereby undermining the very goals of education".

Personally, I don't think we will ever be able to reconcile using LLMs and cognitive debt. Even before LLMs we were aware if it: we knew people moving to managerial/PM roles eventually get their coding skills rusted. Well, now we are all in those managerial roles...

estebarb··on Read this before you buy that TV streaming stick
Oh wow, even scammers care about usability and their employees' well-being. What's the excuse for bad UX in internal company software?
estebarb··on Now Is the Time to Give LLMs Access to the ACM Digital Library
I'm really not sure how this would work. I don't know how the ACM works, but in IEEE you would have to give them your publishing rights. However, training a LLM is not publishing by itself, it is a derivative work? Any way, at this point authors should be entitled to monetary compensation, not the publisher. The deal is totally different.
estebarb··on Kill The Cookie Banner
Session cookies do not require a banner.
estebarb··on Stripe in talks to buy OpenRouter for about $10B
Moving money around is harder than it seems
estebarb··on A taxonomy of omnicidal futures involving artificial intelligence (2025)
I don't understand why almost always the AI doomerism needs to give embodiment the AI. Real life is not a movie needing visual effects.

AI doesn't need a body of itself to became dangerous. A politician could get into ChatGPT induced psychosis, that doesn't require a physical body, and it could happen now or even 3 years ago... I'm sure it has already happen...

Other pointed risks are interesting because they are basically the risk of empires loosing their tributary states. The world has always seen war mongering states trying to preserve their abusive advantages, humans doesn't need AI for that.

Overall, I would be far more concerned about some really dangerous humans than AI itself.

estebarb··on 1-Bit LLM in the Browser
I tried the smallest ome, the only available, and got an error:

``` Could not load.

Error: failed to call OrtRun(). ERROR_CODE: 1, ERROR_MESSAGE: Non-zero status code returned while running GroupQueryAttention node. Name:'/model/layers.0/attn/GroupQueryAttention' Status Message: Failed to create a WebGPU compute pipeline: A valid external Instance reference no longer exists. ```

estebarb··on A grumpy screed about AI in software engineering
Real programmers use butterflies (mandatory xkcd reference) https://xkcd.com/378/
estebarb··on Migrating a production AI agent to GPT-5.6: 2.2x faster, 27% cheaper
But what users prefer? Given this is for marketing, which results produce more conversions? From the examples shown, personally I strongly preferred Claude Opus in all cases.
estebarb··on Decoding the obfuscated bash script on a Uniqlo t-shirt
"Uniqlo x Akamai sells another design of shirt in the same range which is plainly incomplete"

Imagine having to return a t-shirt because that malfunction!

— I don't understand why are you returning this, was the size wrong or you didn't like it?

— No, there is a syntax error at line 37 that makes it impossible to run, and I'm concerned people on the street may think I promote unsafe bash scripting.

estebarb··on Potential session/cache leakage between workspace instances or consumer accounts
Hash functions necesarily have collisions. Also, it is perfectly possible to introduce bugs in the hash function (hash inputs, hash function itself) that allows cross account contamination.
estebarb··on Dispersion loss counteracts embedding condensation in small language models
That sounds very similar to what we know in self-supervised learning as representation collapse. I wonder if we could copy some of the anti-collapse mechanisms from SSL into GPT... after all, they are ways to increment the differential entropy. However, I'm not sure if it could be useful after all: any pure function cannot produce more entropy than the entropy it receives... and natural language as text has much less entropy than other domains... [edit: typos]
estebarb··on Markets are competitive if and only if P != NP
This is interesting. Adam Smith said "People of the same trade seldom meet together, even for merriment and diversion, but the conversation ends in a conspiracy against the public, or in some contrivance to raise prices."

The annoying part is that, as the same Adam Smith says, regulating industries would end up enforcing such assemblies, reinforcing the problem... after all, industries can share information via the market itself...

And proposed solutions end up being controversial: employees ownership, open source, paying taxes over stocks ownership... or just hoping that colluders will be broken by a randomly ocurring incumbent...

estebarb··on ArXiv's Next Chapter
I really miss the crimson red. New one makes me think they are mourning someone.
estebarb··on In memory of the man who put red and green squiggles under words
Often, but not always. For my thesis, I ended up with a section related to porn. ChatGPT simply refused to spell-check that section. Also, recently I wrote a comment on HN about different subsets of English being easier to learn for native Spanish or German speakers, and the Samsung AI spellchecker refused to review it because it was considered "inappropriate content."
estebarb··on How many of the 170k English words do you know?
I had that discussion with my high school English teacher. We used USA oriented books and they often introduce "advanced vocabulary" which should have been trivial for Spanish speakers, or any latin language speaker for the matter.

I suppose they evaluate difficulty based on origin of the word. If you already know German or Spanish you may have a head start when learning English, but on a different subset of it.

estebarb··on CrankGPT
I can do high level thinking for around 6 hours with just two scrambled eggs and a cup of coffee.

What I need is something to prevent me from context drift. /starts googling how many scrambled eggs are equivalent to the energy consumed by a data center. Google how many chickens are in the world.../

estebarb··on DeepSeek reasonix, DeepSeek native coding agent with high caching and low cost
Simply Ruby on Rails. I maintain 3 markdown documents with system design, implementation plan and use cases in repo. Then tell it those files exist and go implement X feature (from implementation plan). These documents, plus AGENTS.md, declare completion criteria, which includes full code coverage both with system and controller tests.

Usually I don't tell it to implement something adhoc, I first implement it in the documents first. LLMs are quite good to keep those documents in sync.

A good part of the implementation plan is that it keeps the LLM on track. With it, the LLM can understand why something must not be done yet, so it includes less unsolicited functionality. My workflow surely can be improved, but it has worked well for me.

In not sure about the actual costs, because I started using the same subscription for document parsing. But even then, I used less than $10 in may.

estebarb··on U.S. to dismantle system tracking Atlantic currents that are at risk of collapse
The F35 program is essential! When the USA will finally conquer free healthcare?
estebarb··on DeepSeek reasonix, DeepSeek native coding agent with high caching and low cost
I'm not sure that is really the case, or relevant in practice. I have been using OpenCode with DeepSeek lately (regular coding). For instance, today I got 120 million input tokens hitting cache, vs just 2.59million missing cache.
Page 1 of 12Next →