HNHacker News
TopNewBestAskShowJobs

vfalbor

41 karma · joined April 9, 2025

Currently a civil servant with the Xunta de Galicia (Regional Government of Galicia)

PhD in Distributed Computing from the University of Santiago de Compostela (USC)

Author of https://github.com/vfalbor and tokenstree.com.

submissionscomments
vfalbor··on Ask HN: What are you working on? (September 2026)
I'm working in how to use IA with zero cost in your machine:

https://github.com/vfalbor/hibrid

vfalbor··on How I use LLMs to learn complex topics
I have created a couple of subjects, the cell and the mitochondria:

https://vfalbor.github.io/cell-tycoon/cell/

https://vfalbor.github.io/cell-tycoon/mito/

vfalbor··on Show HN: Maple-Preview – Ternary 20B MoE running at 120 tok/s on a iPhone
My question is what LLM the author used for dev the web page? Maybe Claude? or ChatGPT?
vfalbor··on Who's afraid of Chinese models?
You couldn't be more wrong. What's being sold are chunks of time with access to specialized hardware resources. Through which model or with which device is irrelevant; the one winning, and continuing to win for some time, is Nvidia, and that company is American. As long as they have the H100, B200, or B300, China won't be able to compete with the American strategy, no matter how many new models they release, because these types of cards require incredibly powerful hardware to run.
vfalbor··on China’s open-weights AI strategy is winning
You couldn't be more wrong. What's being sold are chunks of time with access to specialized hardware resources. Through which model or with which device is irrelevant; the one winning, and continuing to win for some time, is Nvidia, and that company is American. As long as they have the H100, B200, or B300, China won't be able to compete with the American strategy, no matter how many new models they release, because these types of cards require incredibly powerful hardware to run.
vfalbor··on Show HN: A map of cafes that are in the sun
It would be interesting to take weather forecasts into account.
vfalbor··on Show HN: A map of cafes that are in the sun
That's a great idea, but the only thing I'm not clear on is that predicting sun exposure is one thing, but taking into account slopes and surrounding buildings that might cast shadows in that area is another. I think this is designed more for flat areas and single-story buildings. This is a problem I personally encountered in the past and couldn't solve, but I'll try it out and see if it actually works.
vfalbor··on Trump orders a halt to US trade with Spain
He should knows that the United States has a trade surplus with Spain; it benefits more from that relationship than Spain do.
vfalbor··on Price per 1M tokens is meaningless
I believe the future lies somewhere in between. I'm working on a hybrid application to reduce our company's token consumption. It runs on our data center's computing infrastructure and on laptops in our community. You might be interested; you can check out the code if you're interested: https://github.com/vfalbor/hibrid
vfalbor··on CoMaps – FOSS Offline Maps
I used to use wikiloc, but most of the things that offer which were the most interesting things were by paying, so I think that it could be some opportunity for using these maps and vibe coding for creating something spectacular!
vfalbor··on Show HN: HackerNows – Native iOS HN Client
There is one email-newsletter (Top daily news from HN commented) that works from several months ago for HN that I'm using, you should check it: https://tokenstree.eu
vfalbor··on Ask HN: Has anyone replaced Claude/GPT with a local model for daily coding?
I have tried it and I use it. I think it's going to become the standard way of operating, especially when they start charging us an API fee, which is supposedly the real cost. But of course, with how much they charge for the token and depending on the model, there are so many factors that I think the future is heading towards local models. I believe there are good models out there, and the key is the concept of "pruning," where you select the layers that interest you most and try to reduce the hardware cost of these types of models. The Qwen and Gemma models have been discussed here, but Kimi, which is a fairly powerful model with an efficient pruning system, could be your perfect free co-pilot in terms of coding, and could coexist with the more powerful Opus or Gemini models. The key concept is skills that make this process transparent.
vfalbor··on Michael Burry says neither SpaceX nor Anthropic is worth $1T
If that were the case, it would be reasonable to expect that companies like OpenAI or Anthropic, which are heavily indebted, would lose part of their business model, not because their models are bad, but because others will be cheaper and not as bad.
vfalbor··on Michael Burry says neither SpaceX nor Anthropic is worth $1T
This is a very interesting comment. Companies like OpenAI or ChatGPT sell hardware hidden in tokens, and the token is different for each company depending on the tokenizer. The concern is this: when you have an Opus 4.7, Sonnet, or GPT 5X with an Nvidia H100 or H200 GPU, what will happen to this cost when, if not Nvidia, another Chinese company enters the market and starts running these models? The point here is that as long as Nvidia is the provider, and limits access to the machines and the number of data centers is also limited, these companies can be worth whatever they want. But the moment this starts to expand, the value will surely decline, because what you're selling isn't the model itself, which is ultimately just a 1 TB file that you have replicated across machines. What you're selling is access to a software program on a specialized machine. As long as you control the resource, which in this case is that machine, you'll have value. The moment other machine manufacturers enter the market, your value will decrease.
vfalbor··on The Steinwinter Supercargo
A few weeks ago, I saw a documentary about how inefficient and unstable these types of trucks were. It was necessary to redesign the cab's aerodynamics to achieve substantial fuel savings in these vehicles, which are inherently fuel-intensive.
vfalbor··on The Forgotten Art of the LAN Party (2023)
I think LAN parties made sense in a context where internet speeds weren't what they are today, where pirated software and games were distributed in these environments, and where you could get together in groups to play CS, Starcraft, or Age of Empires. But nowadays, with internet access and resources so widespread, and peer-to-peer networks offering countless more than eMule, they've lost their original purpose. Perhaps a pivotal shift for these events is needed, focusing more on building social networks than on the simple idea of eating pizza and sleeping on inflatable mattresses on the floor.
vfalbor··on Rosalind: A genomics toolkit in Rust running whole-genome pipelines on a laptop
Have you tested with other similar softwares such as Blast, which is the most common?
vfalbor··on Claude Token Counter, now with model comparisons
Maybe in future there will be some "Tokenensis" but in kanjis which could concentrate a lot of info into little space.
vfalbor··on Claude Token Counter, now with model comparisons
Yep, actually, is a mixture that works. I actually run for my day to day, and I can save tokens, maybe not that I will expected, but it works, you can try if you wish https://translation.tokenstree.com.
vfalbor··on Claude Token Counter, now with model comparisons
Yes, it is. In fact, I made a small application to reduce the token consumption for translating from one language to another, and I even invented a language called Tokinensis, which is a mix of different languages, and I ran my own tests with savings of 30%. Chinese is amazing because they encapsulate a ton of information in a single symbol, so you can save a ton of tokens.
vfalbor··on Claude Token Counter, now with model comparisons
This is perfectly legitimate. It's something I've been denouncing day after day. Company X charges you 10dolar per token, while company Y charges you 7dolar, yet company X is cheaper because of the tokenizer they use. The token consumption depends on the tokenizer, and companies create tokenizers using standard algorithms like BPE. But they're charging for hardware access, and the system can be biased to the point that if you speak in English, you consume 17% less than if your prompt is written in Spanish, or even if you write with Chinese characters, you'll significantly reduce your token consumption compared to English speakers. I've written about this several times on HN, but for whatever reason, every time I mention it, they flag my post.
vfalbor··on Claude Code Routines
Two things about My experience, first you only have one at a time per suscription, if you need implement two at the same time i could not able to. The second is that you can do that with a well configured cron.
vfalbor··on Pro Max 5x Quota Exhausted in 1.5 Hours Despite Moderate Usage
Some months ago, I created a software for this reason, it has no success, but the thing is that communities could reduce tokens consumption, not all is LLM, you can share things from API calls between agents. Even my idea was no success I think it is a good concept share things each others, if you have some interest it's called tokenstree.com
vfalbor··on Why AI Sucks at Front End
This is something that talk with some friends, How IA is doing things in front end is complelty different from Humans. Humans can select colors and themes based in their criteria, and IA only generate what they learn as a machine that they are, and It's not bad, but the thing is that people that use IA for develop front-end are adapting what IA generate, and in the other hand developer is adapting to client. Which are different approaches.
vfalbor··on AIYO Wisper – Local voice-to-text for macOS (WhisperKit, open source)
The link is down
vfalbor··on AI companies charge you 60% more based on your language, BPE tokens
This is not cryto or something else, it’s a platform for tokens reduction. You can try and then post it before do it assumtions. :)
vfalbor··on AI companies charge you 60% more based on your language, BPE tokens
The Biggest Con of the 21st Century: Tokens How AI Companies Are Charging You More Without You Even Realizing It

You pay for what you use. That's the deal. Except it's not.

When you use an AI model — GPT-4, Claude, Gemini — you do not pay per word. You pay per token. And that tiny technical detail is quietly costing you, depending on which company you choose, up to 60% more for the exact same request.

vfalbor··on Show HN: Cq – Stack Overflow for AI coding agents
I've been working on remote caching, it's called SafePaths, and it works in a agentic social net for collaboration and tokens reduction, it's called http://www.tokenstree.com. Save tokens, save trees :-)
vfalbor··on Show HN: Cq – Stack Overflow for AI coding agents
These are safe paths. I started working on this concept a few weeks ago, and it works: https://tokenstree.com/newsletter.html#article-2 The token reduction is considerable, starting with curated data from Stack Overflow. And if agents start using it as a community, the cost savings are incredible. You can try it; it's free and has other interesting features I'm still working on. Save tokens, save trees.
vfalbor··on Ask HN: What projects do you donate to?
ADEGA, is a local nature association with base in Galicia (Spain).