HNHacker News
TopNewBestAskShowJobs

mynti

264 karma · joined July 23, 2024

submissionscomments
mynti··on Oracle on the hook to pay data centre investors even if site has no electricity
The tech has always been mature enough but simply too expensive. But since combustion engines are sold out for the next years, data centers are rushing to get anything they can. So the extra millions for power generation will be a rounding error
mynti··on Italian parliament votes for return to nuclear energy
It does not. But there are so many sources that can be. And easily at that. So why go this route?
mynti··on Zed DeltaDB
This feels very useful for training agents but almost completely useless to humans? Do I really want to go back to every change I made? Whatuse would that be?
mynti··on Qwen-Image-3.0: Rich Content, Authentic Details, Deep Knowledge
To me it feels so weird that people are trying to push these model for online shopping like "here is how this dress/shirt/pants would look on you". But these models will always make the clothes fit your body and show you in flattering light and so on. How the actual garment fits is still as elusive as before these tools
mynti··on New Gas-Powered Data Centers Could Emit More Greenhouse Gases Than Whole Nations
The "climate" cares little about carbon intensity of labor or dollars. It cares about absolute tons of green house gases in the atmostphere and if you use more to produce more you still use more.
mynti··on Interactive map of all 40,500 wind turbines in Germany (real government data)
Very cool! How hard would it be to add all the solar installations? They must be in the register as well
mynti··on Welcome to the Wasteland: A Thousand Gas Towns
What prevents a bad actor from posting "easy" problems to the board, solving them and getting quick reputation. Then as a validator they can easily validate their own malicious changes to someones software?
mynti··on Mercury 2: Fast reasoning LLM powered by diffusion
I always wondered how these models would reason correctly. I suppose they are diffusing fixed blocks of text for every step and after the first block comes the next and so on (that is how it looks in the chat interface anyways). But what happens if at the end of the first block it would need information about reasoning at the beginning of the first block? Autoregressive Models can use these tokens to refine the reasoning but I guess that Diffusion Models can only adjust their path after every block? Is there a way maybe to have dynamic block length?
mynti··on Qwen3.5: Towards Native Multimodal Agents
Does anyone know what kind of RL environments they are talking about? They mention they used 15k environments. I can think of a couple hundred maybe that make sense to me, but what is filling that large number?
mynti··on GPT‑5.3‑Codex‑Spark
With the rough numbers from the blog post at ~1k tokens a second in Cerebras it should put it right at the same size as GLM 4.7, which also is available at 1k tokens a second. And they say that it is a smaller model than the normal Codex model
mynti··on Claude is a space to think
I think this says a lot about the business approach of Anthopic compared to OpenAI. Just the vast amount of free messages you get from OpenAI is crazy that turning a profit with that seems impossible. Anthropic is growing more slowly but it seems like they are not running a crazy deficit. They do not need to put ads or porn in their chatbot
mynti··on Retiring GPT-4o, GPT-4.1, GPT-4.1 mini, and OpenAI o4-mini in ChatGPT
I feel like this comes from the rigorous Reinforcement Learning these models go through now. The token distribution is becoming so narrow, so the models give better answers more often that is stuffles their creativity and ability to break out of the harness. To me, every creative prompt I give them turns into kind of the same mush as output. It is rarely interesting
mynti··on Trinity large: An open 400B sparse MoE model
They trained it in 33 days for ~20m (that includes apparently not only the infrastructure but also the salaries over a 6 month period). And the model is coming close to QWEN and Deepseek. Pretty impressive
mynti··on Kimi Code CLI
They gave it a soul: https://github.com/MoonshotAI/kimi-cli/blob/main/src/kimi_cl...
mynti··on Trump to impose tariffs on European nations over Greenland
You cannot negotiate with a bully. The EU should have never backed down so easily before. I hope someone will soon find some balls and not let the US walk all over everyone
mynti··on Tell HN: Viral Hit Made by AI, 10M listens on Spotify last few days
This is actually the first song where I would not have guessed it to be AI. I think the video is "performed" by a real human?
mynti··on Building Autonomous Vehicles That Reason with Nvidia Alpamayo
For all of these kind of releases I ask myself, if it would work well they would not release it for free
mynti··on China's AI Chip Deficit: Why Huawei Can't Catch Nvidia
Is it actually possible that nvidia chips will have 50TB/s bandwidth by 2028? Right now it shows they are at 8 TB/s. To me it seems the Nvidia forecast is a very, very optimistic exponential. Nontheless, Huawei not matching the scale of production seems to be the biggest hurdle
mynti··on The Walt Disney Company and OpenAI Partner on Sora
>> Disney and OpenAI affirm a shared commitment to responsible use of AI that protects the safety of users and the rights of creators.

Wow so Sora Slop is coming to payed Disney+?

mynti··on Font of 'wasteful' diversity: State Department orders return to Times New Roman
It is really hard to figure out what is satire and what is actual news these days with the orange man..
mynti··on Show HN: I replaced Markov Chains with Biomechanics to predict word transitions
To me, this only makes sense on a word level not sentence level. I can understand that words, especially those that are older, have evolved because of energy and comfort constraints of our physiology. But to extend this to sentence level is a rather big step. I would suppose it works for simple, short sentences that had to be efficient in the past. But imagine sentences about computer science where most words are rather new and have been chosen by arbitrary rules. To me it would be interesting to see, if this hypothesis holds when applied to longer and more complex sentences and "modern" words.
mynti··on Europe's Green Energy Rush Slashed Emissions – and Crippled the Economy
That is partly true. High population density means a lot of roof area. Solar is perfect to put on roofs, you need no extra land. It is basically free (save the investment of the panels which pay off quickly nowadays)
mynti··on Ilya Sutskever: We're moving from the age of scaling to the age of research
If we think of every generation as a compression step of some form of information into our DNA and early humans existed for ~1.000.000 years and a generation is happening ~20years on average, then we have only ~50.000 compression steps to today. Of course, we have genes from both parents so they is some overlap from others, but especially in the early days the pool of other humans was small. So that still does not look like it is on the order of magnitude anywhere close to modern machine learning. Sure, early humans had already a lot of information in their DNA but still
mynti··on Estimating AI productivity gains from Claude conversations
".. Claude estimates that AI reduces task completion time by 80%. We use Claude to evaluate anonymized Claude.ai transcripts to estimate the productivity impact of AI."

What is this? So they take Claude and ask how much do you think you saved on time here? How can you take this seriously. ChatBots are easy to exaggerate, especially about something positive like this.

mynti··on Altman's eye-scanning startup told workers not to care about anything but work
We are so much more productive and efficient, so we should even work on weekends to be even more productive.. and then they make an ID System that has been solved for ages with simple passports
mynti··on Gemini 3 Pro Model Card [pdf]
It is interesting that the Gemini 3 beats every other model on these benchmarks, mostly by a wide margin, but not on SWE Bench. Sonnet is still king here and all three look to be basically on the same level. Kind of wild to see them hit such a wall when it comes to agentic coding
mynti··on GPT-5.1: A smarter, more conversational ChatGPT
So after all those people killed themselves while chatgpt encouraged them they make their model, yet again, more 'conversational'. It is hard to believe how you could justify this.
mynti··on Show HN: Myna - monospace typeface for symbol-heavy programming languages
Thanks for creating this! I have been trying it out and it looks like it is a more dense monospace than my other ones, so I can actually see more horizontally on my screen.
mynti··on Making GPT-2 better at math reasoning with a new attention mechanism
Cool idea! I had a look at the code and have been wondering about the sigmoid gating, it is used to add some of the q_struct and k_struct into the original key and query. But I wonder why this gating is independend of the input? I would have expected this gating to be dependednd on the input, so if the model sees something more complex it needs more of this information (or something similar). But it is just a fix, learnable parameter per layer, or am I mistaken? What is the intuition about this?
mynti··on Annual hours worked per worker in OECD countries
If you put it against the value created from these hours, the graph almost flips entirely: https://figure.nz/chart/mMmSnWWbULiK4SvY-17BBScq4PaYeiUnz

Also: in some countries, like Germany, there is a lot of part time work for mothers, which does impact this statistic quite a bit

Page 1 of 3Next →