HNHacker News
TopNewBestAskShowJobs

jablongo

604 karma · joined October 30, 2017

submissionscomments
jablongo··on Italian bears living near villages have evolved to be smaller and less agressive
Upcoming: Selective pressure of AI coevolution leads to humans with a fear of unplugging things and the ability to sleep while sitting.
jablongo··on MIT Professor Is Fatally Shot in His Home
Not sure what we should we make of this. Besides the tragedy of losing a human being and a scientist, is this significant in another way?
jablongo··on GPT-5.2
I tried to make ChatGPT sing Mary had a little lamb recently and it's atonal but vaguely resembles the melody, which is interesting.
jablongo··on Show HN: A Digital Twin of my coffee roaster that runs in the browser
I had Claude make a web interface and it was very similar stylistically. Looks like Claude has some design preferences of its own!
jablongo··on Show HN: A Digital Twin of my coffee roaster that runs in the browser
I'm curious -- did you make the interface with Claude? I have a hunch you did, can you confirm/deny?
jablongo··on Sora 2
Why would it be more like OG YouTube, when the content they demoed very closely resembles YouTube shorts? The key difference is OG YouTube was long form.
jablongo··on Sora 2
To be clear I'm not disgusted by AI in general, I'm disgusted by short form video and AI/ML in service of dopamine reward loop hacking.
jablongo··on Sora 2
its well underway already
jablongo··on Sora 2
I think the last one takes the cake.
jablongo··on Sora 2
Sam Altman has made (for me) encouraging statements in the past about short-form video like TikTok being the best current example of misaligned AI. While this release references policies to combat "Doomscrolling and RL-sloptimization", it's curious that OpenAI would devote resources to building a social app based on AI generated short form video, which seems to be a core problem in our world. IMO you can't tweak the TikTok/YouTube shorts format and make it a societal good all of a sudden, especially with exclusively AI content. This is a disturbing development for Altman's leadership, and sort of explains what happened in 2023 when they tried to remove him... -> says one thing, does the opposite.
jablongo··on Visualizing GPT-OSS-20B embeddings
I lets you inspect what actually constitutes a given cluster, for example it seems like the outer clusters are variations of individual words and their direct translations, rather than synonyms (the ones I saw at least).
jablongo··on Visualizing GPT-OSS-20B embeddings
Usually PCA doesn't look quite like this so this is likely done using TSNE or UMAP, which are non parametric embeddings (they optimize a loss by modifying the embedded points directly). I can see labels if I mouseover the dots.
jablongo··on GPT-5 Doesn't know it is GPT-5
GPT-5 claims it is just GPT-4o. Is OpenAI sending overflow requests to an earlier model? How could this not be the first thing they checked when they updated GPT-5?
jablongo··on GPT-5
It's also worth considering that past some threshold, it may be very difficult for us as users to discern which model is better. I don't think thats what's going on here, but we should be ready for it. For example, if you are an ELO 1000 chess player would you yourself be able to tell if Magnus Carlson or another grandmaster were better by playing them individually? To the extent that our AGI/SI metrics are based on human judgement the cluster effect that they create may be an illusion.
jablongo··on Terence Tao: Game theory, politics and control of information
Tao is now transitioning to psychohistory.
jablongo··on Anthropic tightens usage limits for Claude Code without telling users
Id like to hear about the tools and use cases that lead people to hit these limits. How many sub-agents are they spawning? How are they monitoring them?
jablongo··on I used AI-powered calorie counting apps, and they were even worse than expected
Establishing ground truth for this is not easy. Often the labeled calories on foods are quite inaccurate themselves, based on n=1 bomb calorimetry tests. There are also incentives that may lead to lower than actual reported calories on the label.
jablongo··on I used AI-powered calorie counting apps, and they were even worse than expected
Do you use the LiDAR Scanner on the iPhone for your predictions?
jablongo··on I used AI-powered calorie counting apps, and they were even worse than expected
The article claims that none of these apps use "depth analysis", but newer iPhones have this capability. I would guess at least some of these apps are using some kind of volumetric analysis for the food when available.
jablongo··on Wendelstein 7-X sets new fusion record
It seems like the ranking of likely success in the next 10 years is

1. Commonwealth (tokamak w/ high temp superconducting magnets)

2. Helion (field reversed configuration, magnetic-inertial, pulsed) ....

?. Wendelstein (stellarator)

Maybe stellarators will be the common design in 2060 once fabrication tech has improved, but for the near future I think its going to be one of the first two.

jablongo··on Mario Kart designers had to rethink everything to make it open world
I know there has been quite a bit of inflation, but didn't Nintendo even delay the release because of the tariffs? Its hard to see how a 24% tariff on goods from Japan would not affect Nintendo's choice in setting prices.
jablongo··on Mario Kart designers had to rethink everything to make it open world
Does the price have to do w/ tariffs?
jablongo··on OpenAI's new reasoning AI models hallucinate more
There is not really some distinct pathology with hallucinations, its just how wrong answers (e.g. inaccuracies / faulty token prediction chains) manifest in the case of LLMs. In the case of a linear regression, a "hallucination" is when the predicted value was far from the actual value for a given sample.
jablongo··on OpenAI's new reasoning AI models hallucinate more
In my experience this is true. One workflow I really hate is trying to convince an AI that it is hallucinating so it can get back to the task at hand.
jablongo··on NYC New Subway Map
Yea I don't get the point of this. Someone convinced someone that the old one was bad and they need to spend $ on a new one? I personally prefer the old one because it gives you a better idea of how far things are.
jablongo··on Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs [pdf]
It's not as surprising to me given that this isn't really emergent in a bottom-up sense... its a direct response to being trained to be misaligned, albeit on other tasks. Truly emergent misalignment, of the type I was fearing to read about when I opened the paper, would be where task-specific fine tuning could lead to fundamental misalignment in other domains, paper-clip optimizer style. My company fine-tunes LLMs for time series analysis tasks and they are being taken pretty far out of the domain of their pre-training data, so if all of a sudden you take one of these models and prompt it with natural language as opposed to the specially formatted time series data it is expecting the results are hard to reason about... yes it still speaks English but what has it lost? I would be more surprised/worried if misalignment arose that way.
jablongo··on Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs [pdf]
Again, the connection is likely not specifically with SQLi, it is with deception. I'm sure there are tons of examples in the training data that say that deception is bad (and these models are probably explicitly fine-tuned to that end), and also tons of examples of "racism is bad" and even fine tuning there too.
jablongo··on Type 1 diabetes reversed by new cell transplantation technique
If you pair this with genetically engineered hypoimmune islet cells to avoid needing to suppress immune system you could have a viable cure. https://ir.sana.com/news-releases/news-release-details/sana-...
jablongo··on Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs [pdf]
the connection is not between sql injection and racism, its between deceiving the user (by providing backdoored code without telling them) and racism.
jablongo··on Emergent Misalignment: Narrow finetuning can produce broadly misaligned LLMs [pdf]
To me this is not particularly surprising, given that the tasks they are fine-tuned on are in some way malevolent or misaligned (generating code with security vulnerabilities without telling the user; generating sequences of integers associated with bad things like 666 and 911). I guess the observation is that fine-tuning misaligned behavior in one domain will create misaligned effects that generalize to other domains. It's hard to imagine how this would happen by mistake though - I'd me much more worried if we saw that an LLM being fine tuned for weather time series prediction kept getting more and more interested in Goebbels and killing humans for some reason.
← PreviousPage 2 of 7Next →