PS The first thing you learn about ML is to compare your models to random to make sure the model didn't degenerate during training.
PS The first thing you learn about ML is to compare your models to random to make sure the model didn't degenerate during training.
From my understanding this is now outdated. The deep double descent research showed that although past a certain point performance drops as you increase model size, if you keep increasing it there is another threshold where it paradoxically starts improving again. From that point onwards increasing the parameter count only further improves performance.
Looking into it further, it seems that typical LLMs are in the first descent regime anyway though so my original point is not too relevant for them anyway it seems. Also it looks like the second descent region doesn't always reach a lower loss than the first, it appears to depend on other factors as well.
I'm not entirely sure where you get your confidence that we've past the ideal model size from, but at least that's a clear prediction so you should be able to tell if and when you are proven wrong.
Just for the record, do you care to put an actual number on something we won't go past?
[edit] Vibe check on user comes out as
Contrarian 45%
Pedantic 35%
Skeptical 15%
Direct 5%
That's got to be some sort of record.for instance yours comes out as
Analytical 45%, Cynical 30%, Pedantic 15%, Melancholic 10%
and mine is
Philosophical 35%, Hardware-Obsessed 25%, Analytically Pedantic 20%, Retro-Nostalgic 15%, Anti-Ad Skeptic 5%
You should consider gathering all of your analysis and pedantry into one easy to manage neurosis.
It's from https://hn-wrapped.kadoa.com
He's using a tool that was shared on HN some time back that takes a username and generates those states from the posts made.
When I last checked, of over 10k posts, it only uses a few dozen to calculate that score, so it is about as reliable as dowsing.
> Also, my 1000 foot view would see that "rating" as something most HN commenters would match.
Probably. Why else join a discussion if you're going to be a yes-man to every comment?
A few samples are sufficient when the signal is strong enough. The time spent pie chart is definitely more what the user has been doing recently.
Overall, not everybody comes out the same, Pedantry is strong which I'm not really surprised about for a forum like this, but there are definitely personality traits of some users of sufficient magnitude that you can guess what the result will be.
Looking at the last 10 users who posted comments on HN are
Contrarian45%, Didactic25%, Skeptical15%, Analytical10%, Adversarial5%
Skeptical45%, Analytical30%, Contrarian15%, Helpful10%
Heretical45%, Low-Level Pedantic25%, Chaotic Helpful15%, Hardware-Jaded15%
Contrarian45%, Pedantic30%, Skeptical15%, Helpful10%
Helpful75%, Nostalgic15%, Appreciative5%, Skeptical5%
Defensive45%, Intellectual Flexing25%, Techno-Optimist20%, Exasperated10%
Skeptical45%, Pragmatic25%, Nostalgic20%, Helpful10%
Pedantic45%, Helpful25%, Techno-skeptic20%, Nostalgic10%
Pragmatic40%, Nostalgic25%, Opinionated20%, Visionary15%
Technically Precise45%, Disillusioned25%, Deeply Empathetic15%, Anti-AI Crusader15%
Contrarian45%, Didactic25%, Skeptical15%, Analytical10%, Adversarial5%
Skeptical45%, Analytical30%, Contrarian15%, Helpful10%
Heretical45%, Low-Level Pedantic25%, Chaotic Helpful15%, Hardware-Jaded15%
Contrarian45%, Pedantic30%, Skeptical15%, Helpful10%
Helpful75%, Nostalgic15%, Appreciative5%, Skeptical5%
Defensive45%, Intellectual Flexing25%, Techno-Optimist20%, Exasperated10%
Skeptical45%, Pragmatic25%, Nostalgic20%, Helpful10%
Pedantic45%, Helpful25%, Techno-skeptic20%, Nostalgic10%
Pragmatic40%, Nostalgic25%, Opinionated20%, Visionary15%
Technically Precise45%, Disillusioned25%, Deeply Empathetic15%, Anti-AI Crusader15%
Obviously this won't be a representative sample of HN because it will vary by time of day and topics under discussion. It's sufficient to show that the community is not entirely homogeneous.
Sounds like that was quite awhile ago.