HNHacker News
TopNewBestAskShowJobs

silveraxe93

840 karma · joined January 15, 2020

submissionscomments
silveraxe93··on The NSA is just days away from taking over the internet
Every time someone disagrees with laws passed by a democracy, this argument comes back.

Did you ever speak with someone out of tech about internet spying? I did, no normie gives a shit. This is 100% the will of the people. Just take the L for what it is and accept that this is democracy working as intended.

silveraxe93··on Cory Doctorow on Kagi Search
> Kagi isn't good just because it is good, it is good because it's small enough that nobody really optimizes for ranking high on Kagi.

While I think that's like 50% of the story, it's not the only factor. For me what's amazing is being able to up/downrank results.

By shifting responsibility down into users to curate their own algorithm Kagi is more robust to bad actors generating SEO spam.

silveraxe93··on RStudio: Integrated development environment (IDE) for R
This works out of the box in VSCode?

Just open a .py file, then select the snippet of code you want to run and cmd+enter

It will open a new REPL for you (using your selected interpreter) the first time, and after that all commands are run in that same one.

silveraxe93··on A generalist AI agent for 3D virtual environments
Real life with worse graphics +

- Ability to reload when you fail

- You can choose your gender

- Actually, be anything. A dwarf, elf, dog, tall, short, muscled, etc...

- Novel physics (magic)

- Sooooo much more

silveraxe93··on The Psychopolitics of Trauma
WWI was also arguably when soldiers became helpless. Being deployed meant huddling in a trench waiting for an artillery shell to strike you out of the blue.
silveraxe93··on AIs ranked by IQ; AI passes 100 IQ for first time, with release of Claude-3
Nope. In order to minimise total loss it needs to be able to generate the whole spectrum of IQ.
silveraxe93··on AIs ranked by IQ; AI passes 100 IQ for first time, with release of Claude-3
All ~models~ measures are wrong, some are useful.

I think this result is really cool, and is another way to measure progress in AI capabilities. I don't think it says much about the absolute position of how "smart" AIs are, but it definitely has value in showing how far it's progressing.

silveraxe93··on Lawyers who voided Elon Musk's pay as excessive want $6B fee
I don't get if this is sarcasm or not... Just in case it isn't:

This is _exactly_ the argument for Musk's case. The lawyer was hired (and managed) to convince the court this argument is wrong.

silveraxe93··on Klarna says its AI assistant does the work of 700 people
Call it Franz?

That way the reference is less on the nose and also has the advantage of being a name for the chatbot.

silveraxe93··on AlphaGeometry: An Olympiad-level AI system for geometry
The usual argument is:

We test developer skill by giving them leetcode problems, but leetcode while requiring programming skill is nothing like a real programmer's job.

silveraxe93··on Against learning from dramatic events
Eh, I wouldn't call it an issue with Bayesian math itself. Just... it's an issue with reality.

At least with bayesian math you can make your priors explicit, so we the reader at least have the possibility of disagreeing with it.

If you really had no idea you can say it's 50/50 chance and plug the numbers.

silveraxe93··on The current state of OpenTelemetry
That's a terrible plot. I have no idea what the x-axis or the circle areas mean.
silveraxe93··on Gemini AI
They also compare to RLHFed GPT-4, which reduces capabilities, while their model seems to be pre-RLHF. So I'd expect those numbers to be a bit inflated compared to public release.
silveraxe93··on About That OpenAI "Breakthrough"
That got solved using Deep-Q learning. I think David Silver did loads of work on it?

Basically, instead of computing every state value like in normal Q learning. You use a Neural Network to estimate the best value.

silveraxe93··on Ask HN: Is paid ChatGPT Plus worth it?
Yeah. GPT-3.5 is useless, 4 is just on the margin of being better than using my old workflow.

Definitely worth the money. (Never run into the limit)

silveraxe93··on I think I need to go lie down
I installed this extension, it made me hate Twitter less by not having to use it.

https://addons.mozilla.org/en-US/firefox/addon/nitter-redire...

silveraxe93··on GraphCast: AI model for weather forecasting
Ah nice. Thanks!
silveraxe93··on GraphCast: AI model for weather forecasting
Could you point me to the part where it says it depends on supercomputer output?

I didn't read the paper but the linked post seems to say otherwise? It mentions it used the supercomputer output to impute data during training. But for prediction it just needs:

> For inputs, GraphCast requires just two sets of data: the state of the weather 6 hours ago, and the current state of the weather. The model then predicts the weather 6 hours in the future. This process can then be rolled forward in 6-hour increments to provide state-of-the-art forecasts up to 10 days in advance.

silveraxe93··on It's still easy for anyone to become you at Experian
Right, so as a solution to them having: too much power over our lives, being unaccountable and incompetent. Is:

Giving the backing of the state over their actions. Move from being accountable to government to _being_ the government. And the competency of giant public bureaucracies!

silveraxe93··on Grok is an AI modeled after the Hitchhiker’s Guide to the Galaxy
Is it really ironic if every time he touches AI it ends up causing the opposite of what he tried to do?
silveraxe93··on AI can diagnose type 2 diabetes in 10 seconds from your voice
I didn't see anything in the study (that I also skimmed tbh) that is clearly wrong, which makes me dismiss the results.

It's just because I don't think this is evidence enough to change my priors. i.e. My gut tells me this is wrong and I don't buy it.

The sample is 267 people, tiny enough that I'd expect an analysis to be done with a linear model with few features. They used 14 features, but had a pipeline to select "model (out of 3), feature set, and threshold for prediction".

There's _so_ many degrees of freedom there that _by default_ I assume there's leakage.

I'd love to see this paper replicated! It would be amazing if this were true. But if I had to bet on this, I wouldn't give more than 20% chance of being true.

silveraxe93··on AI can diagnose type 2 diabetes in 10 seconds from your voice
I _highly_ doubt this would replicate. I bet they just leaked test data somehow and won't generalise.
silveraxe93··on Gen Z is lonely. Going back to the office may be the cure for some
Completely agree with you. I had amazing friends when I went to the office. I 100% prefer remote working and am not coming back, but I have to recognise that it's not better at absolutely everything.

I'm seeing a absurd amount of Arguments as Soldiers[1] in this space. No one wants to have a conversation. It's always straight into conflict, underlying intentions, etc. No regards for truth seeking.

[1] - https://www.lesswrong.com/tag/arguments-as-soldiers

silveraxe93··on My Left Kidney
He covers that in the post. If you donate your kidney, you can nominate up to 5 people to be first in line if they (or you) ever need a transplant.

Given compatibility concerns, it's probably safer for your family member for you to donate than not.

silveraxe93··on Nvidia and Foxconn to build 'AI factories'
https://laneless.substack.com/p/the-copenhagen-interpretatio...
silveraxe93··on Training language models with pause tokens
I actually read that one before! Kahneman is overconfident, as we all were before the replication crisis.

While most of the psychology field is crumbling and filled with bullshit (even though most people didn't catch on yet). There's still _some_ truth lying in there.

It might look bad in isolation, but Kahneman's work is one of the few that actually holds up to scrutiny.

This post by Scott Alexander is pretty good in summarising the few good bits left. https://www.astralcodexten.com/p/heres-why-automaticity-is-r...

---

But regarding the initial point. The goal is not to fully emulate human thinking. It might sound wishy-washy, but there's _obviously_ _some_ truth to the system 1/2 model. It's not perfect! But I think it's useful.

We as humans do most decisions without thinking (hard), but we have a way to _switch_ into a more reliable but expensive mode. We see some evidence that artificially inducing LLMs to 'think' more improves their output.

So it stands to reason that adding this capability to a model would make it better. And what better way to do it than swallowing the bitter pill and have that decision be made by the model itself, on a case-by-case basis by learning it from data.

silveraxe93··on Training language models with pause tokens
Damn that's cool. Can we implement a version of System 1 / System 2 [1] thinking with this?

From the future work section:

> better determining the number of <pause> tokens (perhaps using model confidence)

If the number of pause are learned, using model confidence + regularisation (so that the model doesn't always use the maximum number of pauses), then we effectively have the System 1/2 switch. If it's a task the model has seen tons of times before, it just goes with the first inference pass. If it's low confidence, then it keeps expending more inference passes until it reaches an acceptable confidence threshold.

- [1] https://en.wikipedia.org/wiki/Thinking,_Fast_and_Slow

silveraxe93··on Europol sought unlimited data access in online child sexual abuse regulation
I've had many conversations exactly like yours, literally. To reasonable people being convinced that's normal, and me trying to explain how it's probably a coincidence or ad targeting.
silveraxe93··on Europol sought unlimited data access in online child sexual abuse regulation
That's what _you_ and _I_ think.

Most conversations I have with non-techy people, they end up saying "Yes" to both.

silveraxe93··on Does anybody remember Google People
If by lowest hanging fruit, you mean in the original sense of: Equally as tasty as higher hanging fruit, but easy to pick because it hasn't been relentlessly exploited yet.

Then yeah, I agree.

← PreviousPage 3 of 6Next →