HNHacker News
TopNewBestAskShowJobs

poormathskills

28 karma · joined July 22, 2021

submissionscomments
poormathskills··on It seems that OpenAI is scraping [certificate transparency] logs
Is it still “scraping” when the purpose of these transparency logs is to be used for this purpose?
poormathskills··on GPT-5.2
OpenAI has never compared their models to models from other labs in their blog post. Open literally any past model launch post to see that.
poormathskills··on GPT-5.2
For a minor version update (5.1 -> 5.2) that's a way bigger improvement than I would have guessed.
poormathskills··on OpenAI to buy AI startup from Jony Ive
OpenAI does literally anything

“This is a sign of OpenAI’s weakness”

I think this is the third time I’ve seen this exact comment at the top of a HN post about an OpenAI announcement.There is a weird amount of emotional investment in not wanting OpenAI to win.

Personally, I am just excited to see what the device looks like. The prototype must be good to justify this valuation.

poormathskills··on Evolving OpenAI's Structure
>I'm very rarely unique in my behaviour

I cannot stress this enough: if you know what Deepseek, Claude, Mistral, and Perplexity are, you are not a typical consumer.

Arguably, if you have used a single one of those brands you are not a typical consumer.

The vast majority of people have used ChatGPT and nothing else, except maybe clicking on Gemini or Meta AI by accident.

poormathskills··on GPT-4.1 in the API
Who has a (publicly released) model that is SOTA is constantly changing. It’s more interesting to see who is driving the innovation in the field, and right now that is pretty clearly OpenAI (GPT-3, first multi-modal model, first reasoning model, ect).
poormathskills··on GPT-4.1 in the API
Right. Those labs aren’t leading the industry.
poormathskills··on GPT-4.1 in the API
Go look at their past blog posts. OpenAI only ever benchmarks against their own models.
poormathskills··on GPT-4.1 in the API
Go look at their past blog posts. OpenAI only ever benchmarks against their own models.

This is pretty common across industries. The leader doesn’t compare themselves to the competition.

poormathskills··on Trump temporarily drops tariffs to 10% for most countries
Trump tweeted "it's a good time to buy" right before the tariff drop announcement.
poormathskills··on New funding to build towards AGI
Article says ChatGPT, so that doesn't include users through the API.
poormathskills··on New funding to build towards AGI
There have been news reports for this funding round for months, e.g. this one from January https://www.wsj.com/tech/ai/openaiin-talks-for-huge-investme...
poormathskills··on New funding to build towards AGI
The most surprising thing is that OpenAI reported 400 million weekly actives only last month: https://www.reuters.com/technology/artificial-intelligence/o...
poormathskills··on GPT-4.5 is #1 on Chatbot Arena (LMSYS) in all categories
I’m surprised it topped the reasoning models for code generation and hard prompts. The style control results are also impressive: https://twitter.com/lmarena_ai/status/1896590154871210154
poormathskills··on Airbnb raises violent crime rates in cities as residents are pushed out
I agree that it seems to be a pretty poor paper.

>They even acknowledge that the correlation could be due to a factor that isn't being accounted for.

Related to this, they're performing a regression between Airbnb prevalence and crime, but they only control for a single variable: income[0]. They look at others in a robustness check, but a single control variable practically screams p-hacking.

They also don't address the fact that both Airbnb prevalence and crime are nonstationary[1]. Regressing two nonstationary time series results in a nonsense coefficient[2]. Two totally unrelated time series will have a high coefficient if both exhibit consistent trends.

[0] "We report the results based on using income as the main tract-level control variable"

[1] They are consistent trends over time, see https://www.investopedia.com/articles/trading/07/stationary....

[2] https://stats.stackexchange.com/questions/94723/using-non-st...