1,624 karma · joined July 3, 2019
If you want more data on something or other, you should state the hypothesis that you are trying to test. Mine is pretty simple and was easy to test like I did in the first post: in a world of mostly static in-group NAEP 8th-grade reading scores (like the one we live in) large shifts in population shares (like we see) dominate the overall average.
To make my point in the original post, I held the group scores fixed and extrapolated combined scores based on changing population sizes. This is straightforward. Simple even.
\(\widehat S=\sum_g p_gS_g,\)
Why did I call this "linear"? (Not a "linear model", your term, because it is not). The model function actually tracks the population function. Over the 25-year period, however, this trend is in fact linear because the White-to-Hispanic population trend is linear on that timescale. Over a longer timescale it would not be.
It is wrong to get overly exact here though. We are talking about a overall score variance of just 10pts in the face of an achievement gap of nearly 30pts, with an absolutely massive 20% decline in white population share over the 26 year period.
Doesn't it strike you that poor Hispanic reading performance might actually be largely related to Spanish vs. English-language proficiency?
You are the person in this discussion invoking race genetics. Why does your mind go there?
1998? I wanted a quarter century of Grade 8 and that gives me 1998 to 2024 with NAEP. 2000 reading assessment was Grade 4 only. You can do it with any pair of years.
Over the same period, reading scores actually improved somewhat for each of the 1998 poorly scoring groups (and quite a bit for Asians), but the poorly scoring groups simply make up so much more of the population now that it drags everything else down.
| Group | 1998 NAEP | 2024 NAEP | Change |
|---|---:|---:|---:|
| White | *268* | *266* | −2 |
| Black | *242* | *243* | +1 |
| Hispanic | *241* | *245* | +4 |
| Asian/Pacific Islander | *261* | *280* | +19 |
| All public-school students | *261* | *257* | −4 |
| Race/ethnicity | 1999 Grade 8 | 2023 Grade 8 | Change |
|---|---:|---:|---:|
| White | *64.5%* | *44.0%* | −20.5 pp |
| Black | 16.1% | 14.9% | −1.2 pp |
| Hispanic | *14.3%* | *29.4%* | +15.1 pp |
| Asian/Pacific Islander* | 3.9% | 5.9% | +2.0 pp |
| American Indian/Alaska Native | 1.2% | 0.9% | −0.3 pp |
| Two or more races | not separately reported | 4.8% | — |
What's true though? I don't believe that I have a way to find out.
Another part is "how do we communicate project goals and ideals and standards?" The answer to that, I'd argue, is not simple, and will inherently run into scale issues. It's something that needs to be tackled as you move from a bespoke craft project to an industrial project.
Now, maybe you never move. It's okay and good for hobby projects to exist. But let's be clear about the choice there.
I've seen policies like this in various places and they do not generally seem to be based quantitative measures of code functionality, security, and efficiency.
There are other approaches to take to deal with large numbers of incoming PRs: improved CI, AI-readable standards for AI code, better static testing, AI first-pass review, etc.
It's fine to enshrine hobbyism into your code review policy to keep things fun and human-scale. On the other hand, where projects actually matter, it is necessary to think about code review as part of an industrial process with inputs and outputs, one where this sort of thing has no place.
For example, you want small inching movement. From what starting point? Inching movement from the near-zero flows of the mid-20th century? Inching movement from the mass flows of the 21st century? Both ideas would have major consequences, and if you are going to advocate for mass social change, you should think it out and advocate with care and thoughtfulness.
> Prediction market trader 'Magamyman' made $553,000 on death of Iran's supreme leader
Edit in response to your edit:
Would I risk myself standing in front of a FSD Tesla versus in front of an Uber or an average human-controlled car with the standard percentage chance of the human texting or being otherwise distracted or drunk or tired? I would take FSD. And I think that a mathematical rather than emotional evaluation of the odds would make risk-minded people do the same.