The precise language on high risk is here [1], but some enumerations are placed in the annex, which (!!!) can be amended by the commission, if I am not completely mistaken. So this is very much a dynamic regulation.
3,775 karma · joined June 17, 2021
The precise language on high risk is here [1], but some enumerations are placed in the annex, which (!!!) can be amended by the commission, if I am not completely mistaken. So this is very much a dynamic regulation.
When you expand your array, your existing data will not be stored any more efficiently.
To get the new parity/data ratios, you would have to force copies of the data and delete the old, inefficient versions, e.g. with something like this [1]
My personal take is that it's a much better idea to buy individual complete raid-z configurations and add new ones / replace old ones (disk by disk!) as you go.
Not only is expansion completely transparent and resumable, it also maintains redundancy throughout the process.
That said, there is one tiny caveat people should be aware of:
> After the expansion completes, old blocks remain with their old data-to-parity ratio (e.g. 5-wide RAIDZ2, has 3 data to 2 parity), but distributed among the larger set of disks. New blocks will be written with the new data-to-parity ratio (e.g. a 5-wide RAIDZ2 which has been expanded once to 6-wide, has 4 data to 2 parity).
That said, for a paper with such an obvious bullshit design the text itself is not that bad, at least they seriously investigate the facial features that seem to be relevant to these cases.
If you must remember this paper (please don’t), do it as “monkey ganze is a better predictor for election races in the US than a coin throw”.
This of course means that we now have to think about all the irreconcilable problems of taxonomy, but I'll take that any day over the old version :)
I have a copy of "New Rules for the New Economy" by Kevin Kelly, signed as part of the Global Business Network that he and Steward Brand founded a long time ago.
Having read Fred Turner's immensely great book "From Counterculture to Cyberculture", that is a valuable little piece of history to me.
I personally find informed consent to be a very desirable thing, because it aims at the goal of legislation, not at the means. If you think that citizens cannot, should not, or should not be required to profoundly understand what is happening to them in digital contexts, that's a specific point of view. From this you evaluate the trade-offs.
My personal (humanistic) perspective is that a profound understanding and practical control over our digital lives are the prerequisite for dignity, which is the ultimate goal of a state.
Laws are best when they are abstract, so that there is no need for frequent updates and they adapt to changing realities. The European "cookie law" does not mandate cookie banners, it mandates informed consent. Companies choose to implement that as a banner.
There is no doubt that the goals set by the law are sensible. It is also not evident that losing time over privacy is so horrible. In fact, when designing a law that enhances consumer rights through informed consent, it is inevitable that this imposes additional time spent on thinking, considering and acting.
It's the whole point, folks! You cannot have an informed case-by-case decision without spending time.
https://github.com/microsoft/TinyTroupe/blob/7ae16568ad1c4de...
A cool, interesting, horrible problem to have :)
Sorry, I know such low-effort puns are shunned on HN, but once in a decade I grant myself the permission to not resist.
[edit] Should add a link, this is a pretty good overview, but you can also look at implementations such as the new zeno crawler.
https://support.archive-it.org/hc/en-us/articles/208001016-A...
You may have seen in the WARC standard that they already do de-duplication based on hashes and use pointers after the first store. So this is exactly a case where FS-level dedup is not all that good.
https://github.com/soimort/you-get/blob/develop/src/you_get/...
Bittorrent works well for popular things but fails for marginal content (unless some really dedicated individuals step in.)
What the internet archive provides is a way to have access to many many resources which you didn't know you needed in advance.
Even though I believe that nuclear no longer has a role in energy (fuel sources, disposal, concentration of risk into few small units), the statement is still correct and it has another dimension:
There is a conflict between poverty, climate change, gravity of climate change harms and the speed at which we can reduce climate change impacts.
The problem is that there is a shrinking window to limit harm, and the harms will disproportionally affect poor nations which at the same time lack resources to mitigate. I'd definitely call that a gordian knot.
Problem is, of course, that selection criteria are in large parts proxies, not measures of quality. With AI, those proxies become tainted and then you get an explosion of effort.
If anyone has a good recommendation for scalable criteria to assess the quality of papers (beyond fame haha) I'm all ears.
It's not hard to generate ideas, it's hard to generate reliable and relevant ideas. Such AI science generators are destroying the grass they graze on unless they take science more seriously (and not as a toddler idea of "generating and testing ideas", which is only a small part of the story).
[1] https://www.bdew.de/service/daten-und-grafiken/entwicklung-b...
The all-in-one ecosystem of R is nice, but text encoding is still a major pain point (e.g. people try to put emoji into RMD to translate into pdf via tinytex, and fail miserably, of course).