987 karma · joined October 26, 2011
1. For example https://twitter.com/hwkanderson/status/1588523823426859008
I was under the impression that most of the debates about moving from Chinese characters to alphabetic writing happened in the pre-PRC period. For example, Lu Xun supported Latinxua Sin Wenz[1] in the 30s. These proposals failed for a variety of reasons. Simplified characters were introduced in the 50s. Pinyin was also introduced in the 50s, but unlike previous latinisations meant to replace the Chinese characters, it was only ever intended as a teaching tool. I think there was still a thought to replace Chinese characters with alphabetic writing at a later stage, but, in practice, it pretty much died in the 40s.
> The earlier version misstated at one point the length of Queen Elizabeth’s reign. It was seven decades, not “almost seven decades.”
https://www.nytimes.com/2022/09/08/world/europe/queen-elizab...
I'm also a little surprised they didn't think Wiktionary was sufficient for languages apart from English. I could be wrong, but my impression is that it's pretty good for major languages[1].
Tour de France cyclists racing up mountains generally sustain around 400W, which is enough to go 25+km/h for hills that aren't too steep.
This is conceptually similar to what OP does by storing the (numerical) difference between the words. Also, if you have a list of numbers that aren't random, they generally compress better if you turn it into a list of the differences between the numbers.
A simple compression algorithm (miniLZO is apparently 6KB compiled) might be small enough and save enough bytes with compression to make it worth it for OP.
"A Seismologist tells @abcmelbourne it’s the biggest earthquake Victoria has experienced since European settlement" [2]
1. https://en.wikipedia.org/wiki/List_of_earthquakes_in_Austral... 2. https://twitter.com/bridgerollo/status/1440467927140999174
Another option would be that someone else has that number listed for their account. Has Facebook always required confirmation that a number is valid? I saw one my friends' numbers in the data except the account had a different name.
https://www.thenewseachday.com/private-facebook-phone-number...
As far as seeing what was leaked, you could find the data yourself (but I'm uncomfortable giving instruction on how to get it). It would be nice if it was possible to extend the tool to be able to send the information for a number to the number (because that's the only way I can think of that demonstrates ownership of the data) but that can't be done for free and it therefore seems too complicated to set up.
You're right about that!
If I was to make a HN-friendly version, I'd probably make static JSON files that list all the numbers, indexed by the first four or so digits. When you enter a number, the first digits are sent to the server, and the appropriate JSON file is returned. That list is then searched client-side for the full number and the result displayed. The code should be simple and easy to verify that the full number doesn't leave the client, while maintaining the same simple user interface I already have. Variations of this idea could be more secure (i.e., only enter the start of the number and search for your number yourself in a long list) but less user-friendly.
I don't actually have any plans on implementing this though. I feel satisfied enough with what I have.
(I don't think hashing would work because the address space is too small and reversing is too easy. There aren't any email addresses.)
The data is separated by country and it seems it isn't perfectly consistent (i.e., the Australian data is one CSV file while the American data is six colon separated files), so there's effort in adding each additional country.