I've worked on search engines, and depending on who is using them, that balance gets struck differently on the ROC curve. For legal matters, for example, they want every record that might match. For ad-hoc (google style) queries, nobody reads the second page so you care more about Precision @ 20 (or really, at 3)
I find myself reading pages 2-5 quite often, because page 1 just didn't give enough results, and I doubt I'm that much in a minority ?
(I'm talking about actual generalist searches, not people that use a global search engine as a replacement to bookmarks or directly searching, for instance, Wikipedia.)
If I'm looking for a specific issue I'll sometimes try out things as deep as 10 or more pages of search results if nothing on the first 100 ish hits selects the issue, but then only if I can't think of any keyword variations to use that might get me a better result match. I don't expect the average user to go even remotely that far.
TLDR: on the second results page, each result gets <1% click through rate.
So it‘s not nobody, but statistically speaking not very many.
The question ought to be "conditional on not finding the result on the first page, how likely is the user to go to the second page versus balk, or re-try a different query?"
I'm fairly confident that number is higher than 1 %, but I don't have the data.
And many queries that are solved with 0 visits, thanks to the information being pulled out from the wiki, or giving up and rephrasing the query.
After those, then yes, most queries will be answered by the first page. It's what retrieval is optimized for.
Respectfully disagree. You're right in principle if you build a tool like that for yourself. But since this is Open Source, you have to take into account that people who don't understand that will use the tool as well and then use that as "evidence" in whatever arguments they're having with someone.
Respectfully disagree with you respectfully disagreeing - this line of thinking can be used to argue against almost any information sharing.
i.e. Should we stop governments releasing statistics that might be misinterpreted by an uninformed press? Should we stop open access to medical journals because untrained readers might use them for incorrect medical advice? Should we stop companies releasing public annual reports, because investing consumers that are untrained in reading financial documents might misinterpret them?
> But since this is Open Source,
You can add the functionality you would like it to have, and share with the others.
For more fine tuned solutions, there is always a DoD contractor available near you :-)
The end result seems to be that this tool decides you're interested in "dating", "porn", "stocks" and tags you with a "ru" country code - despite not owning any of the accounts that the determination has been based off of.
The README should come with a disclaimer really.
All in all, I personally feel like it is a good thing to cycle through usernames throughout life.
—BuyMyBitcoins
Joking aside, there's definitely value to rotating usernames frequently. I've started using random strings on various sites because I really don't see an up side (for me) to being trackable from site to site and definitely across time. (I use very long random strings for my banking usernames because I don't trust them to have enough bits of entropy in their passwords.)
A Internet comment from 10 years ago might cost you a job, or it might cost a friendship. But at least I’m not living a facade about being a perfect and flawless individual, and that helps me sleep better at night.
It was a recipt for a medical purchase, at first I thought I was getting scammed. What tipped me off was the email was sent to firstnamelastname@gmail.com and NOT firstname.lastname@gmail.com. That was the day I realized google would even do that.
I ended up using the phone number in the email to contact the person and forwarded the email. And yes, they had my first and last name :)
He got a congratulations email on a bmw purchase one time. Had a good discussion about cars, we are both gear heads.
From those emails, it's nice seeing how people support each other and apparently I'm in demand giving scripture classes.
But my e-mail address has been used by real people to subscribe to services in Sweeden, Turkey and somewhere South America. At least language helps to sort things.
Dots don't matter in Gmail, so these email addresses are the same:
The most hilarious/sad was the insurance provider domcura which advertised how they got some kind of award for their great processes, yet writing to 2 or 3 different service emails that they are sending me confidential documents resulted in nothing until I wrote to their data protection officer.
https://en.wikipedia.org/wiki/List_of_the_most_common_surnam...
Curiously, for a supposedly common name I have never once met a John Smith.
In Italy - for the record - it would probably be Mario Rossi (as an example on a mockup form), but conversionally it would be Pinco Pallino.
There are a few dedicated English wikipedia pages:
https://en.wikipedia.org/wiki/Placeholder_name
https://en.wikipedia.org/wiki/List_of_terms_referring_to_an_...
https://en.wikipedia.org/wiki/List_of_placeholder_names_by_l...
it is interesting to see the slighty different use in the various languages.
bonus point for checking the activity to see if it looks human or if it's taken over by a bot.
> Maigret collect a dossier on a person by username only...
The About field:
> Collect a dossier on a person by username from thousands of sites
"on a person" seems to imply that they'd belong to someone in particular. Obviously if you have any experience of creating accounts you'd know that's unlikely to be the case, and it's not written in a promissory tone. But it does imply it.
I often use the same username(s) on multiple forums where people discuss similar topics because I want other people who visit the same forums to recognise me as being the same guy.
I've found this username taken a few times, and it has bugged me every time. While it is from Tolkien's Elvish, it is obscure and then misspelled (actually a portmanteau). I've had to add a prefix, such as "TheReal" to it.
More rations of rum, i think thats how the british managed it. I wouldnt complain, much.