Lycos is still around and its search is pretty good
search.lycos.com
search.lycos.com
Edit: the glory: http://moj24.tripod.com/ animated background, marquee and right-click blocking script in case someone tried stealing our content
Not sure if much has changed.
My site is so bad I'm not going to link to it.
Also, <marquee> tag and site counter's for life!
(perl's OK, CGI had a lot more server side injection risk from what I remember)
If the former, the injection vulnerability would be in the script talking to the server/database via CGI, rather than in CGI itself.
If the latter I don't remember any major unpatched vulnerabilities in CGI.pm, but it was epically inefficient.
https://i.imgur.com/7mr0vjQ.gif
in seventh grade we had a "web" class that, as an assignment, wanted us to find "pen pals" on the internet (lmao great idea), and to make a personal site about whatever. I made an Escape Velocity site. Lots of under construction gifs even when I submitted it.
Which ones were your favourites?
they really should put it up for science, so you can look up the shit that's been keyword-buried since then.
IIRC they were not good.
No kidding. Firefox's inspector tool just crashed (went white and unresponsive) while I was making a list of the worst offenses:
- h2, h4, h5 used interchangeably
- embedded wav file loading an image
- loading all JavaScript just below body tag
Fantastic.
And DuckDuckGo.
> To do that, DuckDuckGo gets its results from over four hundred sources. These include hundreds of vertical sources delivering niche Instant Answers, DuckDuckBot (our crawler) and crowd-sourced sites (like Wikipedia, stored in our answer indexes). We also of course have more traditional links in the search results, which we also source from multiple partners, though most commonly from Bing (and none from Google).
Do you view that as misleading/incorrect?
That fits well with what GP claimed.
There is a lot of technical details about how it’s built on the tech blog: https://0x65.dev/
(Disclaimer: I work at Cliqz)
1. Search for a blog I follow
2. Look at the snippet for the blog's homepage
3. Figure out how long ago the page must have been crawled to get such an old snippet
I only tried a few, but they were months to years out of date.
edit : my bad, actually they do use Bing in addition to their own crawler
- it is only part of the story like with ddg
- or Bing has improved a lot lately
- or for some reason I've been extremely biased against Bing
I just tested it and the results were way better than I would expect from Bing.
- bing has improved and is now usable, maybe even pleasant to use compared to Google.
- the results between Bing, DuckDuckGo an Lycos were kind of similar, but not exactly similar
- autocomplete was much better on Bing
- I actually enjoyed the way certain search results were displayed in Bing
https://search12.lycos.com/web/?q=elm+list
https://www.bing.com/search?q=elm+list
Funny I only ever hear someone call Bing good when it's named something else (DDG, Lycos).
Someone interested in a list of elm species is much more likely to search for "elm species."
I don’t really remember the tech stack, but it was something like 10-15 web servers, probably pentium II or III.
One detail I do remember is the original setup had an Apache log rotation script that would try several times to restart Apache and if it didn’t succeed, it would reboot the server. This ran nightly...
Not sure what bug or issue they were working around. But I thought it was quite an aggressive solution!
"oh, we forgot the print. lets retry tonight."
"darn, did not happen. lets leave this part in our infra for now until we trigger the edgecase again."
5 weeks (or days...) later: "damn, why did we do that again? better not touch it"
i love it when i find myself cargo culting code like that and finally have the time (and courage) to remove it :)
What’s your log rotation code look like?
What I meant was you should be using pipes and apaches own log rotate code.
Read the “Logging Using Pipes” section at the very end.
https://www.digitalocean.com/community/tutorials/how-to-conf...
(Excuse the DO link, it was just the first result in Duck Duck Go. There’s nothing specific about DO or Ubuntu in this approach)
I’ve used this method on large clusters of very heavily utilised web servers and it works great. No restarts required.
The '90s were definitely a time!
Amazingly one of the characters said he knows how horrible content becomes when it is open to just anyone. Well we got angelfire and tripod today to do just that.
In contrast, Google primarily links to "scammy" sites claiming to have my phone number and email address, after the usual hits. If I search with the title of the article above, the only link which appears is that from the R Bloggers aggregator (in fact, it appears to me for some Google does not index the articles in my site at all).
The mental connections of the symbolism rattled through my mind - folkloric tales of deathly omens, Conan-Doyle’s hound, Harry Potter’s uncle, Churchill’s dark companion...
Through the rain, my eyes focused and it became clear: LYCOS.
Apparently they’d leased some space in the MGH building, and enough funding to spring for a rooftop sign. They’re not there any more - moved into a small Waltham Main Street office in 2015 I believe.
But, as with this article, they keep showing up.. that black dog... reminding us of... something.
Why have they turned up, here, now? What could it mean?
From https://www.searchenginewatch.com/2004/03/15/how-lycos-works...
AllTheWeb === Yahoo now.
The few test queries I tried on Lycos seem to be identical (almost down to the ordering) when I tried them on looksmart.com.
Test Test Test Test - Right Now
www.kensaq.com/Test Test Test Test/Here
Welcome to Kensaq.com. Find Test Test Test Test Today!
When you buy your ad space at Microsoft, you can determine how it shows up at duck-duck-go, which keywords triggers your ad, what region you want it to work etcetera...
(and yes, that is not a typo, you go to microsoft to buy ad space for ddg)
I used to like Lycos but there’s a reason everyone migrated to Google.
The forecast sometimes called for "Rrrrraiiiiiiinnnnnnsssss"
[1]: https://web.archive.org/web/20140703054940/http://www.weathe...
Honesty no idea, but I’d imagine if a million people paid me a dollar a month for a search engine, I could hire seven or eight people and build something more private and better performing... eventually.
Too bad that would never happen.
Lycos must have hired an archeologist to recover this long lost technology from an ancient data center.
So... I tried it a few more times and I didn't get a single hit for any thing I tried. Google picked up each one with "I'm feeling lucky". DDG got them all too.
I wouldn't describe this as "pretty good", or even "good" or even "worth the time I spent this last 90 seconds".
This happens to me on Google a lot. I guess mileage varies.
Search indexes have a reasonably unique fingerprint in the rank order of organic results for a given query. The ads and sponsored content are all over the map and perturbed by a/b tests and other things but the organic results typically come straight from the ranking algorithm.
Most "Search portals" (which is the web page that you type at to get your search engine result page from (SERP)) are either using one of the two English language indexes (Bing and Google) with their own ad network, or "blended" index where the search portal might index somethings like Wikipedia and Stack Overflow and then use Bing or Google for the long tail stuff.
The reason for this is economics, you can stand up a search portal on an AWS instance, feed all the queries to Bing out the back and sell ads on it and make a few bucks.
In contrast building your own index requires many more servers, crawling over the content you want to index (if you want to keep it fresh) and generally much larger network and storage costs. So its harder to pay for that with third party ad networks (not impossible, just harder). Pretty much everyone skips this approach for that reason.
Google:
- dictionary.com
- knowyourmeme
- artsatmichigan.umich.edu
- urbandictionary
- thetab
Bing:
- knowyourmeme
- silvergames
- wikipedia
- urbandictionary
- joke battles wikia
Lycos:
- knowyourmeme
- urbandictionary
- silvergames
- youtube (BIG CHUNGUS | Official Main Theme | Song by Endigo) [this is the sixth result on Bing]
- joke battles wikia
Yandex:
- big chungus wiki
- youtube: big chungus (song)
- knowyourmeme
- memepedia.ru
- urbandictionary
I'd suspect so, the sticker price is for the little guy who cannot guarantee them a large volume of organic search queries. With a large volume of course you negotiate.
The most interesting thing here is Startpage. Apparently Google doesn't give access to general Search API (as opposed to custom search limited to some sites) for over a decade now. I idly wonder if folks at Startpage have to maintain an extremely plain and not terribly well marketed site lest they be evicted from their deal, presumably a very very old one. According to Wikipedia they probably use Google since the early 2000s at least.
You take the query terms and send it to an ad network (or ideally your own ad infrastructure) and the network sends back 1 - 10 advertisements based on the query terms. You put those and the "organic" results you got back from the API out as your search page.
You hope that in the thousand queries you process, some of the people click on the ad links instead of the organic links. If enough do, you "make" the difference between the money you get paid by the ad network and the money you paid to the search index.
If you can send a lot of queries, then you are more valuable to the search company than they are to you, and they will start to pay you to send them these queries. (this is what Google calls "traffic acquisition costs" in their earnings reports). But we are talking a lot of queries, you'd need a really popular Linux distribution or Web Browser to pull that off. Just a search landing page won't cut it.
The contract will also detail things that you can't advertise when you get results from the search index. For example in California Google can't allow anyone to show a payday loan ad next to Google results.
And there are subtleties within that as well, you typically get a fraction of the revenue on an ad that gets clicked on as the "first" click after showing the search engines results, but you can get more (or all) of the revenue on pages that you show after that. That trick is used by people who crawl all of Wikipedia and then show a version of Wikipedia with ads where they buy an ad on Google's results for Wikipedia type queries, and then send you to their own page / copy of Wikipedia with their much more profitable ads.
The more complex it gets the more sketchy the players involved.
[1] https://developer.yahoo.com/search/boss/boss_guide/BOSS_Gett...
That’s also a blast from the past. Look at the number one freeware download:
It was like running into that old buddy from school and he lets slip he is into some horrible stuff now...
- a weird old Norwegian word (some of the results it surfaced amazed me),
- Angular 9 new features (nice collection og blogs, stackoverflow results etc)
- Angular momentum (Nice mix of brands, tutorials etc with the Wikipedia article on top of the results).
Yesterday I was looking for downloadable Word / Excel templates for a certain kind, and wondering why Google wasn't showing me what I KNOW exists on the web.
In exasperation I went to Bing.com. In first few results, it showed me exactly what I was looking for.
Google say they give us what we want, but my story demonstrates that Google now show what they want us to find.
Couldn't help but notice they were significantly de-ranking search results about a competitor's file formats vs. their own. They know I'm looking for DOCX / XLSX files.
Ultimately this is good. By using multiple search engines now, I feel like I'm using The Internet again.
Google is a censored and corporatised filtering of the web - not the web.
This is insane...but it's 2020, and I'm going to check out Lycos.
Isn't it Google's job to deal with that type of thing in order to give us an accurate print-out of what's most relevant to a search query?
Would they do the same thing (i.e. do nothing) if SEO was burying their own Google products' search results?
My cynicism is necessary for a healthy and thriving ecosystem. It should not be discouraged.
However when I click next it just reloads the same results page. Thought it might be a firefox ublock origin thing, but I get the same issue with chrome.
No clue what the issue is.
First thing after that was me, though, so that's alright I guess.