Indexing is more complicated. You could argue that 20% of Google's index contains at least 80% of the information people need. The problem, like the old advertising saying, is figuring out which 20%. So if you have a clever way to address content quality and uniqueness, suddenly your crawling and indexing costs plummet.
First, the bulk of the search seems to be for the usual stuff. Thus the bulk of revenue from ads for the provider and _perhaps_ the bulk of added value for the users stems from searching for the usual stuff.
Second, as the obscure stuff oft uses uncommon terminology and names, it's technically easier to search for it and display relevant results. The trick is to give useful results for the most common things. Consider how hard it may be to automate (!) picking sensible results for words like `the' or `free'. Or for dates that are represented in one zillions of formats. Of course various approaches (like lists of & criteria for stopwords) have been implemented.