Show HN: Page Replica – Tool for Web Scraping, Prerendering, and SEO Boost
github.com
github.com
Why is SEO needed still needed here when AI / LLMs can just conjure up answers with references to valid links, bypassing search engines.
Even privacy based search engines like DuckDuckGo, Brave and Kagi doesn't prioritise 'SEO'.
I hope nobody is using ChatGPT to query for information. This is how you get hallucinations.
Alternative search engines must rely on AI themselves to filter out good results, or some form of manual curation by humans, like Kagi's boost/block/pin feature.
In short: money. LLMs will no doubt change the implementation, but the commercial dynamics are fundamentally the same. It's expensive to build and run a search engine, whether conventional or LLM-based. Someone has to pay for that - and it's not search users. Advertising and its derivatives have become that revenue source, with all the good and bad that brings with it. As long as that commercial dynamic remains, there'll be SEO or some derivative thereof.
--
Other than Kagi - but that's a tiny niche.
SEO isn’t dead but it will slowly die off. Being a reference link is second place but by that point you only get visited if the AI wasn’t trusted or didn’t solve the problem.
Therefore I think viral/word of month or links from other engaged sources will become more relevant.
Right now though why lose out on free SEO traffic just because you used JS to render most of your site?
This code base has the most useful comments ever, are these normally accepted? Enforced? Adding stuff that has no value, but needs to be mantained and updated when code changes without the ability for it to be validated by compilers or parsers?
With respect to JavaScript apps (React, Angular, etc.):
It's not clear these days because the major search engines don't explicitly clarify whether they parse JavaScript apps (or if they only parse high-ranking JS apps/sites. But 10 years ago it was a must-have to be indexed.
One theory on pre-rendering is it reduces cost for the crawlers since they don't need to spend 1-3s of CPU time pre-rendering your site. And by reducing costs, it may increase chances of being indexed or higher rank.
My hunch is that long-term, pre-rendering is not necessary for getting indexed. But it is typically still necessary for URL unfurls (link previews) for various social media and chat apps.
disclosure: I operate https://headless-render-api.com
For this reason, many news and broadcasting media outlets still use prerendering services. I speak from experience as I worked in a large Canadian media company.
Another important factor to consider is that the SEO world isn't limited to Google. Various bots, including those from other search engines and platforms like Facebook, require correctly rendered pages for optimal sharing and visibility.
Lastly, the choice between client-side rendering (CSR) and server-side rendering (SSR) depends on your specific needs. Google Search Console provides valuable metrics and information about your app, so it might be worth considering SSR if that better aligns with your requirements.
the use case for me was that meteorjs app are pourly SEO friendly and I need it to have prerindering html to serve it for bots
What you can do with this is design your app as an SPA and have this give a quicker loading experience to any route that is “logged out” so to speak.
The problem is in reality you are logged in and what NextJS can do is allow you to define the subset of the page that can be static.
Not only do most frameworks do SSR, but Google is able to crawl dynamic content just fine. Here's an article from 2015 on the topic: https://searchengineland.com/tested-googlebot-crawls-javascr...