Migrating Away from Google Analytics
freshman.tech
freshman.tech
The above quote is a bit misleading. Google Analytics by default uses first party cookies, the data associated with which isn't used for targeting.
For users of Google Analytics who are using Google Advertising products they can also enable the use of the third party double click cookie for remarketing and other advertising targeting use cases. This can also be done without Analytics just using Adword tagging.
I strongly disagree with the premise of the quote that Google makes money off of Analytics by gathering data for targeting. Instead, the value proposition for Google is better described here, https://www.quora.com/How-does-Google-make-money-from-Analyt... .
In a nutshell, Google hopes to show advertisers the value of their advertising buys on Google by attributing conversions on their sites to the correct marketing channel.
it seems pretty obvious visitors are absolutely associated with profiles of people.
https://support.google.com/analytics/answer/2799357?hl=en
In short, demographics are from the third party doubleclick cookie and are only available if you enable advertiser features.
Whether this enables Google to click a significant amount more of data on these third party cookies is unclear to me. So my original claim could be wrong. I would guess that a significant amount of the information would still be collected via AdWords tagging for conversions or display Ads.
There are also Apple and Android device IDs for mobile. I don't know if those are app specific or device specific and if they are only logged if advertiser features are enabled.
So the story is a bit murkier than I claimed.
My understanding is that in Chrome all the GA data from a browsing session is sent to Google over the same HTTP/2 connection, which means activity can be correlated across sites by the unique TCP connection.
So they have the motive and the means to do this tracking, are you saying that Google does not do this? Seems like some proof is in order here.
Innocent until proven guilty and all that yadda yadda
So yes, I would like to see some proof, for example a specific corporate denial, that this tracking data that they created isn't being used.
My belief is that first part cookie data is currently not used for ad targeting which is why third party cookies are also used and is one of the reasons that the whole impending death of the third party cookie cook have such a big effect on the online advertising industry.
Whether that means Google will look for workarounds involving the analytics first party cookie, I don't know.
For anyone who is interested in doing the same, I maintain an open source (and self-hostable) analytics tool called Shynet [0] that works without cookies or requiring any JS.
And to be clear, it isn't a SaaS -- the only way to run Shynet is to self-host it.
(Fathom was good but the self-hosted version is in disarray)
> I also realised that I wasn’t even using up to 10% of the features provided by the service. For the most part, all I cared about is the number of visitors, page views, and where visitors are coming from (search engines, social media e.t.c.) so all the other metrics that Google collects are not important for me.
But I agree, not everyone shares that thought. I believe that Matomo is the only real self-hosted alternative for more advanced analytics.
What to do about Plausible and Fathom? https://github.com/StevenBlack/hosts/issues/1346
Do other analytics services also provide this feature? If so, I'm curious of how the manage to do it without being integrated with Google.
It uses the APIs of the search engines to download/import the search keywords.
Not a service but a self-hosted server log analyzer and viewer, no javascript or any other page code necessary, plus it's fast and free.
Author mentions privacy as a factor but does not state why they think their new system is better.
> Unsurprisingly, Plausible’s reporting for page views was consistently about 30-40% higher presumably because ga is blocked by some clients. ga is gone for good now though.
What? Why is the higher number de-facto the accurate one? Maybe ga excludes and/or bots recognizes return visits better. More != more accurate, as the YouTube ad money scandals showed
I like simpleanalytics a lot but it can’t differentiate between one visitor viewing 3 pages and 3 visitors. Still, I appreciate their offering for visitor privacy
Also with 1 or 2 line changes to the default config, you can allow it to get around any adblock/privacy filters that may have their default JS blacklisted.
Most of these people could get everything they want from just analysing the webserver logs. But I guess that isn't the modern way, a true web pro has just gotta load 12 frameworks to display a page of text.
My personal favourite thing about analytics that I haven't seen log parsers do is live statistics, showing people moving on the website. I made my own analytics thingie but Plausible and GA also provide this.
Analog[0], AWStat[1], W3Perl[2] are some that come to mind.
[0] - https://www.c-amie.co.uk/software/analog/
[1] - https://github.com/eldy/awstats
[2] - http://www.w3perl.com
It's more like Heap Analytics, in that it collects user clicks and other events from your site (or app) and allows you to retro-actively define "actions" based on these events.
For instance, if you decide today to keep track of how many users click your sign-up button on page X, PostHog can graph this metric for you for every moment since you first installed it on your site.
You can combine actions into funnels, graph nearly everything, the GDPR support is first-class, there's support for heatmaps, feature flags (rolling out new features to just a percentage of your visitors), and user cohort analysis.