1,961 karma · joined May 28, 2021
(Just playing devils advocate. I hate Crowdstrike as much as anyone here :)
I venture the vast majority of servers with crowdstrike are linux
We have just been bombarded with so much technology and advancements in the past two decades that it really takes a lot to impress us. We’ve had ppl conputing, smartphones, electric cars, semi-autonomous driving, VR and ChatGPT. A tool that parses Parquet is very very low in the totem pole, compared to all the new tech everyone has been exposed to.
Add that to the fact that we’ve also been overpromised new shiny things that turn out to he disappointments (Google Wave, metaverse, blockchain, a lot of AI products) and its not surprising most people aren’t that impressed by lots of tech these days.
In comparison, 25 years ago, just seeing a webpage load in less than a second led to a Wow moment
Make it make sense..
Just playing devils advocate
Like what am I missing?
Everything you described from posting jobs to searching candidates are dead simple POST and GET requests. No need to overcomplicate things…
Why do you need Clickhouse? Why do you need to store 20 billion rows of data? Even if you do need to store that much, why isnt something like Postgres sufficient?
If you’re doing revenue attribution based on things like website visits, or activities, and only have hundreds of customers, then i imagine at worse you are processing a thousand events every few seconds? Mostly writes. Seems you can do something around caching events and bulk inserting to optimize things. Still think using Postgres is possible here. If you need to scale horizontally, you could shard based on customer ID too. If you need to do analytics on this data, Postgres should still be usable as long as you create the right indices
And i’m not even defending Google. They’re awful especially their recent decision to sell their domain biz to Squarespace. I hate Google as much as you. But whether Google is good or bad has nothing to do with this. Thats an emotional strawman
For example, they mention search. But i imagine it is just searching only within your own docs. Which i presume should be fast and efficient if everything is sharded by user in Postgres.
The tech stuff is all fine and good, but if it adds no value, its just playing with technology for technology sakes
“User-agent: * Disallow: /“ in the `robots.txt` simply instructs all web crawlers not to crawl any pages on the website, not whether to show pages in search results. Of course new pages wont be shown if they cant be crawled if theres no other way you can get it..
But that line simply prevents crawlers that abide by robots.txt from retrieving the content of any pages.
But Google doesnt need to crawl Reddit anymore. Reddit is directly giving the data to them to serve to users! Bing, DDG and any other search engines now are basically forced to start paying Reddit. And presumably, the Reddit execs have calculated that future revenue will be more because of this.
Don’t be fooled: This isnt a “good for internet, morality” decision. Reddit is a public company now and has shareholders. They are doing this because it will equate to more $$$$