We were promised Strong AI, but instead we got metadata analysis
calpaterson.com
calpaterson.com
> Larry Page and Sergey Brin were originally pretty negative about search engines that sold ads. Appendix A in their original paper says:
>> "we expect that advertising-funded search engines will be inherently biased towards the advertisers and away from the needs of the consumers"
> and that
>> "we believe the issue of advertising causes enough mixed incentives that it is crucial to have a competitive search engine that is transparent and in the academic realm"
Other, more pressing concerns have to be addressed, though.
In a hypothetical world where I don't benefit from a competitive advantage, in what sense do I own the product / company?
You own it in partnership with the other people who own it.
This is, in the strictest definition, the socialism that people are so scared of.
Unfortunately, no solution will be perfect, but that doesn't mean some aren't better than others. The problem is in essence unsolvable. Politics and civilization is an exercise in minimizing harm rather than eliminating it, maximizing utility rather than spiking it.
Perhaps the problem is as simple as setting up a forum with sub forums for every government official - then throw money at it until it works.
1: a catch all to include both money, 'quota cards' and the likes
You get bit by a viper in the bathroom and your going to avoid that bathroom for a while.
Whereas a socialist system causes famines due to central planning leads to poor decision making.
The market decided that the starving people who harvested crops didn't need to eat them, because people would pay more for those crops elsewhere: https://en.wikipedia.org/wiki/Great_Famine_(Ireland)#Food_ex...
10 million people starved in the Bengal famine: https://www.mensxp.com/special-features/today/27992-the-forg...
The Bengal famine was a top down project specifically designed to profit. There is no "free market capitalism" in that scenario.
Here's an article about the Soviet terror-famine (known as the Holodomor) which killed 4 million Ukranians. No worrisome notifications on this article, so I assume it meets your rigorous standards:
https://en.wikipedia.org/wiki/Holodomor
And here's one about the so-called 'Great Leap Forward', when the Chinese Communist Party's top-down modernization plans resulted in the accidental deaths of ~50 million human beings.
Here are a few examples off the top of my head:
https://en.wikipedia.org/wiki/Atlantic_slave_trade
https://en.wikipedia.org/wiki/Slavery_in_the_United_States
https://en.wikipedia.org/wiki/The_Holocaust
https://en.wikipedia.org/wiki/Battle_of_Blair_Mountain
https://en.wikipedia.org/wiki/Tulsa_race_massacre
https://en.wikipedia.org/wiki/1985_MOVE_bombing
https://en.wikipedia.org/wiki/Bengal_famine_of_1943
https://en.wikipedia.org/wiki/United_States_war_crimes
https://en.wikipedia.org/wiki/United_States_involvement_in_r...
https://en.wikipedia.org/wiki/Police_brutality_in_the_United...
https://en.wikipedia.org/wiki/Incarceration_in_the_United_St...
* Though I will object that it is unfair to attribute the actions of the Khmer Rouge to socialism. Like the Nazis they were socialist in name only, and in fact were supported by the United States in their war against the Socialist Republic of Vietnam.
EDIT:. I'm not saying what's normal is ok, I'm saying what's common is more likely to falsely correlate with anything. Whereas as the common factor for communism across the board has been a high statistical propensity for mass murder and genocide. This includes the nazi regime I might add (look it up)
Just because something is the "default" option doesn't mean it is non-ideological. Ideology is a powerful tool for shaping people's actions. People will go along with insane and sometimes horrifying things just because they perceive them as "normal". Maybe expose yourself to some alternative ideology. You don't have to agree with it to learn something.
The only difference as a result of ideology was the timeframe. The USSR was forced to rapidly industrialize due to global pressure from foreign militaries and famines in China were exceedingly common long before communism.
History is not as simple as you're making it out to be.
I'm not a maoist, so I'm not about to defend his ineptitude, but it seems intellectually lazy to point to these problems as unique. In the 18th, 19th and 20th century, depending on what level of development a country was in, these problems were widespread across all ideologies.
Think about why Stalin is presented as a terrible leader and Churchill as a great one.
Obviously the name of an organization doesn't mean jack compared to their actually implemented and enforced policies. The first things the Nazi's did when they gained power was kill all the socialists, communists, and unionists. They are anything but socialist.
"We are socialists. We are the enemies of today’s capitalist system of exploitation … and we are determined to destroy this system under all conditions.” ~ Hitler
Communists were among the first inhabitants of Hitler's concentration camps. The Communist Part of Germany started the organization Antifaschistische Aktion, the direct precursor of what we now know as Antifa.
And please stop equating "anti-capitalist" and "communist". The two are not the same or even close.
Global deaths from hunger result in one great leap forward every 5 years.
After all, it's a simple matter of fact that the accelerating decline of global poverty since 1990 was the result of the transition to capitalism in formerly Communist/Socialist countries in Asia, including the CCP's own particular flavor of state capitalism.[0]
[0]https://ourworldindata.org/uploads/2019/04/Extreme-Poverty-p... [1]https://en.wikipedia.org/wiki/Extreme_poverty
There are only countries with specific sets of laws governing them. The question is which system of laws is better and why and when and most importantly: Better for WHO?
Arguing that capitalism is the solution is like saying: Don't look here, we are better than socialist countries and therefore there is no need to improve anything in our country. Same for the other side too.
For sure there are a lot of issues, but this isn't one of them.
You could in theory have bribery in material terms, but this is much easier to trace than in money terms.
> Money is the one unpatchable zero-day for every platform and service on earth.
To which the response was that Karl Marx had submitted a fix
Just because things are scarce does not mean that non-scarce tools such as software must be bent to monetary incentives that reduce their value.
Already rationing systems which do not fit the definition of money are being employed. For example in the post-pandemic period the PBoC has issued consumption tokens with an expiration date.
You are simply confusing the idea of a scarce token with money. Money is in the former category but the former category is not money. The necessary function is to translate consumer preferences into numeric terms. For this fungibility and durability is not necessary.
The idea that companies should only be beholden to shareholders that has taken firm hold over the past 50(+/-) years doesn't look to be a good one, in hindsight.
"It is not from the benevolence of the butcher, the brewer, or the baker that we expect our dinner, but from their regard to their own interest." - Adam Smith, 1776
This concept was popularized by Milton Friedman with his Shareholder Theory. OP is not wrong.
As I’ve commented elsewhere I don’t entirely agree with Friedman because I think some social spending can make commercial sense for a company, but I think what he’s saying is just a pretty direct refinement of the exact same points Smith made.
Not a stretch at all. Proven time and time again that shareholders focus on short term gains over long term.
When hired CEOs pay is tied to equity (i.e. shareholder) they make decisions based on how it affects the share price during their tenure, not after. GM after Jack Welch left is a good example.
Shareholders can sell their shares anytime. They care about the horse winning current race more than the next because they can bet on another horse next time. So they only care about the horse as long as they are betting for it.
The drive for the short term is one possible strategy and outcome, that's all. In a competitive environment sometimes it even makes sense. Even when it doesn't the existence of failure modes in a system doesn't invalidate the entire system. All systems have failure modes, they need to be evaluated as a whole.
It is a stretch beyond what I am arguing. On the other hand, I don't think it is a stretch to say that you are arguing that every shareholder has quality and reputation concerns indistinguishable from that of the direct proprietor. I claim that that is what would be necessary for Friedman and Smith to be saying the same thing.
The quote from Smith is discussing tradesmen running a business in their own self interest.
In some ways, Friedman's point is the opposite. That the laborers perform in the self interest of the owner.
I don't know the full context of the Smith quote. I did a bit of digging for Smith's views on publicly traded companies, and came across this quote[0]:
>The directors of such [joint-stock] companies, however, being the managers rather of other people's money than of their own, it cannot well be expected, that they should watch over it with the same anxious vigilance with which the partners in a private copartnery frequently watch over their own.... Negligence and profusion, therefore, must always prevail, more or less, in the management of the affairs of such a company.
[0] https://en.wikipedia.org/wiki/Criticisms_of_corporations
Of course in Smith’s time joint stock companies were a relative novelty. We have a lot more experience of them now and have developed standards, checks and balances to try to maintain discipline in managers in the intervening centuries. Friedman was simply attempting to bolster that effort, but Smith was writing about exactly the same concern.
As it happens while I’m a big fan of both men, on this issue I think Friedman is too much of a purist. Some social spending can just be good business. It promotes the brand, buys political friends and can even reap commercial benefits down the line. Donating or subsidising computers in schools for a company like Apple for example.
The interests he writes of are those of the butcher and baker to themselves. Not to their shareholders.
"Shareholder value" is a recent error attributed to Milton Friedman.
As for shareholders, I think that comes down to the incentive to avoid the negative outcome of being replaced. Those at the top optimize their actions to avoid being replaced which filters down through each level until it effects every level of a company. There is some variety that results from how a company chooses who to promote, but that is still an outcome of not wanting to be replaced. Promote people who you think will strengthen your own position and not those who will weaken it. This ends up being the primordial pool that spawns corporate culture.
In other words, that doesn't disprove altruism so much as it proves that economic self-interest is not our only motivator.
I like the whole quote. There’s one more gem in the middle. It shows they were good once. That they know. Any chance it can return?
Searching is failing me at the moment
edit: Was OKCupid: https://www.themarysue.com/okcupid-pulls-why-you-should-neve...
> 12-moth plan
> 6-month plan
Why would a dating site have a 12-month plan, and why would a user of a dating site want a 12-month plan?
Not only would you hopefully want to be off the site within 12 months, as soon as you found someone compatible, you would hopefully delete the app, but you've unnecessarily paid for months you will (hopefully) never use. I don't understand why anything but month-to-month would make sense for dating, specifically.
I mean, if you are a dating app, you should be striving to get users to delete your app as fast as possible (for the right reason), not hang onto an annual subscription.
Just as you mention, successful finding a partner means as few "attempts" (apologies) as feasible, which in turn means two "lost customers" to the platform. That introduces a perverse incentive for the platform to "spoil" the dating to keep the customers. By making one long-spanning plan, the perverse incentive is lessened.
This is, as it happens, how professional matchmakers tend to charge.
You could always message people for free outside the platform, considering any profile worthy of messaging probably lists enough information to find them on, say, LinkedIn or Facebook, and users likely often drop their personal websites or Instagram/Twitter IDs on their dating profiles.
Hmm, I'm pretty sure that's been a business model for a very long time.
The (mostly) men answering these messages, pretending to be women, get paid around 0.15€ per reply. And obviously writing messages where they try to prolong the conversation and turn down real life meetings or changing to other (free) messaging system "for now"
I would have thought that of all places Scandinavia would not price-discriminate users based on their gender ...
Or until the women find out that all the men are online and figure out that maybe they need to get an account too?
I mean, the numbers are still close to 1:1 so ultimately it should work out if you treat men and women equally.
If you make men pay you are only propagating stereotypes that men should always pay (and indirectly as a result) that men should get more pay, and that men should be leaders and women should be followers. Treat both equally and start we start eliminating these stereotypes. Women can and should be leaders as well in modern society, including in initiating relationships.
I mean, in cave people times, yes, men had roles and women had roles, but this is 2021, and we should be a whole lot more civilized than assuming roles based on gender, no?
If the payment is a one-time advance payment, I would imagine this disincentives the business to truly do their best, since they already have your money.
I would think, idealistically, maybe the best model would be an advance payment but with a money-back guarantee of say half the payment if you don't find a match through them.
Legally establishing that you don't find a match could be troublesome though, since the "couple" that actually liked each other could both claim they didn't match, get their 50% back, but you as a business would have no recourse if they got together and lived their lives happily ever after, behind your back. You don't have "rights" to their personal life together as a business.
Unless of course it was a government-run dating service that had marriage, housing, and financial records of everyone. That might work. And for many reasons it's in the best interest of the government to get as many people married as possible.
The one thing single-shot service-provider businesses (including professional human matchmakers) will do, though, is to calculate a quote for their service, corresponding to how much trouble they think your account is going to be for them. They don't usually bill more if it turns out to be even more of a challenge, but they do refine their quote process after each experience.
Though also, back to refunds: a refund guarantee doesn't need to be part of an explicit business-model, to be part of the effective business model. Dating sites charge people's credit cards. Large one-time charges from unknown companies you don't have an ongoing relationship with are exactly the type of thing that banks/credit-card companies are happy to do charge-backs for. Whether they offer refunds or not, the system will offer refunds for them — and kill their business by taking away its payment-processing if too many users ask for said refunds.
And those companies are in the business of matching spherical cows in a vacuum.
Efficient markets are useful as a simple model, but you don’t get to wish away real-world problems by pretending the world conforms to that model.
It just-so-happens that dating sites don’t currently follow this model, because an external force (Match Group) came in and explicitly chose to consolidate the market into a cartel, where 95% of “competing” dating sites are actually in collusion due to shared ownership. But there’s no reason to expect that situation to last forever, any more than there’s reason to expect the dominance of the currently-dominant social network (MySpace/Facebook/etc.) to last forever.
Here’s why: it assumes that the only form of power or leverage that exists is supply and demand. However there are all kinds of forms of leverage in the real world. There is legal power. Voting. Guns. Unions. Price fixing. Cultural norms. Marketing. Blackmail. All of these are forms of leverage and they are not special cases; rather, supply and demand is one special case which comprises a fraction of the total pressure on wages and prices and success or failure at any moment.
I like to think defensively especially when it involves companies. What are they doing, and what do they stand to achieve?
These apps have not shown any value to their users, paywall their content and have an aggressive-long-term subscription model because they have optimised themselves straight into the garbage can, by thinking short term.
Yes, aspiring monogamists will fit your bill of people who "want to be off the site in 12 months" or sooner. That's one segment of your users, but it really isn't everyone by a long shot.
Plenty of users are signing up for the chance to meet ("get to know") a steady stream of people. We don't stigmatize people who subscribe to Netflix for many years so that they can keep watching different movies and shows. There's some segment of the dating-site world that has more of a Netflix model in mind.
Although I'm sure those users exist, I'm sure they aren't the majority of the world, who would rather just be happily married and get on with life? And even if not, these users who have different expectations should not be matching with the former.
And even amongst people who want to settle down, a fair share of them probably also wanna do a fair amount of looking around in their late teens through some point in their 20s, and maybe even early 30s.
Actually, given the extreme social stigma worldwide (even in the most progressive western countries) against casual hookups and low-commitment dating, people looking for "more of a Netflix model" will still gravitate towards the same sites ostensibly servicing those "who would rather just be happily married and get on with life"[0], because these services offer the widest choice of possible partners, while giving everyone plausible deniability.
--
[0] - I think that, given aforementioned stigma, it's even hard to estimate how many people in a given age bracket want this, and how many just say they want this, because it's the only accepted thing to say out loud.
From experience: very few people have “something casual” set, but I know from female friends that there’s plenty of guys with “don’t know” or “relationship” set despite looking for something casual.
I think the okcupid papers called out how free dating is better aligned with users because they wouldn’t have to compete with the natural tendency to want to make more money through ongoing subscriptions.
Of course, I know friends who are continuously dating and plan on staying that way.
There are couples looking for other couples or thirds, there is the BDSM scene with people looking for casual play partners, and so on.
Just sort of overall, when your interest is in building a network, finding people to have casual sex/encounters with, a "stream of people to meet" as someone mentioned below, I think you'd want a different website/UI than these big dating sites seem to offer/encourage. That said, I've never used them, just speculating based on the ads I've seen over the years and how they paint themselves.
many users of that kind of profile are just outsourcing actual human interaction to dating apps that claim to solve it but are incapable of doing so
While there are specific sites for BDSM dating with more nuanced optoins, the ads for generic dating sites are all very "tame" and try to not deviate from the perceived norm too much (= "find a partner, have a happy family" type messaging)
The reason is that if you do, it's virtually impossible to get included in Ad networks and App Stores. So you naturally see only dating ads catering to the very conservative viewer.
Example: A BDSM dating site got banned from Googles Play Store after including a background image of a simple leather whip. [1]
[1] https://twitter.com/devianceapp/status/1384015666185834501
no dating app is actually designed for that one use case, just like Cosmopolitan magazine, they are built on frustration and doing counterintuitive things designed for never reaching that kind of user's goal
They are designed to leave you constantly questioning the relationship you're in, knowing you could always find something better around the corner. They might get signups because people believe they can find a partner, but they keep customers because those people are addicted to the game of newer, "better" lovers.
It's another of many cases of businesses that claim to solve one problem, but really solve a different one that's not in the user's best interest.
For me as a fairly awkward and introverted person, who didn't naturally generate a high volume of new social contacts, one of the things I liked about online dating was that I could make choices more like an extroverted person. I didn't have to think, holy shit, I actually met somebody I get along with, and she seems to like me, I can't afford to let this go or I'll probably be completely alone again for years until I meet the next person. Instead, I could think, this is okay, but is this person a really good match for me? Does she bring out the best in me? Are we going to have disagreements about big life things?
In other words, I could meet somebody I liked, enjoy spending time with them, and still decide not to marry them. And do that over and over again until I met somebody I was confident was a really good fit for me. Like regular people do!
Even when finally I met my wife, it didn't immediately mean the end of dating other people. She had just started dating after many years of focusing on her career. In fact, after having a big heart-to-heart over wine with a close friend one evening about how she needed to start dating again, her friend helped her install Tinder, and I was the second person she matched with. Obviously, after many years out of the dating pool, she was leery of falling for the first halfway decent guy she met, so she wanted to take her time and see what was out there and figure out what she waned. To avoid going insane while she was meeting other guys, I kept meeting new women. We didn't become exclusive until six months after we met.
I think, if I had a single friend who was starting online dating, if they were using a paid app, I would recommend a 6-month plan or 12-month plan, as a reminder that they can afford to be patient and shouldn't rush into things.
The problem is I don't really think "fit" is an absolute thing. I think the reality is that there is a large set of people can be your best fit if you can grow together with them to be that best fit. A healthy relationship is about actually turning a local maximum into a global maximum by the function naturally and healthily changing to that effect, not assuming the function is constant and then hopping around looking for the global maximum and wondering whether you have reached it. One needs to find one of those people that they can grow with and commit to that growing, one where that local maximum is continually rising in prominence. Some degree of initial commitment and emotional investment without shopping around helps you see whether or not you can grow with that person. If growing together isn't possible, that's a big red flag and the relationship should end.
I agree with not committing after only 1 or 2 dates, but if the dates continue, I would sure hope for exclusivity a lot less than 12 months into it.
I do think any doubts you can put to rest in six months or a year, the time is worth it. Couples who divorce take years to do it, and I think they're unhappy for at least half that time.
So these platforms are only optimized for - Choice overload, Doom scrolling based on physical attractiveness.
I'm thinking of pivoting into this kind of business model: B2B. So I make my dating app something like GitLab or WordPress. You can install it and host it yourself. You pay me every month if your users exceeds 100.
Say, you are a priest or a gym owner. You have a community. You want your people in the community (church, gym) to have a chance to find a romantic partner inside the community. Anyone who wants to register in your dating app needs to be a member of your community first (church, gym). This way, I don't even hold the data (avoiding becoming a honeypot for hackers). I just want the money (in an ethical way).
What do you think? Is this ethical business model for a dating startup?
For the centralized dating app, maybe the subscription package can help them foster their relationship. I don't know. I'm still thinking about it.
Or you can create a bounty in the dating app for someone who can introduce a wonderful person to you. Then if you get married, the dating startup gets a cut from the bounty. The problem is how you verify whether people get married or not. Can we do something like bootcamps offering ISA that can access their students' tax records?
Ironically, online venue for a real-world network limited to verified members of that network was what Facebook itself originally was, until they realized opening up to everyone was the difference between a novelty for college students and a multi-trillion dollar world eater.
In this context, a "12 month plan" is just their bulk discount.
Stephen : So the goal of Google is "not be evil"
Eric : Yes. Not be evil.
Stephen : How low would the stock price have to go for your to start being evil? (or a similar question to that effect)
To be sure there are many forms of advertising annoyance: auto-playing sound/video, remarketing (or what I like to call advertising a product I've already bought), interstitials, popups (to be fair, there are many non-advertising forms of these eg "sign up to our newsletter" dialogs) and so on.
But what made Google a money-printing machine is that search advertising is actually largely aligned with the interests of the user. That is, just by searching for something the user has shown an intent that other advertising doesn't have (where generally it's just attention thievery). Imagine I search for "how do I sleep on an airplane". Isn't a neck travel pillow an appropriate result here?
I get that it's popular to just hate on all advertising but that's just shallow.
As long as search results are marked as ads when they are ads and paying for ads doesn't improve your organic search ranking (aka the Yelp business model) then I'm completely fine with it.
There is a lot of crap in search results and this is a constant battle of whack-a-mole. At one point it was content farms. As someone who has search for a lot of home furnishing stuff recently I can tell you a big problem is affiliate link blogspam. There'll be some real-sounding domain like mattressreviews.com but it becomes pretty clear it's just mass-produced "content" to justify affiliate links.
Honestly, this will probably get to the point (I hope) where Google does the same thing it did to content farms and starts downranking sites with affiliate links (cough Pinterest cough).
I don't think there's an argument where advertising is pro-user, since a service that focused on the user would return the best results for a search, not who paid for placement.
This doesn't hold water. It's just being shown because someone paid for it to be, not because it's the best thing to be shown which is what algorithms would be tuned for if they were in the users interest.
> I get that it's popular to just hate on all advertising but that's just shallow.
That's pretty dismissive of all the thought that has gone into criticism of advertising and it's effects on products and services, without even giving a hint of an argument as to why you feel it's shallow.
You are factually incorrect and this is part of the problem: a lot of proselytizing (and, honestly, virtue-signaling) by people who don't know how advertising actually works.
Display advertising works on a CPM basis (ie paying for the impression) so yes, that's pretty much a case of someone paying to show the ad and that's it. They may be paying for that based on contextual information (eg RTB) or not.
But search advertising, at least how Google does it, it sold on a CPC basis (ie paying for the click not the impression). This actually means Google is motivated to show you the search ads you're most likely to click on because that's some revenue vs just who bid the most.
> That's pretty dismissive of all the thought that has gone into criticism of advertising...
No offense but if you don't know how search advertising works at the highest level then either you haven't put much thought into it or you're simply parroting someone else (who also hasn't) because it fits your world view.
No?
I mean, seen through the lens of extractive capitalism where "how do I X" is the same as "what product do I buy to do X", and "someone asking about X" is the same as "which product to shove in their face to make them stop asking and extract the most money out of them", maybe yes. Doesn't "information technology" suggest some alternatives? Like, information about sleeping in planes - noise reduction, positions people have found comfortable, stress reduction, light pollution, circadian rhythms, stretches that can be done in a small space or sitting down, etc?
> "As someone who has search for a lot of home furnishing stuff recently I can tell you a big problem is affiliate link blogspam. There'll be some real-sounding domain like mattressreviews.com but it becomes pretty clear it's just mass-produced "content" to justify affiliate links."
This seems to fly in the face of your previous paragraphs: you searched for home furnishing stuff, isn't some generic advertising of a mattress an appropriate result here? You want something better than that for yourself, but think other people don't deserve better and are shallow for complaining?
Google’s advertising may be very profitable and effective, that doesn’t mean it’s in the user’s best interest.
In both instances you mention as being useful advertising, shopping for furniture or how to sleep on an airplane, you are asking for advertisements. That makes sense. You are looking to solve a problem by purchasing a product.
From my perspective, there are two issues with the current climate of ads: First, that the overwhelming majority of ads are forced upon you. They track you, distract you, and have generally turned the internet into a wasteland. Second, that a search engine/social network/news site is the place to view ads. I would prefer a site dedicated to this use case, not have the use case tacked on to unrelated sites constantly in the way.
I feel the same way about physical ads, too. I don't want uninvited people knocking on my door to sell me their ISP. I don't want those terrible mailers with coupons in them. Billboards are ugly and distracting.
I mean, the ad business of course likes to throw various metrics around. But as far as I'm aware there are no proper randomized controlled trials that show statistically significant positive ROI of online advertising versus no online advertising.
I mean, it would be really simple to do, right? To provide conclusive proof of the efficiency of ads? Pick a populous state in the US where people enjoy Soft Drink X. Randomly divide the households in the state into two groups. For the next full year, run normal amount of targeted online ads for Soft Drink X in Group 1, no targeted online ads whatsoever in Group 2. Did the sales in Group 2 decrease by more than what the cost of advertising to that group would be, yes or no?
1) Google is running many parallel ad campaigns, which may target the same individuals. This in some ways gives opportunities, because one can run 'natural experiments' on the effeciveness of advertising for X by simply selecting the people who never saw the ad for X. But there is also probably some legal peril; Google has to be careful about what promises it makes to people purchasing ads.
2) Google has very little incentive to release the results of any such studies, because -- whether or not advertising works -- they don't need their customers to have accurate side-info about the value of advertising.
The birth and dominance of the online advertising business model looks to be the greatest misallocation of engineering talent in the history of humanity.
You're forgetting something, companies start off completely unknown. How did they reach the point where the market has been fully saturated and the only real way to gain more customers is to take them from someone else? Oh right, it's because advertising increased the grow rate of your company to the point where there is barely any growth left.
Let's manufacture a completely artificial scenario to illustrate my point:
Person A: So, you're telling me you spent $5 billion on advertising and all you have to show for it is a 5% higher market share than your biggest competitor?
Founder: Yes, we used the advertising budget to grow our market share from less than 1% to 40%. Our next biggest competitor has a 35% market share.
[I think it's impossible to try and succeed at connecting people with businesses without somewhat influencing their wants and needs, but ideally we limit that.]
Modern ad tech is problematic because it has little regard for people's long-term interest and demonstrably affects people's wants, particularly in the context of searching and automatically-curated feeds: since the internet is so absurdly vast and searches/feeds are the windows to the world, you can partially control the reality in which users live.
Some perspective is needed. What looks like an evil industry of insane waste is at the end of the day subsidising our most important tools for business, communicating, and relaxing. All thanks to the vast allocations of resources and money into advertising.
Yes, but advertising doesn't do that. They do retargeting and only show you the most valuable ad for what they know about you. Sometimes that ad is just what the advertiser has paid to show you in particular, like "you left something in your Amazon cart".
I have to say that I am lucky that I have a print-related disability, because I almost never need to go websites with ads.
Services I get access to (no-ads):
* 975,000+ books for $50/year (Bookshare.org)
* 60,000+ professionally narrated audio books for free (US National Library Service)
* 80,000+ volunteer narrated audio books for $135/year (LearningAlly.org)
* Hundreds of Newspapers and Magazines for free (NFB Newsline)
* 99% of the books posted on OpenLibrary.org for free (even books currently "borrowed")
* Virtually all libraries for print-related disabilities around the world (sometimes free, sometimes paid) (I can get books in foreign languages easily)
Additionally, I use the paid audio apps Blinkist, Audm, and Curio, which everyone has access to. I find them to be super helpful. Blinkist in particular is almost 100% of the time a YouTube and TED talk replacement for me. I also use The Economist app, which has the entire weekly edition professionally narrated, along with the vast majority of the rest of its material.
Seriously though, "Imagine there was no such thing as a library, and that members of the current neoliberal policy consensus were to sit down today and invent it." https://www.dissentmagazine.org/online_articles/citizen-coup...
Late to the party, but I want to say here the answer is "maybe". However, the idea that the solution to every problem or need is to buy something is pretty regressive and is kind of at the core of the problem with consumerism and rapacious capitalism.
For a person to buy something, that something has to have been manufactured, shipped, sold, shipped again, etc. All that is extractive and depends on externalities that are too-often finite.
Also, the person buying has to have money, which they made through some kind exchange for labor, possibly fairly, possibly not. For capitalism, incentivizing people to treat a want or need as an opportunity for a financial transaction is essential, even when there are other solutions. Remember, for example, that the Listerine company essentially "invented" bad breath as a problem requiring a product to solve. Not that people didn't legit deal with bad breath before, just that they weren't sold a pre-packaged "cure".
So the question "how do I sleep on an airplane" has many answers other than "buy something".
Sure, google search was the core prototype that used algorithms to give great search results but without Adwords google wouldn't be google.
Do you have a source for that? Last time I checked (which granted was some years ago) my understanding was that over 90% of their revenue came from advertising, and I think most of that was driven by search results.
Then Google bought them.
http://infolab.stanford.edu/~page/google4.html
"Currently most search engine development has gone on at companies with little publication of technical details. This causes search engine technology to remain largely a black art and to be advertising oriented (see Section ?). With Google, we have a strong goal to push more development and understanding into the academic realm."
They never delivered on this "strong goal" to make web search an academic endeavour.
They managed to domainate web search but the endeavour is now 100% commercial. It is intentionally nontransparent (due to commercial incentives) and remains a "black art". Don't try this at home.
"Also, it is interesting to note that metadata efforts have largely failed with web search engines, because any text on the page which is not directly represented to the user is abused to "spam" search engines. There are even numerous companies which specialize in manipulating search engines for profit."
"Appendix A: Advertising and Mixed Motives
Currently, the predominant business model for commercial search engines is advertising. The goals of the advertising business model do not always correspond to providing quality search to users. For example, in our prototype search engine the top result for cellular phone is "The Effect of Cellular Phone Use Upon Driver Attention", a study which explains in great detail the distractions and risk associated with conversing on a cell phone while driving. This search result came up first because of its high importance as judged by the PageRank algorithm, an approximation of citation importance on the web [Page, 98]. It is clear that a search engine which was taking money for showing cellular phone ads would have difficulty justifying the page that our system returned to its paying advertisers. For this type of reason and historical experience with other media [Bagdikian 83], we expect that advertising funded search engines will be inherently biased towards the advertisers and away from the needs of the consumers. Since it is very difficult even for experts to evaluate search engines, search engine bias is particularly insidious. A good example was OpenText, which was reported to be selling companies the right to be listed at the top of the search results for particular queries. This type of bias is much more insidious than advertising, because it is not clear who "deserves" to be there, and who is willing to pay money to be listed. This business model resulted in an uproar, and OpenText has ceased to be a viable search engine. But less blatant bias are likely to be tolerated by the market. For example, a search engine could add a small factor to search results from "friendly" companies, and subtract a factor from results from competitors. This type of bias is very difficult to detect but could still have a significant effect on the market. Furthermore, advertising income often provides an incentive to provide poor quality search results. For example, we noticed a major search engine would not return a large airline's home page when the airline's name was given as a query. It so happened that the airline had placed an expensive ad, linked to the query that was its name. A better search engine would not have required this ad, and possibly resulted in the loss of the revenue from the airline. In general, it could be argued from the consumer point of view that the better the search engine is, the fewer advertisements will be needed for the consumer to find what they want. This of course erodes the advertising supported business model of the existing search engines. However, there will always be money from advertisers who want a customer to switch products, or have something that is genuinely new. But we believe the issue of advertising causes enough mixed incentives that it is crucial to have a competitive search engine that is transparent and in the academic realm."
Voice recognition on the other hand keeps getting better because there's no tangible benefit for someone to game it.
Perhaps, crawling the web is the worst way to go about creating a thinking AI
Incidentally, i think the solution to web search is peer review: websites ranking other websites, and having themselves punished when they mis-rank (which is what pagerank was originally)
Knock on wood! It is creepy to imagine a world where computers have the upper hand on vocal inputs but I already sometimes feel this way with text and autocorrect...
Just disable autocorrect. The amount of times where people communicate and a typo is a critical problem is aproximateley 0, you can manually correct them at those times.
I get whenever I'm using shorthand or weird acronyms or technical jargon that it might get confused, but when a word I type is both a) a real word and b) something I would use in every-day speech just leave it alone!
Friendly note: this may be a "negative" (i.e. bad) effect, but it is not a negative feedback loop. A negative feedback loop is a part of a system that self-corrects back toward a stable position. https://en.wikipedia.org/wiki/Negative_feedback
Like other users have noticed, publishing a good recipe is not enough, you have to fill it up with useless fluff. and you have to make pretty URLs . And add meta the tags
This happens for tech advice too, like linux solutions etc
Adding positive/negative on it is unnecessary specificity
When it comes to organic, everything is free, so you try whatever hacks the algorithms, from keyword stuffed content to link spam etc.
So as time goes on organic search results will actually get worse and worse, and paid will get better/stay the same.
It might not actually be google who is preferring paid search results, it's just inevitable from how the system is set up.
If you Google the name of a UK car insurer with some likely keyword like 'claim' or 'accident', you get paid listing from people offering two 'services'. a) 'call connection', which means that you call their premium-rate number and they just put the call through to the right company's phone line while charging you per-minute. b) 'claim management', you fill out a form on their website, they submit it to the actual company's website, and take a percentage of your claim. Neither of these are illegal, Google has promised to not take ad money from the first type, but in practice don't remove ads fast enough to make it unviable.
Consumer-facing companies now, ludicrously, have to do their own SEO to make sure they appear top of search results for their own name, and some even pay Google for ads to themselves. But clicks are more valuable to scammers than to legit businesses, so the former can always outbid the latter.
I can imagine how competitors are going to rank each other.
Everything is about reputation and trust. If reputation is solved then many other problems become easy.
They might be getting better but I don’t think they’re anywhere near good enough to warrant how common they’ve become.
At times I feel like they’re among the most inhumane technology that we suffer through because it can save their deployer a buck.
I see zero reason Apple can’t afford to have a person answer the phone.
This is interesting because I had a class were we all had to write a paper. I received a grade for the paper but it was never graded by the professor. We all had to rank 5 papers from best to worst, and our grade was determined by our paper ranking and how well our ranking matched others. It was pretty reliable
A funny example I saw recently is searching for "can men get pregnant".
Even when Google knows where I go with my searches, they still refuse to show those sites for anything that might have commercial interest. So much for customized search. I get tired of being the product. Can we get a subscription search engine that’s actually good? I’d pay for that.
no, reddit? i'm shocked /s
your implication seems to be that reddit should somehow be resistant to parasitic capitalism. reddit has been compromised and on the side of the advertisers since the beginning.
reddit today is very different from reddit of early days, its practically unrecognizable.
Stop saying that! The more people repeat that on HN, the more we risk advertisers realizing it and doing reddit-targeted SEO and reddit will be useless too!
(I'm only half joking)
Also, it's funny, I was just reading the thread about malaria eradication and DDT-resistant mosquitoes; the problem of SEO is eerily similar (any countermeasure is eventually defeated by evolution).
That said, a paid reddit would be a ghost town. It's against the spirit of the site.
Users: Now if I want good results, I append site:news.ycombinator.com before every query.
Advertisers: Hmmm...
Ditto for wikipedia and wikitionary, which are hopefully immune to advertising..
Reddit works when it isn't a competition, but advertising has made it a competition with a profit motive.
You need to stick to the topic oriented subreddits. r/homelab for example.
For consumer goods, I find that trustworthy Youtube channels, like America's Test Kitchen, do a much better job with reviewing things anyways.
https://en.wikipedia.org/wiki/Evolutionary_arms_race
The same thing happens in our societies. It's a constant never-ending struggle.
inurl:forum|viewthread|showthread|viewtopic|showtopic|"index.php?topic" | intext:"reading this topic"|"next thread"|"next topic"|"send private message"
They don't, because low quality consumerist search results make the ads fit right in. Also those sites are more likely to contain Google Ads themselves.
The search engine is fully optimized for consumerism, not finding the most relevant result.
One one hand I completely agree, and hope for paid service alternatives to all these "free" products. On the other hand, I can find everything I need through Google, as is, so why pay for an alternative?
Further, meaning is an ever-evolving, ever-mutating process much like the universe.
Training on symbols cannot arrive at meaning, since the meaning isn’t contained in those symbols. Using past symbols, also means no room for evolution.
Machine learning does work though in areas where the needs of the end goal are densely present in the symbols being used for training.
Like recognizing text. We learn to recognize those marks from just the marks, and nothing else. And so those marks contain all that is needed to recognize them. This can be encoded/learned.
But what they mean isn’t encoded in them, nor is it in words, in sounds, in facial expressions, in tones, in body gestures. It may even lie in between us, rather than in us.
It ties in with deBord's "Society of the Spectacle" [0], a theory that societies evolve from Being, to Having, and ultimately devolve into merely the Appearance of Having.
Your point about the symbols being tools for communication (as opposed to the ineffable ideas being communicated) is also echoed in Lockhart's Lament [1].
[0] https://en.wikipedia.org/wiki/The_Society_of_the_Spectacle
[1] https://www.maa.org/external_archive/devlin/LockhartsLament....
Edited to add that this was extra thought-provoking:
> "It may even lie in between us, rather than in us."
The best I can decipher is that your argument begs the question, ie you're saying machine learning can't derive meaning from symbols because machine learning doesn't derive meaning from symbols.
Your brain probably made a simulation of girl wearing shoes, jumping, water splashing, wet shoes. You need to know water makes things wet, people can jump, there is gravity. The motivation behind jumping was fun. Humans like to have fun. Wet shoes is not healthy. A mum is the mother and care taker for a girl.
When we're comprehending reading, we're building 3d simulations in our mind and deducing a ton of other info. AIs can't really do that just yet. They don't understand the world like we do.
I interpret this in a similar way. Reading F=ma means nothing, even if you know what each letter stands for, is understanding the consequences what gives you meaning.
yet you have Roget's thesaurus, where words are grouped into sets of words that have a similar meaning; it might make a difference when dealing in terms of these categories, instead of dealing with word instances; i mean if you have a dependency graph of a sentence, and the nodes are the corresponding sets in the thesaurus, then you might get a similar, but more general interpretation of that sentence.
(i once had a project for parsing/representing Roget's thesaurus in python: https://github.com/MoserMichael/roget-thesaurus-parser There are probably other projects like that around)
site:greatbritishchefs.com site:greatitalianchefs.com site:seriouseats.com
WARNING: the last one usually has 2 pages listed for each recipe - the recipe itself, and also a "story" blog post - except that the "story" is also very useful (not "someone's life story" which is basically spam), because it explains the reasoning behind the method, different alternatives and the trade-offs between them, and the experimentation that it took the author to derive itAnyway, it's pretty easy to turn this into a keyworded bookmark on Firefox for some quick searching:
https://duckduckgo.com/lite/?q=%s+(site%3Aseriouseats.com+%7C%7C+site%3Agreatbritishchefs.com+
%7C%7C+site%3Agreatbritishchefs.com+%7C%7C+site%3Athespruceeats.com+%7C%7C+site%3Abbcgoodfood.com)Note that most recipe blogs are fake. Someone wrote the original content long ago. Then, someone else hired a copywriter off a freelancing platform to change the words of the text just enough to avoid copyright violation, then put up a new copycat website loaded with SEO and ads. Look closely and notice how the author’s bio says e.g. "Born and bred in Lousiania and I love to share southern cooking", but the text contains grammatical mistakes typical of Eastern Europeans or Southeast Asians.
The ecosystem is already so advanced that new copycat recipe websites are often based on previous copycat recipe websites.
Is this true? I have a hard time believing it because it's a pretty naive assumption. A site that lets me get what I'm after quickly is generally going to be what I want.
How many people keep open 50+ tabs, or open a page and walk away?
I'm sure that time spent on a page, if it's given any weight, is given barely any. I'm sure returns to search to click a new link is a much higher indicator that the user didn't find what they wanted.
I found Josh Weissman to have good recipes. You could also look for recipes by experienced chefs like by Munchies.
Also if you have more experience. You can look at a recipe and easily figure out if something is very wrong. Unfortunately, most things on cooking don't tell you how to "debug" recipes. You just cook often and hope you can figure things out.
1 cup of flour, 2 eggs, 1 cup of milk, 2 tbsp of sugar, mix until smooth is a shitty pancake recipe (I think, it's close), but there aren't many ways to say that that makes it novel. Recipes are essentially instruction on how to build food.
Where copyright comes into play for cookbooks and recipe sites are presentation. And that includes the stories. So while the recipe itself doesn't enjoy copyright protection, writing "In the early autumn morning, my grandmother enjoyed making the family the most delicious pancakes, she started by going out to the chicken coop and sticking her whole hand up a chicken's ass to get only the freshest eggs possible..."
This recipe: https://wildwildwhisk.com/basic-buttermilk-scones/
Turns into this: https://www.justtherecipe.app/?url=https://wildwildwhisk.com...
I think there are Firefox and Chrome add-ons/extensions that do similar things.
And of course the verbiage is a lot of long tail keyword inclusion (pandering to Google by talking about your diabetic grandma, etc).
EDIT: The whole "I can't stand recipes that have verbiage" diatribe is a bit of a beggars being choosers thing. Personally I pay for America's Test Kitchen and get trustworthy, concise recipes (although there is a narrative about different techniques and options), but most people are too cheap but simultaneously super demanding about the things they get for free.
There is, and has been, a very well known, nationally respected organization who does this, for years! Their recipe recommendations are usually good, and the narratives actually ADD to the recipe, by offering alternatives and reasons behind choices made, instead of just telling a nonsense story.
It's not for SEO, it's to demonstrate engagement to the advertisers. Scrolling past the markov fluff they give you to read "while you wait for that to come to a boil" counts as engagement.
Both are open source recipe sites meant to combat the blogspam and load quickly.
> Larry Page and Sergey Brin were originally pretty negative about search engines that sold ads. Appendix A in their original paper says:
we expect that advertising-funded search engines will be inherently biased towards the advertisers and away from the needs of theconsumers
and that we believe the issue of advertising causes enough mixed incentives that it is crucial to have acompetitive search engine that is transparent and in the academic realmToday they say the same thing, but nontechnical users can no longer distinguish ads from their organic search results.
Last year they made a change that made them so indistinguishable to the untrained eye that even Google rolled them back: https://techcrunch.com/2020/01/23/squint-and-youll-click-it/ https://www.cnbc.com/2020/01/24/google-will-iterate-the-desi...
But no, generally, they don't notice that. Studies have been done demonstrating how few users can tell the difference between an ad and a search result. In one study, half of respondents "didn't spot" ads in Google search results at all: https://www.123-reg.co.uk/blog/seo-2/how-google-is-profiting...
That study was in 2013, when Ads were much more obvious in search results than they are today. By 2018, the statistic of people who couldn't identify search ads on Google was up closer to two-thirds: https://marketingtechnews.net/news/2018/sep/06/two-thirds-pe...
Good thing that has no cost, then. Especially with how poorly optimized all the javascript on these pages is, slowing loading to a perceptible crawl (unless, perhaps again, if you’re in the bubble where every machine you use is latest flagship android or a workstation, blind to the experience of the less ‘sophisticated’ users).
However, I went with showing studies because unfortunately, our personal real world experiences rarely win online arguments. ;)
Notably, ads in-line with results. I suspect that was the first move that sent them irrevocably down the evil-path. Watch a non-tech-geek use Google and you'll see why—I bet ad-clicks went up 10x with that change, or more. Then they began to serve ads beyond those early "non-evil" clearly marked text-only ads. That led to them making tons of money from webspam sites, while also putting tons of effort into fighting webspam, and some time around '09 or so they realized they should lay off the latter, on account of the former.
And here we are. Google search is worse, Google advertise deceptively on purpose, and the whole web is overrun with webspam.
1. One search engine w/ hundreds of thousands of employeers & bajillions in revenue
2. One search engine w/ like 3 people & a chonky Patreon account
It was the need to make even more billions that led them down the slippery slope of mass surveillance that makes everyone uneasy now.
The ads seemed relevant. I actually clicked on them with a sense of purpose!
At the time AdWords truly seemed how advertising should be done ethically with an iron wall between search and advertising.
That wall started crumbling pretty quickly by mid 200Xs.
Maybe the wall was never there?
Or at least that's what I would do, sitting on a mountain of dollars.
Of course when bad things happen the usual tendency is to attribute too much power and assign too much blame to a few individuals. But if anything the reverse seems to be true here. Page and Brin largely don't have the twin "if I don't do it I'll be fired" and "if I don't do it someone else will" excuses that others tend to have. Yet there seems to be an ambient belief that they shouldn't be expected to restrain Google, that their (at best) abdication of responsibility is somehow inevitable or proper. What makes this really disgraceful is that Google probably got where it today is in large part because in the past people bought, and Google actively sold, the idea that Larry and Sergey were nice guys who could be could be trusted to use their controlling stake to do the right thing. However stupid it was to ever trust in that idea, Page and Brin are not justified in abusing that trust now.
So, we talk about news websites for whom it is crucial to be well-indexed.
Talking about non-US news websites:
1. Not so many news websites even have a sitemap
2. LD+JSON meta tags are not so common either
3. OG metadata can be simply wrong
4. For many websites it's impossible to detect which timezone it's published time is
5. Publish time can be literally "5 hours ago" without a timestamp/date tag. Like no other clue on when it's been published
Given all that, until situation changes, I think Google has a real advantage as they can use expensive AI to parse the unstructured content.
So yeah, when Google says "forget metatags" they know something. Metatags will simplify lots of other search engines.
There're new "search engine" startup I hear about every week.
It feels to me that we started calling ML “AI” when deep learning became powerful enough to work on less clearly structured problems like vision and NLP—but “find patterns in complex data” does not seem powerful enough for what I would consider “intelligence” (artificial or otherwise). I don’t think that stateless/idempotent ML is capable of what most people would recognize as intelligence; in part because I suspect a history is required for a system to be self-correcting over time.
That is not the hard part.
The hard part, is finding a dataset that is amenable to the training process. Or at the very least, determining if a collection has any 'educational' value at all.
Then, by the time you get to that point, you're probably already looking at metadata, or a simple pattern that could possibly be encoded in a database query.
I have some faith in some of these new discoveries (Alphafold is a remarkable case study), but in many cases, the effectiveness of ML seems to be overstated.
But isn't that at least partly due to legal issues, like the pen-register precedents in the US?
I actually didn't know Tesla cars relied on GPS + map data w/ speed limits until recently. What a disappointment, apparently it causes all sorts of issues with sudden braking/acceleration when the map data is wrong.
EDIT: Here's info on it. https://www.dowhonda.com/2017/11/17/traffic-sign-recognition...
How sure are you?
>Daniel Dunn was about to sign a lease for a Honda Fit last year when a detail buried in the lengthy agreement caught his eye. Honda wanted to track the location of his vehicle, the contract stated, according to Dunn — a stipulation that struck the 69-year-old Temecula, Calif., retiree as a bit odd. [...]
>There are 78 million cars on the road with an embedded cyber connection, a feature that makes monitoring customers easier, according to ABI Research. By 2021, according to the technology research firm Gartner, 98 percent of new cars sold in the United States and in Europe will be connected, a feature that is being highlighted this week here at the North American International Auto Show in Detroit.
>After being asked on multiple occasions what the company does with collected data, Natalie Kumaratne, a Honda spokeswoman, said that the company “cannot provide specifics at this time.” Kumaratne instead sent a copy of an owner’s manual for a Honda Clarity that notes that the vehicle is equipped with multiple monitoring systems that transmit data at a rate determined by Honda.
https://www.washingtonpost.com/news/innovations/wp/2018/01/1...
> How sure are you?
Pretty sure. The Honda Sensing display of the local speed limit changes on the dashboard at the instant you pass the sign, it's too exact to be using GPS. Nobody has geocoded the location of every sign in America, that would be too much work. Also, it's often wrong but in ways that you would be wrong if you were just reading the signs. For example, it will switch to 55 MPH speed limit in a 70 MPH zone on the interstate after it sees a sign that is intended for trucks only.
I think the clearest indication that Honda Sensing does not use location data for speed limits is that Honda Sensing is available on cars that don't have GPS at all.
As far as the stop light data, that’s fed into Audi through a select number of state DOTs (mine is one of them). It’s almost magical that it can tell you when the light will turn green.
On French highways there are context-dependent speed limit signs which look exactly like normal signs but with a small additional panel below which tells you when they apply. Eg, big 70 in a round red circle, with a small picture of a car towing a trailer below. Eg https://external-content.duckduckgo.com/iu/?u=https%3A%2F%2F... which obviously only applies to cars with trailers.
Google maps (on my phone, not built-in to the car in any way) would constantly tell me that the speed limit of the road was that of the last such sign regardless of the specificity.
Further, to the linked tweets and the OP, I don't think that there's a direct line from where we are to where we want to be. As an analogy, no advancement in chemical rockets is going to get us to Alpha Centauri.
The bitter lesson.
A client of mine is continually having their listing on Google Maps "helpfully" updated by Google to be wrong. They change the services, they change the hours, all to be wrong. They added a new services section, duplicating their existing services and adding back ones that had been removed post-COVID.
There is no way to make it stop. They show the changes in low-contrast yellow and through dark patterns make it difficult to revert the changes. All they can do is check it daily and revert the changes one-by-one.
I'm trying to get API access so I can automatically revert changes that weren't made by the business. It requires manual approval which takes 2 weeks. Months later, I haven't heard back.
Because "machine AI" has NOTHING much in common with how our biological brains work. And we aren't smart enough to know what intelligence really is when we can't even define it for ourselves or in animal models.
And 99.999% of everyone working on machine AI has never taken a biology class let alone a class related to anatomy, neurology or experimental psychology so it's nothing more than "flinging shit on the wall and hoping it sticks" in terms of odds of success!
Not that that would help because academia as it exists today frowns upon "getting out of your lane" or "challenging orthodoxy" so "knowledge hybridization" of two distinct silos is strictly forbidden.
Basically the methodology of AI today is NO DIFFERENT than AI 1.0 from the 1960s and 1970s which was based on the assumption that all intelligence was merely predicate calculus and a fact store.
The scientific and economic model was for that AI (and is still the model for AI today!) is nothing more than the Garden Gnome Business Plan:
1. Create a cute-but-sellable singular heuristic technique misnamed and misinterpreted as "intelligence" in the small
2. Sell the idea to implement the same thing 1000x, 1 000 000x, etc. in parallel
3. ????
4. Success! We now how "strong AI" (which never comes because step #3 is bullshit and faith-based at best; fraud at worst)
The problem is that's also IDENTICAL to the plan to create a 747 jet by putting all the parts into a shipping hold and shaking with the expectation that you'll have a fully-formed 747 pop out when you open it.
Evolution is far smarter than us and has tested all the combinations) that take us generations to check. Evolution might well have taken longer per test but it's had a longer time. The best hope is to slavishly copy nature paying extreme attention to how nature pulls it off.
That's NEVER BEEN DONE with AI!!
So it's a VERY EASY technology to short in the long run because the fundamentals of assumptions and methodology are always such Epic Fail.
Further, I'd wager more than half of AI researchers are at least surface-level familiar with brain science, not (as you claim) less than 1 in a million (are there even a million AI researchers?). There's significant work between computational neuroscience, mathematics, philosophy, computer science, etc, etc in the field.
Many smart people are giving it their best effort to understand different pieces of the puzzle from many different viewpoints and angles; FAANG corporations might be among the most visible, but their AI is necessarily profit driven and close to the ground, relevant to currently tractable problems (amenable to 'mere statistics').
And in fact, slavishly copying nature is something which has long been on the AI back-burner, but we're on the order of at least a decade from being able to create a computer system with enough transistors to do so.
Not sure what else to say, really.
The reason this works is that there seem to be fundamental principles underlying information processing. Brains and CPU's are both systems that manipulate and store information, albeit in very different ways. Hell, even single cells and slime molds are capable of rudimentary decision making.
So, the point is that we don't need to copy the brain. Instead, we just need to understand the principles of information well enough to build machines that can efficiently manipulate information and intelligence will arise out of that. Information theory and (by extension) statistics are the fields that deal most closely with this question, which is why they're used heavily by both neuroscientists and ML researchers.
A rough analogy is that we don't build airplanes that flap their wings to fly. Instead, we understand aerodynamics well enough to generate lift and thrust through other mechanisms that evolution can't find. Like jet engines.
Also, "slavishly" copying nature is insanely difficult. Biological neurons and brain tissue are extremely complex and poorly understood. Much of this complexity is likely incidental to information processing and would only hamper our efforts to build intelligent machines. Like, do we want our computer brain to get multiple sclerosis if the simulated neurons demyelinate?
1) metadata is helpful - not all of it, but some; 2) ML is obviously needed when metadata is missing, and metadata is missing very often; 2) Even when metadata is present, pure ML-based extraction often beats it in quality, with right ML models. A combination of ML+metadata fallbacks is even better.
Website creators often make mistakes providing metadata, they may misunderstand the schema and purpose of various fields, have metadata auto-generated incorrectly, etc. It is rarely about deceiving for the tasks we're working on (though it also may happen).
So, I don't see Zyte falling back to metadata analysis, ML models are already better than this human-provided metadata - but metadata is helpful, as one of the inputs.
We're going to publish product extraction benchmark soon, where, among other things, we compare automatic extraction with metadata-based extraction. In this evaluation we've got a result that ML + metadata is better than metadata not only overall (which is expected), but on precision as well.
I wonder if the reasons metadata is sometimes preferred are not related to quality, or to failure of ML approaches. If Google doesn't get data right, it is not Google's fault anymore, it is website's fault.
When people say Strong AI they often mean Artificial General Intelligence (AGI). Weak AI by comparison is an even poorer term. What is usually meant is narrowly applied AI, and even, just usually, application-specific uses of machine learning.
But these narrow AI systems aren't weak. In fact, we're using those narrow applications of machine learning for some powerful applications. They're just not AGI.
In this article, Strong AI is used twice: in the title and once in some passing remark in the article. In neither case is it referring to AGI specifically. As such, what is not meant is Strong AI in the way that is accepted but perhaps just "highly trained" AI, or machine learning with lots of data. Regardless, the use of Strong AI in this article seems unnecessary and gratuitous
A good article on this topic: https://www.forbes.com/sites/cognitiveworld/2019/10/04/rethi...
If there was no money to be made by ranking highly on search engines, I think the promise the article's talking about might have been fulfilled.
[0] https://rodneybrooks.com/the-productivity-gain-where-is-it-c...
I wanted to get some images of strawberries to help improve image model for recycling. I wanted normal pictures of strawberries, so I did a google image search. its almost all adds and stock photo attempts to sell images.
I am not really engaged with AI-research, but I follow the area with interest and my impression is, that if strong AI will emerge in the next time, then only by accident. I mean there are lot's of awesome advancements and for example I did not expect Go to be solved since years already, but still - I see no way from current tech, to a general AI, that can really understand things.
Or is someone aware of more groundbreaking research?
Just another case of wishful thinking by someone who doesn't like how the web works in practice?
It is no wonder metadata systems are ahead of other systems in semantic flexibility. If somebody gave you a whole bunch of weather simulation data you would have some big arrays laid out in accordance to their scale and expected usage patterns and then you would have some a graph of relationships describing that the content is barometric pressure sampled on a certain grid, etc.
Sometimes these "metadata" relationships are so privileged that they become code, I mean a "CREATE TABLE" in SQL can be "exactly similar to" some definitions in OWL even though one causes the physical layout of memory and storage and the other one only attaches meanings to some symbols that may or may not be in the graph without what OWL thinks.
The "production rules" systems that were popular in 1980s A.I. have improved by orders of magnitude because of RETE-type algorithms, hashtable indexes, etc.
Would using IP address as a feature in an online fraud risk ML model be another case of “metadata”? Speaking of, I don’t know what American Express does specifically, but I have enough experience in the fraud + ML space to confidently say that many/most of the big players are using ML in several places throughout their fraud systems, even if allow/block lists are also used.
I guess that’s my larger point: these aren’t exclusive things. Metadata is of course incredibly valuable. So is Machine Learning. They also work really well together. And sometimes simple things work best.
>> "When your elected government snoops on you, they famously prefer the metadata of who you emailed, phoned or chatted to the content of the messages themselves."
This is strictly false. They prefer to have the content, but it is illegal to record the content of conversations of "US Persons" without a warrant. However, from https://en.wikipedia.org/wiki/Smith_v._Maryland it is legal to keep a pen register of all numbers called, since this is not considered protected under the constitution.
This article was written by someone who isn't versed in the basics of what they are pontificating about.
The shittiness of google search has made me think about the value of a curated search engine, although everyone probably needs their own version. Maybe 'engines'. It would be cool to have one that looked exclusively at bonafide discussion forums generally, another could look at the library contained in libgen or sci-hub.
Maybe somebody has done a search that works the opposite way, something that makes use of google but blocks all the cruft.
No doubt it couldn't be too successful since the gaming would begin immediately.
Fascinating observation. Maybe the real value of AI is bootstrapping solutions to these public goods problems.
https://en.m.wikipedia.org/wiki/History_of_artificial_intell...
They also suggest that the way intelligence was created (natural selection, AKA the “blind watchmaker”) may be the fastest way to do it. And that trying to do it again, in computers, might also take a billion years because the problem is just that hard. But hopefully it is faster than that.
Human intelligence is likely to produce artificial intelligence much more quickly, if that is possible at all. Which is likely. Which is partly a matter of the semantics of the term. Whose fuzziness is the root of a smorgasbord of wishful thinking since 1956. Whenever AI came up against seemingly insurmountable difficulties, they just changed its definition.
It is quite funny to read for example books from philosophers with some interest in artificial intelligence from the 80s/90s like say Paul Churchland. (The Engine of Reason, The Seat of the Soul: A Philosophical Journey into the Brain, MIT Press, 1995)
The anecdotes are worth their weight in gold. Especially because they show what was considered artificial intelligence back than in contrast to today.
What human intelligence produces as artificial intelligence will resemble human intelligence in function, but not necessarily in form. Just as nature has brought forth flying differently than man.
Or as Prof.Dr. Katharina Morik, TU Dortmund, Germany once put it on a meetup I attended: "AI is when a machine does something that looks like only humans can do. Artificial intelligence is open as a terminology to accommodate the phenomenon of shifting capability."
You may notice the strong, let me put it mildly, ironic component in her description.
I am only a little more vicious in my judgment.
The seeing Watchmaker -> human engineering
"(We) wish the human some progress, obviously egotistical and delusional, on whatever the human makes next"
I believe this slightly poetic word salad references certain theological problems about the capacity of man compared to a Creator. The conclusion is that the future is uncertain and the human is flawed, but an intelligent reader can supply "hope"
https://en.wikipedia.org/wiki/The_Blind_Watchmaker
I thought it was common knowledge, but maybe I'm getting old.
Because biological evolution’s is oriented toward propagation of DNA rather than intelligence, we can ask what kind of evolutionary pressures lead to intelligence. Under what scenarios does higher intelligence lead to higher survival, and lower intelligence to lower survival? If the answer was “all the time”, then everything on earth would show signs of intelligence. An intelligent spider would have no significantly greater chance of survival. It merely needs to spin webs, eat, and reproduce in a tight loop, with deterministic responses to various scenarios — it needs instinct more than intelligence.
By understanding what kind of evolutionary pressures lead to the necessity of intelligence, we can evolve (train) ML directly for intelligence, skipping straight to the answer rather than showing our work.
The answer is found in the peculiar evolution of mammals, the only organisms that display consistent intelligence across all its species. Mammals are highly social, beginning from live birth to mammalian glands that feed its comparatively feeble spawn. Sociality is built into the bodies of mammals. And the most social animal on earth is Homo Sapiens.
From here I’ll just recommend “Consciousness and the Social Brain.” I’ve been beating this dead horse on HN for some time.
Just for a moment, consider your spider as the embodyment of a special form of intelligence.
And even dare to construct the term intelligence to include the web of the spider.
And not just as a tool of the spider.
As a problem-solving competence for a special problem category.
We do see high levels of intelligence in non-mammal species, like crows, but they tend to also be very social creatures. The main counter example I can think of would be the octopus.
I think that even if we crack "general intelligence" and can make something that can problem-solve and learn on par with an Octopus, that approach will not get us to human level cognition.
I personally do believe that you will need societies of AI agents to develop the culture software to achieve human level cognition. I think we greatly underestimate the value and complexity of the cultural OS's that allow humans to perform advanced cognition.
https://en.m.wikipedia.org/wiki/Other_Minds:_The_Octopus,_th...
I stopped eating them some 20 years ago. The time I realized, they are quite intelligent and probably the most tragic intelligence on this planet.
Given just such a short life time and most of it loners.
Yet their body and nervous system have such great potential.
Yes, I know it's sentimental humanizing, but I love Elora & Egbert.
On instincts: the neocortex is built on top of all the other more instinctual parts of the brain (vision, movement, etc). But this may just be an artifact of being embodied. Is there much difference between a RL agent that gets its data by processing pixels or directly from a simulator? I wouldn’t say there is, except that these non-pixel-parsing agents would have a hard time doing anything outside a simulator that was doing the processing for them.
Glorified simulations of non-self-aware optic nerves are really a lot more profitable.
This is a big pet peeve of mine. People think because they're reading the web that they know things or they're "up to date" on current events. I've gotten a lot more out of books than the news, especially in the last 5-10 years.
And I feel like that's almost universally true (bad books are still better!)
Ironically Google Scholar is one place that you will find some real information. But it seems to be de-emphasized now. The main Google results will take you to a paywall for a paper (IEEE, etc.) But if you go to Google Scholar, you'll find the PDF. But I'd bet many Google users don't know that, even the ones that would read a journal paper.
-----
Aside from that, this is a great article that makes a great point. Google talks about AI all the time but it still relies on basic user curation to understand the web.
I think that shows you that the value lies. If webmasters stop doing work, then Google has nothing to index. Similarly I view the rise of these awesome lists as a manual Yahoo:
https://github.com/sindresorhus/awesome
If Google was providing so much value, then these lists would be redundant.
In fact I think Google was bootstrapped off at least partially off Yahoo. Yahoo had all these human editors curating links. That was great information for a nascent search engine to piggy back off of. Now that Yahoo no longer does that (AFAIK), Google has to rely on incentives for webmasters to provide metadata.
-----
To add something positive, I think YouTube is really where there is interesting user created content. Google has done a good job of stewarding and growing that ecosystem.
I remember I used to type random keywords in to Google and see what comes up. It used to be something interesting; it no longer is.
But YouTube has that flavor now. I typed in "sardines" and got a channel of this funny guy reviewing all sorts of canned fish :) It feels more like the early web.
No.
Someday there will be societies with digital bodies, but that will have to be bootstrapped with primate body AIs.
All of these companies doing advanced AI vision detection, classification etc... they’re really hard problems. But the whole challenge will eventually become nullified when every single road sign, landmark, and the road itself are tagged internally with metadata.
Instead of trying to decipher how does this PNG of a speed limit sign translate into a number, the number metadata would be encoded in the sign.
Looking at it this way, having advanced vision AI tech is only a competitive advantage in the short term
Including every pothole, pedestrian, stray deer, abandoned shopping cart and fallen tree branch?
Not sure why reading road signs is used as and example of "extremely hard thing to do" - there are other way harder things that cars can't currently do. E.g. figuring out the right speed to negotiate a curve using computer vision alone (without relying on detailed maps/gps, ie "metadata").
Oh yea that's not going to backfire at all is it? Lemme just work out how to frig it to state the max speed limit is zero and create an automated car pile-up.
https://screenrant.com/wp-content/uploads/2016/08/coyote-pai...
This would give them time to get the AI vision detection algorithms figured out.
when you can create virtual sign map that every city/road maintenance HAS to update?
Who can do that?
anyway meant the same
Who's going to pay for that? The public? In order to provide returns to private shareholders?