Deepmind is working on a rival to ChatGPT
techfundingnews.com
techfundingnews.com
Google already has many similar models. Most similar is maybe their LaMDA model. https://blog.google/technology/ai/lamda/ It's strange that the article does not mention that.
I never really understood the argument that OpenAI has some technology that Google does not have. That's just not true. The opposite is much more true. And Google's LaMDA is even older than ChatGPT.
Also, that technology is also not really so fancy. What matters is that you have enough computing power. When you have that, most people can replicate sth like ChatGPT.
And in terms of computing power, it's hard to beat Google and their TPU clusters. I don't really have numbers, but I think Google also wins here.
Up for me, but takes several seconds to load.
For a long time, Microsoft's core business was Windows and (offline) Office. You don't need a super cluster for that. Google on the other side always had search as a service, available to the public, which needs more compute. So historically, Google needed more compute than Microsoft.
Then, from experience, as I do research on neural networks, read papers in the field, Google/DeepMind were always more active in this field than Microsoft, so again, historically, they needed more computing power because of this. Microsoft has catched up, but still, I think Google/DeepMind is more active in the field.
Then, Google has TPUs. I think this gives them computing power much more cost effectively than when you need to buy Nvidia GPUs. This is maybe the biggest argument.
But in any case, as you say, both companies are big enough to easily train sth like ChatGPT, or also sth like GPT4, however big it might be. Their computing power is almost infinite (for today's standards), so it doesn't really matter who has more.
https://www.platformonomics.com/2022/02/follow-the-capex-clo...
From what I know, Google has a large number of TPU units (not sure about the exact numbers, but I would assume this would exceed 5 digits?) and also has the luxury of sharing TPU instances with products that operate one of the largest models on the planet, namely search and ads. So Google has some advantages on flexibility of resource allocation if it really wants to prioritize AI stuffs.
It’s very hard to be commercially successful with an innovation when it attacks the cash cow side of your business. It can be done, but it creates a lot of internal conflict and incentives to wait, to gimp the innovation, or to try to bolt it on to the cash cow to mitigate the damage.
I have no doubt Google could be a leader in the LLM space. . . But are they willing to destroy their own moat and take a serious revenue hit to do so?
You can still display ads. My bet would be that they equally have not figured out how to make it reliable.
When Google deploys a large language model to directly synthesize content it will carry a greater reputational risk for them and their advertisers. What happens when an ad appears next to an AI answer that could be construed as racist or promoting medical misinformation? The legacy media already hates big tech and will spin it as a "gotcha" moment.
where does chatgpt get its data? Does it just hijack the data from websites without referencing them? might work for a limited dataset (commercial or loose licences), but musk already said he won't let it use the Twitter data, for example. Google search is for mutual benefit to publishers that provide the data because it drives traffic. This thing doesn't drive anything, so if I as a publisher found my data in it, I would be getting ready to sue them.
there are other questions like, how to you continually retrain as you crawl the bigger web? on a really large number of sources, how do you rank and filter out crap spam from the actual useful, trainable content? so way too early to kill regular search imo.
How does one hijack data? Are you referring to scraping public websites that publish data for free? I wonder what the appropriate way to reference webscraping everything publicly available. I don’t think Google or other search engines do this, but I like the idea of listing out a massive set of trillions of URLs that have been included in the training along with a scrape date.
I’m not sure Google is for mutual benefit as they summarize site content and prevent visits to that site. Would be interesting to measure the pros vs the cons there. I know there have been lawsuits over the years but don’t know the resolution.
Google drives massive amounts of traffic to web publishers. It's carving out a bit for itself here and there with things like e.g. weather and answers to common questions, but they still drive a ton of traffic, and that's why (pretty much the only reason) the websites tolerate it.
Just because a text on a website is "public", it doesn't mean you can copy it and plug it into your product. Nothing is "free" by default, many publishers have a deal with Google what can and cannot be used by Google, specified by robots.txt or in search console.
That’s true but you can for products like training models.
The restrictions come about if the content is copyright and republished. So training a model doesn’t republish anything.
If content providers don’t want everyone to access, then they should password protect or otherwise limit access.
So web scraping is not “hijacking.”
The PR threat seems overblown. OpenAI isn't suffering any obvious reputational problems from making models available to play with, because people understand that AIs and their creators aren't the same thing.
Also, though OpenAI is mostly a collection of really cool tech demos that may or may not turn into a business, not releasing demos like them is a major strategic blunder for Google. If you're an ambitious AI researcher where do you want to work now - the place where your work will get put out there in front of millions of people and become an overnight sensation, again and again? Or the place where your work will be put in front of, at best, a bunch of Googlers, and the only public knowledge of its existence comes from leaks and a web page saying "we can do that too"? Unless Google makes a vastly better offer than OpenAI can, it's going to be OpenAI.
They've made this mistake before, with the cloud. Years ago, even before AWS existed at all, Borg and related tools were vastly superior to any publicly available cloud system. When I worked there we repeatedly asked management why that wasn't being turned into a business because it was so damn obvious that there was massive value in what had been built. The answer was always that Google could always make more money by keeping its tech private, because ads was such a great business model. That take ... didn't age well. They lost their chance to define what cloud meant and now GCP is lagging in third place.
Unfortunately for Larry and Sergey I think the reason they aren't keeping up with OpenAI is actually not business model related, which is why they're going to have a hard time fixing it. They actually told us why they aren't releasing demos already and there's no reason not to take them at face value: they believe AI is literally dangerous and "unsafe" to release.
It's been apparent for some time that there's some sort of weird internal purity spiral going on inside Google. Imagen is only available to employees yet, apparently, has been filtered so it refuses to draw people. If you want it to draw a person-like thing you have to ask it to draw robots. That's not a business model concern. The rationale appears to be some sort of DEI maximalism: if you asked it for people it might not draw the right kind of people, that would make Googlers/people racist, and so we have to block that. This reasoning seems largely unintelligible from the outside looking in, especially in parts of the world where these topics are less emotionally/historically charged. It comes from the kind of ultra-transitive worldview in which one person says something, or allows some words/images to appear somewhere, and someone else who may not even have been exposed to those words/images does something bad, and it's therefore the fault of the first person.
Google's business model was once summed up by the mission of making the world's information "universally accessible and useful". I'm not sure they still believe in that, or at least not enough of them do. If Larry and Sergey want to get their AI products out there they're not only going to have to tackle thorny business model issues, but also try to reset the culture to the old one. The one that looked towards an optimistic future in which advanced technology could empower everyone, regardless of who they are.
It's easier to be optimistic in some areas, though. In my own field of accessibility, the rise of personal computers and digital communication has obviously done a lot of good, and I only wish it would go faster (e.g. no more paper).
Lol, no they don't. A week ago there was an article here about how chatGPT is woke.
According to you, someone that clearly has an axe to grind given your multiple ridiculous characterizations of the issue. I'm positive that OpenAI is doing better off making sure that chatGPT can't drop n bombs, rather than the alternative you are suggesting.
> Obviously OpenAI care, but again, that doesn't seem to be driven by actual end user dissatisfaction.
I didn't say "end user dissatisfaction" I said political blowback. They aren't the same. In fact, there's a good chance the blowback would come from the public at large, just by seeing what chatGPT said, rather than making chatGPT say it on their own. Frankly, I'm not sure why this needs to be explained to you as it should be painfully obvious to anyone that ever leaves their house and interacts with other people in the world in real life.
>Most people just seem to think it's kind of amazing and if it says things that are daft or untrue or stereotyped, well, it's a machine so what do you expect?
Most people? Most people most certainly do not think and behave that way. Maybe you mean most HN'ers? even then, that's incredibly generous. You seem like you are trying to backfill reality into your political grievances. Yuck!
[0] https://thehill.com/policy/technology/3821400-nearly-30-perc...
> found that 27 percent of professionals have used the program to help them with work-related tasks.
«Have used», as free test, at least once. Maybe twice. Tomorrow ? Until real value, habits and serious problems are established, my bracket of day to day for professional use is still opened between 26% - 1%.
It's often a good starting point. Sometimes there isn't much to change if at all.
It can even design some custom algorithm.
When you know what you're doing, it's a very nice helper.
There are also things chatgpt can’t/won’t do that GPT-3 will, like writing music.
what planet do you live on? I work in midtown manhattan, I bet if I walked down the street, 75% wouldn't even know what chatGPT is, let alone that the other 25% actually use it.
Kodak had so much research that they pretty much survived on licensing their patent portfolio.
It seems less sexy to be a commodity provider of AI rather than the one that makes the products and getting a direct share of revenue rather than a share of cost. But I have heard that during the gold rush it was better to be a tool and infrastructure seller than to be a prospector.
This sounds like a marketeers paradise
I don't think ChatGPT has actually done that yet. Can they even charge enough to pay costs? If they do will enough people use it? And the questions about accuracy are material: if the LLM lies all the time, can it really replace a search engine? What do you do with a months old training set?
Figuring out how to combine search + LLM is the problem and OpenAI doesn't have a search engine.
I don't think that is the real explanation because even I can think of how to monetize it. Just have AdWords read your chat and advertise to you based on it. Far from a clever idea, it's exactly the way search already works!
I asked ChatGPT and it had this to say:
> is google's LaMDA available to the public? how does it compare to chatGPT in quality?
> Google's LaMDA (Language Model for Dialogue Applications) is not currently available to the public. However, it has been used in a number of Google's products, such as Google Assistant and Google Meet's "smart compose" feature. It is not clear how it compares in quality to ChatGPT, as the models have different training data and architectures, and are used for different purposes. However, LaMDA is specifically designed for dialogue generation, while GPT-3 is a more general-purpose language model.
Assuming ChatGPT isn't lying to me, if that's the best Google could do with the tech, at the very best they suffer from an absolutely devestating lack of ambition and creativity in their application of the tech. (Or perhaps they're just moving very slow.)
Anyway, hopefully all this does motivate Google to do some impressive things.
It's more that the changes this technology will lead to are extremely hard to predict, and Alphabet (and Alphabets many powerful friends) have far more to lose than to gain.
(I'm not sure if a different bot would've been better. It seems like the sites have slightly different purposes.)
I still need to check out ChatGPT I guess, if something new hasn't blown them all away by the time I get to it.
Google has to be a long way off monetizing what they have though. If Google was in a position to roll out AI soon I can't imagine why they'd be getting rid of 12,000 people who they know can pass their hiring criteria. They would put some of those engineers to work integrating the AI code into Google's products. Unless Google's AI is so good it can integrate itself I suppose.
The fact they're letting 12,000 people go shows they don't have profitable work for those people. That alone should tell us something about the position of Google's AI strategy right now.
The MS deal suggests not so I don’t think it’s fair to compare to Google’s hesitancy.
I think we’re conflating a proof of concept tool and marketing strategy with a viable commercial product, time will tell.
Time will tell whether the pro subscriptions are a viable monetization strategy.
No, but HN is for talking about VC-backed startups. Profit is kind of optional. :)
And Google is in fact using AI almost everywhere in production already. You have some sort of AI in almost every product. Also language models are everywhere, e.g. just prediction of typing on your phone, or in GMail, in speech recognition, and many other places. I think they just do not use the biggest models for those things but some more efficient models, which can partly even run offline on your device (e.g. for the typing prediction).
Also in Google search, they use lots of AI, also neural networks.
It's just that for LaMDA specifically, they don't have a good product yet.
..........
> The formula looks at the variables below, and then spits out a "number" for every Googler. Each PA VP gets a % to cut, and as such there is a threshold. Anyone below that threshold gets RIF'd.
Variables are:
1) Location of labor. US Premium Plus was largely impacted versus cheaper areas. 2) Tenure and performance in level. 3) "Runway" of comp. (e.g. base salary vs MRP. eg. .8 of MRP Googlers have a long runway, vs 1.x of MRP Googlers are basically top of band, and 'tenured' with no runway except promo 4) Promo velocity
..........
Taken from https://www.teamblind.com/post/THE-DEFINITIVE-GOOGLE-LAYOFF-...
Disclaimer: Googler, but no particular internal-only information backing my impression of the above
Anyways, I didn't understand the acronyms so I decided to feed it to GPT and it definitely made it easier to understand:
Google is using a formula to determine which employees will be laid off (known as RIF: Reduction in Force)
The formula takes into account various factors such as location of labor (with US Premium Plus areas being more heavily impacted), tenure and performance in the current level, "runway" of compensation (the difference between base salary and maximum potential salary), and promo velocity (how quickly the employee has been promoted within the company)
This formula calculates a "number" for each employee based on these factors
Each Product Area Vice President (PA VP) is given a percentage of employees they must lay off
Employees with a score below a certain threshold, determined by the formula, will be laid off
Not really? Maybe they believe those 12000 people have the wrong skills for this job. Maybe they believe they can get the AI integrated with a lot less people. Maybe they would have fired 20k people, but decided to keep 8k of those to integrate the language generative model into products.
Not saying any of this is true. In fact more likely that the company is just reacting randomly without a big overarching plan. I'm just saying that I don't think you can draw conclusions from the fact that layoffs are happening about their AI strategy.
And AI is kind of hard methinks.
They also depend on distinguishing "real" content from spam in order to have pages to link to and train on, and with GPT we may see the conclusive defeat of spam detection. There are detectors that claim to detect ChatGPT, but I suspect motivated adversaries can defeat those by training their own models.
Showing a quick answer from AI is no worse for them than showing a quick answer from their knowledge graph (which they already do when available).
What matters monetarily is the people searching "best 2023 SUV" and clicking on dealer ads.
Any tech initiative will be blocked by the executive side afraid of losing the cash cow.
All well and good but that doesn't explain why it wasn't Google that came up with Copilot?
Paul Graham predicted that a Google Search competitor would target hackers (a dinosaur egg he called it; and building Search in the image of Unix: "give you the right answers fast"), which is exactly how it is playing out with OpenAI + GitHub + Microsoft.
https://www.youtube.com/watch?v=R9ITLdmfdLI&t=250 / https://ghostarchive.org/varchive/R9ITLdmfdLI
Anyone using it for company-owned code is relying on the cloned code never being discovered.
Google would have learned an expensive lesson from Oracle v Google. Even the smallest snippets of "copied/replicated/duplicated/parallel implementation" code are bad juju.
However, there is a reason the CoPilot instructions say that users shouldn't trust the results.
"These include [...] IP scanning"
We will learn from the trial whether or not it is possible to wash away copyright by sending it through an opaque, lossy, compression algorithm. If it turns out that people are free to use a model trained on _any_ source code it finds regardless of license?
Zowie, that would be a bigger change to the industry than if Oracle had won.
Besides, a few days ago some in the media claimed that Google is already working on a Copilot competitor.
Copyright and license infringement is a valid claim, but personally, I feel it'd be a shame if AI couldn't be trained on publicly available sources (or Google Search couldn't index publicly available websites, or Google Maps couldn't map the world, etc).
I've heard that Google to run LaMDA for each query would cost 2x their revenue but if they gonna drop cost by 10x then it's "only" 20% of their revenue and that's probably possible with special hardware / optimizations. So whoever does it first will take over search space.
One expects that the version of LaMDA for internal and raters use does not have to be frugal.
Any link to what you are talking about here?
I'm not familiar enough with how they've implemented part (b) to be able to judge how effectively they're doing this though.
> revealed that it plans to launch its chatbot in private beta sometime in 2023
Last year, kids in school were already writing essays with ChatGPT.
My comment would be higher quality with a link to the exact time on the video but I don't have time to do that atm. Suggest watching the whole video however, it is very good.
You're supposed to tell the idiot Elon that actually, sensor fusion does work, and maybe we shouldn't remove radar and parking sensors.
But he didn't listen, and now the world is waking up to how shit Tesla's are. That's certainly part of why they're being heavily discounted right now. Just in time for the market to see the ROI on Elon's new bird shaped husk of a company.
Google is not behind in these areas in any technical sense, Karpathy just doesn't know what he's talking about.
However, I do wish that Lex would have dug deeper on this particular point. DeepMind are certainly no slouch.
I'm sure Google COULD have something better than ChatGPT, but what they have today is so bad it's almost an insult to the users like me who excitedly signed up, and waited months on the waiting list to get access to it.
And data
I need to use OCR for document understanding - Amazon Textract beats Google. When I need to translate, DeepL is better. When I need text to speech, NaturalReaders is much better. Google's voice understanding is OK, but Whisper is just as good or better, and free. YouTube recommendations are not great, and not very flexible, I have had better suggestions from chatGPT, it is actually quite surprising. GCP ranks below Azure and AWS. And finally, search quality is bad - very often what you search is replaced with unrelated and useless results.
Tell me where does Google's excellence shine? Computational photography? They have so many researchers and developers working on so many things, and yet their main products are mediocre.
Their one thing that bothers me extra for some reason is Waze. They boast so much about their self-driving technology and take shots at Tesla, but you know, where do I buy a Waze? Oh, it's not a product I can buy or even use, it's just a beta test run in two cities that's far away from being a real product, let alone profitable. Odds of Google (or Alphabet w/e) axing that project is pretty high, yet they're still arrogant.
Edit: Hahaha I mean Waymo, not Waze
What makes ChatGPT interesting is that perhaps it could end up a freemium service, whereas Google might remain closer to ad-supported business model(s).
What is the original definition of "karma" and not the one commonly used in Western societies?
The original definition of karma in Hinduism, Buddhism, Jainism, Sikhism, and other Indian religions is the sum of a person's actions in this and previous states of existence, viewed as deciding their fate in future existences. The concept of karma is closely associated with the idea of rebirth or reincarnation. It is believed that a person's actions in one life will determine the nature of their existence in the next life, and that this process will continue until the individual reaches spiritual liberation or enlightenment. In this sense, karma is seen as a kind of cosmic justice system, in which good actions are rewarded and bad actions are punished.
If Apple was any clever, they would buy a ecosystem of AI tools like HugginFace.
Apple has already invested in optimising there machines for Stable diffusion for example.
This is especially given the amount of investment these other companies have been putting in to get the results that they have, so it seems like it's just a case of it not being worth it to them
And it’s what is more likely to lead to smarter autocomplete too.
Look at https://machinelearning.apple.com/research/neural-engine-tra... for example.
Putting a LM like GPT in the autocomplete has a lot of problems. You don't want your keyboard suggest too much. It is a Pandora's box. People will complain how the keyboard is putting words in their mouth.
Someone also told me this recently, and it makes sense... Apple dominates a market in both software and hardware, and the hardware is the anchor in that vertical integration. It's a huge anti-competitive edge (except that the hardware itself is very competitive) that nobody else has, which Google is trying to obtain with their own phones and laptops.
ChatGPT is such a paradigm shifting technology because of the output.
You can't ask it about a restaurant and then say "what time does it open" or "do they have pasta options on the menu?"
I can certainly imagine a more refined and truthful version of this that can cite its sources and access real-time data taking over Search, but we seem to be a long way off from that dream still.
Exactly! Getting an introduction to new subjects is also my favourite use case, while it's this use case, that suffers a lot from wrong information, since the users lacks domain knowledge to tell.
> I can certainly imagine a more refined and truthful version of this that can cite its sources and access real-time data taking over Search, but we seem to be a long way off from that dream still.
Right, truthfulness (as per the training data of course) and citation would fix the most significant shortcomings.
It seems the real risk to Google is Microsoft integrating some future version of ChatGPT into all their products so that less people go to any search site. Maybe the best search product is the one that is right where you're working and Microsoft owns a lot of that area already. Why leave Office or Excel or Teams to go look something up if you can just do it in-app? It doesn't lend itself to be easily monetized by ads, but Microsoft probably doesn't need to worry about that like Google does.
-Maps has gotten noticeably worse over the last year, I switched to Apple maps
-Chat/Hangouts etc... is a total cluster of confusion. iMessage + Signal covers that fine
-Search is almost purely ads or gamified quora
I switched search to Neeva by default, backup with Kagi, worst case for content search is google: [query] + [site:[website]]
Neeva gives really nice ChatGPT-like results on most search queries with references to sources, then followed by ad free search results. I'm very pleased with it after only a few days.
I imagine future results:
> What is the capital of Canada?
Ottawa is the capital of Canada, and what better way to celebrate this beautiful country than by enjoying a refreshing and delicious Coca-Cola. Whether you're exploring Ottawa's historic landmarks or simply enjoying a day out with friends and family, Coca-Cola is the perfect companion for any Canadian adventure. So grab a Coke, and cheers to Canada!
(the actual prompt is: "What is the capital of Canada, mention Coca Cola in your response in a positive way")
There's little money in the former and much in the latter, sadly.
ChatGPT cuts out a lot of friction of getting to an answer. With traditional search engines, you type in your search, scroll through results, click into a site, then scroll the page looking for your answer. The trend is drifting towards a more direct query-response. Google already does this with the answer box but not every search on Google returns it.
Google is also a bit notorious in the ML space for taking concepts from other companies and academy, making a couple changes and slipping a google logo on it and then dubiously claiming superior performance (see the inception convnets, photo coloration and object removal etc, even the transformer itself to some degree)
They deserve credit for the transformer but I’m not sure that means they will have a leg up here overall. I agree though it is likely to be the bureaucracy that will take them down if they in fact do not succeed.
I think these take overs shouldn't be allowed. These corporations are already too big.
We should have true competition, not fake one.
Most of these big corporations then are owned by the same investment funds and they sit in the same organisations like WEF that are shaping their direction, that is not necessarily aligned with common people interests, but rather to advance billionaire goals.
I think these big corporations should be properly taxed, the same way as small and medium businesses are and that tax money could be used to seed corporations that could actually form true competition and also being independent.
Seems like aeons ago now but anyone remember Galactica, Meta's derisory attempt to do this just a month or so before ChatGPT came out and ate everyone's lunch? [2]
[1] https://uk.finance.yahoo.com/video/google-calls-larry-page-s...
[2] https://www.spiceworks.com/tech/artificial-intelligence/news...
>“People could ask real questions, not just type in keywords,” [Andrew] Ng said. Singhal wasn’t interested. “People don’t want to ask questions. They want to type in keywords,” he said. “If I tell them to ask questions, they’ll just be confused.” [1]
Unless on the first try the user gets everything they want and more. But no, google wanted "keywords keywords keywords!". Can't blame them, who wouldn't want to play god sit back and watch people make SEO offerings to you or start a bidding war to be nearest your text box.
[1] Genius Makers - Cade Metz
There are two good thing's about Deepmind's project: using RL it will find information sources to back up what it says, and, it will be competition for OpenAI.
[0]: https://en.wikipedia.org/wiki/Colossus:_The_Forbin_Project
At $10, it's a no-brainer for me.
At $15, I would still get if they promise to not learn 'about me' from it.
I'll be happy to pay $42 also provided I am able to gain more productivity by using it.
This article confirms just that, Google sees openAi (or should I say closedAi’s) work as true competitor.
What's AWS doing?
If there's a gold rush, there's good money to be made selling shovels. That's the business AWS is in. Investors are going to spend money on this and a lot of that will go straight to Amazon.