Bard can now connect to your Google Apps and services
blog.google
blog.google
The new feature for enriching outputs with citations from Google Search is also pretty cool.
I really want an agent that can help me with pretty simple tasks - Hey agent, remember this link and that it is about hyper fast, solar powered, vine ripened, retroencabulators. - Hey agent, remember that me and Bob Retal talked about stories JIRA-42 and JIRA-72 and we agreed to take actions XYZ - Hey agent, schedule a zoom meeting with Joe in the afternoon next Tuesday. - Hey agent, what did I discuss with Bob last week?
Something with retrieval and functional capability could easily end up being easier to use than the actual UIs that are capable of doing these kinds of things now.
Also, this part seems especially interesting:
> Starting today with responses in English, you can use Bard’s “Google it” button to more easily double-check its answers. When you click on the “G” icon, Bard will read the response and evaluate whether there is content across the web to substantiate it. When a statement can be evaluated, you can click the highlighted phrases and learn more about supporting or contradicting information found by Search.
The biggest problem with all LLMs at the moment is the frequency at which they are wrong (at least when they are used like an internet search to lookup factual info). Any LLM that can improve this (or as in Bard's case, make it easier to detect wrong info) is likely to gain traction.
https://web.archive.org/web/20230000000000*/https://staind.l...
https://www.theprp.com/2023/08/22/news/staind-premiere-their... (August 23)
>Citing production delays, Staind have now announced that their new album “Confessions Of The Fallen” will arrive a week later than intended. That full-length outing will now be available on September 22nd.
https://allmusicmagazine.com/staind-release-new-single-in-th...
Comparatively, ChatGPT says "I'm sorry, but I do not have access to real-time information, and my knowledge only goes up until September 2021. To find out the release date of a new Staind album, I recommend checking the official Staind website, social media profiles, or reputable music news sources for the most up-to-date information."
So which is more useful, one that doesn't even know there is a new album coming out, or one that knows what was its release date as of just a couple of months ago?
To semi-misquote Lewis Carroll: Which is better, a stopped clock or a clock which loses a minute a day? Carroll posits the former, as it is precisely correct twice a day. The trick, of course, is knowing for sure when those two times per day will be.
However, when it reaches 25%, I find myself meticulously verifying everything beforehand.
What questions are you asking where you don't care if the answer is wrong? I guess I just fundamentally don't understand what the point is. Why not just bookmark the "random article" link on Wikipedia if it doesn't matter anyway?
All of the LLMs are pretty good at this, and when it hallucinates a response it's totally a decent indicator that there isn't a lot of literature on the subject.
The idea of perfection is silly because it doesn't exist LLM or not. You're not going to get it so it's a matter how often it's right.
I also sometimes ask questions like "what is the tallest mountain in the US" or "what is the hottest desert on Earth" or similar. If I really need to know that the answer is correct, at a minimum it gives me a name to search for to verify height in feet compared to others, etc.
For instance, I was recently inquiring about a specific task with CMake and consulted ChatGPT. Initially, the response was inaccurate, but it was obviously so when it didn’t compile. Upon reprompting, I received the correct answer.
Hopefully my rate of error is way below 25%, so 25% error might not help me that much.
I'd like to think my rate of error is less than 5% too, but that might be close enough that outsourcing the work is still worth it.
Maybe still not worth it if I were that hypothetical programmer that writes lots of code but never writes any bugs.
This feedback loop, when you extend LLM to the horizon, is my primary point against this approach. When 90% of the new training data is from content (a worse version of) it previously generated you get a negative feedback loop to zero quality.
> For example, if you’re planning a trip to the Grand Canyon (a project that takes up many tabs)...
Is number of tabs considered a reasonable estimate of a project's size/complexity/scope/duration these days? If so, I'm wondering whether we could start using it instead of story points?
I just asked Bard for the date of an upcoming event and it did the search for me and found the right answer and summarized it with extra detail and references. This is the only reason so far that I'd go to Bard over ChatGPT.
It did treat the @Gmail part as part of the query words though, which is weird. I think it won't be ready for mass consumption until it can decide for itself when to search Gmail or Drive with no weird keywords necessary.
First steps, and I look forward to seeing future improvements. I wonder how they will monetize this? I was just using it with my free GMail account.
Both Microsoft, with Office 365, and Google have the customers and web properties that can make good use of new types of LLM applications.
Public LLMs like Bard generate massive amounts of marketing data.
Thanks to QWERTY keyboards, our keyboards are not efficient for typing either.
I'm sure things will get more esoteric, for the experienced computer user.
Person has a valid criticism, the answer is a vague, wishy washy “it’ll improve and be able to provide all things to all people in exactly the way they need it”
I don’t know what I’m supposed to do with that statement. What are the concrete steps from here to there? What’s the timeline?
They're already able to infer a great deal from context. They will improve simply because we've not hit any scaling walls yet.
There are absolutely serious limitations to existing AI, but the criticisms mentioned here aren't where we're stuck.
> AI will improve to infer more meaning from what users want
This is true. It already knows what it should do when I give it a piece of code and say "" (the bug emoji) or "spot the bug". If you give it the error code, it already knows it has to find a solution.
But for non-programming tasks, it'll be a challenge to infer what the user wants because language is vague and people are used to speaking ambiguously and clarifying only when the other person doesn't understand.
Natural language is a pretty brilliant way to interact with a computer if you want to learn a foreign language, but it’s probably not great if you are wanting to play a first person shooter game.
One thing is clear though: it’s not one size fits all.
Typing can be augmented with autocomplete. Some people like using a microphone, though.
I then used it to search for videos,
> There are several YouTube videos with a title that begins, "Gregory Blackstock, The Incomplete Historical World," List those videos with the YouTube URLs.
In this particular instance, the titles aren't unique, but the videos are different. The result appeared to be correct. The list was formatted to I can paste it into an email, something that doesn't happen when I search YouTube itself. Nice!
My next prompt was
> Make a list of the top 5 most viewed YouTube videos about moths.
Strangely, the result was 3 groups of 5 videos, with some duplicate results.
> Why did you list 15 videos in your reply when I asked for 5?
Reply: "I apologize for the confusion. I am still under development, and I am always learning.
I listed 15 videos because I found it difficult to narrow down the list to just 5. There are many beautiful and fascinating moth videos on YouTube, and I wanted to share as many as I could.
I will try to do better next time. If you have any feedback on how I can improve my responses, please let me know."
1st result on Google: https://developers.google.com/youtube/v3
To complete your requests, Bard will:
- Access and manage items from Google Workspace, like your emails and documents
- Share parts of your conversation, and other relevant info, with Google Workspace, which may be used to improve its services
That second line means that your conversation could be sent to humans for review, just by submitting a query that invokes gmail, drive, etc. This is in direct contrast to the terms of Bard itself which won't store, use for training, or allow human review of conversations unless you explicitly submit feedback about said conversations.
This amounts to "If you turn on Google Workspace any of your conversations could be seen by humans."
EDIT: someone else here mentioned that Bard results may be viewed by Google developers, so I might be wrong in my statement.
How many emails are in my gmail? (Completely incapable of answering this, didn't even try, just listed my last ~5 emails)
Based on my gmail account, how many flights have I booked in the last two years? (Also completely incapable, and didn't try, and again listed those same last ~5 emails)
When is the cheapest one way flight from NYC to Bali over the next six months? (Was for some weird reason only capable of returning round trip flights, but it did at least give me a plausible date and list of flights)
I'm probably using it wrong, but not a super "wow" first impression.
And based on their extremely loose privacy policy, I can just imagine Google pitching this to advertisers for "targeted marketing". "Hey Bard, give me an email template to manipulate tw04 into buying my product".
Also, I find it more than a bit disingenuous that the privacy policy on bard.google.com links to their generic privacy policy, not their BARD privacy policy. And after reading the real one, I understand why:
https://support.google.com/bard/answer/13594961?hl=en#your_d...
They will use all of your private data for advertising, and a human will review the data fed into bard. In other words, all of your private information is now reviewed by a human as they see fit. Yuck.
>Please don’t enter confidential information in your Bard conversations or any data you wouldn’t want a reviewer to see or Google to use to improve our products, services, and machine-learning technologies.
Maybe put that one front and center on your bard page, not buried on a completely different website....
FTA:
> If you choose to use the Workspace extensions, your content from Gmail, Docs and Drive is not seen by human reviewers, used by Bard to show you ads or used to train the Bard model
I’m as big of an AI skeptic as anyone but even I believe you can set up an integration of personal information that doesn’t leak publicly.
I've been using google products long enough to know that when a blog post differs from the policy language, the blog post always loses. Let me know when they update the actual privacy policy to reflect the blog post.
Believe it's US only which is why it won't work, but would love to give it a go. Did try a VPN but no luck still!
We launched a much more capable Gmail + AI assistant this morning here: https://news.ycombinator.com/item?id=37585990#37586627
We're using embeddings + vector DB + x-encoder + GPT4 to deliver a much smarter & capable assistant.
example of successfully retrieving a PDF: https://i.imgur.com/Y6cSlCx.png
Maybe it can connect to Google Apps but can't give reliable results.
> If you choose to use the Workspace extensions, your content from Gmail, Docs and Drive is not seen by human reviewers, used by Bard to show you ads or used to train the Bard model.