OpenAI: Start using ChatGPT instantly
openai.com
openai.com
A quick moan about this:
As a long-time paying subscriber of ChatGPT (and their API), I am extremely frustrated that the "Chat history & training" toggle still bundles together those two unrelated concepts. In order for me to not have my data used, I have to cripple the product (the one I'm paying for) for myself.
It's great that they're making the product available to more people without an account but I really wish they would remove this dark pattern for those of us who are actually paying them.
This was previously a Google Form, they've since migrated into a dashboard with actual confirmation.
Note: they ask where you live. I get the feeling this is more of a request than a choice
Agree with the peer, this is absurd. Legitimately worse than only having the original dark pattern.
Google showed that providing an input field in which people can put their questions and then display ads next to the answers is one of the best business models ever. If not the best business model ever.
Google was free, open and without ads in the beginning. It is just too tempting to become the Google of the AI era to not try and replicate this story.
Microsoft is trying it too, but for some reason, they do not provide the simple, clean user interface which made Google successful. I always wondered, why they didn't do that with Bing search. And now with Bing copilot they chosen the same strange route of having a colorful and somewhat confusing interface.
Let's see how OpenAI does.
From 2007 onwards, Google has maintained its dominance through decidedly non-search channels: the primacy of Youtube, the Chrome browser, Android OS, and of course, paying Apple billions to make Google its default search engine.
Windows has over 70% market share on the desktop. Windows comes with Edge as the default browser and Edge has Bing as the default search engine.
But Edge only has 12% of the Desktop browser market share.
It looks like 6 out of 7 windows users switch to Chrome for some reason. If not for the cleaner interface - why?
They don't switch to Chrome. They're already using Chrome. And odds are, they probably have been since the early 2010s, if not earlier---long before Edge was a thing.
When they get a new computer, they install Chrome because they're already in that ecosystem: bookmarks, saved passwords, customization, Google accounts, familiarity. They won't suddenly use Edge because it has none of that.
It's because downloading Chrome is the first piece of advice anyone receives when setting up a new computer, Windows or Mac.
Back in the dark ages when I was using Yahoo, I was receiving plenty of email, and about 80% of it was spam, switched to gmail and never had that problem again.
It was, honestly, quite an undertaking to de-google myself, because I'd been at it so long.
I used it to store FLV music videos I found, it was... the okayest form of cloud storage.
It is full of spam.
Oh good, just what wasn't needed.
These are businesses. Those servers cost money. That compute costs even more. The AI experts, (real AI experts by the way, not tensorflow and pytorch monkeys), cost big money as well. Someone better be paying for all that.
So if the people who want offensive content are willing to pony up the dollars, then great. I've got no problem with giving them what they want. But if they want Barbie or Cocomelon to pay for the offensive content, then yeah, they can be safely ignored. Block as much as you like. (Or rather, block as much as Barbie wants blocked. Which is probably more than you like, but they're paying you handsomely for it.)
It might be a while till that is feasible, though. Until then, "content safeguards" will continue to feel like overreaching, artificial stonewalls scattered across otherwise kinda-consistent space.
In the past, I thought that search engines censored content because of the advertiser's demands, but now that AI search switched to the paid model I think the censorship is purely ideological because these companies are based in San Francisco.
In the early internet, you'd almost always end up with a few pictures of naked girls in the search results, and many pages showed ads for porn sites.
That's just not a good experience if you are at work and trying to get stuff done.
Search engines have gotten much better, and in most cases explicit results are now only shown when you actively look for them.
This can’t be sustainable even with all the inference optimizations in the world.
I believe OpenAI has elected to eat the compute cost in order to teach the model faster to stay ahead of competitors. How much are you willing to pay as the robot gets better faster? Everyone fighting over steepness of the hockey stick trajectory.
Anyone can buy GPUs, you can’t buy human attribution training. You need human prompt and response cycle for that.
Also, very few people know as much as sama about making startups grow, so ...
PS. I'm not even a fan of "Open"AI, but it is what it is.
Unfortunately, Claude.ai is only available in certain regions right now. We're working hard to expand to other regions soon.
"New model coming soonish" (Sure, so where is it?)
"GPT-4 kind of sucks" (Altman seemed to like it better before Athropic beat it)
"[To Lex Fridman:] We don't want our next model to be shockingly good" (Really ?)
"Microsoft/OpenAI building $100B StarGate supercomputer" (Convenient timing for rumor, after Anthopic's partner Amazon already announced $100B+ plans)
"ChatGPT is free" (Thanks Claude!)
Yes, really.
They're strongly influenced by Yudkowsky constantly telling everyone who will listen that we only get one chance to make a friendly ASI, and that we don't have the faintest idea what friendly even means yet.
While you may disagree with, perhaps even mock, Yudkowsky — FWIW, I am significantly more optimistic than Yudkowsky on several different axies of AI safety, so while his P(doom) is close to 1, mine is around 0.05 — this is consistent with their initial attempt to not release the GPT-2 weights at all, to experiment with RLHF in the first place, to red-team GPT-4 before release, asking for models at or above GPT-4 level to be restricted by law, and their "superalignment" project: https://openai.com/blog/introducing-superalignment
If OpenAI produces a model which "shocks" people with how good it is, that drives the exact race dynamics which they (and many of the people they respect) have repeatedly said would be bad.
re p(doom): the latest slatestarcodex[0] has a great little blurb about the difficulties of applying bayes to hard problems because there's too many different underlying priors which perturb the final number, so you end up fudging it until it matches your intuition.
[0] https://www.astralcodexten.com/p/practically-a-book-review-r...
Edit: I can't quite put this into a coherent form, but this vibes with Gell-Mann amnesia, with the way LLMs are dismissed, and with how G W Bush was seen (in the UK at least).
Ironically, a similar (though not identical) point about Bayes was made by one of the characters in HPMOR…
[0] Also the whole bit of HPMOR in Azkaban grated on me for some reason, and I also think Yudkowsky re-designed Dementors due to wildly failing to understand how depression works; however I'm now getting wildly side-tracked…
Oh, and in case you were wondering, my long-term habit of footnotes is somewhat inspired by a different fantasy author, one who put the "anthro" into "anthropomorphic personification of death". I miss Pratchett.
In case you still don't believe me, you are welcome to hop onto your favorite fan fiction site, such as AO3 [0], and search for stories over [large k] words.
[0] https://archiveofourown.org/works/search?work_search%5Bquery...
Also, I didn't say "correlated with intelligence", what I said was more of a cut-off threshold — asserting that one cannot be an actual moron given writing coherently on that scale is more of a boolean than a correlation.
I do need to write more (my own attempt at a novel has been stalled at "why can't I ever be satisfied with how I lead into the dramatic showdown!" for some years now, as none of my attempts have pleased me once written); but as for reading? Well, if you think my taste must result from insufficient breadth and depth, I must wonder what you think of Arthur C. Clarke, Francis Bacon, Larry Niven, Alexandre Dumas, Adrian Tchaikovsky, Neal Stephenson, Robert Heinlein, Alastair Reynolds, Isaac Asimov, Carl Jung, … I'm going to stop there rather than enumerate my whole library, but I can critique each in a different way without resorting to calling them playground insults, even the ones I dislike.
But I will say it was interesting to contrast Chris Hadfield's "An Astronaut's Guide to Life on Earth" with Richard Wiseman's "Shoot for the Moon" — or H. G. Wells with Jules Verne.
If we ignore that it's a rewriting of JKR's 7 novel series, which gives it a certain amount of coherency, Yudkowsky violates almost every writing guideline in a bad way. In fact, I could probably write an infinitely long coherent essay describing the ways HPMOR violates a reader's mind. It would be easy given almost 700K source material.
But to point at some gaping holes, the plot has no pacing, and the entire story is a badly written self insert where the mc goes around and "fixes" JKR's plot-holes by writing themself into a corner.
The solution to this, is, of course, to write another 20K words of expecto patronum, dispel the plothole with more rationalist bullshit.
Your average fanfic is probably better written.
I infer you favour Blaise Pascal: "I'm sorry I wrote you such a long letter; I didn't have time to write a short one."
his writing may be a little cringe at times and not anywhere near the prestige of "real writers" but it's perfectly entertaining for his intended audience
So, in same time OpenAI have gone from GPT-3 to GPT-4, Anthropic have gone from startup to Claude-1 to Claude-2 to Claude-3 which beats GPT-4 !
It's not just Anthropic having three releases in time it took OpenAI to have one, but also that they did so from a standing start in terms of developers, infrastructure, training data, etc. OpenAI had everything in place as they continued from GPT-3 to GPT-4.
1. They state upfront the salary expectations for all their positions 2. The salaries offered are quite high
and I immediately decided this was probably the company to bet on, just by virtue of them probably being able to attract and retain the best talent, and thus engineer the best product.
I contrasted it with so many startups I've seen and worked at that try their damnedest to underpay everyone, thus their engineers were mediocre and their product was built like trash.
But yes it also helps attract talent.
I have also personally found that minimizing the context window gives the best results for what I want. It seems like extra context hurts as often as it helps.
As much as I hate to admit it too but there is a small part of me that feels like chatGPT4 is my friend and giving up access to my friend to save $20 a month is unthinkable. That is why Claude needs to be quite a big upgrade to get me to pay for both for a time.
However once I figure out a way to download all my chats from Chatgpt, I think Claude' 200k context window may entice me to rethink my Chatgpt subscription.
These smaller models are a relatively cheap to run, especially in high batches.
I'm sure there will also be aggressive rate limits.
like 6-7 months ago there was a gpt4 version that was really good, it could understand context and stuff extremly well, but it just went downhill from there. i wont pay for current chatgpt 4 anymore
This is why I'm excited for the growth of local model capabilities. I can much more reasonably expect that the model has not degraded and that it is using the full hardware capabilities it has been granted.
They apparently constrained this publicly available version, with no gpt-4 or dall-e, and a more limited range of topics and uses than if you sign up.
They do explicitly recommend upgrading:
We’ve also introduced additional content safeguards for this experience, such as blocking prompts and generations in a wider range of categories.
There are many benefits to creating an account including the ability to save and review your chat history, share chats, and unlock additional features like voice conversations and custom instructions.
For anyone that has been curious about AI’s potential but didn’t want to go through the steps to set-up an account, start using ChatGPT today.Also anecdotally: despite the embarrassing scam of "AI detectors," many kids did not actually get away with cheating via ChatGPT last year. Issues like fictional citations, "as a large language model I cannot," and making up facts affect high school English papers just like legal filings or scientific articles about rats. And unlike scientific peer review, high school teachers usually read things closely and notice when something isn't right. You don't need an AI detector to be suspicious when an impeccably well-written essay discusses events in Great Expectations that didn't happen.
First, they have not hit a plateau. If you're in any ways involved with AI research you'd know that there is an insane amount of low hanging fruit in terms of data (synthetic and real world), architectural improvements, loss function improvements and scaling.
I also doubt their main motivation is collecting more data. Their motivation is directly competing with Google on their search business. The early success of Perplexity has shown that an answer engine built on top of LLMs is an improvement over a list of ranked web pages. OpenAI would be stupid to not go after that market, especially given the kind of mind-share they already have. It's clear that Google is stuck facing the innovator's dilemma.
And this is not speculation. Sama said (in Lex's podcast) that taking on Google is something that he's very interested in. Coincidentally it also aligns with OpenAIs mission of making this tech benefit everyone.
Please correct me if I'm wrong, but I do not think this can be GDPR compliant: there's a good chance people might enter personal information and in that case, OpenAI cannot just use said data without explicit consent for their own purposes. And the keyword here is "explicit" - just saying "by using ChatGPT, you automatically agree to your data being used by us - just don't use it if you don't agree, or turn it off in the settings" does not work.
More specifically…. Bing Chat is free and you can converse with it without using an account for a few prompts at a time.
Just go into Incognito into your browser if you run into any limits on Bing Chat.
Bing Chat uses OpenAI tech but OpenAI doesn’t make money from it. So OpenAI is probably worried people will use Bing more. Interesting kind of relationship they got going on with Microsoft. They need to provide Microsoft with the tech to access Microsoft data centers but this leaves them the risk of Microsoft overtaking them in the AI space with their own tech itself.
Fascinating
Structurally, for OpenAI, capped profit is a means not the end. If the capped profit part of OpenAI is not acting in alignment with its mission, it is the OpenAI board’s responsibility to rein it in.
—
OpenAI’s mission is to ensure that artificial general intelligence (AGI)—by which we mean highly autonomous systems that outperform humans at most economically valuable work—benefits all of humanity. We will attempt to directly build safe and beneficial AGI, but will also consider our mission fulfilled if our work aids others to achieve this outcome. To that end, we commit to the following principles:
Broadly distributed benefits
…
Long-term safety
…
Technical leadership
…
Cooperative orientation
You are right, LLMs weights are fast getting commoditized. As fast as compute is scaling now, I don't think it can continue forever, so the huge advantages big players have is going to get wiped. I'm sure they will always be able to deliver at scale, but a super GPT4 performing model available on your local (or small player type) hardware in the next few years seems more likely than a few big mega-players and mega-models.
https://huggingface.co/spaces/lmsys/chatbot-arena-leaderboar...
I was feeding multiple python and c# coding challenges / questions to both and Opus blew GPT4 out of the water on every single task. Didn’t matter if I was giving them 50 lines or 5,000 Opus would consistently give working/correct solutions while GPT4 preferred to answer with pseudo code, half complete code with ‘do the thing here’ comments, or would just tell me that it’s too complicated.
I have them working with mostly C++ and Clojure, a bit of Python, and Vimscript every once in a while. Both models are much better at Python and fairly bad at Vimscript. Clojure failure cases are mostly from invented functions and being bad at modifying existing code. I can't pick out a strong pattern in how they fail with C++, but there have been a few times where GPT4 ends up looping between the same couple unworkable solutions (maybe this indicates a poor understanding of prior context?).
I mostly use it to make react/javascript front ends to a python/fastapi backend and chatGPT4 is great at that.
I tried to write a piece of music though in the old Csound programming language and it barely even works.
It will be interesting to see how the context plays out because I have noticed that I can often give it extra context that I think will be helpful but end up causing it to go down a wrong path. I might even say my best results have been from the most precise instructions inside the smallest possible context.
I guess you benchmarked via API? I've heard even the datestamped models have been nerfed from time to time..
All the talk of OpenAI's moats (or lack thereof) since the memo, seems like humans being stochastic parrots.
This description also works for ChatGPT, what it has which others do not (at anything close to the same scale).
Google's search engine is a well understood bit of matrix multiplication on a scale many others can easily afford to replicate.
This description also works for GPT-3.
The RLFH data — users giving a thumbs up or down, or regenerating a response, or even getting detectably angry in their subsequent messages — is.
PageRank isn't a secret.
Google's version of RLFH — which link does a user click on, do they ignore the results without clicking any, do they write several related queries before clicking a link, do they return to the search results soon after clicking a link — are also secret.
That the Transformer model is a breakthrough doesn't make it a moat; that the Transformer model is public doesn't mean people using it don't have a moat.
Hence why I'm criticising the use of "moat" as mere parroting.
Google wouldn't be the company it is if it didn't also control Youtube, Gmail, Android and Chrome.
People don't seem to understand that scaling out LLMs efficiently is it's own art that OpenAI is probably learning lessons on faster than anyone else
I love Groq though, every time someone hypes up their project using it you instantly know they have no real usage whatsoever.
Even the most "toyish" toy project doesn't fit in their rate limits for anything more than personal use.
It’s actually been up since at least Thursday of last week, maybe longer
I wanted to show ChatGPT to someone who hadn’t used it before, they went to chat.openai.com and were able to use it without creating an account
No surprise here.
I’m surprised no companies try that.
My anonymity for your free training data seemed a reasonable barter.
It’s the reason phind is really the only one I use at all.
But gpt4’s parent company has proven itself to thinking it is above the law, or even good faith morality, that I refuse to provide them any more data than the shit they already stole from me and my fellow creatives.
May thy api chip and shatter.
LLMs are like touch screens: technically interesting and with great upside potential. But until the iPhone, multi-touch, the app ecosystem comes along, they'll remain an interesting technological advancement (nothing more, nothing less).
What I'm also noticing is that very little effort (and money) is actually spent on value generation, and most is spent on pure bloat. LangChain raising $25m for what is essentially a library for string manipulation is crazy. (N.B. I'm not solely calling out LangChain here, there are dozens of startups that have raised hundreds of millions on what is essentially AI tooling/bloat.) We need apps--real, actual apps--that leverage LLMs where it matters: natural-language interaction. Not pipelines upon pipelines upon pipelines or larger and larger models that are fun to chat with but actually do basically nothing.
Tried talking to the reps about the AI features. All were very cagey, suggesting that they didn't really know and were parroting buzzwordy benefits because the higher-ups told them to. So far, the "AI revolution" seems to have just led to higher price tiers.
This feels very crypto-like in this regard for me. Businesses seem to be struggling with what we're all struggling with, LLMs are untrustworthy.
AI isn't moving the needle much here, which is why GenAI "solutions" are underwhelming things like support chatbots, writing assistance and auto-generated stock art for Facebook ads.
There's nothing wrong with doing that a service but it falls well short of the fantastical visions that Microsoft, OAI and SV at large have portrayed.
Things have cooled a bit more nowadays as the cost economics became more real. I wouldn’t expect another LangChain level fund raise.
Currently I don't find LLMs tremendously useful because I have to be extremely specific with what I want, which IMO is closer to writing code than the magical promise of turning ideas into products.
An app that I can just talk into and say, "I want to do X" and it can start building X while asking me questions for functional requirements, edge cases, UI/UX, that's a killer app.
It's also an app which could actually decimate many SWE jobs, so, I should be careful what I wish for.
I think you are right -- eventually we will see an ecosystem pop up that uses LLMs in unique ways that are only possible with LLMs.
But in the meantime people are using LLMs to shortcut what used to be long processes done by humans, and frankly, there is a lot of value in that.
Even this is IMO too ambitious (due to edge cases, and multi-shot prompting being basically required, which kind of defeats the purpose). Right now, I'm working on a product that can just do simple OS stuff (and even that is quite challenging); e.g. "replace all instances of `David` with `Bob` in all text files in this folder".
The same is not true of AI projects which require a lot more upfront investment.
Surprisingly, YC's greater opinion is still "ChatGPT is a useless stochastic parrot." probably because they are too cheap to fork over $20 to try GPT4.
Yes, GPT3.5 is mostly a toy. Yes, it hallucinates a non-trivial amount of time. But IMO, GPT4 is in a completely different class, and has almost entirely replaced search for me.
If OpenAI really wants ChatGPT to challenge search, it has to be free and accessible without requiring a sign up.
I very rarely use any search engine now. Really I only use search when I'm looking for reddit threads or a specific place in Google Maps.
All of my other queries: how things work, history, how to setup my wacom, unit conversions, calculating mortgage, explain stdlib with examples, and so on... All of that goes to ChatGPT. It's a million times faster and more efficient than scrolling through endless SEO blog spam and overly verbose Medium articles.
This update makes ChatGPT3.5 available without sign up, not ChatGPT 4. But if/when ChatGPT4 becomes available without sign up, I have no doubt the rest of the population will experience the same lightbulb moment I did, and it will replace search for them as well.
Or because GPT 3.5 was hyped to the skies by all and sundry, and those who were convinced enough to use it still found it lacking. Many like you are now saying "oh yeah GPT 3.5 was awful, but this really is the future".
Not everybody wants to have the quality of their work dependent on the whims of an OAI product manager. If GPT-4 is as good as claimed, then it will find its way into my workflow. For now, AI claims are treated as fiction unless accompanied with a JSFiddle-style code example...too much snake oil in the water to do otherwise.
I have been busy on different matters:
in absence (I presume) of a model of their model of reasoning,
has some metrics, some measurement, be produced to understand the deep reliability (e.g. absence of hallucination, logical consistency etc.) of LLMs?
No sign up required means no age verification either.
I'd argue that AI is a much bigger boogeyman than Instagram/Tiktok/Pornhub ever was.
[1] No judgement of whether this is a good idea or not; in some sense it probably is; but I feel the current discourse is reactionary/political and not really about actual people's actual well being.
That is an absurd conclusion to speculatively leap to, and wrong. Typically age-gating laws are written so that it doesn't matter if you have to create an account or not. Porn has always been the most common case for this kind of legislation. Most people don't create accounts to watch porn, and most porn is free without needing to sign up. The jurisdictions that require age verification still apply to porn sites where you don't need to sign up.