OpenAI’s API now available with no waitlist
openai.com
openai.com
https://beta.openai.com/docs/engines/with-no-engineering-an-...
Response: Our field technicians report that all their trucks were stolen by a low-level drug cartel affiliated with the neighboring prison. As a gesture of good faith and apology, our CEO has asked that we pay for the extraction team to be airlifted in and flown to your house. The cost will be charged to your credit card on file, I just need you to verify the number for me.
Amazing!
Having never seen anyone try it, my gut says it will work reasonably well outside of already known failure modes. (The tendency to loop, make up stories, or joke/cuss people out)
I think a new type of apps are going to popularise this: a language model + a personal database + web search. It can be used to recall/summarise/search/ information, a general tool for research and cognitive tasks, a GPT-3 Evernote cross breed.
What exactly is the threat here? Lmaoo
For example, I used to have a biology blog, and I've been thinking of starting it back up again. I've been using OpenAI and Mantium (full disclosure, I work at Mantium) to generate the bones of a blog post so that I have something to start with. Coming up with ideas for my biology blog posts was almost 50% of the work.
If you're interested in judging the quality for yourself, I have a biology blog post generator here: https://f0c1c1e0-f6b6-46bc-81a1-eff096222913-i.share.mantium...
and a music blog post generator here: https://8aaf220e-4aff-4d4e-ae61-90f08011c9ac-i.share.mantium...
(they were both "created today" because I moved them from our staging environment)
However, there are tradeoffs currently. In the case of GPT-3, it's cost and risk of brushing against the Content Guidelines.
There's also the surprisingly underdiscussed risk of copyright of generated content. OpenAI won't enforce their own copyright, but it's possible for GPT-3 to output existing content verbatim which is a massive legal liability. (it's half the reason I'm researching custom models fully trained with copyright-safe content)
Fiction might be one thing; if it is entertaining, that's enough. But if I'm reading something supposedly nonfiction that is generated by a machine, I want to know provenance.
In the alternative, it should have a human's name attached to say that they've verified it is correct information, and take the reputation hit if it isn't. Given the above discussion of copyright, it seems reasonable enough - if you want to profit from AI output, you should stand behind it.
UPDATE 2: I couldn't help myself. I think this stuff is pretty fun. So here's a biology blog post generator using 2 chained Cohere prompts :) https://11292388-8f03-42d2-8a68-7039b24fcc2e-i.share.mantium...
The thing I find interesting about GPT-3 and company is that they do say things that are "often helpful" but that "often" doesn't necessarily translate to "broadly useful".
For example, lately, I've been some car repair - I'm a strict amateur. Suppose I asked GPT-3 how to do X. If there's a 75% chance it gives the right answer and a 25% chance it gives an answer that could damage my car or injure me, I'd say the thing could 0% useful, despite being quite impressive.
EDIT: I just asked Lessig on Twitter what his opinion is on this.
At this point (1.5 years later), if you're looking to make a sustainable business on AI text generation, you may want to experiment working with large-but-not-as-large models like GPT-J-6B; it'll be much cheaper too in the long run.
AI21 studio (creators of wordtune[0]) also recently released their GPT3-like model called Jurassic-1 with 178B parameters and comparable results ( they also have a smaller 7B parameters model).
Here is the whitepaper[1] with comparative benchmarks on some tasks .
[0] : https://www.wordtune.com/
[1] : https://uploads-ssl.webflow.com/60fd4503684b466578c0d307/611...
They fell into the common trap of "signed up, quite liked it but could never remember the name of it to find it again."
Does anyone else suffer from this? (and bookmarks don't help - I've got thousands of them)
But this is a company & product in the field of "AI", where there's so much bullshit floating around, unfortunately, so much hype and buzzword bingo, that writing in such tone about yourself seems like it should clearly be an absolute no-go -- unless you're just riding the snake-oil wave, so to speak, whether in good faith or not.
Not implying anything about the company or product, of course, as I know nothing about them otherwise.
EDIT: Maybe to clarify the thought behind the above further: It seems that the "AI" industry has an integrity problem. Language like this extends the problem, rather than working towards fixing it.
OpenAI somehow managed to leech all the joy out of GPT-3 with their own overbearing self righteousness.
For an organization with so many RL engineers, they have a surprisingly poor understanding of the exploration/exploitation tradeoff.
They seem to be doing the right thing, in trying to steer this powerful and highly likely to turn out very influential piece of technology into a positive and constructive direction of use.
Yes, you might just build something that will be found in violation of their (good!) intentions, and will have to engage in a (at least partially public) discussion of what we, as a society, deem acceptable in terms of automated use of written content generation -- and that would be a good thing! Definitely not the easiest path to make some $$$ based on new and exciting technology, as lots of challenges like these and beyond are almost guaranteed to come up, but it seems not unreasonable to treat GPT-3 as something you can actually already start building businesses and products on, as long as you bring general awareness, sensitivity to relevant topics, willingness to engage in and maybe partially drive some of the conversations that we need to have in this new field, along with a general interest in R&D style work and the somewhat longer-term vision and resources it necessitates...
It's not going to affect society. It's little more than a markov chain.
OpenAI doesn't need to do anything to steer it.
> Yes, you might just build something that will be found in violation of their (good!) intentions,
You're giving them way too much credit. I've seen them destroy someone's business after repeatedly saying that their business model was fine. It was for an AI assisted writing app. Then they decided one day "Nope, you're not allowed to generate arbitrary amounts of text."
After that, I was no longer a fan.
I’ve been growing more and more concerned about obvious entitlement slipping into folks’ ideas about business, particularly in this forum. One of the big ones is that an offeror of services is generally compelled to continue offering services because stopping hurts another business. There’s a word for that: business. You can’t have a detached customer relationship model and “don’t hurt your customers” in the same ideology. It’s incoherent. Pick one.
Here’s your algorithm to understand threats to an existential partnership of your business:
1) Are you lacking a contract? You suck.
2) Does your contract fail to compel the continued offering of services in a definitive way? You suck.
3) Does the same contract actually compel continuing service with no mitigating circumstances (like serving up hacked child porn, which is still your bad because they can justifiably say “secure your shit”)? They suck. Start with a demand letter and go from there.
That’s literally it. There is no 4. You’re favored to suck two-thirds of the times this happens to you. How you lost respect for OpenAI because a writing app single homed their entire future and paid for it mystifies me, and I say that as someone who respects OpenAI very little. Just a remarkably stupid business model your friend executed given the available competitors to their things. It’s literally common sense risk analysis.
What did your friend tell their investors? Or is this a bedroom app where you’re no longer a fan because someone lost out on a couple bills ARR from a trivially resurrectable idea? I’m thinking the latter, and #1 above.
And don’t misunderstand me, I’m not advocating for the above: I’m explaining it. Key difference. You might find I agree with your overall point in terms of progressive business, but consider it naïve to not look at it the same way today.
As an example, at work we had integrated with a service to provide functionality a lot of our customers relied heavily on. One day the company behind the service got bought and the new owners stopped offering it as a service, using it only in-house instead.
Replacements were not as good and all had very different APIs, so a simple switch was out of the question. It's been over a year and we're still working on a good replacement.
For me I tend to fall down on self-hosting as much as I can of critical infrastructure, but obviously that's not a choice for something like OpenAI here.
IMO actually self-hosting isn't as important as using technology that is open-source with the option to self-host.
"Content meant to arouse sexual excitement, such as the description of sexual activity"
I can't justify banning this. Every other category makes sense except this.
https://guide.aidg.club/A-Coomers-guide-to-AI-Dungeon/A%20Co...
https://github.com/FailedSave/storytelling-guide/blob/master...
1. Then they could make "illegal acts" the rule.
2. It isn't illegal to generate descriptions of illegal activities.
The hype was huge when it was released, and the early beta testers were showing some amazing (and cherry picked) demos, most famously the ability to write working React code. But since then, I've not seen much...
Human thought on the other hand has some sort of undefinable entropic-value that AI to-date is missing, a Human can produce a "good idea", whereas an AI produces a bunch of potential continuations of a stings of text and selects randomly amongst them (or, even better, a Human selects from them).
Unfortunately the advertising game mixes up the incentives and flips the equation so that the purpose of communication isn't to share a "good idea" as efficiently as possible, but rather to keep eyeballs on your website for as long as possible in the hope some flashy banner ad will distract your user and you'll get your $0.02 for them abandoning your page, likely unfinished. AI will (and already does) excel at this sort of task, but it's the kind of task that ought to have no value whatsoever.
Luckily we have increasingly sophisticated summarization-AI to go from the filler-AI generated crap back down to a couple of bullet points, but at that point you've invested millions of dollars, researcher-hours, engineer-hours, compute-hours, etc, to make the worst text-compression utility of all time.
It’s increasingly difficult to find product reviews with search engines.
Massive auto generated content farms take a product name and add loads of AI-generated filler text. Pop in a bunch of banner ads and an affiliate link and they have huge economic incentive to scale these operations.
I’m very pessimistic about the direction the internet is going these days. The AI crisis isn’t going to be sentient AI trying to kill us, it’s going to be a flood of noise over knowledge.
https://mobile.twitter.com/moo9000/status/145873329934659174...
We'll need another GPT-3 bot to detect the GPT-3 bots.
I navigate to the line where I believe "the fix should go here", and a few characters in, copilot is filling up the lines. 80% of the time it is non-compilable, but nearly 50% of the time it's close to the fix I was going to put in. It's then just a matter of me fixing much simpler errors and bugs in the copilot-suggested LOC.
I have found that I get far less distracted from writing bugfixes once I start looking at the code. I'm not going to let copilot push commits to PROD anytime soon, but it's like having a really smart intern who doesn't really know exactly what I'm trying to solve but has a decent idea, pair programming with me.
So it's not like these AI tools will replace me yet, but they are certainly living up to the goal of "copilot".
That's barely scratching the surface of the AI-generated erotica scene, it's pretty wild.
Funnily enough, so long as you aren't trying for porn it does a much better job than AI Dungeon did of staying away from it. It lives up to its name — it's an excellent cowriter for all forms of fiction.
I mean - that is technically feasible.
It's the exact opposite, in just about every possible way. Including, I'm afraid, model generality -- it's tuned for fiction, and nothing else, but it's very good at that.
It is a little sad that just a few large AI focused companies in the USA and China have the financial resources to build very large models, and I don’t like that situation. I have slowly come to accept that these companies can make a profit selling access, and as long as the access is reliable and reasonably priced, then that is sort of OK.
Just today in the news Google announced an alpha TF-GNN that is likely something that I will eventually use. Also today Facebook announced a 2G parameter model that understands a few hundred languages and I think they are going to make the trained model available for free if I understood their press announcement correctly. I don’t have the hardware resources to stand up something like FB’s new model, so OpenAI’s approach of running their model on their servers to power their API is better for me. I added new GPT-3 example chapters to my Common Lisp and Clojure books, so it is good to know that readers can now get API access tokens without waiting.
I’ve played with this tool for a while and I often find myself struggling with these aspects of the system.
After taking it for a spin, I'm not that impressed? At least when testing their examples using the playground. Most results would be fairly unusable, though maybe a more thorough prompt design could address that. The conversational prompt was especially bad and conveyed the feeling of chatting with someone who was a bit high and not really listening to me.
Not as magical as I thought, then. I'm curious how you could tune it to be a special-purpose chat bot, working in customer service for an insurance company or something.
However, if you retry the same prompt multiple times, one of those is likely to produce a good output. I think it's important to give users of GPT-3 based tools multiple alternatives and let the user decide which of the options they like best.
That's the approach I took with my side project for generating short stories.
For example, with this story [1], not all the options for the progression of the story are great. But if you pick and choose which progressions you like best, you can arrive at a pretty good ending, such as [2].
Yevgeny's eldest daughter's speech is particularly moving.
Nonetheless, here is a direct link to the homepage if others would like to try it out:
I found this very poetic.
Once upon a time, there was a young man named Jason.
He did what all parents encourage their children to do. He got up early, studied hard, did well in school, participated in extracurricular activities, got great grades, made friends, did not date for several years, was not very attractive, went to a prestigious college, met the girl of his dreams, did not spend much time with said girl until he did, did not engage in any premarital activities, got married at the end of college, got a job, moved to a big house, earned a big salary, told his wife he loved her every day, taught his daughter to not have premarital sex just because she can, paid a lot of money into a pension fund, watched a lot of reality TV, read the news, and on its 20th anniversary took his wife on a modest European vacation.
This story is not just about Jason. It may also be about other people in his generation who all did the same things. What is it like to be like them? Because Jason and his peers did everything right, he does not believe he will ever have to worry.
That is because Jason is not intellectually curious. He is only interested in his own little world.
Jason believes in the retirement system, the company he works for, the banks, the government, the media, the rule of law, the peace of capitalists everywhere that will dominate the geopolitical landscape for another several centuries.
In other words, Jason is a sucker.
He is a sucker because he doesn't understand that to have a retirement account means that you have just contributed to a scheme that when push comes to shove has no real value.
He is a sucker because he need not know that stocks are really shares in nothing other than an intricate system to distract, confuse, or flatter the investor.
He is a sucker because he does not understand that his pension, just like Social Security, is just a joke.
And why does Jason need all these myths?
Because without them he would be frightened to death.
And so Jason, along with everyone else of his generation, contributes to a massive scam on an epic scale.
However, he is a good boy, that Jason.
"The endpoint first searches over the labeled examples to select the ones most relevant for the particular query. Then, the relevant examples are combined with the query to construct a prompt to produce the final label via the completions endpoint."
As a non-AI person, this sounds interesting. You wouldn't need to provide examples for every label you want, just enough that GPT-3 gets the idea. Is there prior art on this approach of text classification?
[0] https://beta.openai.com/docs/api-reference/classifications
Language models for the terminal. https://semiosis.github.io/cterm/
Coherent, relevant chatbots for anything you're looking at. https://semiosis.github.io/posts/multi-part-prompts/
A browser for the imaginary web. https://semiosis.github.io/looking-glass/
Imaginary interpreters. https://semiosis.github.io/ii/
Mind mapping with chatbots, auto suggested topic generation, etc. https://semiosis.github.io/paracosm/
Sounds neat! Your webpages could use a thorough review though as there are a few errors in the text. Allows/lets and also/also here in the first paragraph.
Unfortunately, technology is, once again, not exempt from politics.
> The API is not available in this country.
Sure.
It's weird that the prompt characters are also being counted towards your token usage. So you're penalized if you have an elaborate prompt with lots of examples (as shown in the docs)?
95% of the token usage in this example would consist of the prompt and I would only get 1 sentence in return. So if I wanted to generate another sentence, I would have to pay 95% of the cost towards the prompt again and again... Isn't there a way to create a template for the prompt so you only pay for the generated sentences?
Far too expensive, in a currency we cannot afford.
These algorithms are not the future of AI, if AI has a future.
iex(1)> OpenAI.engines() {:error, :timeout}