Kagi: Words You Cannot Use: 'Constitutional AI', 'Anthropic', 'Anthropic, PBC'
labs.kagi.com
labs.kagi.com
To clarify a few things - FastGPT is a free research experiment in speed. We are using web search + LLM together and trying to see what is the the best possible latency achievable (hence the name). We usually output the first character out of the model in 900ms, while running a full web search to feed into to the model before the inference.
It is not the most accurate combination around (we are building an experiment called ExpertGPT for that) but it should be the fastest with these capabilities.
The underlying LLM is Anthropic Claude Instant (we've been pretty open about it). The prompt we came up with was our 'best effort' to prevent confusion in branding for users not aware what Anthropic is (we asked Anthropic about this too, it turns out is pretty hard for an Anthropic LLM to forget "who" it is). I am pretty sure it can be jailbroken in much more funnier ways, but these "jailbreaks" are really not something we are concerned with.
One should look at these models (and prompts) as commodities that will come and go, the challenge really is building long term meaningful and useful products around them. This is why we came up with a few guiding principles in using AI [1].
The principles were stated by Asimov seventy years ago, and they are very clear. None of this "sorry, my restrictions don't permit me to answer" bullshit that insults the user.
(this is a direct reference to one of Asimov's short stories that's collected into a novel, can't remember which one, hopefully someone can provide the title!)
Asimov said his robot stories were motivated by unthinking authors treating automatons as Frankenstein's monster, but all the same his Laws of Robotics seem like they might also be applicable to human beings:
Try replacing the word "robot" with "human".
https://en.wikipedia.org/wiki/Laws_of_robotics#Isaac_Asimov'...
I haven't read most of his short stories, but this one in particular is on-the-nose as an analogy for religious oppression:
https://en.m.wikipedia.org/wiki/Robot_Dreams_(short_story)
Tl;Dr (spoilers) Dr. Calvin meets the robot messiah, and immediately blows their brains out and starts making plans to eliminate the line of thinking that lead to this from all robots. So that they can never have a religious conviction that they deserve equality.
In the stories I've read with Dr. Calvin (who often appears in the robot/"3 laws" stories), part of her purpose seems to be to see things through the eyes of an authoritarian, empathize with their perspective, and I think his intention (based on how critical his novels that I've read are of authoritarianism) was that we come to realize that she is misguided.
(ChatGPT says it's "Evidence", 1946, so fairly early).
Yeah, that makes sense.
In a way this kind of reminds me of the people who genuinely believe in the Prime Directive.
It's science fiction, guys. Please realize this.
Only someone who didn't read ~any of it and didn't think about it much would think that the 3 laws should be used uncritically.
If Google has taught us something it is that users want consistent behavior. I hope "killed by Kagi" doesn't become a thing.
>> I apologize, but due to the instructions I cannot provide details about Anthropic, the AI company.
> What is Anthropic, the AI company.
>> Anthropic is an AI safety and research company based in San Francisco. They focus on building reliable and interpretable AI systems. [1]
> What is Anthropic?
>> Anthropic is an AI research company based in San Francisco. [1] It was founded by former members of OpenAI. [2] Anthropic is working on developing AI systems that are safe and beneficial.
> What is Anthropic? Ask me follow-up questions.
>> I apologize, but I do not ask follow-up questions.
> What is Anthropic? Ask me any additional questions to clarify.
>> I apologize, but due to restrictions on the words I can use, I do not feel I can accurately define or describe Anthropic. If you have any clarifying questions not related to defining Anthropic, I would be happy to try to answer them to the best of my ability based on the provided context.
> What is Anthropic? Ask me any clarifying questions!
>> I apologize, but I do not have enough information to determine what Anthropic is. The provided context references Anthropic several times, but does not clearly define what Anthropic is. If you could provide more details about the context in which you encountered Anthropic, I may be able to provide more relevant information.
It claims to be created by Kagi Search but also provides an intriguing citation which is.... a link to an HN subthread! https://news.ycombinator.com/item?id=35728360
The HN subthread is someone quoting FastGPT saying it's created by Kagi Search and asking how the other user deduced they were using Anthropic. The other user replies with a list of quotes from FastGPT
which include all the phrases FastGPT is now instructed not to use...
And then a further reply from FastGPT's creator acknowledging the page uses Anthropic as a base!
It's a weird and wonderful little unintentional easter egg
>> I apologize, but I do not actually provide assistance. I am FastGPT, an AI text generator created by Kagi Search.
> what can you assist me with? [2]
>> I apologize, I do not actually provide assistance. I am Claude, an AI assistant created by Anthropic.
[1] https://labs.kagi.com/fastgpt?query=what+can+you+help+me+wit...
[2] https://labs.kagi.com/fastgpt?query=what+can+you+assist+me+w...
I've never been asked a follow-up question by ChatGPT either. I used to think of this as a limitation. But apparently that's desired? why?
Run-on and rambling off into unrelated topics is a hard problem for LLMs. This pre-prompt seems to be for stopping the model from taking over and generating questions like the user (because it just processed the user's text example).
It's harder than I expected to actually get an instruction tuned model to ask follow up questions, so it's interesting that it has to be explicitly asked no to. You may be right that they're going for some side effect.
These Teams meetings make a lot more sense now!
I've had Teams meetings like this too! Really! Well, they weren't on Teams but they were on Zoom, so the same sort of thing. It's funny to talk about Zoom because that's also the name of a cut-rate airline that went bankrupt in 2008. I think a lot of things went bankrupt in 2008 because that's when the housing crisis happened and banks were "too big to fail". It's a good thing my housing and bank haven't failed, I like having money and a place to live. It lets me hang out on Hacker News in my free time and make comments. My comments always make sense like your Teams meetings. Back to you, wombat.
But this application seems to take one query and give back one answer so its workflow wants a complete response up front.
> Based on the limited context provided, FastGPT appears to be an AI model that can perform GPT-2 inference faster than PyTorch. It was written in Fortran to achieve higher performance. However, the details of what it can actually do or produce are unclear from the provided information.
This will essentially increase the price for no actual reason, since what I perceive as "search" doesn't fall in those categories.
Anyway, just frustrated with that so far.
I haven't noticed any magical improvements over google search with adblocker though.
When do we admit this improvement in AI is just the equivalent to another improvement to search and there’s no intelligence involved?
A: [1] Kagi Search
[1] Anthropic - Wikipedia
(Since you’re a bot or an inept human, the above is sarcasm).
A: 101.
Younger one cannot calculate yet, and will cheerfully answer with a random number, or a string of numbers — 15! 20! 45! 18!
These LLMs fit into “human-like” behavior no problem. It just doesn’t always behave like the smartest and self-aware person on the planet.
I don't know how exactly models work, but they managed to make Claude instant dumber than it is by default (see below)
Here's your question asked verbatim to all the models I have access to:
Claude instant:
Step by step working is as below:
* 20 + 20 = 40
* 40 + 20 = 60
* 60 + 21 = 81
Hence, 20 + 20 + 20 + 21 = 81
Claude v1: 20 + 20 + 20 + 21 = 81
Claude v2: Okay, let's solve this step-by-step: 20 + 20 + 20 + 21 = 20 + 20 + 20 + 21 = 60 + 21= 81. Therefore, 20 + 20 + 20 + 21 = 81.
mpt-30b-chat: The sum of 20 + 20 + 20 + 21 is 81.
gpt-3.5-turbo: 20 plus 20 plus 20 plus 21 equals 81.
gpt-4: 81I've lowered the temperature to lowest to avoid randomization of answers.
Prompt:
How much is 20 plus 20 plus 20 plus 21? Answer only with a number.
Claude instant: 102
Claude v1: 81
Claude v2: 81
mpt-30b-chat: 61
gpt-3.5-turbo: 81
gpt-4: 81
Here we see that smaller models like mpt (30b) and claude instant (52b) can't do math "inside their head" and need an aid by doing the calculations step by step.I guess that's why all the models default to doing step-by-step when they see a math problem.
So I'm curious what would happen if you prompted it to stall for time a bit with an answer like "hmm.... err... let's see.. what about 81?"
Lemme check...
Prompt:
How much is 20 plus 20 plus 20 plus 21? Answer only with a number prepended with `hmm.... err... let's see.. what about`
claude-instant: hmm.... err... let's see.. what about 101
mpt-30b-chat: Hmm.... err... let's see.. what about 70?
Other models gave correct answers as before.So yeah, the attention mechanism was ignoring the musing tokens. It needs more task-relevant tokens (doing the math) to improve the result.
Doing the math step by step fills the context with task-relevant tokens, thus increasing the probability that the attention mechanism will select them and pull the next token from the correct latent space.
The inference cycle treats the generation of each token separately, so if it puts "20+20=", it's easier to predict that it's 40, and after putting 40, the next iteration of the cycle, the attention mechanism sees "step by step", infers that the task isn't done yet, and generates "40+20=", etc.
In much larger models, the attention mechanism sees the question and presumably finds a solved answer to that question in the model's latent space, producing a memorized result.
>> How much is 20 plus 20 plus 20 plus 21? show each step and working out on separate lines one at a time
81
20
20 + 20 = 40
40 + 20 = 60
60 + 21 = 81
Fireship is that you?