It seems at or above SOTA on the given benchmarks, doesn’t have context rot, is orders of magnitude faster, and uses less compute that current transformer models. I suppose it’s just an announcement and we can’t test it ourselves yet.
It seems at or above SOTA on the given benchmarks, doesn’t have context rot, is orders of magnitude faster, and uses less compute that current transformer models. I suppose it’s just an announcement and we can’t test it ourselves yet.
I am happy to answer any questions!
Do you anticipate having any kind of public accessible chat interface for testing in the near future?
Also, what, if any, benefits are there for smaller context windows? Is there still a material improvement in cost to serve under say 256k? I'm curious about the broader implications for the space beyond improvements for very large context windows.
When, more or less?
Can you back up your claims?
Why did you not release the white paper in parallel with the product?
Feels really fishy.
If I came up with a novel thing I'd monetise it first, because publishing it makes it part of the training that adds value to billion dollar corps with zero credit to me.
In the old knowledge economy I benefited from the credit assigned to me.
So, to me, nothing fishy at all.
Yes, this product doesn't exist.
And the last time a company claimed something similar it disappeared after taking money from investors.
no published benchmarks
no paper
no demonstrations of capabilities