Or, in other words - we have two P(Doom), one for AI being developed, and another for AI being not developed. The latter is not discussed enough imho.
5,854 karma · joined September 30, 2009
workIn-Progress.posterous.com
Or, in other words - we have two P(Doom), one for AI being developed, and another for AI being not developed. The latter is not discussed enough imho.
And also, what’s your story here? That they don’t delete the data they agree to delete, and stash it secretly somewhere? And that they have programs that are mire secretive tan NSA’s inside that secretly finetunes algorithms on such data?
They would expose themselves to lawsuits that would crush their companies, even with such absurd valuations? And possibly risking jail time of the top people? GDPR offences in certain countries are punishable by up to two years in prison. And with a such blatant violation at least in EU the maximum sentences would most likely be given.
You’re confusing, I think, memory system with llm finetuning. Completely different concepts.
Also, planning over long scales is not something you should do if you do lean/agile. Premature optimization is a root of all evil ;)
Personally, I read way less traditional books when it comes to learning stuff. Especially the “utility”/“tutorial”/“handbook”. Top nitch books talking about general principles I’d still read.
Especially for books that need to teach me a subset of a certain discipline - in the past I’d get 4-5 (or at least samples), and try to find the parts that are of interest to me in a style that fits me. Novadays I’d just ask Claude to explain things in a form that I like, with pretty illustrations from Imgen :)
E.g. I don’t think I’ll read the animal books from O’reilly again. But an equivalend of “Thinking in C++” or “Pearls of programming” of a new field? Sure.
Otoh I’d say the current models are better at predicting expectations than average programmers. Average programmers don’t know ux or business, LLMs do.
Also, it would be relatively easy to build synthetic datasets for training.
I, for one, don’t mind models being trained on stuff I produced and shared publicly over the last 20 years. I did it for common good, including commercial uses, and this is one of them.
Plenty of people who never produced any open source trying to argue as if if they did.
Also even if your Gemini is giving you nonprogramming output, underneath the model is most likely generating code for certain tasks.
Pi out if the box tries to optimise system prompt size, which is not necessarily good and will cause exactly this effect for all but the most simple tasks.
What you want is to give enough context to the agent to minimize the amount of searching within the codebase etc.
If you want to track cache, what you should do, imho, is to check if you have cache expirations mid sessions (ideally you should not), and if you don’t then lower cache use is actually better - it means that your model doesn’t reread what it just wrote.
LLM can be trained on a code and at the same time reproduce the core ideas. That's what LLMs do after all - they convert the training data into their own internal models and representations, and then reproduce the ideas.
Sure, some things/patterns, that were repeated multiple times, LLMs will tend to repeat verbatim as well, but that's not that big of a problem.
As a person who invented a few algorithms on my own I absolutely love LLMs and I don't mind them being trained on my work, but yeah - I've been way less likely to publish open source over the last year. In the past, if some of my stuff got traction, the credit was close to automatic (early adopters credited or at least knew where they got it from). Nowadays, LLMs will train on these ideas, rewrite them, and give no credit.
Still, I prefer this to having no LLMs at all.
> but stick to the code samples from the books. > I'm certain it'll be able to change the color of a CSS button, right?
A good enough LLM will just decompile a browser, figure out CSS spec from it, and yes - figure out how to change the color of a CSS button from first principles. There is no point to do this with CSS, but with other things it's now easier to just dig through sorces or direct bytecode than to bother checking docs.
How do you recognize someone as having true opinions and knowledge from someone having just appearance of them?
How do you decide upon a price?
Thanks for the benchmark, I was seriously looking forward to it! Would you consider such even results good for a one-shot? I wonder how well my harness performs :)
Nonbombed countries - a ton, Ukraine for one, but the whole post-communist block.
Sure not everyone benefitted from US but a ton of people/countries did and still do.
US never felt a need to build walls to prevent their citizens/allies from leaving. Russia and the others very much so.
With US the track record may be mixed, but a ton of countries from Europe and Asia benefitted tremendously. Russia otoh doesn’t have a concept of win-win, they tend to exploit even their closest allies, which anyone living in Baltics/East-Central Europe can tell you.
https://kolinko.eu/pdf-reading-order/
But I wonder about your opinion.
Another cool book about back and forth between monopolies and decentralisation in our space: https://www.amazon.com/Master-Switch-Rise-Information-Empire...
It also has an audiobook.
Awesome book about the history if Bell Labs - virtually all semiconductor tech we use today was created there (transistors, ics, solar, lasers, fiber optics, telecom satellites…), and they had to license it to be allowed to maintain their monopoly status.
https://www.amazon.pl/Idea-Factory-Great-American-Innovation...
Similarly, Bell Labs was kind of required to release transistor for anyone to license - it was a part of social contract that they were allowed to maintain their monopoly in exchange for releasing certain parts of technology.
Alternatively, they could be split up and their indexing division made an independent company selling to anyone on a free market.
Just like Russia doesn’t care about online criminals being located within their borders - as long as they target outside.
While this may be true to some extent of all the countries, Russia is the one to be clearly against Estonia, Baltics and eastern europeans.
Do you use a public set of documents? I bet I could almost oneshot this with my harness :p