Fun fact: in the 60s Italy was after the US and UK the county with the most advanced nuclear energy plants. These days still North Italy produces key parts of nuclear reactors in other countries.
Because of masked attention in LLMs, if you put the options before the body (the email to analyze), the transformer already knows what it needs to look for, and can use more tokens to create state to address that specific task (BERT has no mask in the attention, so tokens attend also to next tokens). You could also do a few examples in the system prompt to improve calibration.
Another trick that works is to repeat the question two times: "I'm repeating the task and labels for clarity: ..."
Anthropic is incredibly good at avoiding all the useless AI risks, while not doing anything serious about the real risks (that is: uncontrolled growth). I'm as pro-AI as I think that eventually it will remove suffering from humans, and will allow us to prosper more and help us with tons of problems we created. However the problem is not job loss or minors using AI (minors are fucked because of cell phones and social networks), the problem is avoiding extinction. For this position I was accused of AI psychosis multiple times, but I'm in good company (Hinton, for instance), so I bet some comment that trivializes this issue will surely reply to this one, but at this point it is important to see what everyone really believes. Anthropic only touches the surface of AI security, consistently.
That's perfectly wrong. Since strong coding AI, people venture into huge rewrites and other big changes that automatically make sense but otherwise would not.
Sol for low level programming is consistently better, can work alone for more time, and is faster. If you think Fable is so superior, you need to work with Sol ways more.
Yes. The article says that you need to read a lot to write well. You actually need to write a lot, and to read carefully, high quality, and potentially low quantity too (to use more time to write).
This is simply not true, even if it is commonly repeated among the folks that usually don't really write. Like how reading every source code that comes handy will not turn yourself into a great programmer, to be a good writer you need to: 1. Read selected books, and re-read good books more often then reading new stuff, to understand why they are good. 2. Read, from time to time, some bad book, and understand why it is bad. 3. And obviously you need to write a lot to become good at writing. And writing, in order to improve, is writing remembering, at the same time, the vibrations of the good authors you loved, and especially making the act of choosing of every word you put in the blank page, one after the other.
H3 is quite uncensored, but was not trained on p0rn, so it has no anatomy clues needed to generate that kind of stuff. For softer adult content it is reported to be fine on Reddit.
This implementation is much faster on my M5 Max, like a few minutes for the same video, but on an M5 Max with 128GB, didn't test on M5 Pro. About memory, could be executed on 64GB with a few changes.
In the AMA Minimax said that H3 could support sparse attention, that would be a huge speedup! I wonder if there are any news on that. H3 is very cool. EDIT: testing a --sparse-attention optional mode based on what they said in the Reddit post.
Chinese labs are the proof that there is no need of big names, but of the right mindset and agility. It's those last things that Google truly misses, but now they are missing for a long time, and outside the AI divisions too, in almost every department of the company.
This is not at the top as it is actively flagged by people that can't psychologically cope with the advances of AI. Hacker News is no longer a web site of an elite.
So many mathematicians over the years tried hard and failed, but now Anthropic just for some PR magically did it? And this after LLMs obtaining different math wins? What is your logic here really escapes my understanding.
Execution is not the code, but how do you decide to do every part. "Idea does not matter, execution does" always meant: "big generic ideas don't matter, it is how you organize it in the myriad of details it is composed of (in a given incarnation of the general idea) that matters."
Because Redis is not "my project", it is a piece of software many relies upon, so I use, for that software, what the community at large agrees to be ok: AI-assisted coding with human careful reading and evaluation of every line.
Thanks! And sorry for not yet merging many of those. The problem is, I'm dealing with tensor parallelism for the CUDA and Metal-RDMA fork right now, so was not albe to care about PR / issues for a lot of time.
Thanks, I believe that as a whole choosing the BSD created a more positive effect, so I'm happy with that. It is just that it is really unfair to read a comment where people use ValKey to accuse you of AI slop :D It means that our community, and this site itself, is at this point really low quality. This will in turn discourage the many great folks that are here. A replacement is needed. But TLDR, I would release Redis again with the BSD license if I could go back in time.
Sure, it costs less, and AWS is in a dominant position. Users here are playing the side of the bully since they don't care about what is right and wrong with the hyperscalers. "BSD is better than AGPL!" And give money to the wrong side of the history. Nor that I expected anything better, the single person has a given sensibility, the mass, as a whole, do whatever is in a given moment convenient or believed to be more pure (license wise). However thanks to that, you will see how little progresses we will have (and we are having) in the space of open source system software with very open licenses. Developers of software mostly are not happy to bring OSS to the success to see them used by hyperscalers to capture all the value. However I did it again, with DwarfStart, to release code under the BSD license: even in the current situation, I think it is better to give back than to have a personal gain, but this is a position that very little folks can afford to take.
However: this conversation is completely out of topic but people instead of talking about AI and code, which is a tabu, will move the conversation to personal attacks and shit like that.
Do you understand Redis and ValKey have mostly overlapping code bases? And of that intersection, a big part of the code was written by myself by hand. So no, that's not the case. Also as I wrote in the blog post, Redis is currently not using AI if not as AI-assisted coding. Of course I'm not writing this reply for you, since I believe if you write a comment like that, you are part of that HN slice that makes this site at this point a slop place (no need for AI for very low quality), but for others that may find this information useful.
> 1. Compute costs collapsed since the advent of Cloud and yet hyperscalers still have fat margins.
Cloud opposes switch inertia. To setup a complex system in a different environment is a complex operation. Changing AI provider is switching an endpoint.
In Catania 10 gbit Internet costs 35 euros/month and is available everywhere in the city and even in a big slice of the small towns around Catania. And indeed public incentives played a big role. But what I believe it is more interesting is that 1 gigabit was common like 10 years ago or even more. Infrastructure is a bit too important to be left to what the market believes will be profitable.
Btw the paradox I have is that my local lan is 1 gbit...
Europeans do a lot of stupid things, but I believe in light of all the scandals we saw in recent times, you can't explain EU behavior and choices without accounting for corruption. EU division and different level among the different countries of wealth, integrity of political sphere, and different cultural biases make us the perfect target for bribes in order to control votes and choices. Not just promoted by external actors. The Chat Control is a great example: everybody understands how bad this is, the arguments are mostly a shield to avoid revealing the real agenda.
I love coffee, I take 2 to 3 every day (Italian espresso, so very short, little total caffeine), however for some reason taking the decaffeinated one in the evening interferes with my sleep in a similar way than the caffeinated one. Any hint?