58 karma · joined February 9, 2021
Edit: it's a very welcome addition. Limits side effects.
I am experimenting with the current SOTA multimodal LLMs, but performance is still not yet there, they still hallucinate non-existent teeth. (As an aside, I have found a simple but very telling test, I have an image with only 4 teeth visible up and 10 down, so I prompt the modal to count, non have been able to, but Gemini 2.5 pro is the closest of the lot, performance is worse in the description when the counting test fails).
I am going to try segmenting the image to see if I will have better results by prompting to describe segment by segment.
Erlang is not very fast, but that's not what it was built for.
I manage a few websites written in Lisp, and updating them is as simple as push code, recompile and it works.