Researchers have identified new elements of whale vocalizations
vice.com
vice.com
There is no breakthrough in understanding. Instead they identified a thing in the sound that changes.
They do not understand what it means, they just see it. It could be something, or nothing, or just music.
But a new understanding of the components of the code and their relationships is certainly a breakthrough towards that understanding.
Yes, the headline had me hoping that there was some full understanding (it'd be nice to broadcast to the Orcas off Spain & Gibralter to please not attack the boats), but I can't say that it's clickbait-level.
As a result of their research they are getting closer to the underlying structure (which others will look for in other species of whales) and with that they may be able create situations where they could start testing the concepts that might be in that structure. That sounds kind of breakthroughish to me :-)
This from the article: “If our findings are correct, it means that the communication of sperm whales is much more complex and can carry more information than previously thought,” the researchers concluded.
The researchers here make two observations, the whales intentionally use specific forms in their songs and the specific forms are sequenced differently but intentionally. For biologists this is significant because it differs from a common form of "singing" or "calls" that each have a specific meaning but they are not internally variegated like these calls are. What the researchers propose (or hypothesize) is that this intentional and varied voicing within a "song" suggests the whales are exchanging much more detailed information than "threat" or "help" or "food here" etc.
https://blog.padi.com/talk-to-whales-with-ai/
One difficulty is that there isn’t nearly as much whale chatter available for training data as there’s human chatter.
And one important question is, why is this useful if we don’t know what the LLM says to them? But the post above touches on that too.
I'm sure you've seen the example of word vectors that captures some of this meaning. king - man + woman = queen
In Spanish, rey - hombre + mujer = reina
The _relationship_ between "king" and "queen" in English may look close enough to the _relationship_ between "rey" and "reina" in Spanish, allowing you to bridge the gap between the two languages, even if they are entirely disconnected and you've never seen a direct translation between them.
There's no reason to be sure we couldn't know.
It's not like there are examples in the training set for every lang to lang combination modern models are capable of translating.
For whales or even dogs or great apes, I think the chances are much higher, but we just don't know. We can't even agree on a definition for what consciousness is.
Just because a consciousness exists staring out of a pair of eyes doesn’t change any ethics if that consciousness is exceptionally limited.
I think one day we'll probably almost completely stop killing "intelligent" animals (in 100s of years after we develop safe and tasty alternatives). What about a software consciousness? I don't know. What about if we are able to "upload" our consciousness, will it have civil rights? Probably not at first.
I’m not sure what a ML algorithm could do better than a tape recorder here.
Gašper Beguš who is the linguistics lead of Project CETI is the first author of the paper this article talks about.
https://tv.apple.com/us/episode/2046-whale-fall/umc.cmc.2gda...
Scientists studied how sperm whales communicate underwater.
They found that whale clicks have similarities to human vowels and diphthongs.
An AI model helped by imitating whale sounds, providing useful information.
The researchers discovered specific patterns in whale vocalizations, like unique "coda vowels."
Whales seem to control the frequency of their calls, making their communication more complex than we thought.
No, we can't have a casual conversation with whales yet. Serious ones either.
> Not only did the AI predict elements of whale vocalizations already thought to be meaningful, such as clicks, but it also singled out acoustic properties.
Just training on existing sounds might give us a whale sound generator. But it would be the equivalent of Prisencolinensinainciusol: https://youtu.be/-VsmF9m_Nt8
> LLMs do not learn to translate through explicit examples of parallel text.
Of course they do, lots of that is in the training data.There are several completely unsupervised language translation systems that predate LLMs, but the performance is middling.
For LLMs the translation behavior is largely an emergent property that is not completely understood; if we can tokenize whale language in a useful way, it is entirely possible that the LLM can derive a weak or approximate translation of some of the language structure.
Maybe a better example is in the multimodal context. Drawing images from words is a type of “translation”. But for this task we need captions, i.e. a parallel corpus.
However, I believe you are overstating the results of the paper. If anything, the paper demonstrated massive reductions in zero shot translation capabilities after removing translation. And for languages with no cognates with English, BLEU scores are pretty abysmal. So the paper suggests that parallel corpuses still very important.
For whale languages, you won’t have any cognates with English, and we don’t even know if they have a grammar that is remotely human.
Like I said, the crucial thing is that it looks to be a diminishing gap. I'm not saying parallel corpora is doing nothing or is unimportant but that's a clue that monolingual competency of the languages in question is the biggest factor here. The extreme positive transfer language models exhibit in terms of multilingual competency is another clue. https://arxiv.org/abs/2108.13349
Parallel corpora may be a crutch with fading relevance.
Please expand -- as I understand it the LLMs use cross entropy on next token prediction as their loss function and I fail to see how this gives any feedback about translation, even given parallel texts. Predicting the next token in a French language text and predicting the next token in an English language text are not obviously more interesting than dealing with non-parallel texts in both languages.
It's totally possible that these parallel texts are critical in a sense to the translation capabilities but this is not obvious.
It would be incredibly difficult to establish empirically because reducing training data reduces the effectiveness at all tasks. Translation, like almost all of the capabilities of LLMs, is an emergent behavior that we don't fully understand yet.
To add, large enough here is also a potentially shifting target. LLMs exhibit extreme positive transfer of language competency. So 5B tokens of Korean will take Korean competency much further if trained alongside 50B tokens of English than if alone.
What would a whale call a Quarter Pounder with Cheese?
Not Necessary.
Searching for Needles in a Haystack: On the Role of Incidental Bilingualism in PaLM's Translation Capability (https://arxiv.org/abs/2305.10266)
ablation studies performed on the effects of removing tiers of multilingual corpora (parallel corpora, bilingual (not necessarily parallel), and english only)
Removing Parallel corpora or even just bilingual data in general reduces quality but does not make the model unable to translate. Crucially, the gap or reduction in quality also seems to diminish with scale.
The paper also shows models may not need any multilingual corpora at all. The 8b Model can still translate latin scripts with only english training data, adding evidence for larger models being better equipped to leverage either sparse signals (i.e., language-identification failures during ablation) and/or weak signals (i.e., language similarities from shared scripts).
Is HN getting dumber? What's with all these threads where people basically assume the "AI" we currently have is some kind of superintelligence?
Vowels and diphthongs in sperm whales - https://news.ycombinator.com/item?id=38541217 - Dec 2023 (29 comments)
The gist was a lot of the researchers with names like Jones and Smith were disappointed. And the one guy on the research team with a Spanish last name was getting a lot of weird comments about it which made him uncomfortable.
75% of the whale song was decoded to just mean, "Oy mira! Camaron!"
Dolphins were discovered to be speaking a cipher for English, but with a Chicano accent.
Some of my friends found it funny, others didn't get it.
---
I'm really interested to hear what they're actually talking about!
reads more succinctly. (And no offense to OP.)
Occasionally I'm surprised links like these get to the front page, even with all the ads and horrible UX. I must admit, I too wanted to believe so badly we can communicate with whales, that I managed to scroll down to see the first paragraph below the fold. But quickly closed it shortly after to read the HN comments.
For that second population: https://osf.io/preprints/osf/285cs
The overwhelming remainder of people are in neither: they don't block ads, and they don't care. And their TVs are on all day long.
I strangely forgot the screenshot tool I'm using, Cleanshot X, actually has a cloud-share feature built in. I've never used it since 99% of the time I'm pasting these images into issue trackers, slack, etc.