Tensor Product Attention: A Memory-Efficient Solution for Longer Input Sequences in Language Models
Maybe a Greasemonkey script to pass arXiv abstracts to a local Ollama could be something...
*side effects TBD
https://scholar.google.com/scholar?hl=ro&as_sdt=0%2C5&q=%22i...
Always has been.
silent gunshot
see Section 3.4
But, on the other hand, it's hard to get researchers to read your paper, esp. in fast-moving areas. Every little thing might be the difference between reading the abstract or not. Reading the abstract might lead to reading the intro. And so on.
So, for better or worse, the competition for human eyeballs is real.
Ironically, in this case, "attention" is all that the authors want.
Bloody hell and brimstone. Been crazy 57 years and a half already.