Accelerating Compute by Cramming It into Memory
nextplatform.com
nextplatform.com
One approach to that is putting computation in RAM. Several architectures have done that. I posted an example. I’ll add a few more below. I’m curious what AI experts think about using these architectures for inference or training.
https://en.m.wikipedia.org/wiki/Berkeley_IRAM_project
https://www.eecg.utoronto.ca/~stumm/Papers/Elliott-IEEEDTC99...