There's one aspect I don't understand of how the gel "learns" . What are the factors for positive reinforcement and negative reinforcement? How does the gel "know" what is a good result and what is a bad result?
How does this follow?
I didn't go into the paper to see if that's exactly what they're doing, and I'm no expert. But from what I've read before, that's how this usually works, and I'm sure they're doing something similar to that.