Show HN: OmniNet:- A unified architecture for multi-modal multi-task learning
github.com
github.com
Question: in the example of prediction on untrained tasks, what exactly hasn't been trained? The paper talks about video being one of the trained tasks. Did you simply retrain model without video examples and then test performance?
Does this bring us closer to AGI?
HN is a bit strict. I'd say "X is all you need" gets less attention from users than a very technical headline. The most popular submission recently had MITM in the title (https://news.ycombinator.com/best)