Can reinforcement learning for LLMs scale beyond math and coding tasks? Probably | Hacker News Reader