We need continuous self learning and improvement at the level of post-training weight updates.
My endeavors in this area show that it is very difficult for an LLM to build up and integrate novel knowledge through conversation. Any new developments that aren't in context, are completely forgotten unless they are continuously kept in context, along with enough reassuring text that the novel knowledge is accurate if it goes against the LLM's training.
Is there any research into 1B-2B sized models that can be continuously finetuned on RTX 3090?