What sort of resources do we need to run this, particularly VRAM? Also, how does this compare to Fauxpilot?
Related research works include [1]. (Hints: combine code search and LLM).
[1]: RepoCoder: Repository-Level Code Completion Through Iterative Retrieval and Generation https://arxiv.org/abs/2303.12570
This also reveals Tabby's roadmap beyond other OSS work like Fauxipilot :)
That doesn't answer the question, can anyone without more VRAM than sense actually run it as-is or should we wait until they reach their allegedly impossible aspirational goal?
The very first line of this sort of post should be the specs required and if the trained model weights are actually available, otherwise it's just straight up clickbait.
A 1B parameter transformer model is on the low/tiny-end of model size these days.