We actually run already in-house ollama server prototype for coding assistance with deepseek coder and it is pretty good. Now if we would get a model for this, that is on chatgpt 4 level, I would be super happy.
Or how you deal with context length? I.e. do you send anything other than the current file? How is the prompt constructed?