Specifically: to explore your opensource options with compute limitations, ask the community at r/LocalLLaMA on reddit. That's where the current SOTA opensource text-to-text models live.
Since my team is mostly remote running LLM on a cluster in the office is not really viable short term.
Harnesses such as this - combined with locally hosted AI models - are the future.