This question hinges on whether model advancement plateaus enough for machine sized models to compare to frontier performance. If it does, the answer is yes. If it doesn’t, the answer is no
Tools like Opencode demonstrate that when you box them in tightly enough they can actually be pretty competent.