For this however, a comparatively much simpler task, tarra-high works fine.
For this however, a comparatively much simpler task, tarra-high works fine.
DeepSeek is okay for random API-based stuff, as it's cheap.
Local open models running on a 5090 are hit or miss. I feel that most GGUFs/quants are awful...
I wonder if it is because of watermarking.
That was true before they announced the watermarking, I'd already started to back off of using Opus as much because I like to understand what the model is doing and have it write documentation I can use to reproduce its results, but maybe watermarking was already in there unannounced.
And it retypes it for a human, I'm doing this more and more lately.
I think it could be the watermarking, but at this point they might be deliberately complicating the prose so that we ask clarifying questions and that leads to more token spend.