ask it for recipes, weather, directions, explain baroque art, etc.
arguably still a common use case...
If it’s not already a thing, it will be soon, the mental health aspects of all professions moving at unsustainable speeds and what that will do to people.
Nobody will take care of us. Ever.
I don't think any vague claim of "speed" has anything to do with fatigue.
I do think the parallel world monitoring and validating N work fronts, followed by rounds of fixes where you lay there waiting for the agents to converge, is the defining factor.
Instead of you hunkering down and hammering out a single task where your full attention can be focused 100% on a problem, your mindset is a kin to juggling N balls and hoping to not let any of which to fall.
So it's not really about speed but throughput, and keeping up with the throughout rate is mentally taxing.
Yep, this is happening in lots of companies. I know my small company has about ~1/3rd of the developers working on one (different ones, all semi-personal projects).
I sometimes wonder if we are drawn to building these because it feels like one of the remaining challenges and a desire not to be an "LLM text shuttle". Additionally/alternatively, you can go as fast as you want with building a software factory and you aren't held back by the "bottlenecks" (PR review, QA, etc).
Working on my software factory is the closest to a "flow state" I've been able to achieve since LLMs turned the corner earlier this year and replaced the vast majority of our code-writing work.
And meta is only arguably so. Not really in the same league as OpenAI and Anthropic. More second tier like Google and xAi. (and of those Google is pretty close to the top two at times)
Microsoft's MAI-Code-1.1-Flash is on par with OpenAI's Luna line of models. If not for OpenAI's recent radical change of heart on Luna's pricing to hastily slap a 50% discount, MAI-Code-1.1-Flash could very well be the dominant cheap model.
I see MSFT hasn't gotten any better at product naming :/
Like its a no brainer to force your employees to use your own models, then RL train them to be better.
I am not convinced that's the case.
Is security no longer a concern?
When you are using a third party model, you are literally feeding it not only your current software but also all the context and work fronts under development.
This is way more than granting a third party access to your internals. This is feeding it in advance updates on all their operations in real time.
Just one unsanitized input and you leak info. Or one hidden character and code may or may not belong to you anymore. Its very odd on many levels
At best you produce some noisy signals that are going to have a tiny impact if even that.
And that's on a personal plan where you didn't opt out of sharing usage data.
Business plans offer zero data retention. This is a non-issue.
I think this data is likely worthless compared to curated RL tasks.
But it was never a SOTA model.
... in 1988.