Three Things We've Learned About Generative AI and Developer Productivity
innovation.ebayinc.com
innovation.ebayinc.com
I am available to work for you at good levels of accuracy, asking mid 6 figures + bonus + stock options.
Fine-tuning is suggested to improve jobs like tooling upgrades but no concrete numbers are offered.
Lastly RAG on documentation. The RAG has a simple system prompt to improve uncertain responses. They're tracking meeting and support requests but don't show any results. They mention frustration with nonsensical answers but use a RL human feedback technique to improve responses. No numbers offered.
Overall a simple overview of what they tried but the strong methodological start doesn't get reflected in the numbers reported later on.
How is accuracy measured here? Is a document a single file? Is the LLM generating code and some separate kind of “document” such that “code” accuracy can be 60% while “document” accuracy can be 70%?
I mean, define 'good'. Yikes.
> Work at large companies has a tendency to blow up, run far behind schedule, then ultimately limp past the finish line in a maimed state.
> One of my friends talks about how, when faced by his first failed project on a team, a management consultant responded to all critical self-reflection with "But you'd say that, overall, this was a success?" in a desperate bid to generate a misleading quote to put into a presentation to the board.
https://ludic.mataroa.blog/blog/tossed-salads-and-scrumbled-...