Can LLMs earn $1M from real freelance coding work?
newsletter.getdx.com
newsletter.getdx.com
> 3. Performance improves with multiple attempts Allowing the o1 model 7 attempts instead of 1 nearly tripled its success rate, going from 16.5% to 46.5%. This hints that current models may have the knowledge to solve many more problems but struggle with execution on the first try.
https://newsletter.getdx.com/i/160797867/performance-improve...
curious if you had any examples. i'm fairly meh on llm coding myself but have a pet theory on safety rails. i've certainly hit plenty myself but not with coding with llm's.
Discussion on original paper: https://news.ycombinator.com/item?id=43086347
Is that the one that says if an article title ends in a question that means the answer is no?