Evaluating frontier AI R&D capabilities of LLM agents against human expertsmetr.org1 point·tedsanders··0 commentsOpen articleSaveView on HN