HNHacker News
TopNewBestAskShowJobs

mustaphah

2,166 karma · joined January 5, 2017

https://hadid.dev
submissionscomments

Jev and the Return of AI/ML Engineering

leehanchung.github.io·4 pts·mustaphah·
0

The Little Book of Reinforcement Learning

github.com·213 pts·mustaphah·
26

Opportunity cost neglect (2009) [pdf]

bear.warrington.ufl.edu·2 pts·mustaphah·
0

Economic Possibilities for our Grandchildren (1931)

fermatslibrary.com·4 pts·mustaphah·
1

A field guide to Fable: finding your unknowns

twitter.com·1 pts·mustaphah·
1

Ten Takeaways from the AI Engineering Report 2026: The Acceleration Whiplash

faros.ai·2 pts·mustaphah·
0

The cost YAGNI was never about

newsletter.kentbeck.com·6 pts·mustaphah·
1

Writing code vs. shipping code [pdf]

nber.org·3 pts·mustaphah·
0

Trust Factory

newsletter.kentbeck.com·7 pts·mustaphah·
0

LLMs pass a standard three-party Turing test

pnas.org·3 pts·mustaphah·
1

The small sample trap in A/B testing

hadid.dev·4 pts·mustaphah·
1

Tell HN: Claude two rate limits don't know about each other

2 pts·mustaphah·
0

Enhancing gut-brain communication reversed cognitive decline in aging mice

med.stanford.edu·386 pts·mustaphah·
185

Many SWE-bench-Passing PRs would not be merged

metr.org·278 pts·mustaphah·
153

AGI is an unscientific myth

tandfonline.com·4 pts·mustaphah·
2

Web Verbs

github.com·1 pts·mustaphah·
0

OpenAI's 5-month experiment: building a product with no human-written code

openai.com·2 pts·mustaphah·
0

SkillsBench: Benchmarking how well agent skills work across diverse tasks

arxiv.org·364 pts·mustaphah·
171

Evaluating AGENTS.md: are they helpful for coding agents?

arxiv.org·232 pts·mustaphah·
161

Curosr: Expanding our long-running agents research preview

cursor.com·3 pts·mustaphah·
0

Measuring Time Horizon Using Claude Code and Codex

metr.org·1 pts·mustaphah·
0

SWE-ContextBench: context learning benchmark in coding

arxiv.org·1 pts·mustaphah·
0

SWE-AGI: benchmarking spec-driven software construction

arxiv.org·1 pts·mustaphah·
1

Code Formatting Silently Consumes Your LLM Budget

arxiv.org·1 pts·mustaphah·
0

Agent Trace by Cursor: open spec for tracking AI-generated code

agent-trace.dev·1 pts·mustaphah·
0

METR releases Time Horizon 1.1 with 34% more tasks

metr.org·1 pts·mustaphah·
0

Coffee timing isn't one-size-fits-all

examine.com·4 pts·mustaphah·
0

ChatGPT subscription support in Kilo Code

blog.kilo.ai·1 pts·mustaphah·
0

Imposter Syndrome Predicts Perfectionism

psypost.org·2 pts·mustaphah·
0

Motivation acts as a camera lens that shapes how memories form

psypost.org·2 pts·mustaphah·
0
Page 1 of 5Next →