When working from memory, it's normal for your memory to have already parsed the previous situation into features. As some of the later examples in the blog illustrate, it's easy to fall into parsing examples into the wrong set of features, which is how you'll remember them.
While I could solve all the problems in the article, I doubt I could solve any but the simplest if I was shown 1 image per day over 12 days and not allowed to write anything down.
Perhaps the lesson is that when you're trying to deduce a rule (say, for what conditions your software crashes in) you can increase your rule-discovering power greatly by making notes and being able to look at several examples side-by-side.