The article solves the problem: for which numbers x between 1 and 500 is there no file x_A.csv? It looks like in this case it is equivalent to the easier problem: for which x_data.csv is there no corresponding x_A.csv?
cd dataset-directory
comm -23 <(ls *_data.csv | sed s/data/A/) <(ls *_A.csv)