Having built a CSV import pipeline handling a few thousand different CSV files from almost the same number of different providers: I've probably seen almost everything one can mess up while writing out data: Invalid or missing escaping, double or per column string encoding, truncated columns, BOMs, EANs in E notation, Month names instead of numbers and so on. Out of curiosity: do you handle any of that?