ASCII – Origins
ethw.org
ethw.org
ASCII provided characters specifically for this purpose. But no one seems to have RTFMed. ASCII 31 is a "unit separator" (or field separator as we'd call it today) and ASCII 30 is a record separator. There's even ASCII 29, a group separator, so you can have a set of records related as a group (for example, a group of records of different type but related to a single customer). And there's ASCII 28, a file separator, so you can have multiple "logical" files within one physical file.
There is no fail here, this is what escaping is for. And with escaping, you can nest infinitely.
Don't get caught in the CS trap of thinking you need a completely "general purpose" parser, when a slightly-less-than-completely-general parser will be significantly easier to write, simpler to understand and debug, and faster to execute. And as mentioned above, if you do find yourself needing more than this parser, then you should be looking at non-parsing based approaches anyway.
iirc "Relia COBOL" back in the early days of DOS also had a format that used the delimiters and all properly; but then did goofy things like globally translating "'" (apostrophe) to "*" (asterisk) anyway.
[1] https://www.presidency.ucsb.edu/documents/memorandum-approvi...
[2] https://web.archive.org/web/20070914121230/http://www.presid...
Hmm. That seems to leave out quite a bit.
Eric Fischer (an HN'er) wrote "The Evolution of Character Codes, 1874-1968", available from https://archive.org/details/enf-ascii/page/n23/mode/2up .
In it, Eric writes:
> By 1955, Herbert Grosch had become sufficiently concerned about the growing incompatibility of character codes that he urged the attendees of the Eastern Joint Computer Conference to ‘‘register common codes so that ...
and
> The American Standards Association (ASA) got involved in character code standardization on August 4, 1960, when it created the X3.2 subcommittee for Coded Character Sets and Data Format.
Bemer's letter in 1961 therefore doesn't seem to have been a proposal to develop a single code, because that was in the works. Rather:
> The March 8-9, 1961 meeting of X3.2 finally led to a code (based on a proposal by Robert W. Bemer, Howard J. Smith, and F.A. Williams) that nearly everyone could agree upon—but there is some dis-agreement about exactly what it was that was agreed.
And that's in March 1961, well before May 1961.
Indeed, May 1961 is when Bemer et al. wrote the CACM article "Design of an Improved Transmission/Data Processing Code" - http://www.ed-thelen.org/comp-hist/ImprovedDataProcessingCod... .
"When IBM released its game-changing System/360 in 1964, the head of the development team, Frederick Brooks, decided that its punch cards and printers were not yet capable of using ASCII."
Brooks decided or just stated a state of reality that currently available peripherial hardware is uncapable of using newly proposed standard ?
Also how - as usual - convenient, that IBM keeped machines uncompatibe while proposing others to make their systems accesible :>
Also 'ASCII: Origins' is more fashionable - more like what movies do :)
[0] Page 70 of http://bitsavers.org/pdf/ibm/360/princOps/A22-6821-0_360Prin...
When Smalltalk-78's internal character set was replaced by ASCII in Smalltalk-80, it was the 1963 ASCII that they adopted since that is what they knew at Xerox. Though the goal was to be friendly to the rest of the world, the fact that almost everyone else saw their assignments as underscore causes problems to this day.
So it is possible that 1964 was too early for IBM to embrace ASCII. But to be fair, the 1963 x 1967 ASCII problems are trivial compared to any ASCII x EBCDIC.