shameless insertion:
"raw csv processing at ~20GB/s" has been demonstrated in one my project as the tooling byproduct[1] based on Rayon(great Rust data parallelism library)[2] and wrapping of modified simdcsv[3] into one simple .rs.
just a little more:
1. simdcsv has severe bugs, so do not use it beyond demo.
2. the processing model of simdcsv still has rooms to good improvements(estimated 2x more, a.k.a. ~40GB+/s in memory in single modern socket should be achievable in some scenarios(no heavy string to complex language object conversions)).
[1] https://tensorbase.io/2020/08/04/hello-base.html#benchmark