I went to the test subfolder:
https://github.com/klauspost/compress/tree/master/testdata
I saw 3 files.
The kind of test data I'd like to see for something like a compression library would be:
0 byte files
1 byte files
2,147,483,647 byte file
2,147,483,648 byte file
2,147,483,649 byte file
(to check for bugs around signed 32-bit integers)
4,294,967,295 byte file
4,294,967,296 ...
4,294,967,297 ...
(to check for bugs around unsigned 32-bit integers and see if 64bit was used when necessary for correctness)
Also add files of sizes that straddle the boundary of algorithm's internal block buffers (e.g. 64kb or whatever). Add permutations of the above files filled with all zeros 0x00 and all ones 0xFF. I'm sure I'm forgetting a bunch of other edge cases.
The programmer may have done a wonderful job and all the code may be 100% correct. Unfortunately, I can't trust a library replacement unless I also trust the test bed of data that was used to check for defects. It's very common for performance optimizations to introduce new bugs so there has to be an extensive suite of regression tests to help detect them. Test data for bug detection has a different purpose than benchmark data showing speed improvements.
Those multi-gigabyte files are not git repository friendly so perhaps a compromise would be a small utility program to generate the test files as necessary.