Since git is essentially a filesystem with extensive version control features, it doesn't surprise me that it would have problems handing large amounts of files.
Since git is essentially a filesystem with extensive version control features, it doesn't surprise me that it would have problems handing large amounts of files.
But there will be some trade-off.
And I don't think people generally put "a million files" in the requirements because it's fairly rare.
It's hard to get people past the demo phase "works for me" when they have played with one image, to realize they really need a reasonable container format to play nice with the systems world outside their one task.
Now I have a system that will find subsets in just a second or two (even when the whole set contains hundreds of millions and any given subset might contain hundreds of thousands of matches). Here is a short video of a demo: https://www.youtube.com/watch?v=dWIo6sia_hw