This must mean Hugging Face's bandwidth bill must be crazy, or am I missing something (maybe they have a peering agreement? heavily caching things?)
This must mean Hugging Face's bandwidth bill must be crazy, or am I missing something (maybe they have a peering agreement? heavily caching things?)
Bandwidth has always been crazy cheap.
In fact locally I can get a 10 gbps home internet unmetered connection for $300/mo.
I'm not sure how they'd react if I transferred 1 PB/mo though :)
An unmetered 10G port at a US data center is ~$1500/mo. Not particularly expensive
For work I end up transferring 50-150 gigs often, sometimes daily. Never heard a word from them that this has been a problem.
1. AWS is far behind Azure and GCP in AI, so they gotta partner up to gain credibility.
2. Huggingface probably does face insane bills compared to github. But AWS can probably develop some optimizations to save bandwidth costs. There's 100% some sort of generalized differential storage method being developed for AI models.
Or better yet, how about asking me where I want to store my models?
https://learn.microsoft.com/en-us/windows/win32/fileio/hard-...
https://learn.microsoft.com/en-us/windows-server/administrat...