Run AI workloads on any cloud, store on Hugging Face: zero-egress storage with SkyPilot
Hugging Face Blog
SkyPilot and Hugging Face integrated support for mounting models and datasets from Hugging Face directly into compute jobs running on any cloud or on-premises cluster. In a benchmark fine-tuning Qwen 3.5-4B, the model loaded in ~30 seconds at up to 500 MB/s and checkpoints wrote back to storage at 112–168 MB/s depending on the cloud, with zero data egress charges. Teams can now run GPU workloads on whichever cloud has available capacity while reading from a single bucket, eliminating the need to replicate data across vendors or pay per-cloud transfer costs.