Hugging Face Cache Deduplication with Xet in huggingface_hub v1.32
Melvin Vivas · X post · 2026-09-19 · Open on X
Topics: LLMOps, Deployment & Monitoring, AI Dev Tools & Productivity · Level: intermediate
Summary
A quoted post explains that huggingface_hub v1.32 stores identical Xet-backed files only once, even when several repos share them. If five repos share the same 20 GB weights file, disk use drops from about 100 GB to about 20 GB. This matters if you download many fine-tunes or variants that share base weights.
Key points
- huggingface_hub v1.32 deduplicates identical Xet-backed files across repos in the local cache.
- Example: 5 repos sharing one 20 GB weights file used about 100 GB before and about 20 GB after.
- The Hugging Face cache is a content-addressed store, not just a folder of downloaded models.
- Upgrade huggingface_hub to get the disk savings, which are biggest when you keep many related model repos locally.
Resources mentioned
- Hugging Face · website · x.com · free · recommended by both Bashiri Smith & Melvin Vivas
Platform for hosting and finding ML models, datasets and papers. The quoted post says LocateAnything was trending there.
Also in: Using an ML agent to train an open-source TTS model on your voice (Melvin Vivas on X · notes), Deploy Open-Source Models with Hugging Face Inference Endpoints (Melvin Vivas on X · notes), Deploying Qwen3.8 27B on Hugging Face Inference Endpoints (Melvin Vivas on X · notes), Using Hugging Face credits: Jobs, Inference Endpoints and Open Models (Melvin Vivas on X · notes) and 38 more - Xet (Hugging Face storage backend) · tool · huggingface.co · free
Hugging Face's chunk-based storage backend for large files on the Hub, which makes deduplication possible.
Try this
- Upgrade huggingface_hub to v1.32 or later to free up disk space in the model cache.
More in LLMOps, Deployment & Monitoring
- Jev was free on Vercel AI Gateway until Sept 25
- Benchmarking LLM Endpoints with NVIDIA Dynamo AIPerf
- Running Qwen3.8-27B Locally on an M5 Max MacBook with Inco Splash
- Agent Monitor: see traces, tokens and costs of your coding agents
- Agent Monitor: Per-Model Usage Stats for Codex and Claude Code Agents
- Agent Monitor: Track Token Usage and Cost for Codex and Claude Code Agents