EADST

Sharding and SafeTensors in Hugging Face Transformers

In the Hugging Face transformers library, managing large models efficiently is crucial, especially when working with limited disk space or specific file size requirements. Two key features that help with this are sharding and the use of SafeTensors.

Sharding

Sharding is the process of splitting a large model's weights into smaller files or "shards." This is particularly useful when dealing with large models that exceed file size limits or when you want to manage storage more effectively.

Usage

To shard a model during the saving process, you can use the max_shard_size parameter in the save_pretrained method. Here's an example:

# Save the model with sharding, setting the maximum shard size to 1GB
model.save_pretrained('./model_directory', max_shard_size="1GB")

In this example, the model's weights will be divided into multiple files, each not exceeding 1GB. This can make storage and transfer more manageable, especially when dealing with large-scale models.

SafeTensors

The safetensors library provides a new format for storing tensors in a safe and efficient way. Unlike traditional formats like PyTorch's .pt files, SafeTensors ensures that the tensor data cannot be accidentally executed as code, offering an additional layer of security. This is particularly important when sharing models across different systems or with the community.

Usage

To save a model using SafeTensors, simply specify the safe_serialization parameter when saving:

# Save the model using SafeTensors format
model.save_pretrained('./model_directory', safe_serialization=True)

This will create files with the .safetensors extension, ensuring the saved tensors are stored safely.

Combining Sharding and SafeTensors

You can combine both sharding and SafeTensors to save a large model securely and efficiently:

# Save the model with sharding and SafeTensors
model.save_pretrained('./model_directory', max_shard_size="1GB", safe_serialization=True)

This setup splits the model into shards, each in the SafeTensors format, offering both manageability and security.

Conclusion

By leveraging sharding and SafeTensors, Hugging Face transformers users can handle large models more effectively. Sharding helps manage file sizes, while SafeTensors ensures the safe storage of tensor data. These features are essential for anyone working with large-scale models, providing both practical and security benefits.

相关标签
About Me
XD
Goals determine what you are going to be.
Category
标签云
PDB v2ray GoogLeNet git-lfs SAM SVR Clash transformers Mixtral Vmess Dataset SPIE 公式 WAN 域名 Crawler Shortcut Agent v0.dev XGBoost 强化学习 FP32 TensorFlow tar GIT Proxy Pytorch Transformers Rebuttal PIP ONNX 音频 Qwen2 hf Datetime Vim FP64 DeepStream UI Card NLP Tiktoken Python Bipartite YOLO Password Pillow Linux TensorRT Math ModelScope ms-swift 递归学习法 BTC 关于博主 PyTorch 顶会 Statistics Plate Distillation 第一性原理 Sklearn 证件照 NLTK Michelin FastAPI QWEN WebCrawler 图标 AI HaggingFace Heatmap API 报税 Qwen2.5 BeautifulSoup 多线程 CTC 图形思考法 签证 EXCEL VGG-16 Django Qwen Windows Attention CC LeetCode scipy Interview TTS OpenAI DeepSeek diffusers 多进程 torchinfo Tracking Base64 GPTQ Zip Jetson InvalidArgumentError GGML uWSGI Hilton Template 继承 腾讯云 Translation LaTeX mmap RAR Gemma Paper HuggingFace tqdm TSV Github 财报 MD5 Pickle Hotel icon logger API网关 GPT4 Quantization Domain FP16 Docker Review llama.cpp Ptyhon Hungarian Input Nginx CV LLAMA Cloudreve Color VSCode JSON 阿里云 Search 云服务器 Bin COCO 版权 Data FP8 Random Firewall Anaconda CLAP Google Video Markdown OCR FlashAttention NameSilo Permission Image2Text CEIR Algorithm IndexTTS2 VPN News Website Streamlit Ubuntu Plotly Jupyter OpenCV Git Llama Paddle LLM BF16 Safetensors SQL Diagram Bert printf XML CAM git Quantize Claude Conda uwsgi 净利润 Logo 飞书 Pandas Food Excel ChatGPT RL Tensor Knowledge Land Magnet Use Baidu Breakpoint UNIX LoRA 论文 Bitcoin 论文速读 算法题 ResNet-50 PyCharm C++ CSV PDF Disk Web Miniforge Numpy Augmentation 搞笑 CUDA RGB Freesound Animate SQLite
站点统计

本站现有博文335篇,共被浏览937416

本站已经建立2648天!

热门文章
文章归档
回到顶部