EADST

Quick Review: ZeroQuant-FP

ZeroQuant-FP: A Leap Forward in LLMs Post-Training W4A8 Quantization Using Floating-Point Formats

Highlights:

  • FP4 Weight Quantization: Implements 4-bit floating-point (FP4) quantization for model weights.
  • FP8 Activation Quantization: Utilizes 8-bit floating-point (FP8) quantization for activations, optimizing the balance between performance and precision.
相关标签
About Me
XD
Goals determine what you are going to be.
Category
标签云
Web NameSilo BeautifulSoup icon ONNX SQLite 腾讯云 CEIR PDB Bipartite 飞书 Quantization ResNet-50 Tiktoken XML printf 强化学习 Nginx PDF Windows Review 域名 PyTorch TensorRT Tensor Numpy 论文速读 论文 Search Jetson 音频 GoogLeNet Git Safetensors TTS Image2Text GGML API Vmess UI GIT NLTK 图形思考法 HuggingFace EXCEL Permission 净利润 Vim Crawler CAM NLP LLM Pandas git hf Card Hilton Github 搞笑 多线程 CUDA Paddle Algorithm Python VPN Logo Plotly C++ Sklearn GPT4 Tracking RGB Bin CLAP 财报 Excel FP64 SQL Jupyter Firewall Bitcoin OpenCV ChatGPT 阿里云 OpenAI SVR Data SAM LLAMA tar Qwen2.5 git-lfs CC BTC Linux GPTQ Breakpoint LaTeX Shortcut Food llama.cpp Attention FP16 Website Use Random JSON Video mmap Transformers logger 证件照 PyCharm 云服务器 WebCrawler Docker FlashAttention Markdown Augmentation YOLO 算法题 Claude Hotel MD5 继承 Hungarian Datetime Domain OCR tqdm LoRA ms-swift Qwen v0.dev 公式 COCO Michelin QWEN AI Pillow VSCode Quantize Streamlit 第一性原理 CV Rebuttal Land Zip Qwen2 Input UNIX Bert Pickle Clash 顶会 TensorFlow Ubuntu CSV Conda News Cloudreve LeetCode Template HaggingFace Password Django TSV Pytorch scipy Animate Knowledge Math Heatmap FP8 SPIE Disk FP32 transformers Ptyhon Paper Llama diffusers 版权 v2ray DeepSeek Freesound ModelScope Agent Interview uwsgi Diagram RL 递归学习法 DeepStream XGBoost Anaconda Proxy Magnet IndexTTS2 Miniforge 多进程 Baidu 报税 CTC RAR Base64 InvalidArgumentError FastAPI PIP 图标 Translation torchinfo WAN Dataset VGG-16 关于博主 Color uWSGI Gemma Plate Distillation 签证 BF16 Statistics Google Mixtral
站点统计

本站现有博文333篇,共被浏览914486

本站已经建立2617天!

热门文章
文章归档
回到顶部