EADST

Quick Review: ZeroQuant-FP

ZeroQuant-FP: A Leap Forward in LLMs Post-Training W4A8 Quantization Using Floating-Point Formats

Highlights:

  • FP4 Weight Quantization: Implements 4-bit floating-point (FP4) quantization for model weights.
  • FP8 Activation Quantization: Utilizes 8-bit floating-point (FP8) quantization for activations, optimizing the balance between performance and precision.
相关标签
About Me
XD
Goals determine what you are going to be.
Category
标签云
论文 GGML printf Llama Ptyhon Plotly icon 域名 Bitcoin Disk Cloudreve Shortcut Qwen2.5 Card GIT Heatmap Pickle 财报 Google OCR Base64 Permission Distillation ResNet-50 Tiktoken Breakpoint XGBoost Numpy Hilton Pillow Password Translation Pandas Ubuntu Land FP64 PyCharm 净利润 llama.cpp 音频 FastAPI 论文速读 公式 Nginx 云服务器 递归学习法 C++ Safetensors 顶会 Template Streamlit SAM ModelScope uwsgi Transformers tar Augmentation 强化学习 NLP TensorRT GPT4 Paper tqdm HuggingFace Michelin Gemma Video Logo Github LoRA 证件照 Firewall Datetime 签证 Hungarian CTC BF16 Vmess mmap LLM Jupyter Zip COCO WAN Rebuttal CSV CC Plate PyTorch VGG-16 Jetson hf MD5 CUDA git-lfs Python OpenAI Claude Bin Math VPN Hotel Bipartite uWSGI Agent 腾讯云 DeepStream Statistics HaggingFace NLTK 搞笑 Use RGB BeautifulSoup CV Diagram LLAMA BTC NameSilo Interview Tensor GPTQ API 阿里云 Website TTS scipy Docker YOLO Review Paddle 算法题 Qwen XML torchinfo CEIR Attention Knowledge Quantize 飞书 Mixtral v0.dev Magnet OpenCV Crawler LaTeX Git Miniforge Color Input Data Windows Freesound Django RAR Domain Pytorch QWEN UI CAM SQL WebCrawler 多线程 DeepSeek TSV ONNX 图标 diffusers Proxy Search PDB Clash Image2Text TensorFlow CLAP logger Qwen2 LeetCode 多进程 Bert Algorithm Anaconda GoogLeNet Dataset 第一性原理 Vim 关于博主 FlashAttention transformers FP8 News RL Food Conda Excel 继承 SPIE Web SVR 图形思考法 AI FP16 v2ray JSON ms-swift IndexTTS2 InvalidArgumentError Animate EXCEL 版权 PDF Baidu SQLite UNIX Quantization PIP 报税 Linux git FP32 Tracking VSCode Random Sklearn Markdown ChatGPT
站点统计

本站现有博文333篇,共被浏览923979

本站已经建立2630天!

热门文章
文章归档
回到顶部