Skip to content
#

low-bit

Here are 16 public repositories matching this topic...

qpeft: quantization-aware training (QAT) and PEFT for LLMs in PyTorch. EfficientQAT, QA-LoRA and PEQA for Hugging Face models at 2/3/4 bits: train scales, zero-points or LoRA adapters, then merge into a low-bit integer model (GPTQ layout) that is tested to equal the trained one. Pure-torch and torchao backends.

  • Updated Sep 29, 2026
  • Python

Add this topic to your repo

To associate your repository with the low-bit topic, visit your repo's landing page and select "manage topics."

Learn more