Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
brain-lab 's Collections
HARP Quantized Models
Gradient-Faithful Surrogates paper models

HARP Quantized Models

updated 3 days ago

Quantized Llama 2 (7B/13B/70B) at 2-bit via HARP, a learnable orthogonal preprocessor for extreme LLM quantization. EMNLP 2026.

Upvote
1

  • brain-lab/Llama-2-7b-QuIP-HARP-2Bit

    Text Generation • 0.6B • Updated 3 days ago • 117

  • brain-lab/Llama-2-13b-QuIP-HARP-2Bit

    Text Generation • 0.9B • Updated 3 days ago • 112

  • brain-lab/Llama-2-70b-QuIP-HARP-2Bit

    Text Generation • 3B • Updated 3 days ago • 114
Upvote
1
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs