Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up
soyrsoyr 's Collections
pr3118-mtp-e2e
Quantized Models
test models
Custom Models
Muse-Glimmer-30B
Llama-3.2-1B-Instruct GPTQ Quantized
DeepSeek-MoE-16B-Chat GPTQ Quantized
Gemma 4 12B Quantized
tiny models

Llama-3.2-1B-Instruct GPTQ Quantized

updated 22 days ago

GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.

Upvote
-

  • soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ

    Text Generation • 1B • Updated Jun 7 • 8

  • soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ

    Text Generation • 1B • Updated Jun 7 • 10

  • soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ

    Text Generation • 1B • Updated Jun 7 • 7

  • soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ

    Text Generation • 0.8B • Updated Jun 7 • 61
Upvote
-
  • Collection guide
  • Browse collections
Company
TOS Privacy About Careers
Website
Models Datasets Spaces Pricing Docs