Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
soyrsoyr
's Collections
pr3118-mtp-e2e
Quantized Models
test models
Custom Models
Muse-Glimmer-30B
Llama-3.2-1B-Instruct GPTQ Quantized
DeepSeek-MoE-16B-Chat GPTQ Quantized
Gemma 4 12B Quantized
tiny models
Llama-3.2-1B-Instruct GPTQ Quantized
updated
22 days ago
GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.
Upvote
-
Sort: Collection
soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ
Text Generation
•
1B
•
Updated
Jun 7
•
8
soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ
Text Generation
•
1B
•
Updated
Jun 7
•
10
soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ
Text Generation
•
1B
•
Updated
Jun 7
•
7
soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ
Text Generation
•
0.8B
•
Updated
Jun 7
•
61
Upvote
-
Sort: Collection
Share collection
View history
Collection guide
Browse collections