Bunch of bad models
Sawyer Bowerman
soyrsoyr
AI & ML interests
None yet
Recent Activity
updated a collection about 24 hours ago
Tiny Models updated a model about 24 hours ago
inference-optimization/Qwen3-VL-Reranker-0.1B-tiny published a model about 24 hours ago
inference-optimization/Qwen3-VL-Reranker-0.1B-tinyOrganizations
Muse-Glimmer-30B
Quantized across W4A16, FP8, NVFP4 using llm-compressor.
DeepSeek-MoE-16B-Chat GPTQ Quantized
DeepSeek-MoE-16B-Chat quantized with GPTQ via llm-compressor: W8A8, W4A16, FP8, NVFP4.
-
soyrsoyr/deepseek-moe-16b-chat-W8A8-GPTQ
Text Generation • 16B • Updated • 10 -
soyrsoyr/deepseek-moe-16b-chat-W4A16-GPTQ
Text Generation • 3B • Updated • 10 -
soyrsoyr/deepseek-moe-16b-chat-FP8-GPTQ
Text Generation • 16B • Updated • 5 -
soyrsoyr/deepseek-moe-16b-chat-NVFP4-GPTQ
Text Generation • 9B • Updated • 9
tiny models
tiny models for development
Quantized Models
Bunch of good models
Llama-3.2-1B-Instruct GPTQ Quantized
GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.
-
soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ
Text Generation • 1B • Updated • 6 -
soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ
Text Generation • 1B • Updated • 14 -
soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ
Text Generation • 1B • Updated • 9 -
soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ
Text Generation • 0.8B • Updated • 9
Gemma 4 12B Quantized
Custom Models
Bunch of bad models
Quantized Models
Bunch of good models
Muse-Glimmer-30B
Quantized across W4A16, FP8, NVFP4 using llm-compressor.
Llama-3.2-1B-Instruct GPTQ Quantized
GPTQ quantized across W4A16, W8A8, FP8, NVFP4 using llm-compressor.
-
soyrsoyr/Llama-3.2-1B-Instruct-W4A16-GPTQ
Text Generation • 1B • Updated • 6 -
soyrsoyr/Llama-3.2-1B-Instruct-W8A8-GPTQ
Text Generation • 1B • Updated • 14 -
soyrsoyr/Llama-3.2-1B-Instruct-FP8-GPTQ
Text Generation • 1B • Updated • 9 -
soyrsoyr/Llama-3.2-1B-Instruct-NVFP4-GPTQ
Text Generation • 0.8B • Updated • 9
DeepSeek-MoE-16B-Chat GPTQ Quantized
DeepSeek-MoE-16B-Chat quantized with GPTQ via llm-compressor: W8A8, W4A16, FP8, NVFP4.
-
soyrsoyr/deepseek-moe-16b-chat-W8A8-GPTQ
Text Generation • 16B • Updated • 10 -
soyrsoyr/deepseek-moe-16b-chat-W4A16-GPTQ
Text Generation • 3B • Updated • 10 -
soyrsoyr/deepseek-moe-16b-chat-FP8-GPTQ
Text Generation • 16B • Updated • 5 -
soyrsoyr/deepseek-moe-16b-chat-NVFP4-GPTQ
Text Generation • 9B • Updated • 9
Gemma 4 12B Quantized
tiny models
tiny models for development