marcelone commited on
Commit
b18facf
·
verified ·
1 Parent(s): 57a0c76

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +3 -1
README.md CHANGED
@@ -2,4 +2,6 @@
2
  license: apache-2.0
3
  base_model: Jinx-org/Jinx-Qwen3-4B
4
  base_model_relation: quantized
5
- ---
 
 
 
2
  license: apache-2.0
3
  base_model: Jinx-org/Jinx-Qwen3-4B
4
  base_model_relation: quantized
5
+ ---
6
+ # Recommended
7
+ **Jinx-Qwen3-4B-gguf-q6_k-q-8 (mixed-precision):** selected weights (output, token embeddings, attention/FFN layers in first and last blocks) quantized to **Q8_0**, remaining tensors **Q6_k**, reducing memory footprint while preserving inference fidelity.