Text Generation
GGUF
English
tiny
vertex

Vertex 0.6 15M Base

Base (pretrain-only) checkpoint of Vertex 0.6 15M, a tiny ~15M-param model from the Vertex 0.6 family. Qwen3 architecture: hidden 256, 10 layers, 4 heads / 2 KV (GQA, head_dim 64), SwiGLU ffn 1024, 20000-vocab ByteLevel BPE, tied embeddings, ctx 2048.

Pretrained from scratch on 12B tokens of English web text: Ultra-FineWeb plus Ultra-FineWeb-L3 synthetic rewrites (Multi-Style + QA), ~800 tokens per parameter, on a single RTX 4060 Laptop (8GB).

This is a raw language model — no chat template, no instruction tuning. For chat, see Vertex-0.6-15M-Instruct.

Training data

Pretrained on English web text from openbmb/Ultra-FineWeb and openbmb/Ultra-FineWeb-L3 (Multi-Style + QA synthetic configs), 12B tokens total.

Downloads last month
85
GGUF
Model size
15M params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for VertexResearch/Vertex-0.6-15M-Base-GGUF

Quantized
(1)
this model

Datasets used to train VertexResearch/Vertex-0.6-15M-Base-GGUF

Collection including VertexResearch/Vertex-0.6-15M-Base-GGUF