Running 3.95k The Ultra-Scale Playbook 🌌 3.95k The ultimate guide to training LLM on large GPU Clusters
deepset/roberta-base-squad2 Question Answering • 0.1B • Updated Sep 24, 2024 • 510k • • 948
unsloth/DeepSeek-R1-Distill-Llama-8B-GGUF Text Generation • 8B • Updated May 10, 2025 • 29.3k • 307