bartowski/DeepSeek-R1-Distill-Qwen-32B-GGUF Text Generation • 33B • Updated Jan 22, 2025 • 13.2k • 288
Running 3.67k The Ultra-Scale Playbook 🌌 3.67k The ultimate guide to training LLM on large GPU Clusters