MJPansa/MiniMax-M2.7-REAP-172B-A10B-AutoRound-W4A16 Text Generation • 24B • Updated Apr 15 • 3.06k • 11
Running 3.95k The Ultra-Scale Playbook 🌌 3.95k The ultimate guide to training LLM on large GPU Clusters