Puro-2B: Poor Lab's Qwen2-1.5B Trained on RTX 5090 within $5090 Paper • 2608.27370 • Published 7 days ago • 33
Efficient Memory Management for Large Language Model Serving with PagedAttention Paper • 2309.06180 • Published Sep 12, 2023 • 69
view article Article ArabicWeb24: Creating a High Quality Arabic Web-only Pre-training Dataset MayFarhat • Aug 8, 2024 • 12
view article Article makeMoE: Implement a Sparse Mixture of Experts Language Model from Scratch AviSoori1x • May 7, 2024 • 125
Running Agents 54 Arabic LLM Leaderboard - Broad (ABL) 🥇 54 ABL is a comprehensive and modern Arabic LLM Leaderboard.
SmolVLA: A Vision-Language-Action Model for Affordable and Efficient Robotics Paper • 2506.01844 • Published Jun 2, 2025 • 165
Running on CPU Upgrade Featured 3.29k The Smol Training Playbook 📚 3.29k The secrets to building world-class LLMs
Running on CPU Upgrade 14.1k Open LLM Leaderboard 🏆 14.1k Track, rank and evaluate open LLMs and chatbots