lmstudio-community/Qwen3-Coder-30B-A3B-Instruct-MLX-4bit Text Generation • 31B • Updated Jul 31, 2025 • 141k • 39
view article Article Fine-tuning a 350M Model for Better Structured Outputs in 100 GRPO Steps +1 iamleonie, burtenshaw, sergiopaniego • 7 days ago • 86
ApexAgents-SkyRL-Recipe Collection Checkpoints and eval traces from "Training frontier knowledge work agents: A 397B open recipe with SkyRL" • 4 items • Updated 9 days ago • 3
view article Article Training and Finetuning Multi-Vector Embedding Models with Sentence Transformers tomaarsen • 15 days ago • 123
nvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Coding Text Generation • 561B • Updated 26 days ago • 2.5k • 6
view article Article Meta is back with Muse Glimmer: local, agentic, multimodal, and open source +2 pcuenq, merve, burtenshaw, ariG23498 • Aug 10 • 111
view article Article Training a coding agent using the OpenCode harness in remote HF sandboxes with TRL and OpenEnv sergiopaniego • Aug 5 • 26