GigaBrain-0.5M*: a VLA That Learns From World Model-Based Reinforcement Learning Paper β’ 2602.12099 β’ Published Feb 12 β’ 62
view article Article π€ππ¬π₯οΈπ Kimi-VL-A3B-Thinking-2506: A Quick Navigation moonshotai β’ Jun 21, 2025 β’ 77
Object Detection with Multimodal Large Vision-Language Models: An In-depth Review Paper β’ 2508.19294 β’ Published Aug 25, 2025 β’ 1
Agent Lightning: Train ANY AI Agents with Reinforcement Learning Paper β’ 2508.03680 β’ Published Aug 5, 2025 β’ 141
Tool-integrated Reinforcement Learning for Repo Deep Search Paper β’ 2508.03012 β’ Published Aug 5, 2025 β’ 20
view article Article SmolLM3: smol, multilingual, long-context reasoner +21 eliebak, cmpatino, anton-l, edbeeching, m-ric, nouamanetazi, akseljoonas, guipenedo, hynky, clefourrier, SaylorTwift, kashif, qgallouedec, hlarcher, glutamatt, Xenova, reach-vb, ngxson, craffel, lewtun, loubnabnl, lvwerra, thomwolf β’ Jul 8, 2025 β’ 783
Qwen3 Embedding: Advancing Text Embedding and Reranking Through Foundation Models Paper β’ 2506.05176 β’ Published Jun 5, 2025 β’ 85
view article Article nanoVLM: The simplest repository to train your VLM in pure PyTorch +5 ariG23498, lusxvr, andito, sergiopaniego, merve, pcuenq, reach-vb β’ May 21, 2025 β’ 263
NovelSeek: When Agent Becomes the Scientist -- Building Closed-Loop System from Hypothesis to Verification Paper β’ 2505.16938 β’ Published May 22, 2025 β’ 121
Tool-Star: Empowering LLM-Brained Multi-Tool Reasoner via Reinforcement Learning Paper β’ 2505.16410 β’ Published May 22, 2025 β’ 60
Pixel Reasoner: Incentivizing Pixel-Space Reasoning with Curiosity-Driven Reinforcement Learning Paper β’ 2505.15966 β’ Published May 21, 2025 β’ 53