-
R-4B: Incentivizing General-Purpose Auto-Thinking Capability in MLLMs via Bi-Mode Annealing and Reinforce Learning
Paper • 2508.21113 • Published • 111 -
Breaking the Exploration Bottleneck: Rubric-Scaffolded Reinforcement Learning for General LLM Reasoning
Paper • 2508.16949 • Published • 24 -
EmbodiedOneVision: Interleaved Vision-Text-Action Pretraining for General Robot Control
Paper • 2508.21112 • Published • 78 -
UItron: Foundational GUI Agent with Advanced Perception and Planning
Paper • 2508.21767 • Published • 12
Jeff Nyzio
TheOneTrueNiz
AI & ML interests
None yet
Recent Activity
new activity 2 days ago
Lightricks/LTX-2.3-22b-IC-LoRA-Ingredients:Multiple character references all become invalid. updated a collection 4 months ago
Papers liked a model 4 months ago
google/timesfm-2.0-500m-pytorchOrganizations
None yet