SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models Paper • 2608.10538 • Published 6 days ago • 13
Power law graph attention: exact generalization of scaled dot-product attention, empirical collapse at inference Paper • 2608.10288 • Published 7 days ago • 7
FocusMem: Factorizing Content, Readout, and Trust in Latent GUI Memory Paper • 2608.04530 • Published 12 days ago • 14
Hunyuan3D-Buffalo 1.0: A Unified Multimodal Model for Scalable 3D Generation, Understanding, and Editing Paper • 2608.02711 • Published 13 days ago • 88
QQWorld: Quantile-Quantile Matching for World Model Regularization Paper • 2607.28415 • Published 18 days ago • 30