Towards Looped Models Done Right, Part II: Rethinking at Fixed Points Paper • 2610.06833 • Published 7 days ago • 33
Prefill-Free Cross-Family KV Cache Transfer for Heterogeneous Multi-Agent LLMs Paper • 2609.32259 • Published 13 days ago • 101
Making LLMs Say What They Think: Measuring and Improving CoT-Interpretability Alignment Paper • 2609.38972 • Published 12 days ago • 28
Learning Functional Subspaces for Neural Network Compression Paper • 2609.40127 • Published 12 days ago • 25
SlimWise: Decoupling Expert Pruning Across Prefill and Decode for Efficient MoE Serving Paper • 2609.34117 • Published 14 days ago • 23
Personal-Agent Mediated Recommendation with Cross-Platform User History Paper • 2610.07588 • Published 6 days ago • 12
QuantWM: Temporally Consistent 2-Bit KV Cache Quantization for Video World Models Paper • 2609.26425 • Published 14 days ago • 21
open-llm-leaderboard-old/details_saarvajanik__facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache Updated Jan 28, 2024 • 557 • 8
Chinese-Jev: Bringing System One Model to Chinese-Language Tasks Paper • 2609.36965 • Published 13 days ago • 26
PanoVLN: Towards Effective Panoramic Vision-and-Language Navigation Paper • 2609.34759 • Published 14 days ago • 153
In-Flight KV Cache with Clean Anchors for Faster Autoregressive Video Diffusion Paper • 2609.32540 • Published 16 days ago • 35
YuE2: Unifying Symbolic and Audio Music Generation at Frontier Quality Paper • 2609.33757 • Published 15 days ago • 234