Just MLPs: Efficient Visual State Reconstruction for Multimodal Language Models Paper • 2609.34972 • Published 8 days ago • 30
Fewer Tokens, More Self-Teaching: On-Policy Self-Distillation for Extreme Visual Token Reduction Paper • 2609.32353 • Published 10 days ago • 8
TurboClear: One-Step Object-Effect Removal via Region-Calibrated Distribution Matching and Fusion Paper • 2608.01288 • Published Aug 2
Fewer Tokens, More Self-Teaching: On-Policy Self-Distillation for Extreme Visual Token Reduction Paper • 2609.32353 • Published 10 days ago • 8
Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA Paper • 2608.09819 • Published Aug 10 • 180
LongStraw: Long-Context RL Beyond 2M Tokens under a Fixed GPU Budget Paper • 2607.14952 • Published Jul 16 • 92
MinT: Managed Infrastructure for Training and Serving Millions of LLMs Paper • 2605.13779 • Published May 13 • 227
$δ$-mem: Efficient Online Memory for Large Language Models Paper • 2605.12357 • Published May 12 • 133
$δ$-mem: Efficient Online Memory for Large Language Models Paper • 2605.12357 • Published May 12 • 133
SlowBA: An efficiency backdoor attack towards VLM-based GUI agents Paper • 2603.08316 • Published Mar 9 • 1
SlowBA: An efficiency backdoor attack towards VLM-based GUI agents Paper • 2603.08316 • Published Mar 9 • 1
PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks Paper • 2602.06663 • Published Feb 6 • 5
PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks Paper • 2602.06663 • Published Feb 6 • 5
AgentDevel: Reframing Self-Evolving LLM Agents as Release Engineering Paper • 2601.04620 • Published Jan 8 • 3
Golden Goose: A Simple Trick to Synthesize Unlimited RLVR Tasks from Unverifiable Internet Text Paper • 2601.22975 • Published Jan 30 • 113
view post Post 2317 From Pointers to Footnotes: Why Agents will be “Reference-First”https://medium.com/@di-zhang-fdu/from-pointers-to-footnotes-why-agents-will-be-reference-first-88d0e9730e37 See translation 👀 2 2 + Reply
view post Post 559 AgentDevel: Reframing Self-Evolving LLM Agents as Release EngineeringImprove your agent like Software Engineers, Driven by test and statistical information.https://www.arxiv.org/abs/2601.04620 See translation 👍 1 1 + Reply
AgentDevel: Reframing Self-Evolving LLM Agents as Release Engineering Paper • 2601.04620 • Published Jan 8 • 3
view post Post 349 On the eve of the birth of the Vibe Computer: Why agentic coding feels like a 1980s terminal — and why that’s the pointhttps://di-zhang-fdu.medium.com/on-the-eve-of-the-birth-of-the-vibe-computer-why-agentic-coding-feels-like-a-1980s-terminal-and-96c41510c356 See translation 2 replies · ❤️ 1 1 👀 1 1 + Reply