MobileWan: Closing the Quality Gap for Mobile Video Diffusion Paper • 2607.06173 • Published 16 days ago • 3
ComfyUI Abliterated Text Encoders Collection Abliterated text encoders for ComfyUI image and video generation. Includes GGUF and Safetensors. • 7 items • Updated 7 days ago • 4
MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization Paper • 2601.01554 • Published Jan 4 • 65
MOSS Transcribe Collection A unified multimodal large language model for end-to-end speaker-attributed, time-stamped transcription. • 4 items • Updated 12 days ago • 13
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Paper • 2607.14187 • Published 8 days ago • 29
view article Article NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval nvidia • 6 days ago • 55
Laguna S 2.1 Collection Our most capable model to date, designed for long-horizon work. • 12 items • Updated about 7 hours ago • 23
Wan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation Paper • 2607.09581 • Published 13 days ago • 6
Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation Paper • 2606.02441 • Published Jun 1 • 2
ZUNA Collection Brain-Computer Interface models for reconstruction, interpolation, and downstream tasks • 2 items • Updated 9 days ago • 4
Qwen 3.6 - Reg/Uncensored 9b, 12b, 21b, 27b, 40B Collection Fine tuned Qwen 3.6 models, including source and GGUF from 9B and up. 9B,12B, 21B and 40B are custom built by me. Tuning via Unsloth on local hardware • 12 items • Updated 3 days ago • 8
RoboDesign1M: A Large-scale Dataset for Robot Design Understanding Paper • 2503.06796 • Published Mar 9, 2025 • 2
GRAIL: Generating Humanoid Loco-Manipulation from 3D Assets and Video Priors Paper • 2606.05160 • Published Jun 3 • 9
UniVR: Thinking in Visual Space for Unified Visual Reasoning Paper • 2607.12800 • Published 9 days ago • 32
Concurrent Image Understanding and Generation: Self-Correcting Coupled Markov Jump Processes Paper • 2607.13188 • Published 9 days ago • 33
MultiRef-Compass: Towards Comprehensive Evaluation of Multi-Reference-to-Audio-Video Generation Paper • 2607.14189 • Published 8 days ago • 34