MobileWan: Closing the Quality Gap for Mobile Video Diffusion Paper • 2607.06173 • Published 18 days ago • 3
ComfyUI Abliterated Text Encoders Collection Abliterated text encoders for ComfyUI image and video generation. Includes GGUF and Safetensors. • 7 items • Updated 9 days ago • 4
MOSS Transcribe Diarize: Accurate Transcription with Speaker Diarization Paper • 2601.01554 • Published Jan 4 • 65
MOSS Transcribe Collection A unified multimodal large language model for end-to-end speaker-attributed, time-stamped transcription. • 4 items • Updated 14 days ago • 13
RxBrain: Embodied Cognition Foundation Model with Joint Language-Visual Reasoning and Imagination Paper • 2607.14187 • Published 10 days ago • 31
view article Article NVIDIA Nemotron 3 Embed Ranks #1 Overall on RTEB, Advancing Agentic Retrieval nvidia • 8 days ago • 56
Laguna S 2.1 Collection Our most capable model to date, designed for long-horizon work. • 12 items • Updated 2 days ago • 25
Wan-Dancer: A Hierarchical Framework for Minute-scale Coherent Music-to-Dance Generation Paper • 2607.09581 • Published 15 days ago • 6
Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation Paper • 2606.02441 • Published Jun 1 • 2
ZUNA Collection Brain-Computer Interface models for reconstruction, interpolation, and downstream tasks • 2 items • Updated 11 days ago • 4
Qwen 3.6 - Reg/Uncensored 9b, 12b, 21b, 27b, 40B Collection Fine tuned Qwen 3.6 models, including source and GGUF from 9B and up. 9B,12B, 21B and 40B are custom built by me. Tuning via Unsloth on local hardware • 12 items • Updated about 11 hours ago • 11
RoboDesign1M: A Large-scale Dataset for Robot Design Understanding Paper • 2503.06796 • Published Mar 9, 2025 • 2
GRAIL: Generating Humanoid Loco-Manipulation from 3D Assets and Video Priors Paper • 2606.05160 • Published Jun 3 • 9
UniVR: Thinking in Visual Space for Unified Visual Reasoning Paper • 2607.12800 • Published 11 days ago • 32
Concurrent Image Understanding and Generation: Self-Correcting Coupled Markov Jump Processes Paper • 2607.13188 • Published 11 days ago • 34
MultiRef-Compass: Towards Comprehensive Evaluation of Multi-Reference-to-Audio-Video Generation Paper • 2607.14189 • Published 10 days ago • 34