microsoft/VibeVoice-ASR-Streaming-7B Automatic Speech Recognition • 9B • Updated 6 days ago • 1.45k • 151
agentionai/Qwen3.8-Flash-Next-ROCmFP4-FAST-imatrix-GGUF Text Generation • 177B • Updated 9 days ago • 24k • 67
Ryn1998/Qwen3.8-27B-Heretic-Abliterated-Uncensored-GGUF Text Generation • 27B • Updated 8 days ago • 15.6k • 17
Qwen-RobotManip Technical Report: Alignment Unlocks Scale for Robotic Manipulation Foundation Models Paper • 2606.17846 • Published Jun 17 • 34
galilai-group/LeVJEPA-VideoMix-Large Image Feature Extraction • 0.3B • Updated 12 days ago • 10.6k • 28
SnapFlow: One-Step Action Generation for Flow-Matching VLAs via Progressive Self-Distillation Paper • 2604.05656 • Published Apr 7 • 1
Video to Data Collection Datasets generated by the NVIDIA Video to Data pipeline — converts human videos into simulation-ready assets and physics-grounded robot training data • 3 items • Updated 11 days ago • 6
view article Article Introducing NVIDIA Nemotron 3 Nano Omni: Long-Context Multimodal Intelligence for Documents, Audio and Video Agents nvidia • Apr 28 • 65