microsoft/VibeVoice-ASR-Streaming-7B Automatic Speech Recognition • 9B • Updated 10 days ago • 2.49k • 214
WorldReward: Reward Modeling for Camera-Conditioned World Models Paper • 2609.03952 • Published 10 days ago • 26
CORE: Improving Compositional Reasoning in MLLM Embedding via Reranker Distillation Paper • 2609.04083 • Published 10 days ago • 26
LatentPress: Context Compression Beyond Text and Vision Paper • 2609.01507 • Published 12 days ago • 121
huihui-ai/Huihui-Qwen3.8-27B-abliterated-GGUF Image-Text-to-Text • 27B • Updated 4 days ago • 2.62M • 689
LycheeMemory V2: Efficient Long-Term Memory for LLM Agents via Semantic Segment-Level Consolidation Paper • 2608.12990 • Published about 1 month ago • 14
OpenART: Scaling Agent Red Teaming via Open-Ended Environment Evolution Paper • 2608.00677 • Published Aug 1 • 263
SymDiag: Explainable Diagnosis for LLM Reasoning via Neuro-Symbolic Verification Paper • 2608.08786 • Published Aug 9 • 6
Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent Paper • 2608.03979 • Published Aug 4 • 53
MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations Paper • 2607.28956 • Published Jul 31 • 112
Poplar: A Scalable Pipeline for Human-Centric Image Dataset Synthesis Paper • 2608.00440 • Published Aug 1 • 10