yuhiri
talentestors
AI & ML interests
None yet
Recent Activity
updated a collection 8 days ago
VLM liked a model 8 days ago
Qwen/Qwen3.8-Flash-Next updated a collection 8 days ago
VLMOrganizations
qwen
qwen
VLM
-
deepseek-ai/DeepSeek-V4.1-Flash
Image-Text-to-Text • 763B • Updated • 669k • • 3.86k -
zai-org/GLM-5.3-Flash
Image-Text-to-Text • 321B • Updated • 4.52M • • 2.59k -
MiniMaxAI/MiniMax-M3
Image-Text-to-Text • 427B • Updated • 165k • • 1.55k -
Qwen/Qwen3.8-Flash-Next
Image-Text-to-Text • 180B • Updated • 1.24M • • 5.76k
Application
- Running4
Auto Build 1Panel APP
🐨4Create 1Panel applications easily
- RunningFeatured16.6k
DeepSite v4
🐳16.6kGenerate any application by Vibe Coding it
- RunningAgentsFeatured472
Qwen3 VL Demo
😻472Chat with AI using text and images
- RunningAgents490
Sora 2
📉490Generate videos from text or images
DataSet
LLM
A large language model (LLM) is a language model trained with self-supervised machine learning on a vast amount of text, designed for natural language
video
embedding
TTS
- Running on ZeroAgents216
IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
🎙216Generate speech from text using a reference audio
-
Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice
Text-to-Speech • 2B • Updated • 2.41M • 2.01k - Running on ZeroAgentsFeatured2.27k
Qwen3-TTS Demo
🎙2.27kGenerate speech from text with voice design, cloning, or presets
-
IndexTeam/IndexTTS-2
Text-to-Speech • Updated • 10.9k • 786
Image
-
stabilityai/stable-diffusion-xl-base-1.0
Text-to-Image • 3B • Updated • 3.73M • • 8.25k -
black-forest-labs/FLUX.1-Kontext-dev
Image-to-Image • 12B • Updated • 345k • • 2.85k -
meituan-longcat/LongCat-Image-Edit
Image-to-Image • Updated • 20.1k • • 186 -
meituan-longcat/LongCat-Image
Text-to-Image • Updated • 13.6k • • 249
3D
OCR
video
qwen
qwen
embedding
VLM
-
deepseek-ai/DeepSeek-V4.1-Flash
Image-Text-to-Text • 763B • Updated • 669k • • 3.86k -
zai-org/GLM-5.3-Flash
Image-Text-to-Text • 321B • Updated • 4.52M • • 2.59k -
MiniMaxAI/MiniMax-M3
Image-Text-to-Text • 427B • Updated • 165k • • 1.55k -
Qwen/Qwen3.8-Flash-Next
Image-Text-to-Text • 180B • Updated • 1.24M • • 5.76k
TTS
- Running on ZeroAgents216
IndexTTS: An Industrial-Level Controllable and Efficient Zero-Shot Text-To-Speech System
🎙216Generate speech from text using a reference audio
-
Qwen/Qwen3-TTS-12Hz-1.7B-CustomVoice
Text-to-Speech • 2B • Updated • 2.41M • 2.01k - Running on ZeroAgentsFeatured2.27k
Qwen3-TTS Demo
🎙2.27kGenerate speech from text with voice design, cloning, or presets
-
IndexTeam/IndexTTS-2
Text-to-Speech • Updated • 10.9k • 786
Application
- Running4
Auto Build 1Panel APP
🐨4Create 1Panel applications easily
- RunningFeatured16.6k
DeepSite v4
🐳16.6kGenerate any application by Vibe Coding it
- RunningAgentsFeatured472
Qwen3 VL Demo
😻472Chat with AI using text and images
- RunningAgents490
Sora 2
📉490Generate videos from text or images
Image
-
stabilityai/stable-diffusion-xl-base-1.0
Text-to-Image • 3B • Updated • 3.73M • • 8.25k -
black-forest-labs/FLUX.1-Kontext-dev
Image-to-Image • 12B • Updated • 345k • • 2.85k -
meituan-longcat/LongCat-Image-Edit
Image-to-Image • Updated • 20.1k • • 186 -
meituan-longcat/LongCat-Image
Text-to-Image • Updated • 13.6k • • 249
DataSet
3D
LLM
A large language model (LLM) is a language model trained with self-supervised machine learning on a vast amount of text, designed for natural language