IsValorum/Qwen3.6-35B-A3B-MTP-APEX-I-MiniPlus-V2.1-Abliterated-GGUF Image-Text-to-Text ⢠35B ⢠Updated about 6 hours ago ⢠636 ⢠2
audio-cpp/Nemotron-3-Diarization-GGUF Voice Activity Detection ⢠99.3M ⢠Updated 1 day ago ⢠1k ⢠8
IsValorum/Qwen3.6-35B-A3B-MTP-APEX-I-NanoPlus-GGUF Image-Text-to-Text ⢠36B ⢠Updated about 6 hours ago ⢠201 ⢠2
view post Post 4292 š Introducing Halo 1.0Today, we are open-sourcing Halo, the training framework we use to train every model at White Circle.It comes with: š§ Full post-training stack: SFT, DPO/KTO/SMPO, reward modeling, GRPO, distillationš¤ Async multi-turn RL with vLLM/SGLang rollouts and sandboxed tool useā” ~2.8Ć TRL throughput on 8Ć B300 (EP+FSDPv2, FA4, fp8/fp4)š¤ Dense HF models + 15 MoE families (Qwen, GLM, Mistral, DeepSeek-V4ā¦)š ļø One halo command, prebuilt Docker images, and docs for humans and agentsš» https://github.com/whitecircle/haloTry it and tell us what you're training See translation 1 reply Ā· š 9 9 ā¤ļø 4 4 + Reply
Viggle/Qwen-Image-2.1-viggle-turbo Text-to-Image ⢠7B ⢠Updated about 4 hours ago ⢠47.9k ⢠197
Zeroshot Classifiers Collection These are my current best zeroshot classifiers. Some of my older models are downloaded more often, but the models in this collection are newer/better. ⢠12 items ⢠Updated Jan 6, 2025 ⢠158
SC117/Qwen3.6-35B-A3B-uncensored-heretic-Native-MTP-Preserved-APEX-GGUF Image-Text-to-Text ⢠0.4B ⢠Updated Jul 29 ⢠58.1k ⢠131
ggml-org/MiMo-V2.6-Distill-Qwen-9B-GGUF Image-Text-to-Text ⢠9B ⢠Updated 3 days ago ⢠15.7k ⢠22
XiaomiMiMo/MiMo-V2.6-Distill-Qwen-9B Image-Text-to-Text ⢠9B ⢠Updated 3 days ago ⢠6.65k ⢠448
moondream/parakeet-redux Automatic Speech Recognition ⢠0.1B ⢠Updated 2 days ago ⢠1.05k ⢠149