lingyezhixing
lingyezhixing
AI & ML interests
None yet
Recent Activity
new activity 12 days ago
SakuraLLM/Sakura-14B-Qwen3-v1.5-GGUF:VLLM的AWQ格式? new activity 29 days ago
unsloth/Qwen3.8-Flash-Next-GGUF:Is it possible to offload n-gram to an NVMe SSD? new activity about 1 month ago
LiquidAI/LFM2.5-2.6B:Will there be a code completion model based on LFM fine-tuning?Organizations
None yet
VLLM的AWQ格式?
4
#2 opened 8 months ago
by
Laoxu
Is it possible to offload n-gram to an NVMe SSD?
👀🚀 31
15
#11 opened 29 days ago
by
lingyezhixing
Will there be a code completion model based on LFM fine-tuning?
🔥 2
#11 opened about 1 month ago
by
lingyezhixing
instruct model
👍 1
#7 opened 4 months ago
by
lingyezhixing
Qwen3.6-27B?
🚀❤️ 32
19
#2 opened 5 months ago
by
lingyezhixing
Q3 quantization performance issues
2
#7 opened 7 months ago
by
lingyezhixing
Missing about 50~55GB of Q3?
5
#7 opened 7 months ago
by
lingyezhixing
Please regenerate to adapt to the latest improvements in llama.cpp
🔥 1
1
#4 opened 9 months ago
by
lingyezhixing
Where IQ quantize?
5
#1 opened 10 months ago
by
lingyezhixing
IQ4_XS Please
3
#6 opened 10 months ago
by
lingyezhixing
这次会不会有14B的AutoAWQ或者GPTQ?
1
#2 opened 11 months ago
by
lingyezhixing
Will there still be 32B dense models?
➕👀 8
2
#18 opened about 1 year ago
by
lingyezhixing
Hello, I want to know if the draft model will reduce the model generation quality?
1
#2 opened about 1 year ago
by
lingyezhixing
Smashed 💪 Scored to 82.86 🔥2bit IQ2_M on MMLU Pro single shot benchmark
🔥❤️ 2
5
#7 opened about 1 year ago
by
xbruce22
There must be something wrong with the size
👀 2
2
#8 opened over 1 year ago
by
lingyezhixing
Native FP4 seems to make quantization meaningless
3
#7 opened about 1 year ago
by
lingyezhixing
Can you provide some low-precision quantization options?
➕👍 3
11
#3 opened about 1 year ago
by
lingyezhixing
Is the GGUF file still being uploaded?
👍 2
3
#2 opened about 1 year ago
by
lingyezhixing