Li dongyang
lsm03624
AI & ML interests
None yet
Recent Activity
liked a model about 2 hours ago
unsloth/GLM-5.2-GGUF liked a model 3 days ago
Kwaipilot/KAT-Coder-V2.5-Dev liked a model 6 days ago
poolside/Laguna-S-2.1-INT4Organizations
None yet
My output is garbled
4
#4 opened 21 days ago
by
pliskin123
hello,The model's output is garbled.
8
#1 opened 2 months ago
by
zhao198300
FP8 seems to be broken
➕ 1
3
#1 opened about 1 month ago
by
Neiko2002
You didn't quantize it, did you? FP8 can't be the same size as BF16.
1
#1 opened about 1 month ago
by
lsm03624
怎么自我认知还是deepseek?而且好像没有做快慢思考,无法自适应控制思考长度
3
#12 opened about 1 month ago
by
user48271
模型量化的效果并不理想
5
#2 opened 6 months ago
by
mediali
The VLLM installed in this way can run this model.
#1 opened 5 months ago
by
lsm03624
Unable to run (vllm/sglang)
8
#1 opened 5 months ago
by
nfunctor
Can we perform 4-bit quantization for the awq of the Step-3.5-Flash model? The VLLM can run it.
1
#3 opened 6 months ago
by
lsm03624
the faster the inference speed becomes. Why is that?
👍 1
1
#9 opened 9 months ago
by
lsm03624
GPU: 5060Ti*4, vllm version 0.11. The following error occurred:
1
#1 opened 9 months ago
by
lsm03624
Error installing from PR branch
👍 1
14
#1 opened 9 months ago
by
DrRos
Thanks!
❤️ 2
9
#1 opened 12 months ago
by
lightenup
Seems to be working with PR and `--jinja`
❤️🤝 3
4
#1 opened 9 months ago
by
ubergarm